Chips, GPUs, data centers, cloud compute, and edge AI hardware
23 stories in the last 7 days
Piyush Goyal Meets Microsoft India Chief To Discuss Data Centres, AI Infrastructure Expansion - KNN India
Piyush Goyal meets Microsoft India chief to discuss AI infrastructure expansion. The meeting focuses on scaling data centers across India…
Empower India bets on sustainable infrastructure for AI computing boom - Prop News Time
Empower India invests in green data centers to support the AI computing boom. The company plans to build renewable-powered facilities tha…
Dell launches new AI PCs and workstations for Indian enterprises with on-device AI focus
Dell launches AI‑optimized PCs and workstations for Indian enterprises. The new lineup includes commercial PCs, tower and rack workstatio…
Palantir Foundry and cuOpt drive NVIDIA supply chain allocation
NVIDIA uses Palantir Foundry and cuOpt to automate hardware supply chain allocation. The system tracks delivery from wafer‑out to first t…
Philippines sets out $34.4 billion blueprint to build an AI infrastructure hub
The Philippines launches a $34.4 billion plan to build an AI infrastructure hub. The masterplan targets 1.5 GW of data‑center capacity by…
Rapidly scaling online storage to serve over 1 billion ChatGPT users
OpenAI expands Habitat into a global storage platform for ChatGPT. The system evolves from a Python library to a distributed platform tha…
Reduce LLM latency with prefix-aware routing on Amazon SageMaker Inference
Amazon SageMaker adds prefix‑aware routing to reduce LLM inference latency. The feature routes requests with the same prompt prefix to th…
Reduce inference cold starts on Amazon SageMaker HyperPod with model caching
Amazon SageMaker HyperPod now supports model caching to cut inference cold starts. The feature pre‑loads model weights and container imag…
Video and image search in Amazon Bedrock Knowledge Base using Marengo 3.0
Marengo 3.0 is a new embedding model now available in Amazon Bedrock Knowledge Bases. It enables fully managed natural language search ac…
Nvidia and Palantir team up to run supply chains with AI, starting with Nvidia's own million-part operation
Nvidia and Palantir team up to apply AI to supply chain management. They will start with Nvidia's own million‑part operation, using AI to…
Powering AI is an architecture problem
Power outages in Ashburn threaten the world’s largest AI data center cluster. A transmission line fault knocked 3 GW off the grid, and a …
Top AI spenders cut per-employee costs by nearly 10 percent in August
Top AI spenders slash per-employee costs by nearly 10% in August. The price per million tokens has fallen 41% since March 2026, prompting…
GitHub availability report: August 2026
GitHub reports five incidents in August that degraded service performance. The incidents caused widespread slowdown across GitHub service…
Deploying Qwen3.8-2.4T-A95B on Amazon SageMaker HyperPod with vLLM
Qwen3.8‑2.4T‑A95B, a 2.4‑trillion‑parameter open‑weight model, is deployed on Amazon SageMaker HyperPod using vLLM. The guide covers clus…
India joins 24 countries in global initiative on 6G technology - News On AIR
India joins 24 countries in a global 6G initiative. The effort aims to set standards and accelerate research for next‑generation wireless…
NVIDIA Brings Real-Time AI to Broadcast, Sports and Global Streaming at IBC
NVIDIA introduces real‑time AI solutions for broadcast, sports, and global streaming at IBC. At the Amsterdam conference, the company unv…
How NeoClouds should choose their next AI region – and why India deserves a closer look - Data Center Dynamics
NeoClouds should consider India as a strategic AI region. The article argues that India offers a growing talent pool, competitive data ce…
AWS is using Qualcomm for AI inference while Qualcomm uses AWS Bedrock to design the chips
AWS partners with Qualcomm to run AI inference on custom chips. Qualcomm designs custom chips for AWS across multiple product generations…
Take on your most ambitious work with GPT-6 Astra on Amazon Bedrock
GPT-6 Astra is a new OpenAI model now available on Amazon Bedrock. It delivers deeper reasoning and sharper judgment for demanding tasks,…
Benchmarking small LLM inference on SageMaker AI: G7 vs G5 and G6
Benchmark compares 30B Mixture-of-Experts models on AWS SageMaker GPU instances. The study evaluates Qwen3-Coder-30B and NVIDIA Nemotron-…
Patagonia has what AI data centers want, including no resistance so far
Patagonia is being eyed as a prime location for future AI data centers. Argentina’s vast, sparsely populated region offers abundant renew…
ASML locks in TSMC, Samsung, and Intel while Huawei races to break its grip
ASML secures TSMC, Samsung, and Intel to adopt larger photomasks, boosting EUV throughput. The move should raise throughput by 40 percent…
Arm launches Total Design for Physical AI and robotics framework
Arm introduces Total Design for Physical AI alongside a new robotics framework. The initiative establishes common standards and targets t…