CategoriesInfrastructure

Chips, GPUs, data centers, cloud compute, and edge AI hardware

23 stories in the last 7 days

Piyush Goyal Meets Microsoft India Chief To Discuss Data Centres, AI Infrastructure Expansion - KNN India

Piyush Goyal meets Microsoft India chief to discuss AI infrastructure expansion. The meeting focuses on scaling data centers across India…

India AI·Sep 12

Empower India bets on sustainable infrastructure for AI computing boom - Prop News Time

Empower India invests in green data centers to support the AI computing boom. The company plans to build renewable-powered facilities tha…

India AI·Sep 12

Dell launches new AI PCs and workstations for Indian enterprises with on-device AI focus

Dell launches AI‑optimized PCs and workstations for Indian enterprises. The new lineup includes commercial PCs, tower and rack workstatio…

ET CIO·Sep 12

Palantir Foundry and cuOpt drive NVIDIA supply chain allocation

NVIDIA uses Palantir Foundry and cuOpt to automate hardware supply chain allocation. The system tracks delivery from wafer‑out to first t…

AI News·Sep 11

Philippines sets out $34.4 billion blueprint to build an AI infrastructure hub

The Philippines launches a $34.4 billion plan to build an AI infrastructure hub. The masterplan targets 1.5 GW of data‑center capacity by…

ET CIO·Sep 11

Rapidly scaling online storage to serve over 1 billion ChatGPT users

OpenAI expands Habitat into a global storage platform for ChatGPT. The system evolves from a Python library to a distributed platform tha…

OpenAI Blog·Sep 11

Reduce LLM latency with prefix-aware routing on Amazon SageMaker Inference

Amazon SageMaker adds prefix‑aware routing to reduce LLM inference latency. The feature routes requests with the same prompt prefix to th…

AWS ML Blog·Sep 10

Reduce inference cold starts on Amazon SageMaker HyperPod with model caching

Amazon SageMaker HyperPod now supports model caching to cut inference cold starts. The feature pre‑loads model weights and container imag…

AWS ML Blog·Sep 10

Video and image search in Amazon Bedrock Knowledge Base using Marengo 3.0

Marengo 3.0 is a new embedding model now available in Amazon Bedrock Knowledge Bases. It enables fully managed natural language search ac…

AWS ML Blog·Sep 10

Nvidia and Palantir team up to run supply chains with AI, starting with Nvidia's own million-part operation

Nvidia and Palantir team up to apply AI to supply chain management. They will start with Nvidia's own million‑part operation, using AI to…

The Decoder·Sep 10

Powering AI is an architecture problem

Power outages in Ashburn threaten the world’s largest AI data center cluster. A transmission line fault knocked 3 GW off the grid, and a …

MIT Tech Review·Sep 10

Top AI spenders cut per-employee costs by nearly 10 percent in August

Top AI spenders slash per-employee costs by nearly 10% in August. The price per million tokens has fallen 41% since March 2026, prompting…

The Decoder·Sep 10

GitHub availability report: August 2026

GitHub reports five incidents in August that degraded service performance. The incidents caused widespread slowdown across GitHub service…

GitHub Blog·Sep 10

Deploying Qwen3.8-2.4T-A95B on Amazon SageMaker HyperPod with vLLM

Qwen3.8‑2.4T‑A95B, a 2.4‑trillion‑parameter open‑weight model, is deployed on Amazon SageMaker HyperPod using vLLM. The guide covers clus…

AWS ML Blog·Sep 9

India joins 24 countries in global initiative on 6G technology - News On AIR

India joins 24 countries in a global 6G initiative. The effort aims to set standards and accelerate research for next‑generation wireless…

India AI·Sep 9

NVIDIA Brings Real-Time AI to Broadcast, Sports and Global Streaming at IBC

NVIDIA introduces real‑time AI solutions for broadcast, sports, and global streaming at IBC. At the Amsterdam conference, the company unv…

NVIDIA AI·Sep 9

How NeoClouds should choose their next AI region – and why India deserves a closer look - Data Center Dynamics

NeoClouds should consider India as a strategic AI region. The article argues that India offers a growing talent pool, competitive data ce…

India AI·Sep 9

AWS is using Qualcomm for AI inference while Qualcomm uses AWS Bedrock to design the chips

AWS partners with Qualcomm to run AI inference on custom chips. Qualcomm designs custom chips for AWS across multiple product generations…

The Decoder·Sep 9

Take on your most ambitious work with GPT-6 Astra on Amazon Bedrock

GPT-6 Astra is a new OpenAI model now available on Amazon Bedrock. It delivers deeper reasoning and sharper judgment for demanding tasks,…

AWS ML Blog·Sep 8

Benchmarking small LLM inference on SageMaker AI: G7 vs G5 and G6

Benchmark compares 30B Mixture-of-Experts models on AWS SageMaker GPU instances. The study evaluates Qwen3-Coder-30B and NVIDIA Nemotron-…

AWS ML Blog·Sep 8

Patagonia has what AI data centers want, including no resistance so far

Patagonia is being eyed as a prime location for future AI data centers. Argentina’s vast, sparsely populated region offers abundant renew…

The Decoder·Sep 8

ASML locks in TSMC, Samsung, and Intel while Huawei races to break its grip

ASML secures TSMC, Samsung, and Intel to adopt larger photomasks, boosting EUV throughput. The move should raise throughput by 40 percent…

The Decoder·Sep 8

Arm launches Total Design for Physical AI and robotics framework

Arm introduces Total Design for Physical AI alongside a new robotics framework. The initiative establishes common standards and targets t…

AI News·Sep 8