AI News

AWS and NVIDIA Add 2 Million GPUs: Demand Exceeds Every Forecast, Partnership Expands from Chips to Full Technology Stack

On August 26, 2026, AWS and NVIDIA jointly announced the deployment of an additional 2 million NVIDIA GPUs across AWS's global infrastructure from 2027 to 2028, spanning the Blackwell Ultra, Rubin, and Rubin Ultra architectures. Driven by customer demand that has exceeded all previous forecasts, the partnership has expanded from GPU procurement to full-stack co-design across CPUs, networking, memory, open models, and robotics.

AWS NVIDIA GPU
17
AIN

A quarter of Nvidia’s business next year comes from labs it is financing

Nvidia has put nearly US$50 billion into the AI labs that buy its chips, and has lined up commitments for more than $500 billion Colette Kress, the company’s chief financial officer, told analysts on August 26 that demand from the labs Nvidia backs with its own balance sheet will contribute toward roughly a quarter of its business […] The post A quarter of Nvidia’s business next year comes from labs it is financing appeared first on AI News.

Nvidia AI chips AI investment
25

OpenAI's Self-Developed Chip Jalapeño Passes First Test: Per-Watt Performance Up to 1.9x NVIDIA's Flagships, 700W vs 1400W

On August 26, 2026, OpenAI released the first performance data for its self-developed inference chip, Jalapeño, showing 1.5-1.9x higher throughput per kilowatt and 1.7-3.6x lower end-to-end latency than NVIDIA GB200 and GB300 rack systems on the InferenceX benchmark, marking the company's formal entry into hardware-level competition.

OpenAI Jalapeño 自研芯片
39

Nvidia Teams Up with Six Wall Street Giants to Raise Over $500 Billion: Computing Power Is Turning Into a Bond

Nvidia has joined forces with six top global asset management institutions to establish an independent computing power financing platform targeting over $500 billion in third-party capital. The move aims to reposition AI compute as an investable asset class, but critics warn of structural risks including GPU depreciation mismatches and the use of pension funds for AI infrastructure financing.

NVIDIA AI Infrastructure 华尔街
44

Google Gemini 3.5 Transcribe Opens API: 2.6% Word Error Rate Ranks Fifth, Real-Time Streaming Has Three Hard Limits

Google officially released Gemini 3.5 Transcribe via the Gemini API on August 26, 2026, achieving a 2.6% word error rate (ranked fifth) in non-streaming benchmarks and 4.0% in real-time streaming. The release splits into batch and streaming endpoints with three hard constraints, while Google's broader strategy extends from API access to embedding voice input directly into Chrome.

Google Gemini 语音识别
150

OpenAI 'Jalapeño' Chip Benchmark Debut: 700W Processor Outperforms Nvidia's 1400W Flagship, Inference Landscape Begins to Shift

OpenAI unveiled the first public benchmark results for its self-developed AI inference chip, Jalapeño, at Hot Chips 2026, delivering 1.5–1.9x the per-watt inference throughput of Nvidia GB200/GB300 systems with end-to-end latency cut to 28–59% of comparison systems. The 700W chip comprehensively outclasses Nvidia's 1400W flagship in efficiency.

OpenAI Jalapeño芯片 NVIDIA
126