AI News

Google Gemini 3.5 Transcribe Opens API: 2.6% Word Error Rate Ranks Fifth, Real-Time Streaming Has Three Hard Limits

Google officially released Gemini 3.5 Transcribe via the Gemini API on August 26, 2026, achieving a 2.6% word error rate (ranked fifth) in non-streaming benchmarks and 4.0% in real-time streaming. The release splits into batch and streaming endpoints with three hard constraints, while Google's broader strategy extends from API access to embedding voice input directly into Chrome.

Google Gemini 语音识别
24

OpenAI 'Jalapeño' Chip Benchmark Debut: 700W Processor Outperforms Nvidia's 1400W Flagship, Inference Landscape Begins to Shift

OpenAI unveiled the first public benchmark results for its self-developed AI inference chip, Jalapeño, at Hot Chips 2026, delivering 1.5–1.9x the per-watt inference throughput of Nvidia GB200/GB300 systems with end-to-end latency cut to 28–59% of comparison systems. The 700W chip comprehensively outclasses Nvidia's 1400W flagship in efficiency.

OpenAI Jalapeño芯片 NVIDIA
24
VB

Orchestration is the new challenge for CX in the age of AI agents

Presented by Tata Communications Enterprises are deploying AI agents, voice AI, and automation across messaging, voice, and digital channels faster than the architecture meant to support it. Most of that deployment has involved attaching conversational AI to legacy systems never built for it, says Gaurav Anand, global head of the Customer Interaction Suite at Tata Communications."In the rush to deploy AI, organizations have largely bolted conversational AI onto legacy systems," Anand s

AI agents Customer Experience Orchestration
38