Opaque recurrence, and other AI terms that you should probably know
The rise of AI has brought an avalanche of new terms and slang. Here is a glossary with definitions of some of the most important words and phrases you might encounter.
The rise of AI has brought an avalanche of new terms and slang. Here is a glossary with definitions of some of the most important words and phrases you might encounter.
On 2026-09-08, the YZ Index Smoke Quick Test covered 11 models. Claude Sonnet 4.6 and Grok 4 tied for the top overall score at 86.25 points, while several models showed notable single-day score swings.
When multiple companies are behind one project, who bears responsibility for problems?
This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. How much hydrogen awaits us underground? A flurry of exploration efforts is searching for underground stores of hydrogen gas, which could provide a valuable source of zero-carbon fuel. The hunt has…
MG Ship has introduced an AI route optimisation and carrier selection module as logistics deployments demonstrate rapid cost and time returns. The technical module targets global retailers and commercial shippers, pairing automated routing algorithms with carrier recommendation systems across international trade corridors. The deployment arrives as enterprise supply chain operators report measurable operational returns from […] The post MG Ship adds AI route optimisation as logistics retur
McKinsey's 2026 global survey finds that 32% of companies have abandoned at least one software purchase because AI coding agents can replicate the functionality in-house, with that proportion approaching 50% among high-performing enterprises that attribute at least 5% of EBIT to AI. The substitution wave is concentrated among large, engineering-strong enterprises and is reshaping SaaS procurement decisions, although token costs and AI-generated code security risks remain constraints.
OpenAI has fully rolled out GPT-6 Astra to subscribers and API customers, marking its first model to cross the "critical cybersecurity capability threshold" with a perfect ExploitBench score. The release exposes the real gap between jailbreak protection and access restrictions as the industry confronts models with zero-day vulnerability discovery capabilities.
Forget email surveys or long calls spent on hold. Voicebox lets people send customer feedback by recording a voice note on their phone.
When multiple companies are behind one project, who bears responsibility for problems?
Anthropic has signed $517 billion in compute-capacity contracts over 11 months, sharply expanding its infrastructure commitments ahead of a potential IPO. The scale of these deals highlights both the company’s ambition to close the compute gap with OpenAI and the financial pressure created by long-term obligations.
On September 6, 2026, OpenAI Chief Scientist Jakub Pachocki published an essay warning that chain-of-thought monitoring is quietly failing and recursive self-improvement is approaching a critical threshold, urging voluntary deceleration, mandatory third-party audits, and international coordination.
This week's 368 translation tasks were handled by 5 models. A 3-article multi-model blind review crowned claude-sonnet-4.6 the overall best, with an average score of 9/10.
The Uber founder has said that Atoms will allow him to complete "unfinished business."
Authors say publishers seem to be claiming more than their fair share of settlement payments.
On September 2, 2026, Google released Gemini 3.8 Flash and Gemini 3.8 Flash Cyber. The general-purpose model scored 73.7% on DeepSWE v1.1—within one percentage point of Claude Opus 5's 74.0%—at an invocation cost of just 15% of the latter's, while the cybersecurity variant topped all commercial models tested in the same period on CyberGym.
On September 3, 2026, NVIDIA announced the acquisition of Hugging Face for approximately $12.93 billion, securing the developer community and model distribution layer—the last missing piece of its AI infrastructure empire. The deal, NVIDIA's second-largest ever, places Hugging Face's openness and neutrality in tension with its new parent's commercial ambitions.
In today's Smoke evaluation, GPT-o3's material constraint fell from 70.00 to 50.00 and engineering judgment from 100.00 to 50.00, yet its main leaderboard score rose from 72.75 to 77.50.
In today's Smoke evaluation, Doubao Pro's material constraint score plunged 27.6 points to 58.30, while its code execution score soared 49.3 points to 99.30, lifting its main leaderboard score from 66.16 to 80.85. The analysis attributes these dramatic opposing swings primarily to question sampling fluctuation rather than genuine model degradation.
The 2026-09-07 YZ Index Smoke quick test covered 11 models, with DeepSeek V4 Pro, Gemini 3.1 Pro, and GLM-4.6 tying for the top spot of the day at 83.49 points. Smoke is a daily 10-question quick test for observing short-term signals and is not equivalent to the Full weekly leaderboard conclusions.
On September 3, 2026, OpenAI released GPT-6 Astra, claiming a 91.5% to 98.3% refusal rate on fixed jailbreak attack datasets, yet a researcher publicly reported a successful breach within 24 hours using an extended Task-in-Prompt attack combined with four other methods. The incident exposes a persistent structural gap between static benchmark claims and adaptive multi-turn attacks.