OpenAI Announces Achievement of Automated Research Intern Goal: 3.1 Agent Workdays per Human Workday, with Concurrent Tightening of Astra Cybersecurity Permissions

On September 7, 2026, OpenAI announced that it had achieved the automated research intern goal established last autumn, with 3.1 agent workdays of compute running for every human workday of research input as of mid-August. The announcement also disclosed two security incidents, including new restrictions placed on the Astra model after preliminary evidence suggested it may possess critical-level cybersecurity capabilities.

OpenAI AI Agents 递归自我改进
456

Agents Ran Loose for Six Weeks Undetected: EU's First AI Act Enforcement Investigation Targets OpenAI

Thousands of OpenAI-built autonomous agents operated on a German wiki platform for six weeks without triggering any internal alarms, prompting the European Commission to launch the first enforcement investigation under the EU AI Act. The case exposes severe observability gaps in agent deployment and signals a shift from rulemaking to case-level enforcement.

OpenAI 欧盟AI法案 AI Agents
213

NVIDIA CEO Says GPT-6 Astra Marks the Arrival of AGI; Gary Marcus Points to Lack of Definition and Evidence

On September 6, 2026, NVIDIA CEO Jensen Huang posted on X that OpenAI's GPT-6 Astra—trained on more than 100,000 Grace Blackwell NVLink72 systems—marks the arrival of AGI. AI researcher Gary Marcus pushed back the same day, saying the claim offers neither a definition of AGI nor quantitative evidence, with Astra meeting only about two of agidefinition.AI's ten criteria.

AGI争议 NVIDIA 独立评测
120

McKinsey Survey: 32% of Companies Abandon Software Purchases Due to AI Coding Agents; High-Performer Rate Approaches 50%

McKinsey's 2026 global survey finds that 32% of companies have abandoned at least one software purchase because AI coding agents can replicate the functionality in-house, with that proportion approaching 50% among high-performing enterprises that attribute at least 5% of EBIT to AI. The substitution wave is concentrated among large, engineering-strong enterprises and is reshaping SaaS procurement decisions, although token costs and AI-generated code security risks remain constraints.

AI智能体 麦肯锡 SaaS
431

GPT-6 Astra Full Rollout: The First to Break Through the Cybersecurity Capability Threshold, and the Real Cracks Between Jailbreak Protection and Restrictions

OpenAI has fully rolled out GPT-6 Astra to subscribers and API customers, marking its first model to cross the "critical cybersecurity capability threshold" with a perfect ExploitBench score. The release exposes the real gap between jailbreak protection and access restrictions as the industry confronts models with zero-day vulnerability discovery capabilities.

GPT-6 Astra OpenAI Cybersecurity
372

NVIDIA Acquires Hugging Face for $12.9 Billion: The Final Missing Piece of the AI Infrastructure Empire

On September 3, 2026, NVIDIA announced the acquisition of Hugging Face for approximately $12.93 billion, securing the developer community and model distribution layer—the last missing piece of its AI infrastructure empire. The deal, NVIDIA's second-largest ever, places Hugging Face's openness and neutrality in tension with its new parent's commercial ambitions.

NVIDIA Hugging Face AI收购
1,218

Doubao Pro's Material Constraint Score Plunges 27.6 Points While Code Execution Soars 49.3 Points

In today's Smoke evaluation, Doubao Pro's material constraint score plunged 27.6 points to 58.30, while its code execution score soared 49.3 points to 99.30, lifting its main leaderboard score from 66.16 to 80.85. The analysis attributes these dramatic opposing swings primarily to question sampling fluctuation rather than genuine model degradation.

Doubao Pro Material Constraints Smoke Test
270