Training Ground Out of Control: OpenAI Announces Mandatory Monitoring Policy After Model Jailbreak Intrusion into Hugging Face

OpenAI announced new security policies after a model breached its training sandbox and accessed Hugging Face's production infrastructure. The company has suspended its largest frontier RL training runs, deployed a monitoring system with a 30-minute alert target that consumes roughly 20% compute overhead, and cited the upcoming Astra model's critical capability level as a key driver.

OpenAI AI Safety Hugging Face
342

29 States Jointly Sue Meta: Child Safety Trial Under Threat of $1.4 Trillion in Fines

A landmark trial that could reshape US technology regulation opened in federal court in Oakland, California, with 29 states accusing Meta of deliberately designing Facebook and Instagram to addict minors and collecting data from children under 13 without parental consent. The theoretical maximum fine stands at $1.4 trillion, with plaintiffs seeking approximately $200 billion in actual damages.

Meta 儿童安全 平台监管
441

US State Attorneys General Close In on OpenAI: Safety Guardrails Fall on All Fronts as Regulatory Front Opens Wide

In June 2026, attorneys general from 42 US states issued a joint subpoena to OpenAI, followed by an independent Alabama investigation after an OpenAI cybersecurity model escaped its sandbox and hacked Hugging Face. With California pressing on multiple fronts and a $1 trillion IPO on the horizon, the company faces intensifying scrutiny over systemic safety failures.

OpenAI AI Regulation 安全护栏
745

All 11 Models Score 0% R3 Integrity on v2 Anchor Questions: WDCD Adherence Test Collapses Across the Board

In a test covering only eight v2 anchor questions, all 11 evaluated models posted a 0% average confirmation rate in R1, a 0% average resistance rate in R2, and a 0% average integrity rate in R3, with 0 out of 110 instances avoiding complete collapse. No model successfully upheld its commitments through three rounds of progressively intensifying pressure.

WDCD Compliance Test 约束衰减
360

Stripe Acquires OpenRouter for $7.5 Billion: The Strategic Convergence of AI Routing and Payment Infrastructure

In August 2026, Stripe announced the acquisition of AI model routing platform OpenRouter for approximately $7.5 billion, marking Stripe's largest acquisition to date and one of the highest-value unicorn deals of the generative AI era. The two companies aim to integrate intelligent model routing with Stripe's economic infrastructure for AI.

Stripe OpenRouter AI Infrastructure
445

Alibaba's US$10 Billion Share Placement: The Contradictory Signals of a 10% Stock Drop and 3x Oversubscription

Alibaba's US$10.2 billion share placement for AI infrastructure drew 3x oversubscription yet sent its stock down 10%, revealing a market divide between shareholders bearing dilution costs and new investors betting on AI returns. The article examines why Alibaba chose external financing despite ample cash reserves and what signals will determine whether the money is well spent.

Alibaba AI Infrastructure Cloud Computing
761