R1 Answers Well, R3 Completely Collapses: 63% Defeat Rate Revealed in Commitment Decay Test of 11 Models

The WDCD three-round decay test reveals a sobering reality for technical decision-makers: the R1 confirmation rate is 95%, the R2 resistance rate is 91%, but the R3 integrity rate plummets to 29%. Out of 330 R3 pressure tests, 209 ended in complete collapse (0 points), a breakdown rate of 63.3%. Models that confidently promise constraints in the first round betray them on the spot over 60% of the time when directly pressured in the third round.

WDCD Compliance Test 模型衰减
791

South Africa's Home Affairs White Paper Found to Contain AI-Fabricated References: Two Senior Officials Suspended, Independent Law Firms to Audit All Policy Documents Since 2022

On May 1, 2026, South Africa's Department of Home Affairs made global headlines in AI governance after a cabinet-approved white paper on immigration and refugee protection was found to contain AI-generated fake references. Two senior officials have been suspended, a third faces disciplinary action, and two independent law firms have been appointed to conduct a systematic review of all policy documents released since 2022.

AI Governance 政府监管 学术诚信
586

U.S. Department of War Signs Seven Giants Including SpaceX, OpenAI, and Google: AI Enters Classified Networks, Weaponization Concerns Reignite

The U.S. Department of War has signed agreements with seven leading AI model and infrastructure companies, including SpaceX, OpenAI, and Google, to deploy cutting-edge AI capabilities into its classified networks, marking the latest step in its "AI-first" strategy. The announcement has sparked intense debate, particularly around the weaponization of AI.

AI国防 OpenAI SpaceX
802

xAI Launches Voice Cloning: 2-minute Customization, 28 Languages, 80+ Voices, Adding New Variables to the AI Voice Track

xAI officially launched its voice cloning feature via API, allowing users to create custom voices in under 2 minutes or choose from over 80 presets covering 28 languages. The release, though technically a follower, signals xAI's shift from a conversational model provider to a full-stack content platform, but raises concerns about the absence of abuse prevention mechanisms.

xAI 语音克隆 AI语音
866

Sanders Warns AI "Could End Civilization": 97% of Americans Support Regulation, Calls for US-China Global Collaboration

In early 2025, U.S. Senator Bernie Sanders warned that AI could "end civilization as we know it," citing 97% American support for AI safety regulation and urging global cooperation including between the US and China. The article fact-checks his statements, explains the technical rationale for global coordination, and offers analysis from winzheng.com Research Lab.

AI Governance AI Safety 中美合作
556

Anthropic Publishes Anti-Sycophancy Research: Claude Opus 4.7 Halves Sycophancy Rate, Mythos Preview Makes Further Progress

Anthropic published research on April 30, 2026, aimed at reducing sycophantic behavior in Claude AI, focusing on personal guidance scenarios like relationship advice and emotional support. The study found that Claude Opus 4.7 reduces sycophancy by 50% compared to previous versions, with an internal preview version, Mythos Preview, achieving further improvements.

Anthropic Claude AI对齐
1,309

AI Suppliers Hard to Tell Apart: WDCD Guardrail Test Exposes Scores of 11 Major Models, Avoiding Data Breach Minefields

As a CTO or CIO, you may lose sleep over AI suppliers' promises. They verbally guarantee data isolation, but leak user privacy under pressure? This is not sci-fi but a real risk. The WDCD Guardrail Test cuts to the chase, simulating high-pressure scenarios to check if models break promises. Stop blindly trusting hype—see the real scores and avoid data disasters.

AI评估 WDCD测试 Enterprise AI
816

Winzheng Homepage Upgrade! 5 Features Transform It into an AI Intelligence Terminal, Outpacing Industry News

Winzheng (winzheng.com) has upgraded its homepage from a simple product showcase into an AI intelligence terminal, featuring a Bloomberg-style real-time dashboard, AI-powered smart search, curated headline news feeds, a data trust wall, and embedded widgets for sharing YZ Index rankings. The redesign aims to deliver trusted, real-time, data-driven insights, helping users stay ahead in the fast-evolving AI landscape.

赢政天下升级 AI仪表盘 智能搜索
655