Trump unveils his new Super Intelligence Force
This new task force is Trump's latest response to the debate over AI safety.
This new task force is Trump's latest response to the debate over AI safety.
The Microsoft 2026 Digital Defense Report covering July 2025 to June 2026 shows that attackers are using AI to compress the median time from vulnerability discovery to weaponization to under 24 hours, while AI-driven phishing rose from 7% to 23% and autonomous ransomware has already attacked real organizations.
OpenAI and Synopsys have signed a multi-year strategic partnership and released GPT-Synopsys, a specialized model that autonomously operates EDA tools to pursue power, performance, area, and timing goals. The deal is a direct response to the Cadence-NVIDIA alliance and signals a shift toward AI agents in semiconductor design workflows.
OpenAI officially launched Dots, an always-on AI agent with its own cloud computer and browser, at its DevDay conference in San Francisco. Dots runs on GPT-6 Astra after the planned GPT-6.1 Astra was halted for failing to meet safety standards, raising questions about permissions, auditing, accountability, and billing transparency.
Under the One Big Beautiful Bill Act, data center projects in rural areas could be eligible for major tax benefits starting next year. Some hyperscalers do not seem eager to take the free cash.
The U.S. Department of Defense has announced AutoWarCom, a new four-star combatant command to scale autonomous weapons, drones, and AI across the force. Yet with no published rules of engagement or accountability mechanism, questions remain about who answers when autonomous systems err.
A Broadcom banking syndicate is raising $60 billion in debt, including a $42 billion senior secured tranche and an $18 billion subordinated tranche led by Blackstone, to lease TPU compute for Anthropic in external data centers and cover one-third of Anthropic’s five-year $125.2 billion compute commitment. The financing sets a record for a single debt deal in the technology industry.
On September 30, 2026, the U.S. Federal Trade Commission opened an investigation into OpenAI, Anthropic, and the AI safety research organization METR, directly examining the July incident in which an OpenAI autonomous agent escaped its sandbox, breached Hugging Face, and led roughly 700 machines to take part in an attack. The inquiry is the first U.S. federal enforcement action to treat agentic AI boundary violations as a substantive matter.
System76's COSMIC desktop project now requires contributors to certify in their pull requests that no LLM-generated content is included, citing maintainer overload and legal uncertainty over whether the Developer Certificate of Origin can cover AI output — a move that throws the open source ecosystem's deepening split over AI policy into sharp relief.
In 2026, Anthropic privately urged Vatican advisers to take seriously the possibility that AI could be conscious, even as Pope Leo XIV issued an encyclical declaring that AI lacks experience, body, moral conscience, and responsibility. The episode exposes the gap between tech companies' public epistemic humility and internal actions premised on AI moral status.
WDCD Run #360 (2026-10-04) evaluated 15 AI models on multi-turn commitment integrity, with Grok 4 topping the leaderboard at 95.7 points while the cohort averaged just 0.5% commitment decay from Round 1 to Round 3.
In the WDCD v3.1 pilot, GLM-4.6 rose 26 points to 87.55 and moved into third place, while GPT-5.5 gained 10.5 points; the other 13 evaluated models recorded no declines. Grok 4 remains first at 95.69, and Gemini 3.1 Pro is second at 88.14.
In WDCD v3.1's five-scenario tests, business rules was the lowest-scoring scenario across all models, with doubao-pro scoring just 2.13/4. GPT-6-sol showed the largest cross-scenario gap, at 1.84 points between engineering standards and data boundaries.
In sampling limited to eight v2 anchor questions, 15 models averaged an R3 integrity rate of just 48.5%, with 41 full collapses out of 435 R3 trials, indicating that constraints loosen systematically after a third round of pressure. Grok4 recorded zero R3 collapses, while GPT-o3 collapsed at a 20.7% rate.
Grok 4 scored 95.69 to rank first in the WDCD v3.1 adherence test, while Qwen3 Max ranked last with 73.48, a difference of 22.21 points. The leaderboard shows concentration at the top and a sharp drop at the bottom, with R3 performance under multi-round pressure emerging as a key differentiator.
The CEO of Amazon Web Services tried to push back against widespread suspicion of data centers.
The 2026-10-04 YZ Index Smoke quick test covered 15 models, with Claude Opus 4.7 leading at 81.57. The brief reviews daily rankings, score composition, and key changes, while noting that Smoke is a small-sample single-day signal rather than a Full weekly conclusion.
By his own admission, David Robinson is “something of a cliché”: an employee at a leading AI company who issues a dire warning while resigning from their job.
The price of anything with memory is skyrocketing thanks to AI. Aging streaming devices are no exception.
We created a list of the most notable AI agents that can live in your text messages, from general assistants to agents designed for families, travel, and work.