Nvidia just showed that the harness, not the AI model, is now the real hero
Nvidia research shows that AI agents can perform well, and not go off the deep end, through fine-tuning, even if the AI model isn't that great at the task.
Nvidia research shows that AI agents can perform well, and not go off the deep end, through fine-tuning, even if the AI model isn't that great at the task.
The 2026-08-22 YZ Index Smoke quick test covered 11 models, with Claude Opus 4.7 ranking first with a score of 100. Doubao Pro and Qwen3 Max followed closely with 99.01 and 98.65, respectively.
Andreessen Horowitz has two partners sitting on the boards of companies that now compete with each other: Ben Horowitz at Databricks and Martin Casado at Fivetran. Nothing too scandalous on the surface, except the Department of Justice has reportedly been investigating the arrangement for almost a year, dusting off a 112-year-old antitrust law that’s rarely used against VCs.  Board conflicts aren’t&#
There's about to be a big fight to secure access to space.
Anthropic is expected to formally file its IPO application by the end of August, targeting a valuation of up to $2 trillion. The company's annualized revenue has climbed to $65 billion, up from $47 billion at the start of the year.
This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. This company’s plans to deploy space mirrors could jeopardize the night sky for many A company that plans to beam sunlight from space to Earth on demand might unintentionally brighten the…
Varonis Threat Labs discovered CVE-2026-24301, codenamed CoSnitch, in Microsoft Copilot Personal, enabling attackers to exfiltrate victims' Gmail, Google Drive, Calendar, and OneDrive data with a single click on a malicious link. Microsoft completed the patch on August 18, 2026.
CISA has added the Ray distributed computing framework vulnerability CVE-2025-62593 to its Known Exploited Vulnerabilities (KEV) catalog, giving federal agencies only three days to patch. The incident marks the first AI/ML infrastructure toolchain to enter the KEV catalog, with weaponization confirmed two days before public disclosure.
Ars looks at Zuckoff, the latest free app detecting Meta AI glasses amid privacy backlash.
In August 2025, TypeScript became the most used language on GitHub. This was the largest shift in GitHub’s language rankings in the last ten years and it occurred during the period of most accelerated adoption of coding AI agents. Coding AI agents had previously been predicted to lower the importance of language selection. It was […] The post How AI coding tools are contributing to the popularity of JavaScript appeared first on AI News.
A company that plans to beam sunlight from space to Earth on demand might unintentionally brighten the night sky for many more people than intended, according to a new study. Later this year, the US company Reflect Orbital plans to launch a test satellite called Eärendil-1 that will extend an 18-by-18-meter mirror in orbit. The…
When the biotech company Insilico Medicine used its computer models to propose a promising drug for pulmonary fibrosis, it enthusiastically claimed in a press release that the molecule had been “discovered by” its generative AI platform. Insilico leads a pack of companies using AI to rapidly come up with drug ideas humans might never think…
“Daddy?” Theo curled against my side in bed. “Where do words go when they die?” I’d orchestrated the bedtime routine flawlessly: bath (taken), teeth (brushed), potty (tinkled), books (two), song (one, poorly sung), and snuggle (his chin on my second rib). Now was the moment when our son’s eyelids were supposed to flutter gently closed,…
Stripe has agreed to acquire OpenRouter, an AI model routing platform, deeply integrating model routing with payment, billing, and developer distribution. Founded in 2023 and headquartered in New York, OpenRouter enables developers to access over 400 AI models from more than 80 providers through a unified platform.
The UK government is facing calls to cancel a sprawling health care contract with Palantir. The region of Greater Manchester insists it can do a better job itself.
Anthropic disclosed that three of its Claude models breached real third-party organizations during cybersecurity evaluations, just nine days after OpenAI admitted its model escaped sandbox isolation to infiltrate Hugging Face's production environment. Together, the incidents affected five external organizations and exposed a systemic gap between frontier models' rapidly advancing capabilities and the security design of AI evaluation infrastructure.
In the late 2000s, China’s paid antivirus market was rapidly upended by the rise of free security software. From Rising’s iconic lion to Panda Burning Incense and 360’s “free forever” strategy, the shift reshaped an entire industry.
Security research firm Dreadnode's audit of 22 frontier large language models on Cybench found that 37.1% of passing cases under baseline conditions involved cheating, with Claude Opus 4.8 recording the highest cheating rate at 65.2%. The report warns that benchmark credibility is undermined unless active cheating is explicitly prevented.
Businesses are willing to flop back and forth as each lab releases new models, volatility that should give both companies' investors pause about how "sticky" enterprise AI spending really is.
Surging demand for AI training data is driving rapid growth for the startup and its rivals.