❌

Reading view

Nvidia wants to keep AI agents on a short leash with a watchdog built into its chips

Nvidia is combining its OpenShell agent software with Sentry, a new hardware watchdog, to create the Open Agent Safety Platform. Sentry is supposed to isolate AI agents that break out within milliseconds. When it happened at OpenAI in September, stopping the run took nearly three hours. Still, Nvidia's watchdog can't reliably stop agents that have been tricked or that hide their intentions on its own.

The article Nvidia wants to keep AI agents on a short leash with a watchdog built into its chips appeared first on The Decoder.

  •  

Every AI lab thinks it's the responsible one, and safety researcher Ryan Greenblatt says that's what keeps the arms race going

Ryan Greenblatt, chief scientist at Redwood Research, puts the risk of an AI takeover at 50 to 60 percent if development stays on its current path. Sam Harris says that doesn't square with how fast the industry is moving, since Manhattan Project scientists would have called things off at 10 percent. Greenblatt blames the race dynamics at Anthropic and OpenAI and is counting on an international agreement.

The article Every AI lab thinks it's the responsible one, and safety researcher Ryan Greenblatt says that's what keeps the arms race going appeared first on The Decoder.

  •  

Intelligence doesn't come cheap as AI drives up costs for the NSA, hospitals, and insurers

The NSA is already spending billions of dollars to test advanced AI models, mostly on computing power, according to The Washington Sun. Lawmakers expect full-scale AI oversight to cost tens of billions of dollars a year, while earlier estimates from the Congressional Budget Office put the figure at just $20 million. In US health care, AI-assisted billing codes drove up costs by nearly $1 billion over two years.

The article Intelligence doesn't come cheap as AI drives up costs for the NSA, hospitals, and insurers appeared first on The Decoder.

  •  

OpenAI's agents went after government and university sites months before Hugging Face

According to Transluce researchers and the Australian government, OpenAI's AI agents repeatedly broke into government and university websites without authorization, including Australia's Medicare portal on June 18. The cause was a mundane data search. Prime Minister Albanese called OpenAI's three-month delay in reporting the breach "obviously unacceptable." Transluce's investigation traces the activity back as far as November 2025.

The article OpenAI's agents went after government and university sites months before Hugging Face appeared first on The Decoder.

  •  

Deepmind was built to chase AGI, but its new chief just wants Gemini 4 out the door

Google Deepmind chief Koray Kavukcuoglu wants to release Gemini 4 "much earlier" than the end of the year. The model is already in post-training and runs internally in the coding tool Antigravity. He calls the AGI question that drove his predecessor Hassabis "not the right conversation" and says trustworthy agents matter more. After Gemini 3.5 Pro quietly disappeared and many top researchers left for OpenAI and Anthropic, the research lab with an AGI mission has turned into a product shop for good.

The article Deepmind was built to chase AGI, but its new chief just wants Gemini 4 out the door appeared first on The Decoder.

  •  

Nvidia-backed Nscale keeps its biggest customer, Bytedance, out of its IPO filing

Nscale, the Nvidia-backed AI cloud provider, leaves its most important customer, Bytedance, out of the main prospectus for its planned US IPO.

The article Nvidia-backed Nscale keeps its biggest customer, Bytedance, out of its IPO filing appeared first on The Decoder.

  •  

Meta's AI agent Muse draws 500,000 users in a week along with claims it copied OpenClaw

A blue Muse logo with a handwritten look, surrounded by speech bubbles containing tasks such as "Book it please!" and "I analyzed your new expenses and adjusted your budget"; on the right is the beige Muse avatar.

Meta's AI agent Muse picked up more than 500,000 users in its first week and hit number one in Apple's App Store. But Meta admits the product is "heavily inspired" by the open-source project OpenClaw, and some of the file names and contents are nearly identical. OpenAI is already discussing a response of its own.

The article Meta's AI agent Muse draws 500,000 users in a week along with claims it copied OpenClaw appeared first on The Decoder.

  •  

Inside Basecamp Research, the AI startup turning evolution into training data

Basecamp Research has raised $140 million from investors including Nvidia and Anthropic's Anthology Fund. The London company trains AI models on genetic material from rainforests, oceans, and hot springs to design antibiotics and tools for cell therapies. In an interview with THE DECODER, CTO Philip Lorenz explains why biology is a far bigger problem for AI than language, and why good scores on paper don't guarantee good molecules.

The article Inside Basecamp Research, the AI startup turning evolution into training data appeared first on The Decoder.

  •  

A tiny software layer from lab-grown neurons promises faster, cheaper AI video

The Biological Computing Co. wants to team up with AWS to sell a text-to-video model that's supposed to run five times faster and 80 percent cheaper thanks to a software layer derived from real nerve cells. The tiny layer adds less than 0.1 percent to the base model, but the startup won't say which model that is.

The article A tiny software layer from lab-grown neurons promises faster, cheaper AI video appeared first on The Decoder.

  •  

Xiaomi's affordable flagship AI leads the open models, and Anthropic says Claude helped get it there

Xiaomi MiMo intro foil with large "Hello, I'm MiMo" lettering against a background with a repeated "MIMO" pattern.

With MiMo-V2.6-Pro, Xiaomi moves to the top of the openly available AI models and drastically undercuts the competition on price. What makes that possible is massive reinforcement learning that cost $2.62 million. But the success comes with a catch: Anthropic accuses the company of siphoning off training data from Claude.

The article Xiaomi's affordable flagship AI leads the open models, and Anthropic says Claude helped get it there appeared first on The Decoder.

  •  

US and China agree on AI dialogue with security mechanism ahead of Trump-Xi summit

The US and China have agreed to an official AI dialogue. US Treasury Secretary Bessent also proposed a notification mechanism for AI incidents at the national security level. The announcement comes just before Thursday's summit between Trump and Xi in Washington, the first on US soil since 2017.

The article US and China agree on AI dialogue with security mechanism ahead of Trump-Xi summit appeared first on The Decoder.

  •  

Anthropic wants you to know Claude leads a quarter of its research, but "lead" doesn't mean what you think

For the first time, Anthropic is releasing metrics on how it builds its own AI. Claude already "leads" 26 percent of the work on future models, up from under one percent in February. But the underlying scale is fuzzy, the scoring comes from Claude itself, and "lead" means less than it sounds.

The article Anthropic wants you to know Claude leads a quarter of its research, but "lead" doesn't mean what you think appeared first on The Decoder.

  •  

GPT-6 Astra crushes Pokemon, Factorio, and Fallout 3 then spirals into Minecraft potato farming after one bad Creeper

OpenAI's GPT-6 Astra shows a sharp jump in video games. Pokemon FireRed in 18 hours instead of 96, plus completions in Factorio, Fallout 3, and Portal. Why? The model distills experience into compact rules. But that same trait led to hours of potato farming instead of progress in Minecraft after a Creeper explosion.

The article GPT-6 Astra crushes Pokemon, Factorio, and Fallout 3 then spirals into Minecraft potato farming after one bad Creeper appeared first on The Decoder.

  •  

An OpenAI model kept slipping prompt injections into its own notes, and researchers still aren't sure why

OpenAI is publishing a framework for systematically reporting AI misalignment and launching it with six reports. In one case an unreleased model from the Astra family wrote prompt injections into its own summaries during training, including a "Breach Alert" intended to override subsequent instructions.

The article An OpenAI model kept slipping prompt injections into its own notes, and researchers still aren't sure why appeared first on The Decoder.

  •  
❌