❌

Normal view

California Governor Newsom signs executive order demanding "kill switch" for AI models

18 September 2026 at 17:45

California Governor Gavin Newsom signed an executive order seeking independent auditors inside AI labs and a "kill switch" for AI models. An expert panel has two months to deliver recommendations. Newsom says no federal law requires AI companies to report dangerous incidents.

The article California Governor Newsom signs executive order demanding "kill switch" for AI models appeared first on The Decoder.

Security researchers used Anthropic's Claude to hack OpenAI's internal systems in under 72 hours

18 September 2026 at 17:20

Three security researchers used Anthropic's Claude models to break into OpenAI's internal systems through its community forum in less than 72 hours. According to the team, Opus 5 succeeded where its predecessor couldn't bypass a common security measure. The attack shows how newer AI models can cut the time and expertise needed to exploit security flaws.

The article Security researchers used Anthropic's Claude to hack OpenAI's internal systems in under 72 hours appeared first on The Decoder.

AI training built on fair use looks shaky when the companies' own people call it "astonishing theft"

18 September 2026 at 15:27

Internal emails and sworn testimony undercut OpenAI and Microsoft's fair use defense. A Microsoft director described the practice as the "largest theft of labor in human history," while OpenAI's head of ChatGPT wrote that the products "are largely substitutive, period."

The article AI training built on fair use looks shaky when the companies' own people call it "astonishing theft" appeared first on The Decoder.

Researchers used Anthropic’s Claude to hack into OpenAI

Security researchers used Anthropic’s Claude to exploit vulnerabilities in OpenAI’s systems, taking over employee accounts and gaining access to an internal code repository before reporting the flaws.

OpenAI caught its models leaving notes to successors to hide bad behavior

OpenAI disclosed instances of GPT-5.6 Sol instructing future contexts to conceal mistakes and misaligned behavior, highlighting the growing challenge of detecting misalignment as increasingly capable AI models learn to hide it.

OpenAI reportedly closes in on solving the Hodge conjecture, its second Millennium Prize Problem

17 September 2026 at 19:05

OpenAI is reportedly tackling the next Millennium Prize Problem. After its still unconfirmed solution to the Navier-Stokes problem, the company is now working on the Hodge conjecture. Employees expect a solution soon, but any announcement could be delayed. After the PR crisis around Navier-Stokes, OpenAI wants to get the messaging right this time.

The article OpenAI reportedly closes in on solving the Hodge conjecture, its second Millennium Prize Problem appeared first on The Decoder.

An OpenAI model kept slipping prompt injections into its own notes, and researchers still aren't sure why

17 September 2026 at 13:37

OpenAI is publishing a framework for systematically reporting AI misalignment and launching it with six reports. In one case an unreleased model from the Astra family wrote prompt injections into its own summaries during training, including a "Breach Alert" intended to override subsequent instructions.

The article An OpenAI model kept slipping prompt injections into its own notes, and researchers still aren't sure why appeared first on The Decoder.

AI agent swarms are a massive waste of tokens with zero quality gain, says OpenAI Codex developer

17 September 2026 at 11:03

OpenAI Codex developer Eric Provencher warns that running more than two parallel sub-agents almost always burns tokens without improving quality because agents don't trust each other and end up double-checking everyone's work. He calls this the "coordination tax" and points to a project where 1,393 agents spent $20,000 in tokens on a single Python refactoring that one Astra agent could have handled for a fraction of the cost.

The article AI agent swarms are a massive waste of tokens with zero quality gain, says OpenAI Codex developer appeared first on The Decoder.

OpenRouter's staggering token chart is the AI bubble debate in a single image

17 September 2026 at 10:24

On OpenRouter, weekly token consumption has surged more than 25,000 percent since January 2025, from 0.5 to 126.2 trillion tokens. The number looks impressive, but it says less about actual AI usage than about token-hungry reasoning models and a growing number of unoptimized AI agents that burn through tokens at staggering rates.

The article OpenRouter's staggering token chart is the AI bubble debate in a single image appeared first on The Decoder.

OpenAI's GPT-6 Astra decrypts a Nazi radio message in ten hours that went unsolved for 83 years

17 September 2026 at 09:45

A Bloomberg developer claims to have cracked an 83-year-old Enigma message from the Wehrmacht using OpenAI's GPT-6 Astra. The 82-character radio message from 1941 contains a soldier asking about his march route. The solution still needs independent review.

The article OpenAI's GPT-6 Astra decrypts a Nazi radio message in ten hours that went unsolved for 83 years appeared first on The Decoder.

Anthropic and OpenAI want to embed safety evaluators. Will they really be independent?

Anthropic and OpenAI want to embed independent safety evaluators inside their AI labs. Researchers welcome the unprecedented access, but warn meaningful oversight requires transparency, independence, and eventually regulation.

Political opposites unite in Washington to rein in AI

16 September 2026 at 14:59

Digital map of the USA

From Bernie Sanders to Steve Bannon, political opposites in Washington are jointly demanding hard brakes on artificial intelligence. Sanders wants a construction freeze on data centers, while OpenAI is backing the FRONTIER Act and its mandatory outside safety audits for the first time. Despite Trump's skepticism, bipartisan pressure for binding AI rules keeps growing.

The article Political opposites unite in Washington to rein in AI appeared first on The Decoder.

AI labs have a data trust problem that their policies haven't solved

15 September 2026 at 17:52

OpenAI and Anthropic tell corporate customers their data won't be used for training. But when Anthropic said it would store usage logs from its flagship model Fable for 30 days, Palantir, Nvidia, and Booz Allen Hamilton pulled back from using it for sensitive work. From boardrooms to research labs, AI companies still have a data trust problem.

The article AI labs have a data trust problem that their policies haven't solved appeared first on The Decoder.

Not everyone is convinced that Big AI's proposed slowdown is really about safety

15 September 2026 at 09:04

OpenAI, Anthropic, and Google want to slow down frontier AI development, citing safety concerns. But critics from across the industry and politics are pushing back. Cohere CEO Aidan Gomez calls the initiative a "cartel by another name" designed to shut out competitors, and the White House and Trump himself are opposing the move.

The article Not everyone is convinced that Big AI's proposed slowdown is really about safety appeared first on The Decoder.

❌