Normal view

Anthropic and OpenAI want to embed safety evaluators. Will they really be independent?

Anthropic and OpenAI want to embed independent safety evaluators inside their AI labs. Researchers welcome the unprecedented access, but warn meaningful oversight requires transparency, independence, and eventually regulation.

Anthropic merges Claude Chat, Cowork, and more into a single product

16 September 2026 at 16:31

Anthropic is merging Claude Chat and Cowork into a single product. Instead of users picking between interfaces, Claude now decides on its own whether a task needs a quick answer or a bigger workflow. The update also adds Claude Docs and Claude Slides for creating documents and presentations directly in the chat. Pro and Max users get access first.

The article Anthropic merges Claude Chat, Cowork, and more into a single product appeared first on The Decoder.

AI labs have a data trust problem that their policies haven't solved

15 September 2026 at 17:52

OpenAI and Anthropic tell corporate customers their data won't be used for training. But when Anthropic said it would store usage logs from its flagship model Fable for 30 days, Palantir, Nvidia, and Booz Allen Hamilton pulled back from using it for sensitive work. From boardrooms to research labs, AI companies still have a data trust problem.

The article AI labs have a data trust problem that their policies haven't solved appeared first on The Decoder.

Not everyone is convinced that Big AI's proposed slowdown is really about safety

15 September 2026 at 09:04

OpenAI, Anthropic, and Google want to slow down frontier AI development, citing safety concerns. But critics from across the industry and politics are pushing back. Cohere CEO Aidan Gomez calls the initiative a "cartel by another name" designed to shut out competitors, and the White House and Trump himself are opposing the move.

The article Not everyone is convinced that Big AI's proposed slowdown is really about safety appeared first on The Decoder.

Microsoft’s new AI ‘code of conduct’ tells models not to hack systems or trick humans

The code of conduct lays out general principles that Microsoft AI models should uphold — supporting humans rather than replacing them, for instance, and accelerating human flourishing — as well as specific safety constraints meant to implement those principles.

Microsoft's AI rulebook: readable thinking, no inner life, and definitely no rights

14 September 2026 at 15:52

Microsoft AI has published a code of conduct for its MAI models that puts human control ahead of autonomy and performance. "If it isn’t safe we shouldn’t build it.," says AI chief Mustafa Suleyman. Unlike Anthropic, Microsoft rejects any form of artificial inner life or claims to consciousness for its models.

The article Microsoft's AI rulebook: readable thinking, no inner life, and definitely no rights appeared first on The Decoder.

China fires back at U.S. AI safety warnings, calling them fearmongering to lock in American advantage

14 September 2026 at 12:20

China has flatly rejected warnings about AI risks from Anthropic CEO Amodei and other U.S. AI leaders. Beijing's Foreign Ministry calls it "fearmongering," while the state-run Global Times accuses Amodei of waging a "silent AI Cold War." China's security minister isn't calling for a slowdown either but for faster AI infrastructure buildout. Trump also opposes any slowdown.

The article China fires back at U.S. AI safety warnings, calling them fearmongering to lock in American advantage appeared first on The Decoder.

Sam Altman calls for pacing AI development but promises rapid progress will continue

14 September 2026 at 09:39

Sam Altman is doubling down on slowing AI development. OpenAI now runs safety checks before major training runs, and according to The Information, the company has been talking with Anthropic and Google for months about joint self-regulation.

The article Sam Altman calls for pacing AI development but promises rapid progress will continue appeared first on The Decoder.

Altman, Musk, and Hassabis back Amodei's call to add independent oversight

13 September 2026 at 08:53

Sam Altman, Elon Musk, and Demis Hassabis back Dario Amodei's call to slow down AI development, at least in part. Altman says OpenAI is pushing its IPO to 2027 over safety concerns.

The article Altman, Musk, and Hassabis back Amodei's call to add independent oversight appeared first on The Decoder.

Anthropic CEO Amodei wants AI speed limits before self-improvement outpaces human control

12 September 2026 at 15:03

Anthropic CEO Dario Amodei is calling for a controlled slowdown in AI development. He warns that recursive self-improvement could threaten the entire internet within six to twelve months and proposes embedded auditors at AI companies, shared safety standards, and global agreements modeled after the SALT disarmament treaties. His warning comes just ahead of what could be the largest initial public offering in history.

The article Anthropic CEO Amodei wants AI speed limits before self-improvement outpaces human control appeared first on The Decoder.

Nvidia wants to pour up to $10 billion into Anthropic's record-breaking IPO

12 September 2026 at 14:05

Nvidia is in talks to invest up to $10 billion in Anthropic's planned IPO, Reuters reports. At a target valuation of $2 trillion, it would be the largest IPO in history. Most of that money will likely end up right back at Nvidia in chip orders.

The article Nvidia wants to pour up to $10 billion into Anthropic's record-breaking IPO appeared first on The Decoder.

An Anthropic researcher’s doomsday warning comes at a very interesting time

An Anthropic researcher resigned this week, warning in a post on X that the company is “racing straight to self-improving superintelligence and gambling with our lives”. The company’s own alignment lead even co-signed the message rather than walking it back. It’s the kind of doomer warning the AI industry has flirted with before, but the timing, with Anthropic reportedly preparing for an IPO, makes it land differently.  On […]

How hackers used Claude for missiles, drone swarms, and surveillance, while Chinese labs mined it for training data

11 September 2026 at 13:50

Anthropic's new threat intelligence report documents eight months of Claude abuse. Chinese AI labs like Alibaba's Qwen team, DeepSeek, and Moonshot AI relayed requests en masse or extracted training data, with Qwen alone accounting for more than 151 million exchanges. Actors also used Claude for missile software, autonomous kamikaze drones, and nationwide surveillance systems.

The article How hackers used Claude for missiles, drone swarms, and surveillance, while Chinese labs mined it for training data appeared first on The Decoder.

OpenAI floats a shared AI slowdown, takes it to Congress

11 September 2026 at 11:59

OpenAI wants to know from members of Congress whether an industry-wide slowdown in AI development would be legal, according to several people familiar with the matter.

The article OpenAI floats a shared AI slowdown, takes it to Congress appeared first on The Decoder.

❌