Normal view
-
AI News & Artificial Intelligence | TechCrunch
- Anthropic is operating a lab that conducts biology experiments
-
AI News & Artificial Intelligence | TechCrunch
- Anthropicβs first embedded evaluator isΒ β¦ Accenture?
Anthropicβs first embedded evaluator isΒ β¦ Accenture?
California Governor Newsom signs executive order demanding "kill switch" for AI models
![]()
California Governor Gavin Newsom signed an executive order seeking independent auditors inside AI labs and a "kill switch" for AI models. An expert panel has two months to deliver recommendations. Newsom says no federal law requires AI companies to report dangerous incidents.
The article California Governor Newsom signs executive order demanding "kill switch" for AI models appeared first on The Decoder.
-
THE DECODER
- Security researchers used Anthropic's Claude to hack OpenAI's internal systems in under 72 hours
Security researchers used Anthropic's Claude to hack OpenAI's internal systems in under 72 hours
![]()
Three security researchers used Anthropic's Claude models to break into OpenAI's internal systems through its community forum in less than 72 hours. According to the team, Opus 5 succeeded where its predecessor couldn't bypass a common security measure. The attack shows how newer AI models can cut the time and expertise needed to exploit security flaws.
The article Security researchers used Anthropic's Claude to hack OpenAI's internal systems in under 72 hours appeared first on The Decoder.
-
AI News & Artificial Intelligence | TechCrunch
- Dario Amodei and other AI leaders want to βPace the Frontierβ butβ¦how?
Dario Amodei and other AI leaders want to βPace the Frontierβ butβ¦how?
-
AI News & Artificial Intelligence | TechCrunch
- Automatticβs 33-Hour Coup, and can AI labs police themselves?
Automatticβs 33-Hour Coup, and can AI labs police themselves?
-
THE DECODER
- Anthropic wants you to know Claude leads a quarter of its research, but "lead" doesn't mean what you think
Anthropic wants you to know Claude leads a quarter of its research, but "lead" doesn't mean what you think
![]()
For the first time, Anthropic is releasing metrics on how it builds its own AI. Claude already "leads" 26 percent of the work on future models, up from under one percent in February. But the underlying scale is fuzzy, the scoring comes from Claude itself, and "lead" means less than it sounds.
The article Anthropic wants you to know Claude leads a quarter of its research, but "lead" doesn't mean what you think appeared first on The Decoder.
-
AI News & Artificial Intelligence | TechCrunch
- Researchers used Anthropicβs Claude to hack into OpenAI
Researchers used Anthropicβs Claude to hack into OpenAI
LLMs respond differently to harmful prompts when AI watermarking is used
In response to a new European Union law, AI platforms are implementing new schemes for watermarking the content they generate. Anthropic recently disclosed its future Claude models will use SynthID-Text, an approach Google created and released as open source. It uses a secret key that subtly changes the process a model uses for choosing the next word in a sentence. Whereas a top next word choice might be βcloudy,β the key might change it to βovercast.β Anyone who knows the key can determine if it was generated by the platform using it.
New research shows that SynthID-Text can change not just word selection but also the tools a model invokes and the chances it will adhere to or disregard safety guardrails it has been trained to follow. The threat can become greater in the face of an adversarial prompt, in which an attacker attempts to cause a model to carry out a harmful action, such as revealing a password or other sensitive information. Instructions that normally wouldnβt be followed will, in some cases, be performed once the watermarking is deployed. The finding underscores the need for developers to thoroughly test how their LLMs and agents behave when watermarking is in place.
Changing safety behavior
βAs compared to the same models without watermarking, it is definitely going to change their behavior, especially when we place it under adversarial conditions, or we make these models call tools when theyβre powering an agent,β Andrea Siposova, an AI security researcher at Lasso Security, told Ars. βWatermarking is made to not be perceptible to a reader, but we know that when we are changing anything about what the model is generating, it is going to cause some tradeoffs, itβs going to show up somewhere.β


Β© Getty Images
Is the AI safety debate about safety or control?
-
THE DECODER
- Anthropic keeps pushing Claude Code toward autonomous coding with new parallel agent workflows
Anthropic keeps pushing Claude Code toward autonomous coding with new parallel agent workflows
![]()
Anthropic has rebuilt Projects in Claude Code. A coordinator now splits tasks across parallel cloud threads that independently open pull requests and run tests. All threads share a common memory. The beta is available to select Pro and Max subscribers.
The article Anthropic keeps pushing Claude Code toward autonomous coding with new parallel agent workflows appeared first on The Decoder.
-
AI News & Artificial Intelligence | TechCrunch
- Google, Nvidia, and Anthropic want Emerald AI to find space on the grid for more data centers
Google, Nvidia, and Anthropic want Emerald AI to find space on the grid for more data centers
OpenRouter's staggering token chart is the AI bubble debate in a single image
![]()
On OpenRouter, weekly token consumption has surged more than 25,000 percent since January 2025, from 0.5 to 126.2 trillion tokens. The number looks impressive, but it says less about actual AI usage than about token-hungry reasoning models and a growing number of unoptimized AI agents that burn through tokens at staggering rates.
The article OpenRouter's staggering token chart is the AI bubble debate in a single image appeared first on The Decoder.
Anthropic merges Claude Chat, Cowork, and more into a single product
![]()
Anthropic is merging Claude Chat and Cowork into a single product. Instead of users picking between interfaces, Claude now decides on its own whether a task needs a quick answer or a bigger workflow. The update also adds Claude Docs and Claude Slides for creating documents and presentations directly in the chat. Pro and Max users get access first.
The article Anthropic merges Claude Chat, Cowork, and more into a single product appeared first on The Decoder.
AI labs have a data trust problem that their policies haven't solved
![]()
OpenAI and Anthropic tell corporate customers their data won't be used for training. But when Anthropic said it would store usage logs from its flagship model Fable for 30 days, Palantir, Nvidia, and Booz Allen Hamilton pulled back from using it for sensitive work. From boardrooms to research labs, AI companies still have a data trust problem.
The article AI labs have a data trust problem that their policies haven't solved appeared first on The Decoder.
Not everyone is convinced that Big AI's proposed slowdown is really about safety
![]()
OpenAI, Anthropic, and Google want to slow down frontier AI development, citing safety concerns. But critics from across the industry and politics are pushing back. Cohere CEO Aidan Gomez calls the initiative a "cartel by another name" designed to shut out competitors, and the White House and Trump himself are opposing the move.
The article Not everyone is convinced that Big AI's proposed slowdown is really about safety appeared first on The Decoder.
Microsoft's AI rulebook: readable thinking, no inner life, and definitely no rights
![]()
Microsoft AI has published a code of conduct for its MAI models that puts human control ahead of autonomy and performance. "If it isnβt safe we shouldnβt build it.," says AI chief Mustafa Suleyman. Unlike Anthropic, Microsoft rejects any form of artificial inner life or claims to consciousness for its models.
The article Microsoft's AI rulebook: readable thinking, no inner life, and definitely no rights appeared first on The Decoder.
-
THE DECODER
- Anthropic eyes Nasdaq listing as a second profitable quarter aims to win over investors ahead of a mega-IPO
Anthropic eyes Nasdaq listing as a second profitable quarter aims to win over investors ahead of a mega-IPO
![]()
Anthropic has told investors it will turn a profit for the second straight quarter,Β but the claim rests on an adjusted metric that leaves out costs like stock-based compensation.
The article Anthropic eyes Nasdaq listing as a second profitable quarter aims to win over investors ahead of a mega-IPO appeared first on The Decoder.
-
THE DECODER
- China fires back at U.S. AI safety warnings, calling them fearmongering to lock in American advantage
China fires back at U.S. AI safety warnings, calling them fearmongering to lock in American advantage
![]()
China has flatly rejected warnings about AI risks from Anthropic CEO Amodei and other U.S. AI leaders. Beijing's Foreign Ministry calls it "fearmongering," while the state-run Global Times accuses Amodei of waging a "silent AI Cold War." China's security minister isn't calling for a slowdown either but for faster AI infrastructure buildout. Trump also opposes any slowdown.
The article China fires back at U.S. AI safety warnings, calling them fearmongering to lock in American advantage appeared first on The Decoder.
Sam Altman calls for pacing AI development but promises rapid progress will continue
![]()
Sam Altman is doubling down on slowing AI development. OpenAI now runs safety checks before major training runs, and according to The Information, the company has been talking with Anthropic and Google for months about joint self-regulation.
The article Sam Altman calls for pacing AI development but promises rapid progress will continue appeared first on The Decoder.