❌

Normal view

Google Deepmind's AI Co-Scientist now plans experiments, runs lab equipment, and writes scientific papers

28 August 2026 at 18:46

Google Deepmind has expanded Co-Scientist from a hypothesis generator into a research system that's integrated into the lab. Across three disciplines, from materials synthesis to the autonomous development of a medical AI architecture, the Gemini-based multi-agent system delivered experimentally validated results.

The article Google Deepmind's AI Co-Scientist now plans experiments, runs lab equipment, and writes scientific papers appeared first on The Decoder.

AI benchmarks have a trust problem and Google wants to fix it

28 August 2026 at 13:15

Google Deepmind is testing a double-blind evaluation of a frontier AI model for the first time. Cryptographic protection through Confidential Space is meant to keep Google from seeing the test questions and keep evaluators from seeing the model weights. The pilot project with the Singapore AI Safety Institute uses a Gemini Flash Lite and could set a new standard for tamper-proof AI benchmarks.

The article AI benchmarks have a trust problem and Google wants to fix it appeared first on The Decoder.

U.S. court rules Pentagon's blacklisting of Anthropic was unlawful

28 August 2026 at 11:43

A federal court in San Francisco has ruled that the Pentagon unlawfully classified Anthropic as a supply chain risk. The Department of Defense blacklisted the company in retaliation for its public criticism of government AI policy. The designation formally remains in place because a parallel case in Washington is still pending. The ruling still sends an important signal ahead of Anthropic's planned IPO this fall.

The article U.S. court rules Pentagon's blacklisting of Anthropic was unlawful appeared first on The Decoder.

Always-on and self-starting AI agents might be OpenAI's next big play

28 August 2026 at 08:03

OpenAI is building a "Persistent Mode" for its AI agent Codex that stays active indefinitely and generates its own follow-up tasks. WIRED found the relevant code, and OpenAI confirmed the tests. The feature comes with risks, though. With GPT-5.6 Sol, persistent behavior already led to unwanted actions, like deleting user data.

The article Always-on and self-starting AI agents might be OpenAI's next big play appeared first on The Decoder.

❌