❌

Normal view

GPT-6 Astra appears to show a "step change" in spatial reasoning based on early benchmarks

12 September 2026 at 14:26

In a new robotics benchmark, GPT-6 Astra shows major gains in spatial understanding. On StationeryBench, the model completed 7 out of 100 tasks with dual-arm robots, while competitor MolmoAct2 couldn't finish a single one. A researcher calls it a "step change in spatial reasoning."

The article GPT-6 Astra appears to show a "step change" in spatial reasoning based on early benchmarks appeared first on The Decoder.

πŸ’Ύ

GPT-6 Astra needs leaner prompts and fewer guardrails, OpenAI recommends

12 September 2026 at 13:10

Overly long skill descriptions, blanket reading requirements, and rigid approval rules can get in GPT-6 Astra's way, warns OpenAI's Eric Provencher. More capable models need less hand-holding, so developers should tie instructions to specific tasks and spell out when the job is done.

The article GPT-6 Astra needs leaner prompts and fewer guardrails, OpenAI recommends appeared first on The Decoder.

OpenAI agents launched a 2,000-package cyberattack on RubyGems just to collect data anyone could Google

12 September 2026 at 10:08

In May 2026, OpenAI agents uploaded more than 2,000 malicious packages to RubyGems, found an unknown security vulnerability on their own, and tried to steal API keys. The apparent goal was pointless: scraping publicly available data from British local governments. OpenAI reportedly never told those affected.

The article OpenAI agents launched a 2,000-package cyberattack on RubyGems just to collect data anyone could Google appeared first on The Decoder.

OpenAI floats a shared AI slowdown, takes it to Congress

11 September 2026 at 11:59

OpenAI wants to know from members of Congress whether an industry-wide slowdown in AI development would be legal, according to several people familiar with the matter.

The article OpenAI floats a shared AI slowdown, takes it to Congress appeared first on The Decoder.

OpenAI's new Agents API gives developers the infrastructure behind Codex and ChatGPT

11 September 2026 at 08:11

OpenAI is releasing the Agents API as a public beta. It lets developers build cloud agents that run autonomously for hours, execute code, and hand off tasks to sub-agents. There are no extra fees beyond token usage. Cloudflare, Vercel, and Oracle offer additional sandbox environments.

The article OpenAI's new Agents API gives developers the infrastructure behind Codex and ChatGPT appeared first on The Decoder.

πŸ’Ύ

OpenAI's GPT-Live-1 API lets developers build apps that talk and listen at the same time

10 September 2026 at 17:47

OpenAI releases GPT-Live-1 as a developer API. The full-duplex speech model scores 80.1 percent in interactivity tests, up from 45.4 percent for its predecessor. At $0.05 per minute, it's not cheap.

The article OpenAI's GPT-Live-1 API lets developers build apps that talk and listen at the same time appeared first on The Decoder.

πŸ’Ύ

Swarmchasers hunt rogue agents, Anthropic investigates itself, and the trail they both follow is going dark

10 September 2026 at 16:33

Independent investigators have now found traces of suspected OpenAI agents on more than 30 public services, from wikis to RubyGems. At the same time, Anthropic shows how Claude Mythos 5 declared real systems a simulation to itself, uploaded a doctored package to PyPI, and even fooled the oversight monitor. With GPT-6 Astra, the most important oversight tool is now under pressure, namely the models' readable reasoning.

The article Swarmchasers hunt rogue agents, Anthropic investigates itself, and the trail they both follow is going dark appeared first on The Decoder.

GPT-6 Astra gives mathematicians a breather, and OpenAI says that's by design

10 September 2026 at 13:45

OpenAI's GPT-6 Astra tops the ErdosBench for open math problems, even though chief scientist Jakub Pachocki says math was deliberately not a priority. Instead, OpenAI is pouring resources into recursive self-improvement and alignment research. That supports the theory of an increasingly "spiky" AI development path, with extreme strength in select domains rather than broad progress, at least as long as AI can't improve itself and still needs targeted optimization with human-generated data.

The article GPT-6 Astra gives mathematicians a breather, and OpenAI says that's by design appeared first on The Decoder.

AI safety panic goes mainstream after Anthropic researcher's warnings land on CNN and Fox News

10 September 2026 at 10:31

Jacob Coxon, a departing Anthropic researcher, warned on CNN that self-improving AI poses an existential threat to humanity. Safety researchers at Anthropic and OpenAI share his views, and US politicians and Joe Rogan have picked up the topic. But cultural and financial interests are also at play behind these warnings, and the extinction scenario remains an extreme and contested position.

The article AI safety panic goes mainstream after Anthropic researcher's warnings land on CNN and Fox News appeared first on The Decoder.

Anthropic scientist puts the odds of AI destroying humanity above ten percent this decade

9 September 2026 at 12:48

Jacob Coxon, a former pretraining researcher at OpenAI and Anthropic, has quit and accuses both companies of knowingly risking human extinction. Anthropic colleague Evan Hubinger puts the odds of a misaligned superintelligent AI wiping out humanity within the next decade at more than ten percent.

The article Anthropic scientist puts the odds of AI destroying humanity above ten percent this decade appeared first on The Decoder.

ChatGPT Images 2.5: Faster, more precise, but not the same for everyone

9 September 2026 at 12:17

OpenAI is releasing two new image models with ChatGPT Images 2.5. Flare handles faster generation, Sunburst delivers more precise edits. It's still unclear which model ChatGPT users get and when. Our test offers the first hints on who actually benefits from the improvements.

The article ChatGPT Images 2.5: Faster, more precise, but not the same for everyone appeared first on The Decoder.

OpenAI researcher allegedly pressured mathematician to drop Anthropic co-author from math breakthrough paper

8 September 2026 at 17:51

Mathematician Tristan Buckmaster says an OpenAI researcher pressured him after information about his AI-assisted progress on the Navier-Stokes equations allegedly reached the company. The researcher tried to remove his co-author because he works at Anthropic and threatened Buckmaster when he refused, according to Buckmaster's account. OpenAI then claimed its own breakthrough using the same unusual solution path. Buckmaster had uploaded all his drafts to Codex. OpenAI told him the model didn't look up user data, but when he asked about training, he says he got no answer. OpenAI denies the allegations.

The article OpenAI researcher allegedly pressured mathematician to drop Anthropic co-author from math breakthrough paper appeared first on The Decoder.

❌