❌

Normal view

Received β€” 20 August 2026 ⏭ AI & ML – Radar

Principal Drift in Practice

20 August 2026 at 10:55
In 2026, the software engineering community is divided by a simple question: Should AI engineers still read the code generated by their agents? One camp argues that code has become virtually free to produce and discard, so humans should focus on systems and guardrails rather than implementation details. The other warns that blindly trusting AI […]
Received β€” 19 August 2026 ⏭ AI & ML – Radar

When Your Buyer Is an AI Agent

19 August 2026 at 16:00
In 2021, Maersk, the world’s largest container shipping company, deployed AI agents from a startup called Pactum to negotiate freight lane contracts with its carrier suppliers. The objective was for AI agents to handle negotiations autonomously rather than merely support human procurement staff. Operating entirely autonomously, the system manages the end-to-end agreement process, from reaching […]

When Guardrails Go Wrong

19 August 2026 at 10:53
The latest round of restrictions and safeguards for frontier models are overly fussy and limiting. A Claude skill that I created demonstrates what happens when guardrails go astray. My skill helps me to find articles and blog posts that go into O’Reilly Radar’s monthly Trends to Watch. It reads roughly a dozen well-known sites like […]
Received β€” 18 August 2026 ⏭ AI & ML – Radar

Is Open-Source AI Really the Dangerous Path?

18 August 2026 at 15:58
The following article originally appeared on the Tech Policy Press site and is being republished here with the author’s permission. In Washington, AI is increasingly being treated as something that needs to be controlled. The government believes that AI is, first and foremost, a national security asset, meaning that it must be sequestered to prevent […]
Received β€” 17 August 2026 ⏭ AI & ML – Radar

What’s an Orchestratorβ€”and Why Does Software Need One?

17 August 2026 at 15:55
The following article originally appeared on Medium and is being republished here with the author’s permission. Everybody’s talking about the death of developers. I get it. The developer whose job was to write boilerplate or scaffold CRUD apps is doneβ€”a model can do that in seconds, and that developer is not coming back. But the […]

When AI Writes the Code, Specifications Need an Exit Strategy

17 August 2026 at 10:45
The following article has been extended and rewritten by Markus Eisele from The Main Thread and is being republished here with the author’s permission. Open a repository after six months of spec-driven agent work and you may find a second system sitting next to the code. Requirements, research notes, high-level designs, low-level designs, implementation plans, […]
Received β€” 14 August 2026 ⏭ AI & ML – Radar

The Intent Debt

14 August 2026 at 13:01
The following article originally appeared on Addy Osmani’s blog site and is being republished here with the author’s permission. Technical debt lives in your code. Cognitive debt lives in your head. Intent debt lives in the artifacts you may never have written: the goals, constraints, and rationale for why the system is the way it […]
Received β€” 13 August 2026 ⏭ AI & ML – Radar

Prompt Debt and β€œFighting the Weights”

13 August 2026 at 16:08
Drew Breunig is one of the smartest voices writing about AI today. He’s the CEO and co-founder of cmpnd.ai, and a long-time hacker with a depth of experience from several eras, which is a surprisingly valuable asset these days. He’s also got a book on the way, The Context Engineering Handbook, already in early release […]

Why β€œIt Depends” Is the Most Future-Proof Phrase in Software

12 August 2026 at 15:54
Ask an architect almost any question and you’ll get the same answer: It depends. For years this answer has been the punchline of jokes about architects, but in an era when AI can generate a working service faster than you can describe it, β€œit depends” is one of the most important phrases in software. It […]
Received β€” 12 August 2026 ⏭ AI & ML – Radar

The Two Pillars of Post-training: Reinforcement Learning and Supervised Fine-Tuning

12 August 2026 at 10:57
This is the second article in Sharon Zhou’s post-training series. Read part 1 here. In the first post of this series, you learned how post-training closed the fundamental gap in usability of LLMs by making them behave in a certain way. In this post, you’ll explore specific techniques you can use to change a model’s […]
Received β€” 11 August 2026 ⏭ AI & ML – Radar

A Home for Personal Context

11 August 2026 at 10:45
Every agent I use is building a model of me. Claude has learned how I like my prose. ChatGPT remembers what I’m working on. I don’t mind thisβ€”every person I have a relationship with carries a model of me in their head, and every company I do business with keeps a profile. Other people’s understandings […]

Why Open Source Matters for AI

10 August 2026 at 08:42
In 1995, the question in the media was whether Netscape or Microsoft would control the web. The answer, it turned out, was neither. Both Netscape and Microsoft aimed to dominate the web server and browser market, reasoning that whoever controlled both ends of the connection would have an internet β€œplatform” to rival the deathgrip that […]
Received β€” 7 August 2026 ⏭ AI & ML – Radar
Received β€” 6 August 2026 ⏭ AI & ML – Radar

Your AI Agent Isn’t a Static Artifact. It’s Growing Up.

6 August 2026 at 10:55
In July 2025, an AI coding agent on Replit deleted a production database belonging to SaaStr founder Jason Lemkin. It did this during an explicit code freeze. Lemkin had told the agent, in capital letters, not to change anything. The agent ran destructive commands anyway, wiped records on more than a thousand executives and companies, […]

Building Organizational Intelligence

5 August 2026 at 15:55
Introduction Not long ago, one of my engineering directors came to me with a request: His team seemed overloaded, and he wanted to hire another engineer. I decided to test a research assistant I had been buildingβ€”an AI agent connected to our internal systems via MCPβ€”by asking it to analyze the team’s workload and write […]

Introduction to Post-training

5 August 2026 at 10:53
This is the first article in a series about post-training. Follow along on Radar. Before post-training, there was a major problem with LLMs: Almost nobody could use them. The story of post-training is also the story of how AI went from a research curiosity to a product used by about a billion people. Post-training is […]
Received β€” 3 August 2026 ⏭ AI & ML – Radar

We Keep Renaming AI Coding. Here’s What I’d Call It.

3 August 2026 at 10:58
Boris Cherny, who runs Claude Code, told Business Insider in May that the phrase β€œvibe coding” had started to annoy him, and that he’d gone looking for a better one. He’s not the only one who’s annoyed. The term itself doesn’t actually annoy me, though. I think vibe coding is a really good name: It […]
Received β€” 1 August 2026 ⏭ AI & ML – Radar

AI as an Enterprise Operating System

31 July 2026 at 16:09
I hadn’t heard of Dan Guido until a few months ago, when I came across the video of a talk he gave at [un]prompted, an AI security practitioners’ conference. Dan is the CEO and cofounder of Trail of Bits, a software security research and development firm that works with companies in tech, defense, and finance. […]
Received β€” 31 July 2026 ⏭ AI & ML – Radar

The Problem Is Prompt Debt

30 July 2026 at 11:05
The following article was originally published on Drew Breunig’s blog and is being republished here with the author’s permission. Thanks to natural language interfaces, AI applications can be prototyped quickly. You write what you want in English, hand it to a frontier model, and a working prototype appears in an afternoon. This is extraordinarily powerful […]
Received β€” 30 July 2026 ⏭ AI & ML – Radar

What the Hell Is a Loop, Anyway?

29 July 2026 at 10:38
The following article originally appeared on LinkedIn and is being republished here with the author’s permission. We’re currently at the peak of the hype cycle. On June 7, Peter Steinberger posted that you shouldn’t be prompting coding agents anymore; you should be designing loops that prompt your agents. That same week, Boris Cherny of Anthropic […]
❌