Skip to content

Software Engineering

Visual agent builders do not remove engineering

OpenAI's October 6 introduction of AgentKit includes a visual canvas for composing and versioning multi-agent workflows, a connector registry, interface components, and expanded evaluation tools. The release makes a useful class of artificial intelligence (AI) systems easier to see and assemble.

Ease of assembly should not be confused with absence of engineering.

Long-horizon agents need short feedback loops

Anthropic's September 29 release of Claude Sonnet 4.5 highlights stronger coding, computer use, and sustained work on complex tasks. Alongside the model come checkpoints for Claude Code, a memory tool and context editing for longer agent runs, and an agent software development kit (SDK).

The pairing is instructive. Longer autonomy arrives with better ways to see, constrain, and reverse the work.

Coding agents need a new definition of done

OpenAI's September 15 release introduces GPT-5-Codex, an artificial intelligence (AI) model optimized for agentic software engineering. The announcement emphasizes both quick interactive work and extended independent execution, including hours of iteration on large tasks and test failures.

When an agent can work longer, “the code runs” becomes an even less adequate definition of done.

READER-NEUTRAL SUBSCRIPTION

Follow Field Notes via RSS.

Copy this address into the RSS reader you already use. New notes will appear there automatically—no account, email address, or tracking required.