Skip to content

2025

Thunderforge will be judged by the planning system around it

Military planning is a coordination technology. It turns incomplete information, command intent, operational constraints, staff expertise, and adversary uncertainty into courses of action that can be compared and executed.

The Defense Innovation Unit's (DIU) March 5 announcement introduces Thunderforge, an initiative to bring artificial intelligence (AI) agents, modeling, simulation, and large language models into operational and theater-level planning. The stated ambition is faster synthesis, course-of-action development, and AI-enabled wargaming.

Coding agents change the shape of engineering work

The February 24 preview of Claude Code is more consequential than another improvement in code completion. The tool can search a repository, edit files, run tests, and use command-line tools to carry out a substantial engineering task.

That changes the unit of delegation. Instead of asking artificial intelligence to suggest the next line, a developer can ask it to pursue an outcome.

Simulation is becoming part of the decision-support stack

Microsoft's February 19 introduction of Muse presents an artificial intelligence (AI) model capable of generating video-game visuals and controller actions. It is trained as a World and Human Action Model (WHAM), learning both how an environment changes and how people act within it.

Gaming is the immediate application. The larger signal is that generative simulation is moving closer to an interactive design material.

AI adoption begins with tasks, not job titles

Conversations about artificial intelligence and work often begin at the wrong level. They ask which jobs will disappear, then argue over forecasts that are too coarse to guide an actual organization.

Anthropic's first Economic Index, published February 10, analyzes how people use Claude across occupational tasks. The report finds usage concentrated in particular kinds of work and distinguishes between automation, where the model performs a task, and augmentation, where people and the model work together.

Frontier safety must be governed as a moving threshold

When a technology changes quickly, a fixed policy can be obsolete while everyone is still complying with it.

Google DeepMind's February update to its Frontier Safety Framework addresses that problem by linking stronger safeguards to capability thresholds in areas that could create severe harm. The details will continue to evolve. The organizational principle should endure: controls should respond to what a system can do, not only to the name or generation printed on it.

The national laboratories are an unusually good test for AI

OpenAI's January 30 announcement makes its reasoning models available to scientists at Los Alamos, Lawrence Livermore, and Sandia national laboratories through a partnership using the Venado supercomputer.

The announcement covers an extraordinary range of possibilities: materials, energy, medicine, cybersecurity, mathematics, and nuclear security. Yet the laboratories are interesting not only because their problems are difficult. They are interesting because their knowledge practices are demanding.

A management-system certificate is a beginning, not a verdict

Anthropic's January 13 announcement reports that the company has achieved International Organization for Standardization and International Electrotechnical Commission (ISO/IEC) 42001 certification, making it one of the first frontier-model companies to certify an artificial intelligence management system against the new international standard.

That is meaningful. It is also easy to misunderstand.

READER-NEUTRAL SUBSCRIPTION

Follow Field Notes via RSS.

Copy this address into the RSS reader you already use. New notes will appear there automatically—no account, email address, or tracking required.