Writing

Essays, working notes, and research reports. For inference and training walkthroughs, see Systems recipes.

Essays and working notes

A Theory on Becoming an Expert

Becoming an expert means deliberately building the mental architecture to judge, question, and understand what models generate.

Building a World Model of Consequence

A proposal to train browser models that predict action consequences and test whether their learned dynamics transfer to new goals.

My Agents Keep Failing. Yours Will Too.

Production agents need shared learning loops where failures become reusable experience across the network instead of one-off human patches.

Everything is Changing...Again

Institutional knowledge becomes dynamic when every diff, decision, and correction is searchable, reviewable, and available at the moment of use.

Research reports

Do Language Models Know When to Change Their Mind?

Models distinguish valid from invalid critique, but reviewer-panel pressure can erase that distinction. Internal readouts transferred; targeted steering lacked specificity.

Frontier Security Agents Lack Restraint

Across 40 simulated episodes per model, eight security agents took incorrect containment actions in 45–97.5% of episodes. Correct actions alone hid over-triggering.