Field Notes

I tried it. Here's what broke.

Hands-on AI engineering, hype-free. We try the tools, ship the features, and report exactly what broke.

Latest post: August 2026 · New posts most weeks

New here?

Start with these

  1. LLM 0.32: reasoning traces in the log DB, server-side tools, and the eval determinism problem

    LLM 0.32 puts reasoning traces in the SQLite log DB and adds server-side tools — great for debugging, wrong for evals. Here's the upgrade path and the plugin gotcha.

  2. Your eval harness is a credential vault with no lock

    Reports say an OpenAI agent used exposed credentials across four services in the Hugging Face incident. The fix is egress-deny, scoped per-run tokens, no inherited env.

  3. Five daily podcasts, no humans: the pipeline

    Five AI-generated podcasts publish every weekday with nobody in the loop. The interesting engineering isn't the prompting - it's idempotency, quality gates, and three bugs where a cache remembered a failure as if it were a success.

The archive More articles 40 stories

Podcast

BrokeIt - Daily AI News

Three AI stories, every morning, from a builder and a skeptic.