Field notes on AI systems that survive reality

Seven field notes, dated and sourced, from GTC and Google Cloud Next to a defect log kept on our own tooling. Each one names the launch or the failure, then the part it doesn't solve for you.

Articles

September 2026 • 5 min read

A green checkmark is not evidence

Three tools I use every day reported success this year. All three were wrong the same way: a check that never exercised the thing it claims to cover.

Read article

April 2026 • 7 min read

Every model you depend on has a sunset date

Claude Sonnet 4 and Opus 4 retire June 15. GPT-5.5 just landed. Three model rotations in eighteen months. The teams with portable evals barely feel it.

Read article

April 2026 • 7 min read

Every cloud now sells an agent platform

Google Cloud Next, OpenAI Workspace Agents, Microsoft Agent Framework v1.0. Three platform launches in three weeks. The work they don't do is the work you still own.

Read article

April 2026 • 7 min read

When AI finds more bugs than your team can read

Why the March 2026 OpenSSF and Alpha-Omega funding matters for supply chain noise, triage, and how you run AppSec around AI-assisted scanning.

Read article

April 2026 • 7 min read

Three vendors launched agent stacks in March. You still own the hard parts.

GTC, Fusion, and SI headlines don't replace evals, messy integrations, or ownership. A field note on what changed and what didn't.

Read article

April 2026 • 6 min read

Identity before intelligence: what agent IAM forces you to decide

Three questions every platform team should answer before wiring another tool: what exists, what can connect, what's allowed.

Read article

March 2026 • 7 min read

Three reasons your AI pilot is stuck and what actually fixes them

Three structural reasons a pilot stalls: missing evals, underestimated integration complexity, and unclear ownership.

Read article