September 2026 • 5 min read
A green checkmark is not evidence
Three tools I use every day reported success this year. All three were wrong the same way: a check that never exercised the thing it claims to cover.
Read articleSeven field notes, dated and sourced, from GTC and Google Cloud Next to a defect log kept on our own tooling. Each one names the launch or the failure, then the part it doesn't solve for you.
Archive
September 2026 • 5 min read
Three tools I use every day reported success this year. All three were wrong the same way: a check that never exercised the thing it claims to cover.
Read articleApril 2026 • 7 min read
Claude Sonnet 4 and Opus 4 retire June 15. GPT-5.5 just landed. Three model rotations in eighteen months. The teams with portable evals barely feel it.
Read articleApril 2026 • 7 min read
Google Cloud Next, OpenAI Workspace Agents, Microsoft Agent Framework v1.0. Three platform launches in three weeks. The work they don't do is the work you still own.
Read articleApril 2026 • 7 min read
Why the March 2026 OpenSSF and Alpha-Omega funding matters for supply chain noise, triage, and how you run AppSec around AI-assisted scanning.
Read articleApril 2026 • 7 min read
GTC, Fusion, and SI headlines don't replace evals, messy integrations, or ownership. A field note on what changed and what didn't.
Read articleApril 2026 • 6 min read
Three questions every platform team should answer before wiring another tool: what exists, what can connect, what's allowed.
Read articleMarch 2026 • 7 min read
Three structural reasons a pilot stalls: missing evals, underestimated integration complexity, and unclear ownership.
Read article