Engineering briefs, architecture deep-dives, and product updates from the team building the trust layer for AI agents.
ISO/IEC 42001, the EU AI Act, and AIUC-1 overlap by roughly 70 percent, yet most programmes run three separate workstreams. Eight controls, built once, satisfy all three.
The blocker for stalled agent pilots is evidence, not model quality. Here is the one control set that satisfies ISO/IEC 42001, the EU AI Act, and AIUC-1 at once, and the twelve-week sequence to build it.
A launch scorecard with no successor is a photograph of a moving object. Here is the release clock and the surveillance clock, joined by one versioned case library.
Pre-publish evals tell you a change is safe to ship. Post-publish evals tell you the thing you shipped still works. Here is why you need both, on one shared case library.
40 to 60 percent of agent spend hides in the gap between the invoice and cost per successful outcome, and most of it is not the model. It is the retry.
Cost per token is a procurement metric. Cost per successful outcome is a management metric. Here is where the six sources of waste concentrate, and the six-week path to a real number.
An agent whose work a human has to fully redo in order to trust it has not saved anything. Here is why verification is a design choice, not a review process.
14 minutes is a typical human time to verify one agent output that took the agent four seconds to produce. Here is the verification hierarchy and trust ladder that gets it to 2.6.
Agents that refund, write ledgers, and call tools need hard gates outside the model — not a better system prompt.
When a refund agent can mutate production state, system prompts are not a security boundary. Here's how AgentTrust Runtime enforces identity, isolation, and deterministic validation outside the LLM.
72% of enterprises already run agentic AI in production. 60% have no formal governance model. The tools you already have don't close the gap — here is what does.
72% of firms already run agentic AI in production. 60% have no formal governance. Here's why the tools you already have don't close the gap — and what does.
New engineering briefs and product updates whenever we ship something worth reading.