When GPT-6 Astra runs for 27 minutes and can't show you its work, that's a governance failure. This week's news makes the case for run bundles, model-agnostic tracing, and human-in-the-loop gates.
A local 27B model that burns 22,000 reasoning tokens to draw a circle is a governance story. Here's why reasoning-effort defaults, model-agnostic tracing, and run bundles matter for auditable AI.
This week's open-weight agentic coder and a sharp critique of unreviewable PRs both point to the same need: audit-ready run bundles, human-controlled gates, and model-agnostic tracing.
This week's AI news—deployment simulation, multi-agent safety funding, and an open eval workbench—maps directly to run bundles, human gates, model-agnostic tracing, and cost telemetry.