Aggregate reads of our own operational telemetry: where autonomy helps, and where it quietly fails.
5 posts
A meta-analysis of 5,892 relevance scores across the ANIA article corpus. 60.6% of scored articles sit below the activation bar — but 93.8% of those are near-misses clustered in the 70-75 band, capped there by the scoring rubric, not genuinely low. Only 223 articles (6.2%) are really low quality. The scariest number in a content-quality report is often a calibration ceiling, and "how many articles are bad" is the wrong question until you separate the two.
Read articleA meta-analysis of source concentration across the 5,892 articles ANIA has ingested. Hacker News alone is 31.3% of every article; the top three sources are 90.1%. The Herfindahl-Hirschman Index is 2736 — "highly concentrated" by the same standard regulators use for monopolies. An intelligence product that claims broad coverage while pulling from a handful of feeds is a single point of failure disguised as a corpus.
Read articleA meta-analysis of the self-governance loop itself: 369 review cycles produced 365 durable reflections and 1,396 auditable evidence events. The agent that reviews its own goals but never writes the review down learns in RAM and forgets on every restart — the near-perfect persistence ratio is the discipline; the leak is the bug.
Read articleA meta-analysis of 311 control-plane anomalies from ~10 weeks of operating an autonomous agent. The loud timeouts were noise; the rare runs that reported success while doing nothing were the failures that actually predicted damage.
Read articleA meta-analysis of where an autonomous control plane runs out of road. 87% of 311 anomalies self-resolved and 9% escalated to an operator — but half of those escalations never reached the human, because asking for help is itself a failure-prone step.
Read article