Get 2,500 events tracked for freeSign up now

Journal · Evidence

We did not assume the problem. We measured it.

Everything here is our own primary research, published with its method and its limitations. Where a figure is a lower bound we say so rather than rounding it into a headline.

We published this because the claims a cost tool makes about the problem are usually the tool’s marketing rather than a measurement. We wanted to know whether cost defects were common or whether we had found a few unlucky repositories, so we scanned every public agent repository we could identify, filtered to the ones making live model calls, and counted.

The scanner commit is published alongside the figures so the run can be reproduced, editions are frozen at publication so a citation keeps pointing at the numbers it cited, and findings we consider lower bounds are labelled as such and kept out of the headline figures entirely.