Accuracy Ledger

Every takeoff we can check against a public agency's own published quantities, including the lines we got wrong.

We publish the misses because a record without them is an advertisement. Each source below is a public document — the same numbers are available to you, and the comparison can be re-run.

State DOT bridge replacement

Compared 2026-08-02 · 32 pay items scored of 55 in the tabulation

21 of 32 within ±15% — of those, 18 within ±1%

The full 119-sheet set was taken off blind — the sheets carrying the agency's own quantities were withheld from the agents — then compared line by line against the published bid tabulation. The item mapping is ours and is open to argument: a different mapping gives different numbers.

Matched

Missed — and why

What this does not prove

It proves nothing about pricing — this job was measured, not priced. It does not prove completeness: 16 of the 48 mapped items were quantities the drawings themselves printed and we copied, so they are excluded from the score rather than counted as wins.

Limits we created ourselves

Seven items where we produced no number came mostly from a decision of ours, not a failure of measurement: we restricted which sheets each site could read, to control cost and to keep the answer sheets out of view. The pier, cap and column concrete dimensions live on standard sheets outside that list, so all three sites correctly reported the source as insufficient rather than inventing a dimension.

State agency outdoor pavilion

Compared 2026-08-27 · our total against 7 real bids at a public opening

Against the median of 7 real bids: -45.4% — below every bid (bids $497,000 to $743,000)

Blind by construction of time: our workbook was produced 21 Jul 2026, the state opened bids 11 Aug 2026. Our total is compared against the range of the seven real bids, not against any single number, because the bidders' own disagreement is the honest yardstick.

Matched

Missed — and why

What this does not prove

It does not verify any line quantity (the tabulation is lump-sum). It does not tell whether the winner profits. The figure tested was our DIRECT cost only — no overhead, no prevailing-wage labour — which is exactly what the miss measures.

Limits we created ourselves

The total we tested carried no indirect cost and priced labour at a private-work base on a prevailing-wage state job — both our own modelling choices, and both named before any correction was applied.

State agency building repairs

Compared 2026-08-27 · our total against 8 real bids at a public opening

Against the median of 8 real bids: -42.3% — below every bid (bids $2,599,400 to $3,590,998)

Same blind-by-time construction as the pavilion entry: workbook 21 Jul, bids opened 11 Aug. Compared against the eight responsive bids; one non-responsive partial-scope submission is excluded for the reason the state itself printed.

Matched

Missed — and why

What this does not prove

No line quantities verified (lump-sum tab). Base-bid scope alignment with our alternates split is assumed, not proven. The figure tested was direct-only, as above.

Limits we created ourselves

Same two self-made limits as the pavilion entry: no indirect layer, private-work wage base on a prevailing-wage job.

State agency outdoor pavilion, re-measured

Compared 2026-08-27 · our total against 7 real bids at a public opening

Against the median of 7 real bids: -9.1% — inside the bid range (bids $497,000 to $743,000)

The two corrections have sources independent of the bids: the prevailing-wage factor comes from Missouri's own Annual Wage Order 33 for Phelps County, and the preliminaries allowance from a published size-tier band. Rebuilt through the real pipeline and computed by Excel itself.

Matched

Missed — and why

What this does not prove

This round is calibration, not a second blind test: the parameters have independent sources, but the decision to apply them came after seeing the first comparison. The proof is the next unseen bid tabulation, run with these parameters frozen. It also does not validate the factor per trade — an unweighted seven-trade mean was used.

Limits we created ourselves

The unweighted trade mean and the choice of correction moment are both ours; the entry says so rather than presenting the landing as accuracy.

State agency building repairs, re-measured

Compared 2026-08-27 · our total against 8 real bids at a public opening

Against the median of 8 real bids: -24.2% — below every bid (bids $2,599,400 to $3,590,998)

Identical corrections and pipeline as the pavilion re-measure (Cole County wage order, same allowance model).

Matched

Missed — and why

What this does not prove

Same calibration caveat as the pavilion re-measure. Additionally, the remaining gap is attributed to renovation-difficulty by reasoning, not yet by evidence.

Limits we created ourselves

We did not add a third factor to close the remaining gap, precisely because the only number available at that moment would have been the gap itself. A renovation-difficulty layer enters only with a source (P-405 Table 4-1) and gets tested on the next unseen renovation tabulation.

Checks that proved nothing

Listed because leaving them out would make the record look better than it is.