Skip to content

All Posts

Browse all blog posts by year and month

202612

September4

  • One Prompt, Thirty-Two Calls, or Seven

    Published:
    • 14 min read

    Batching an LLM checklist is a blast-radius decision before it is a performance one: a seeded simulation of three call shapes, and why a call count is arithmetic while a speed-up is a measurement.

  • A Detector With Perfect Precision and Zero Recall

    Published:
    • 14 min read

    A document check cleared every genuine file and every forgery in a labelled set; a seeded synthetic run shows why specificity on genuine inputs is free and recall is the first number to ask for.

  • The Constant Nobody Could Re-Measure

    Published:
    • 15 min read

    A magic number survives because there is nothing to re-measure it with. The calibration data is often already in a log you throw away, and the same data sets the noise floor that makes the result readable.

  • The Forecast Metric That Hides the Failure

    Published:
    • 15 min read

    A volume weighted error of 7.6 percent and a net bias of 0.2 percent can describe a forecast that loses to a seasonal naive on 77 of 116 slow moving series. A seeded panel shows what to measure instead.

August1

June1

May1

April2

February2

January1

20253

November2

October1