- What changed.
- The long-context corpus was cut using a characters-per-token estimate. Measured against the runtime’s own token accounting, English rungs land within ~6% of nominal (30.0k at the “32k” rung) but Russian at ~65% (21.1k). Per-item measured counts now ship beside the corpus, and every description states nominal versus measured.
- Why it matters.
- A depth ladder that overstates its own depth overstates the result.
- Effect on published numbers.
- None. The published depth claims are English-only. The corpus file itself is unchanged and still matches its frozen hash — the error was in describing it, not in the evidence.