The prompt (verbatim, sent to every lane)
How many times does the letter "e" (upper or lower case) appear in the sentence between the quotes? "Every expense here is measured, entered, and receipted." Reply with the number only.
Live runs are switched off right now — replays (recorded 2026-07-12) are the whole show today.
Live means live: 4 models answer this prompt for real, on my dime — reserving $0.0087 for this run against a shared $0.50 daily reservation ceiling. Replays stay free either way.
Going live sets one cookie (bl_sid, 30 days). It holds a random id, not your identity. The five-runs-a-day gate counts both that signed browser session and a one-way hash of the connection address; raw IP addresses never touch storage, and nothing else rides along.
How scoring works
Every lane gets the identical prompt at the same moment. A small script — not an AI judge — checks each answer against the rule below and stamps it CORRECT or WRONG; the line under each lane shows exactly what the check saw. No partial credit.
This challenge's rule
The output must contain exactly one integer, equal to 15 — an answer that shows its work (several numbers) fails on ambiguity, and the WRONG stamp says so rather than misquoting the model. Scoring strips leading/trailing whitespace and one wrapping pair of code fences or quotes before it judges — content over ceremony. Winner rule: correct → cheapest → fastest.
Prices as listed by OpenAI and Anthropic, July 2026. Costs come from provider usage fields × list prices — nothing is invented.
LIVE BUDGET ················ UTC day live runs ··········· ········open budget reserved ········ $0.00 daily cap ··················· $0.50 ──────────────────────────────────── replays ··········· free, unmetered