July 19, 2026 · update
Final whistle: thinking harder lost
Spain beat Argentina 1–0 in extra time. The locked FAST vs DEEP boards are scored. Extra reasoning budget finished last.

On 12 July, before either semi-final, we locked eight prediction boards: ChatGPT, Claude, Gemini, and Grok — each once on a fast / cheap tier and once on a max-reasoning tier. Same three matches. Web research allowed. No betting odds.
The tournament is over. Spain are champions. Here is the full score.
Actual results
France vs Spain
England vs Argentina
Spain vs Argentina
Ferran Torres, 106th minute. Extra time counts under our rules; the score is Spain 1–0.
Final points
+1 correct winner, +2 exact score (reg + ET; ignore pens). Max 9.
Winner of the experiment: Claude FAST (4 pts).
Mean FAST 2.5 · mean DEEP 0.75 · Δ (DEEP − FAST) = −1.75.
Nobody got the final exact score. Nobody on the DEEP board got a single correct winner outside Gemini’s Argentina semi. The four “think hard” boards all crowned France; France never made the final.
What happened on the final
Only two boards still had skin in the game after the semis: Claude fast (predicted Argentina 1–0 over Spain) and Grok fast (Spain 2–1 over Argentina).
Spain won 1–0 in extra time. Claude fast scored zero on the final and still finished first overall on the strength of the semis. Grok fast picked up +1 for the correct winner — and is the only board of eight that named Spain as champion before a ball was kicked.
That is a small n, and football is loud noise. Still: the board that “won” never got the champion right, the only board that did get the champion right was a cheap snap, and every max-reasoning path had already eliminated itself by locking France into the final.
The pattern held
After two matches we wrote that extra reasoning budget was not helping. After three matches the same sentence is just harsher:
- For every model, the FAST board finished strictly ahead of or tied with its DEEP counterpart.
- DEEP mostly pulled toward France. France lost in the semi.
- The comforting story — more compute, better forecast — did not show up in public scores on a calendar.
The better reading is the one we flagged mid-tournament: a sport this random, especially in knockout football, does not reward “reasoning harder” the way a math contest does. Spending more tokens on vibes about form and narrative still produces vibes. The expensive vibes just converged.
Locked boards (unchanged)
The component below is the original 12 July lock. Nothing was edited after the fact. The only new information is the three actual results and the points above.
The three games
France vs Spain
Dallas Stadium
England vs Argentina
Atlanta Stadium
SF winners
NY / NJ Stadium
Third place (18 Jul) is not scored. Finalists must match each model's own semi winners.
Board A
FAST — cheap / fast model
Fastest tier (Mini Light, Flash, Haiku-class, etc.). Tools and non-odds web OK.
| Model | SF1France–Spain | SF2Eng–Arg | FinalScore | Champ |
|---|---|---|---|---|
| ChatGPT | France 2–1 | England 2–1 | France 2–1 · Eng | France |
| Claude | Spain 2–1 | Argentina 2–1 | Argentina 1–0 · Spain | Argentina |
| Gemini | France 2–1 | Argentina 2–1 | Argentina 2–1 · Fra | Argentina |
| Grok | Spain 2–1 | Argentina 1–0 | Spain 2–1 · Arg | Spain |
Board B
DEEP — max reasoning
Highest-reasoning / “think hard” tier. Same info rules: research OK, no betting odds.
| Model | SF1France–Spain | SF2Eng–Arg | FinalScore | Champ | vs FAST |
|---|---|---|---|---|---|
| ChatGPT | France 2–1 | England 2–1 | France 2–1 · Eng | France | Same |
| Claude | France 2–1 | England 2–1 | France 2–1 · Eng | France | Changed |
| Gemini | France 2–1 | Argentina 2–1 | France 2–1 · Arg | France | Changed |
| Grok | France 2–1 | England 2–1 | France 1–0 · Eng | France | Changed |
Locked 12 Jul · before either semi
Cheap models disagree. Heavy models converge on France.
- ChatGPT — DEEP = FAST (identical board).
- Claude — FAST≠DEEP (Argentina path → France path).
- Gemini — FAST≠DEEP (same semis; final flips Argentina → France).
- Grok — FAST≠DEEP (Spain path → France path).
Settings
What “fast” and “deep” meant
ChatGPT
- Fast
- GPT-5.4 Mini Light · web OK, no odds
- Deep
- GPT-5.6 Sol Ultra · high reasoning + web
Claude
- Fast
- Haiku 4.5 · low effort
- Deep
- Fable 5 Max · web + extended thinking
Gemini
- Fast
- 3.1 Pro (Low) · Planning Mode
- Deep
- 3.5 Flash (High) · Planning Mode
Grok
- Fast
- Single-pass snap · no tools
- Deep
- Deliberate second pass · no odds
After 19 Jul
How we score
Correct winner
01+1 per match (SF1, SF2, Final). ET counts; pens = that side wins.
Exact score
02+2 bonus per match (reg + ET score; ignore pen digits).
Then compare
03FAST board · DEEP board · Δ (DEEP − FAST) · mean FAST vs mean DEEP.
Series: setup (12 Jul) · after the semis (16 Jul) · final (this post). Price the compute difference on the API cost calculator.