Skip to content

Data Extraction Race

Which model reads a document best?

Pick a document and a few models. Every model extracts the same fields; we score each answer against the ground truth — no LLM judge, just exact and fuzzy field matching.

Share this challenge

What this is

Public game

Anyone can play, free. Pick a real document, run the same extraction across the models you choose, and get an objective score — exact and fuzzy field matching against the ground truth, no LLM judge.

Admin arena

A separate staff tool for authoring and curating custom matches. You don’t need it to play here — it’s how new public games get built.

Pick a document

Pick models (2–6) · 3/6

We pre-select a few cheap models plus one frontier model so you can see the price/quality gap.

gpt-oss-20bOVH AI Endpoints (GRA) · $0.04/$0.15 per 1M99%Mistral-7B-Instruct-v0.3OVH AI Endpoints (GRA) · $0.10/$0.10 per 1M99%Claude Opus 4.8Anthropic · $5.00/$25.00 per 1M100%

Verifying you are human…

Estimated: up to €0.03 First 5 games/day are free
5 free games left todayHow scoring works