Skip to content

One prompt. Every model. One verdict.

One model might hallucinate, miss context, or just be wrong. A council of models catches what a single answer wouldn't.

Click for more details about pricing

Ship flawless AI images, every time.

5 AI models inspect every image for the flaws humans spot first: extra fingers, broken shadows, impossible physics.. etc.

91%
defects caught with council
~68%
with one model alone

BETA 2026-07 · LOKI-35 + Real Control Photos · Not a Product Guarantee.

AI-generated image with a realism defect
DEFECTAI-generated
Real control photo, no defects detected
CLEANreal photo
Council:Fable 5Opus 4.8Gemini 3 ProGPT 5.5 HighGemini 3.5 Flash

3 of 5 saw it. One model alone would have missed it — hence a council.

Judge verdicts

5,682 evaluations across 97 models — counts only, no customer prompts

⚖️Most endorsed: Claude Opus 4.6 (99% accurate)

Sample data

Fastest response times — Scientific Reasoning

  • 01Mistral Large 3

    780ms

  • 02Claude Sonnet 4.6

    920ms·

  • 03Llama 3.3 405B

    1.18s

  • 04Gemini 2.5 Pro

    1.42s

  • 05GPT-5o

    1.64s·

  • 06Claude Opus 4.7

    1.82s

Sample · methodology pending

how we test →
Pricing

No fee on single calls. You only pay the fee on consensus.

Ask one model and you pay just its tokens plus a small tier margin — no platform fee. The per-call fee applies only to multi-model consensus checks. 100 consensus checks free every month, no card needed; bundles from €10/month for 500 calls. Every token itemised, nothing hidden.

Free

€0/mo

100 calls/mo

token use: provider +5%

Starter

€10/mo

500 calls

token use: provider +4%

Most popular

Studio

€25/mo

2,000 calls

token use: provider +3%

Scale

€50/mo

5,000 calls

token use: provider +2%

Founders prices, locked through 2027 · PAYG also available · "token margin" = the small % we add on the model provider's own token price, lower on higher tiers

Single-model call
What you pay: tokens + margin
Details: No call-fee — only consensus checks carry the per-call fee. You pay the model provider's token price plus your tier margin (+2–5%). Example: a small model on ~4k tokens ≈ €0.001.
Consensus call
What you pay: call-fee + tokens + margin
Details: The call-fee varies per package (PAYG founders: 2c/proposer + 3c/judge, a 3+1 council = 9c; bundles: counts against your monthly quota; over quota: 1.5c/call). On top: the model provider's tokens + your tier margin.
Bring your own key (BYOK)
What you pay: call-fee only
Details: On consensus you pay only the per-package call-fee — your own key bills the provider directly, so no token cost and no margin from us. A single-model BYOK call costs nothing.

No per-seat fee. No single-call fee, ever. Every consensus receipt is itemised per model, per token, in and out.

Every cent, itemised

illustrative example
model                 in      out     cost
──────────────────────────────────────────────────
claude-haiku-4.5      812     540     €0.0041
gpt-4o                812     610     €0.0072
gemini-2.5-flash      812     498     €0.0029
judge (gpt-4o)        240     €0.0038
──────────────────────────────────────────────────
orchestration                         included
total                                 €0.0180

Accurate to the last token · your real receipt contains your exact counts

Estimate your cost

500
1005k

€10.00

Bundle price — overage at 1.5c/call above quota

€10.00

estimated / month

How we test

Real prompts, real latency, real scores. Three-tier framework so cost stays under control without compromising transparency.

Tier A

Full coverage

Speed + intelligence test daily across all four languages.

Tier B

Speed only

Latency and uptime sampled four times per day.

Tier C

Health ping

Up/down check every fifteen minutes.