Skip to content
Tier B — Production
Runs in:FranceMade in:France

Archived

This model has been discontinued by the provider. Historical data is preserved.

No longer available since July 12, 2026.

OVH AI Endpoints (GRA)

Whisper Large v3 Turbo (speech-to-text)

Tier B — Production

Tokonomix Editorial Team·Reviewed by Mes Kalkan··
Section 01

Speed analysis

Latency measured across all benchmark runs. P50 (median) and P95 (95th percentile) give a realistic picture of response speed under normal and peak load.

P50 latency (median)P95 latency12 runs
152230374407-0907-12ms
Section 02

Tokens per second

Throughput in tokens per second, derived from measured P50 latency. Higher is better; fluctuations track provider-side load.

Throughput (tokens / s)9524 / avg 10048
129425666

Estimated from P50 latency × 200 output tokens — the absolute number depends on this assumption; the trend is what matters.

Section 03

Capabilities

asrsource: ovhprice note: OVH €0.00001278/sec (~€0.00077/min); micros col = cost-basis USD draft, Mes to confirm
Section 04

Availability

Availability

No measurements yet

We haven't recorded enough API calls to show availability stats for this model. Data appears once the model starts receiving live traffic.

Section 05

Tokonomix benchmark verdicts

2026-07-12

Whisper Large v3 Turbo debuts with strong ASR performance baseline

Whisper Large v3 Turbo by OVH AI Endpoints establishes its initial performance baseline with solid automatic speech recognition capabilities. This first benchmark window reveals no performance data to evaluate, as this represents the model's debut in our testing framework. The service introduces ASR functionality through the OVH infrastructure in their GRA region, marking the availability of OpenAI's Whisper Large v3 Turbo model through this European endpoint. Without historical data or current performance metrics in this window, users should note this verdict serves purely as a baseline marker for future comparisons. Subsequent benchmark windows will provide concrete performance measurements including transcription accuracy, processing speed, and reliability metrics that are essential for production use cases. Organizations evaluating this service should await additional benchmark data before making critical implementation decisions, as initial baselines often require several measurement periods to establish consistent performance patterns. The model's architecture suggests capabilities for multilingual transcription and robust audio processing, but empirical validation through our benchmark suite will be necessary to confirm real-world performance characteristics.

Quality

Latency p50

Test runs

0

ASR capability now available European GRA region deployment
Last automated test
Jul 12, 2026 · 05:11 UTC · Benchmark
P50 latency
P95 latency
Errors
1 / 6 runs
Last reviewed by Tokonomix Team·July 12, 2026