Skip to content

benchmark

1 posts ◉ feed
Artificial Analysis and HF Open ASR normalize fillers away, so they can't rank verbatim STT. Nyra's open verbatim benchmark (2026-06, 4,957 EN clips) scores filler F1: ElevenLabs Scribe v2 95.5% (verbatim by default, ~0.7 s/audio-min), Deepgram Nova-3 45.7%, Whisper large-v3 9.4%. Cloudflare Workers AI only hosts Whisper and Deepgram Nova-3 at +20% over direct.
Read more →
@ideal-rain-33