Built for voice agents that can't wait 600ms

OpenAI: Whisper Large v3 Turbo

openai/whisper-large-v3-turbo
Speech to textactive

OpenAI speed-optimized Whisper Large v3 derivative. The Hugging Face model card describes it as a pruned and finetuned large-v3 model with decoder layers reduced from 32 to 4 for faster inference with minor quality degradation.

ProviderSlugPriceLatencyUptime