Currently listed through these providers:
Model details
Whisper Large V3 Turbo
This model is designed for automatic speech recognition and speech translation. Its family was trained on five million hours of labeled data with a weakly supervised approach, helping it generalize across many datasets and domains without task-specific training. That broad training base makes the model relevant to varied speech-recognition workflows rather than a single narrow audio setting.
The Turbo variant is a fine-tuned version of a pruned Whisper Large-v3, reducing the decoder from 32 layers to 4. This architectural change substantially improves inference speed while introducing only minor quality degradation, creating a practical balance for transcription and translation workloads where responsiveness matters. Its publication through the Transformers ecosystem and an NVIDIA Riva catalog entry also indicate multiple established paths for integration and deployment.
Quick Info
Powered by- Provider
- Groq
- Model key
- whisper-large-v3-turbo
- Release date
- Oct 1, 2024
- Last updated
- Oct 1, 2024
- Input modalities
- Output modalities
- Capabilities
Limits
- Output tokens
- 0 tokens
- Context window
- 0 tokens