Sulat.com
AI models
Groq logo

Model details

Whisper Large V3 Turbo

This model is designed for automatic speech recognition and speech translation. Its family was trained on five million hours of labeled data with a weakly supervised approach, helping it generalize across many datasets and domains without task-specific training. That broad training base makes the model relevant to varied speech-recognition workflows rather than a single narrow audio setting.

The Turbo variant is a fine-tuned version of a pruned Whisper Large-v3, reducing the decoder from 32 layers to 4. This architectural change substantially improves inference speed while introducing only minor quality degradation, creating a practical balance for transcription and translation workloads where responsiveness matters. Its publication through the Transformers ecosystem and an NVIDIA Riva catalog entry also indicate multiple established paths for integration and deployment.

Groqwhisper-large-v3-turbowhisper

Quick Info

Powered by
Provider
Groq
Model key
whisper-large-v3-turbo
Release date
Oct 1, 2024
Last updated
Oct 1, 2024
Input modalities
Output modalities
Capabilities

Limits

Output tokens
0 tokens
Context window
0 tokens

Latest news about Whisper Large V3 Turbo

Videos about Whisper Large V3 Turbo

More models around Whisper Large V3 Turbo