Currently listed through these providers:
Model details
GPT-4o mini Transcribe
GPT-4o mini Transcribe is a speech-to-text model in OpenAI's audio lineup, introduced alongside its larger sibling as a successor to the earlier Whisper V3 system. It is positioned as a lightweight option within that family, accepting spoken audio and returning written text, and it belongs to OpenAI's proprietary, API-distributed model portfolio rather than the open-weights ecosystem. Its design intent centers on converting recordings, voice notes, and other audio inputs into accurate text transcripts across many languages, leveraging the same broad audio training approach that powers the broader gpt-4o generation of audio tools.
In practical terms, the model suits developers who need reliable transcription embedded directly into applications, such as captioning services, meeting transcription, voice-driven analytics, and localization pipelines that benefit from stronger multilingual handling than legacy automatic speech recognition. Because it is delivered through OpenAI's developer platform and listed within the Voice and Audio section of the official model catalog, teams can integrate it alongside other OpenAI services for hybrid audio-plus-language workflows. The compact footprint of the mini variant makes it attractive for high-volume or cost-sensitive transcription workloads where the larger sibling's quality margin is not essential.
Quick Info
Powered by- Provider
- Vercel AI Gateway
- Model key
- openai/gpt-4o-mini-transcribe
- Release date
- Mar 13, 2024
- Last updated
- Mar 13, 2024
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $1.25
- Output token cost
- $5.00
Limits
- Output tokens
- 0 tokens
- Context window
- 0 tokens