Currently listed through these providers:
Model details
Voxtral Mini 3B 2507 (Amazon Bedrock, US)
Voxtral Mini 3B 2507 is an open-weights audio-language model from Mistral AI that extends the Ministral 3B text backbone with native speech understanding. Positioned as a compact 3-billion-parameter release in the Voxtral family, it targets use cases such as speech transcription, translation, and direct audio question answering, combining ASR-style capabilities with on-model reasoning rather than relying on separate transcription and language pipelines. The model is delivered through a public Hugging Face repository under the Mistral AI organization, signaling an open distribution model alongside its text and audio inputs.
In practical terms, the model is built for long-form and multilingual spoken interactions. It supports a 32k token context window that can absorb up to roughly thirty minutes of audio for transcription or forty minutes for understanding-style tasks, and it ships with a dedicated transcription mode that automatically detects the source language before producing text. A built-in Q&A and summarization path lets users ask questions directly over audio and generate structured summaries, while automatic language detection covers widely used languages including English, Spanish, French, Portuguese, Hindi, German, Dutch, and Italian. Function-calling from voice inputs further broadens its fit for assistant and workflow applications that need to act on spoken intents while preserving the text capabilities inherited from its Ministral lineage.
Quick Info
Powered by- Provider
- Eden AI
- Model key
- amazon/mistral.voxtral-mini-3b-2507@us
- Release date
- Jul 15, 2025
- Last updated
- Jul 15, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.04
- Output token cost
- $0.04
Limits
- Output tokens
- 32,768 tokens
- Context window
- 128,000 tokens
Latest news about Voxtral Mini 3B 2507 (Amazon Bedrock, US)
No articles yet. Fetch the latest news to show it here.