LowRouter
Artificial Analysis rates Ministral 3 8B at 5 on the Intelligence Index, marking it below average among comparable open-weights non-reasoning peers. Output throughput is reported at roughly 99.9 tokens per second, faster than the class median, while verbosity is flagged as somewhat high at around 20M tokens generated across the intelligence evaluation. Pricing sits at $0.15 per 1M input and $0.15 per 1M output tokens, with a 90% cache discount. The model is described as supporting text and image inputs with text outputs, a 256k-token context window, 8B total parameters, and an Apache 2.0 license, with weights available on Hugging Face. The page notes a possible separate reasoning variant while explicitly profiling this snapshot as the non-reasoning configuration, and compares the model within the small open-weights peer class.
