Vercel AI Gateway
Meta reports that Muse Spark achieves its reasoning capabilities using over an order of magnitude less compute than Llama 4 Maverick, its previous mid-size flagship.
Model details
This model is a 17B parameter Mixture-of-Experts architecture built to balance performance with operational efficiency. Designed as part of the Llama 4 series, it functions as an image-text-to-text system that excels at visual recognition, image reasoning, and detailed captioning. Its design intent focuses on providing a capable, streamlined tool for developers who need to integrate both text generation and multimodal processing into their applications, ensuring high-quality outputs for creative writing and conversational tasks.
The model utilizes a compressed tensor quantization approach, specifically employing FP8 precision to optimize its footprint while maintaining strong performance. Beyond its core conversational and visual capabilities, it is engineered to support advanced workflows such as synthetic data generation and model distillation, allowing users to leverage its outputs to refine other systems. This focus on versatility and efficiency makes it a practical choice for building AI assistants that require both deep reasoning and the ability to interpret complex visual inputs.
A provider subscription or plan supersedes token-based pricing for this model.
Vercel AI Gateway
Meta reports that Muse Spark achieves its reasoning capabilities using over an order of magnitude less compute than Llama 4 Maverick, its previous mid-size flagship.
This exact model name is also listed by 4 other providers.