Currently listed through:
Model details
Compound
Groq Compound is presented in Groq's documentation as an AI system rather than a single standalone model, composed of openly available models that the platform orchestrates intelligently and selectively in response to user queries. The system is designed to leverage built-in tools as part of its answer pipeline, with web search and code execution explicitly named as capabilities that Compound can invoke when a query requires grounded information or programmatic reasoning. This tool-routing design positions Compound as a higher-level abstraction that sits above individual open-weight base models, letting users tap into retrieval and computation alongside language generation through a single endpoint.
On GroqCloud, Compound is surfaced as a featured system with a reported token speed of roughly 450 tokens per second, which reflects Groq's emphasis on low-latency inference even for tool-augmented workloads. The combination of open-model composition, selective tool use, and fast response throughput makes Compound well suited to interactive assistants and developer-facing applications where quick, grounded answers matter more than maximal depth. Practical fits include research assistants that need live web lookups, coding helpers that benefit from in-line code execution, and prototyping environments where teams want a single interface to mix language understanding with retrieval and computation without managing multiple model calls themselves.
Quick Info
Powered by- Provider
- Groq
- Model key
- groq/compound
- Release date
- Sep 4, 2025
- Last updated
- Sep 4, 2025
- Input modalities
- Output modalities
- Capabilities
Limits
- Output tokens
- 8,192 tokens
- Context window
- 131,072 tokens
Latest news about Compound
No articles yet. Fetch the latest news to show it here.