Currently listed through these providers:
Model details
Command R7B
Command R7B is positioned by Cohere as the smallest and fastest member of its R family of enterprise-focused large language models, built around a compact architecture aimed at high-throughput, latency-sensitive workloads such as chatbots and code assistants. Its small footprint is intended to unlock dramatically cheaper deployment, including on consumer GPUs and CPUs, making on-device inference a practical option for production teams that need responsive generation without large infrastructure overhead.
The model is documented with a 128,000-token context window and a 4,000-token maximum output, paired with capabilities that include Tool Use, Structured Outputs, Citations, Multilingual support, Safety Modes, Reasoning, and Image Inputs. Cohere highlights retrieval-augmented generation as a flagship use case, where grounding model outputs in external data sources improves factual reliability, and the combination of tool use and structured outputs makes it well suited to agentic pipelines and enterprise assistants that must act on retrieved context.
Quick Info
Powered by- Provider
- Eden AI
- Model key
- cohere/command-r7b-12-2024
- Release date
- Dec 2, 2024
- Last updated
- Dec 2, 2024
- Knowledge cutoff
- 2024-06-01
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.0375
- Output token cost
- $0.15
Limits
- Output tokens
- 4,000 tokens
- Context window
- 132,000 tokens
Latest news about Command R7B
Videos about Command R7B
More models around Command R7B
This exact model name is also listed by 3 other providers.