Cohere
pip install cohere[oci] — one constructor swap runs Command A, Embed v4, and Rerank 3.5 inside your OCI tenancy. Full OCI auth, compartment governance, zero data egress.
Model details
Command A is a large-scale open-weights model designed for enterprise deployments that demand strong agentic capabilities, retrieval-augmented generation, and multilingual fluency without requiring massive infrastructure. At 111 billion parameters, it is engineered to deliver maximum performance with minimum hardware overhead, reportedly deployable on just two GPUs, making it practical for organizations that need private, on-premises AI. The accompanying technical report describes its development and benchmarks, and the model is released under a CC-BY-NC license, giving enterprises access to a research-tier version of Cohere's flagship architecture.
Beyond its standalone release, Command A has been integrated across major cloud platforms, appearing natively on Cohere's proprietary dashboard, Amazon SageMaker and Bedrock, Microsoft Azure, and Oracle Cloud Infrastructure's Generative AI service, giving teams flexible deployment paths. Cohere has continued expanding the broader family with variants such as Command A+, Command A Translate, Command A Reasoning, and Command A Vision, indicating an ongoing evolution toward unified reasoning, multimodal understanding, and tool orchestration. For practitioners, this positions Command A as a practical choice for building secure enterprise agents, multilingual applications, and automated workflows backed by an actively developed ecosystem.
Transparent token rates
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Cohere
pip install cohere[oci] — one constructor swap runs Command A, Embed v4, and Rerank 3.5 inside your OCI tenancy. Full OCI auth, compartment governance, zero data egress.
Cohere
Release of Command A, a performant model suited for tool use, RAG, agents, and multilingual uses, with 111 billion parameters and a 256k context length.
Cohere
Independent benchmark aggregator Artificial Analysis lists Command A among three Cohere models currently tracked, reporting an Intelligence Index of 8, output speed of roughly 57 tokens per second, time-to-first-token latency around 1.71 seconds, a 288,000-token context window, and a blended price of about $3.25 per 1M For developers choosing between Cohere endpoints, the dashboard frames Command A as the balanced general-purpose option: cheaper and faster than Command A+ but with a larger context window, and substantially more capable than North Mini Code (Intelligence Index 20, 24 t/s, 85.65s latency) on speed and latency. The inde
This exact model name is also listed by 4 other providers.