Cohere's first reasoning model designed for enterprise customer service and automation. 111B parameters with tool-use capabilities, supports 256K context and 23 languages including English, French, Spanish, Japanese, Arabic, and Hindi. Optimized for document processing, scheduling, data analysis, and more.
Added Aug 22, 2025
Context Window
256.0K
Max Output
8.2K
Input Price (Auto)
$2.50/1M
Output Price (Auto)
$10.00/1M
Cache Read (Auto)
$1.25/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Cohere Command A (08/2025) with similar models from the same provider or model family.
Cohere: Command R+
cohere/command-r-plus-08-2024104B parameter model that performs conversational language tasks at a higher quality, more reliably, and with a longer context than previous models. It can be used for complex workflows like code generation, retrieval augmented generation (RAG), tool use, and agents
Qwen3.5 27B Claude 4.6 Opus Reasoning Distilled Derestricted
Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled-DerestrictedQwen3.5 27B Claude 4.6 Opus Reasoning Distilled Derestricted is a community creative finetune for reasoning, multimodal chat, expressive writing, and roleplay.
Qwen3.5 27B Claude 4.6 Opus Reasoning Distilled Derestricted Lite
Qwen3.5-27B-Claude-4.6-Opus-Reasoning-Distilled-Derestricted-LiteQwen3.5 27B Claude 4.6 Opus Reasoning Distilled Derestricted Lite is a lighter-tuned community finetune for responsive multimodal chat, expressive writing, and roleplay.
Gemma 4 31B Claude 4.6 Opus Reasoning Distilled
Gemma-4-31B-Claude-4.6-Opus-Reasoning-DistilledArliAI-hosted Gemma 4 31B reasoning-distilled finetune for structured scene planning, dialogue, and multimodal chat.
Nvidia Nemotron 3 Nano Omni
nvidia/nemotron-3-nano-omni-30b-a3b-reasoningNvidia's Nemotron 3 Nano Omni 30B-A3B reasoning model. It accepts multimodal context on supported providers and returns text responses for perception and agentic workflows.
Gemini 2.5 Flash Lite Preview (09/2025)
gemini-2.5-flash-lite-preview-09-2025Deprecated compatibility alias. Requests route to the stable Gemini 2.5 Flash Lite model.