Command A is Cohere most performant model to date, excelling at tool use, agents, retrieval augmented generation (RAG), and multilingual use cases. Command A has a context length of 256K, only requires two GPUs to run, and has 150% higher throughput compared to Command R+ 08-2024.
Pricing
- Input Tokens: $2.500 /M tokens
- Output Tokens: $10.000 /M tokens
- Cache Read: $0.000 /M tokens
Input Modalities
- Text
Output Modalities
- Text
Try this model
Python
