An open-source, efficient hybrid Mamba-Transformer MoE model that supports a context length of one million tokens and excels at agent reasoning, programming, planning, and tool invocation.
Pricing
- Input Tokens: $0.110 /M tokens
- Output Tokens: $0.550 /M tokens
- Cache Read: $0.028 /M tokens
Input Modalities
- Text
Output Modalities
- Text
Capabilities
- Thinking
- Tools
- Tool calling
- Structured outputs
- Long context
Try this model
Python
