AI21 Models
All 2 AI21 Models
Open in model list| aihubmix-Jamba-1-5-Large | $2.20$8.80/M | — |
| o1-global | $15.00$60.00/M | $7.50/M |
AI21 on AIHubMix
Which AI21 model should I start with?
aihubmix-Jamba-1-5-Large at $2.20/M input — the cheapest entry here that declares a token price. Move up to o1-global when answer quality matters more than cost.
Why are there several entries for the same model?
Because each row is a route you can call, not a model release. Some IDs name an upstream (azure-, alicloud-, cc-), and some differ only in capitalisation, kept so older integrations keep working.
The catalog does not carry a field saying which of those a given row is, so this page does not sort them into buckets it would have to invent. Every row shows that route’s own price, context and speed — compare those directly, and open a model to see the upstreams that serve it.
How is cached input billed?
The Cache read column is the rate for input tokens served from the prompt cache — for example o1-global bills cache hits at 50% of the input rate. Cache write is the surcharge for putting a prompt into the cache in the first place, and only a few upstreams bill it separately. A dash in either column means the catalog carries no cache rate for that model, so plan on paying the full input rate.
Do I need a separate AI21 account?
No. One AIHubMix key covers every model on this page, and switching between them is a change to the model string — billing, rate limits, and logs stay in one place.
Start calling AI21 in one line
One key, one endpoint, 844 models across 35 providers.
