Meta Models
Usage
244M
11.1K
3 of 3
- muse-spark-1.2
- muse-spark-1.1
- muse-glimmer-30b
- muse-spark-1.2
- muse-spark-1.1
- muse-glimmer-30b
Which models that traffic went to
- Muse Spark 1.280.4%197M
- Muse Spark 1.116.4%40.1M
- Muse Glimmer 30B3.2%7.8M
- Muse Spark 1.283.6%9.3K
- Muse Spark 1.115.6%1.7K
- Muse Glimmer 30B0.7%81
All 3 Meta Models
Open in model list| Modalities | ||||||
|---|---|---|---|---|---|---|
| muse-spark-1.1 | Takes text, vision, audio, video, returns text. | 1.05M | $1.38$4.67/M | — | 122 tok/s | 7.75 s |
| muse-spark-1.2 | Takes text, vision, audio, video, returns text. | 1.05M | $1.38$4.67/M | — | 65 tok/s | 19.81 s |
| muse-glimmer-30b | Takes text, vision, returns text. | 131K | $0.35$1.50/M | $0.04/M | 119 tok/s | 3.41 s |
Meta on AIHubMix
Which Meta model should I start with?
muse-glimmer-30b at $0.35/M input — the cheapest entry here that declares tool calling, and it carries a 131K context. Move up to muse-spark-1.1 when answer quality matters more than cost.
Why are there several entries for the same model?
Because each row is a route you can call, not a model release. Some IDs name an upstream (azure-, alicloud-, cc-), and some differ only in capitalisation, kept so older integrations keep working.
The catalog does not carry a field saying which of those a given row is, so this page does not sort them into buckets it would have to invent. Every row shows that route’s own price, context and speed — compare those directly, and open a model to see the upstreams that serve it.
How is cached input billed?
The Cache read column is the rate for input tokens served from the prompt cache — for example muse-glimmer-30b bills cache hits at 11.43% of the input rate. Cache write is the surcharge for putting a prompt into the cache in the first place, and only a few upstreams bill it separately. A dash in either column means the catalog carries no cache rate for that model, so plan on paying the full input rate.
Do I need a separate Meta account?
No. One AIHubMix key covers every model on this page, and switching between them is a change to the model string — billing, rate limits, and logs stay in one place.
Start calling Meta in one line
One key, one endpoint, 859 models across 37 model authors.

