Before your first top-up
No credit card required, and trial calls never expire.
Call 48 genuinely free AI models — GPT, Gemini, GLM, Kimi, MiniMax and more — through one OpenAI-compatible API. No credit card required: every account starts with 10 trial calls, and a one-time $1 top-up unlocks daily quotas permanently.
No credit card required, and trial calls never expire.
Top up once — any amount from $1 — and every free model switches to daily quotas with no expiry, reset every day.
Tokens
15.6B
Requests
155K
Models in use
47 of 48
Tokens per day
Requests per day
Share of 15.6B tokens. 4 models with traffic report no token counts and are only in the request view.
Share of 155K requests.
Rate limit shows the daily and per-minute request caps for topped-up accounts. Some models carry a higher request weight — one call counts as several requests — which is why their caps are lower. All free models share one pool of 1M tokens per day. Before your first top-up, the 10 trial calls apply instead. Context "—" means the value is not published in the catalog. Modality icons show what a model takes as input; output formats are listed on each model page.
Email or OAuth, no credit card required. 10 trial calls are available immediately.
One key works for every model on the platform, free and paid alike. Open the API keys page →
Point any OpenAI-compatible client — Cursor, Cline, Cherry Studio, LiteLLM — at the gateway and pick a -free model ID.
POST /v1/chat/completions
A -free ID like coding-glm-5.2-free — the most-called free model this month — is the same upstream model as coding-glm-5.2, served as a subsidized free variant. Swap the suffix off when you need paid throughput; nothing else in your code changes.
Yes. All 48 models bill $0 for input and output tokens. AIHubMix subsidizes the inference cost; usage is bounded by quotas instead of a bill.
10 trial calls, shared across all free models, with no expiry. When they run out the API keeps answering with a note that the trial is used up — your integration never sees a hard error.
Any top-up of $1 or more permanently switches your account to daily quotas: 10 requests per minute, 100 requests per day and 1M tokens per day, shared across the free catalog and reset every day. Models with a higher request weight count as several requests per call — the exact caps are in the table. Past a cap the API returns 429 until the next minute or day starts.
Each call to a higher-weight model counts as several requests against the daily 100. The Rate limit column shows the resulting effective caps per model.
They are the same upstream model. The -free ID runs as the subsidized free variant with the limits above; the paid ID bills per token with no daily cap. Switching between them is a one-line model-name change.
Yes. Free models sit behind the standard OpenAI-compatible endpoint, so the official SDKs, LangChain, LlamaIndex and any client with a custom base URL work unchanged.
Moderate production traffic is fine. The pattern that works: paid models on the critical path, free models for auxiliary work like batch summarization, drafts, and log analysis.
The full catalog — 849 models across chat, image, video and audio, with per-token pricing.
Browse all models →Fill in auto as the model ID and the gateway routes each request to the optimal model by cost and quality.
Ready-to-use apps in the AIHubMix app store — and model discounts beyond the free catalog.
Browse the app store →
AIHubMix© 2023 - 2026 AIHubMix, LLC