Models

Filter
857 models
Models API
Output Modalities
All
Text
Image
Speech
Video
Transcription
Embeddings
Rerank
OCR
Categories
All
Featured
Coding
Free
Discount
Model Authors
AllOpenAIAnthropicGoogleGrokQwenDeepSeekZ.AIByteDanceLlamaAI21MicrosoftCohereMistralYiMoonshot AIStepFunNvidiaMinimaxPerplexityBaichuanIdeogramJina AIStable diffusionHunyuanBaiduFluxMeituanInclusionAIBAAIXiaomiKLingMetaPoolsideAgnesLiquidDots Studio
Refine
Newest
Input Modalities
Reset
autoauto:balancedauto:quality_firstauto:latency_critical

AIHubMix's in-house automatic model router. Not every request needs the strongest, most expensive model — drawing on internal evals and public leaderboards, we bring together a range of today's mainstream models of varying capability and route each request automatically through a small model we trained. Set model to auto and requests are dispatched to the right model based on their content, lowering cost. Currently in beta; we keep updating and improving it.

Charged at the actual routed model's price — routing itself is free.Learn about the LLM Router
icon
New
up to 10% off·00:00–23:59 UTCCopy ID
Input:$ 1.1268$ 1.01412 /M
Output:$ 3.9438$ 3.54942 /M
Context:1M
TTFT:4.975 s
Throughput:29 tok/s

GLM-5.3 is Z.AI’s coding and agentic reasoning model, built for complex software engineering, long-running agent tasks, vulnerability analysis, and other demanding workloads. Building on GLM-5.2, it incorporates further post-training improvements to deliver stronger coding performance, better task execution, and greater token efficiency. We currently offer the production-ready GLM-5.3 API with unlimited concurrency, making it well suited for high-throughput workloads, coding agents, and large-scale automation. For a limited time, GLM-5.3 is available at 10% off.

Input:$ 0.06 /M
Output:$ 0.21999996 /M
Context:-
TTFT:-
Throughput:-

GLM-5.3 is Z.ai’s reasoning model for coding and agentic workflows, designed for complex software engineering, long-running agents, and vulnerability analysis. It uses the same base model as GLM-5.2, with scaled post-training improving coding, task execution, and token efficiency. This model is a limited-time preview version of GLM-5.3, intended for testing and evaluation only. Service stability is not guaranteed, and we do not recommend using it in production environments. We’re waiting for the official commercial API release and will integrate it as soon as official support becomes available.

  • Input: $ 0.75 /M
  • Output: $ 3.75 /M
  • Web Search: $0.014/request
  • Cache Storage: $1/h/M tokens
  • Input Audio: $1/M tokens
  • Input Video: $1/M tokens

Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web development, and knowledge work. It supports a 1M-token context window and adjustable thinking levels. Compared with Gemini 3.6 Flash, it improves coding, tool use, multi-step planning, and instruction following.

Input:$ 0 /M
Output:$ 0 /M
Context:128K
TTFT:-
Throughput:-

GLM 5.2 is a large-scale reasoning model developed by Z Ai that supports text input and output. Featuring a 128,000-token context window, this model is designed to handle complex, data-heavy tasks. It is exceptionally well-suited for long-horizon agent workflows and project-level software engineering.

Input:$ 0 /M
Output:$ 0 /M
Context:512K
TTFT:-
Throughput:-

Dots3-Note Preview is an open-weight mixture-of-experts model developed by Dots Studio, featuring 16B active parameters out of 280B total. As the lightest model in the Dots 3 family, it is designed for efficient performance while supporting an expansive context length of 512,000 tokens. This preview version provides an accessible way to experience the capabilities of the Dots 3 architecture.

  • Input: $ 0 /M
  • Output: $ 0 /M
  • Web Search: $0.014/request
  • Cache Storage: $1/h/M tokens
  • Input Audio: $2/M tokens
  • Input Video: $1/M tokens

Gemini 3.7 Flash free version: fFree model resources are limited and provided only for trial use; stability cannot be guaranteed, and you may encounter 429 errors during use. If you need to use it in a production environment and require unlimited concurrency with absolute stability, please choose the official version: gemini-3.7-flash

icon
up to 30% off·00:00–13:59 UTCCopy ID
Input:$ 1.1268$ 0.78876 /M
Output:$ 3.9438$ 2.76066 /M
Context:1M
TTFT:0.92 s
Throughput:40 tok/s

GLM-5.2 is Z.ai’s flagship model for the era of long-horizon tasks. With a truly usable 1M-token context window, it can handle project-level engineering context, execute long-running tasks more reliably, follow engineering standards more consistently, and complete the full development workflow from requirements to multi-platform deployment in a single task.

Input:$ 0.142 /M
Output:$ 0.284 /M
Context:1M
TTFT:1.011 s
Throughput:98 tok/s

DeepSeek-V4-Flash-0731(deepseek-v4-flash-0731) is an open-source MoE large language model developed by the Chinese AI company DeepSeek, with support for a million-token context window. It is designed for coding, complex reasoning, tool use, agentic workflows, and long-document processing. Its advantages include strong performance with fewer active parameters and improved efficiency through DSpark speculative decoding. Compared with DeepSeek V4-Flash Preview, it offers significantly stronger coding and agent capabilities, while outperforming DeepSeek V4-Pro Preview on several benchmarks with fewer active parameters.

Input:$ 0.6918 /M
Output:$ 2.0754 /M
Context:1M
TTFT:1.682 s
Throughput:60 tok/s

DeepSeek V4 Pro 0813 is DeepSeek’s high-performance general-purpose reasoning and agent model, designed for complex reasoning, coding, long-document analysis, and agentic workflows. It supports thinking and non-thinking modes, a 1M-token context window, up to 384K output, tool calling, and the Responses API. Compared with V4 Flash 0731, Pro prioritizes capability on complex tasks, while Flash focuses on speed, cost efficiency, and high concurrency.

icon
Copy ID
  • Input: $ 2 /M
  • Output: $ 6 /M

Grok 4.6 is xAI’s (SpaceXAI) flagship multimodal reasoning model for coding, long-running agents, knowledge work, and interactive application development. It supports image understanding, a 500K context window, tool calling, and structured outputs. Compared with Grok 4.5, it offers stronger multi-step execution, self-verification, coding, and visual project generation.

Input:$ 2 /M
Output:$ 8 /M
Context:256K
TTFT:6.669 s
Throughput:42 tok/s

MAI-Thinking-1 is Microsoft’s first inference model in the MAI series, built for enterprise-scale workloads. With excellent reasoning, mathematical, and general intelligence capabilities, combined with superior cost-effectiveness, it makes high-throughput, 24/7 AI workloads economically viable.

icon
up to 50% off·00:00–23:59 UTCCopy ID
  • Input: $ 5$ 2.5 /M
  • Output: $ 30$ 15 /M
  • Web Search: $0.01/request

GPT-5.6 Sol (limited-time 50% off) is OpenAI’s frontier reasoning model for complex coding, professional knowledge work, deep research, and long-running agents. It supports a roughly 1.05M-token context window, image understanding, and extensive tool use. Compared with Terra and Luna, Sol prioritizes capability and reliability on demanding tasks.

  • Input: $ 2 /M
  • Output: $ 0 /M

Seedance 2.5 is ByteDance Seed’s next-generation unified multimodal audio-video generation model, designed for filmmaking, advertising, education, simulation, and long-form content creation. It generates up to 30 seconds of synchronized audio and video in one pass and supports multi-round extensions. Compared with Seedance 2.0, it offers stronger storytelling, smoother transitions, richer reference support, improved realism, and more precise editing.

  • Input: $ 0.2 /M
  • Output: $ 1.2 /M
  • Web Search: $0.01/request

GPT-5.6 Luna is designed for cost-sensitive, high-volume workloads. It roughly corresponds to the nano model tier used in earlier GPT-5 families.

  • Input: $ 5 /M
  • Output: $ 30 /M
  • Web Search: $0.01/request

GPT‑5.6 Sol sets a new standard for both intelligence and efficiency, achieving state-of-the-art results across coding, knowledge work, cybersecurity, and science while outperforming previous and competing frontier models with fewer tokens and at lower estimated cost.

  • Input: $ 2 /M
  • Output: $ 12 /M
  • Web Search: $0.01/request

GPT-5.6 Terra is designed for workloads that balance intelligence and cost. It roughly corresponds to the mini model tier used in earlier GPT-5 families.

Input:$ 0.03 /M
Output:$ 0.15 /M
Context:512K
TTFT:3.931 s
Throughput:43 tok/s

Agnes 2.5 Flash is Agnes AI’s fast and efficient language model, an upgraded, fully available model based on Agnes 2.0 Flash. It continues to use an OpenAI-compatible Chat Completions interface and has been optimized for coding tasks, agent workflows, tool calls, multi-turn dialogue, reasoning, and image understanding.

Input:$ 0.45 /M
Output:$ 0.9 /M
Context:1M
TTFT:18.56 s
Throughput:42 tok/s

Agnes 2.5 Pro is Agnes AI’s paid inference model and the commercially stable version of the Agnes 2.5 Pro Alpha ranking model, suitable for advanced coding, scientific reasoning, long-context analysis, agent workflows, and multimodal understanding. The model is accessed via an OpenAI-compatible Chat Completions API.

Input:$ 0.45 /M
Output:$ 0.9 /M
Context:1M
TTFT:4.931 s
Throughput:11 tok/s

Agnes 2.5 Pro Alpha is Agnes AI’s paid inference model, suitable for advanced coding, scientific reasoning, long-context analysis, agent workflows, and multimodal understanding. The model is accessed via an OpenAI-compatible Chat Completions API.

Scroll to load more

Popular models

Qwen3.8 Max Preview · Kimi K3 · Qwen3.8 Max · Qwen3.7 Flash · GLM 5.2 · Grok 4.5 · Claude Opus 5 · Claude Sonnet 5 · GPT 5.6 Luna · Gemini 3.6 Flash · DeepSeek V4 Flash · GPT 5.5 · Gemini 3.1 Pro Preview

Browse by provider

OpenAI (133) · Anthropic (28) · Google (83) · Grok (26) · Qwen (141) · DeepSeek (36) · Z.AI (66) · ByteDance (46) · Llama (50) · AI21 (2) · Microsoft (14) · Cohere (19) · Mistral (10) · Yi (6) · Moonshot AI (34) · StepFun (4) · Nvidia (17) · Minimax (31) · Perplexity (4) · Baichuan (5) · Ideogram (9) · Jina AI (13) · Stable diffusion (1) · Hunyuan (5) · Baidu (23) · Flux (5) · Meituan (2) · Xiaomi (14) · InclusionAI (7) · BAAI (5) · Agnes (4) · Meta (3) · KLing (2) · Poolside (2) · Dots Studio (1) · Liquid (1)

51 free models — no credit card required →

All models (852)

auto · glm-5.3 · coding-glm-5.3 · gemini-3.7-flash · glm-5.2-free · dots-3-note-preview-free · gemini-3.7-flash-free · glm-5.2 · deepseek-v4-flash-0731 · deepseek-v4-pro-0813 · grok-4.6 · mai-thinking-1 · gpt-5.6-sol-disc · doubao-seedance-2-5-260628 · gpt-5.6-luna · gpt-5.6-sol · gpt-5.6-terra · agnes-2.5-flash · agnes-2.5-pro · agnes-2.5-pro-alpha · agnes-image-2.1-flash · grok-4.5 · lfm-2.5-2.6b-free · minimax-h3 · muse-glimmer-30b · nemotron-lightning-3.5-30b-a3b · qwen3.8-2.4t-a95b · claude-opus-5 · gemini-3.6-flash · ling-3.0-tiny-free · nemotron-3.5-lightning-free · qwen-image-3.0 · qwen-image-3.0-pro · qwen3.8-max · claude-sonnet-5 · kimi-k3 · ling-3.0-flash-free · muse-spark-1.2 · qwen3.8-max-preview · gemini-3.1-flash-lite-image · gemini-3.5-flash-lite · gemini-3.5-flash-lite-free · gemini-3.6-flash-free · glm-5.2-fast-preview · muse-spark-1.1 · qwen-audio-3.0-tts-flash · qwen-audio-3.0-tts-plus · claude-fable-5 · jina-reranker-v3.5 · claude-opus-4-8 · hy3 · doubao-seed-2-1-pro · doubao-seed-2-1-turbo · mai-image-2.5-pro · gemini-3.5-flash · grok-build-0.1 · mai-image-2.5 · mai-image-2.5-flash · coding-kimi-k3 · happyhorse-1.1-i2v · happyhorse-1.1-r2v · happyhorse-1.1-t2v · coding-glm-5.2-free · coding-kimi-k3-free · gemini-3.1-flash-image · gpt-oss-20b-free · kimi-k2.7-code · kimi-k2.7-code-highspeed · gemini-3-pro-image · gpt-4o-transcribe-diarize · gpt-audio-1.5 · hy-3d-3.1 · kling-v3-omni · kling-video-o1 · longcat-2.0 · nemotron-nano-9b-v2-free · hy3-preview · minimax-m3 · nemotron-nano-12b-v2-vl-free · qwen3.7-flash · qwen3.7-plus · step-3.7-flash · claude-opus-4-8-think · nemotron-3-super-120b-a12b-free · nemotron-3-nano-omni-30b-a3b-reasoning-free · nemotron-3-ultra-550b-a55b-free · qwen3.7-max · gpt-image-2 · nemotron-3.5-content-safety-free · coding-glm-5.2 · ernie-5.1 · gemini-3.1-flash-lite · gemini-3.1-flash-lite-nothink · grok-4.3 · happyhorse-1.0-i2v · happyhorse-1.0-r2v · happyhorse-1.0-t2v · happyhorse-1.0-video-edit · north-mini-code-free · gpt-5.5 · gpt-5.5-pro · laguna-xs-2.1-free · deepseek-v4-flash · deepseek-v4-pro · gemma-4-31b-it-free · command-a-plus-05-2026 · doubao-seedream-5.0-pro · ernie-5.0 · kimi-k2.6 · laguna-s-2.1-free · qwen3.6-max-preview · xiaomi-mimo-v2.5 · xiaomi-mimo-v2.5-pro · claude-opus-4-7 · claude-opus-4-7-think · gpt-chat-latest · nemotron-3-nano-30b-a3b-free · qwen3.6-27b · qwen3.6-35b-a3b · qwen3.6-flash · cohere-rerank-v4.0-fast · cohere-rerank-v4.0-pro · gemma-4-26b-a4b-it-free · grok-4-20-non-reasoning · grok-4-20-reasoning · qwen-image-2.0 · qwen-image-2.0-pro · coding-minimax-m3-free · doubao-seedance-2-0-260128 · doubao-seedance-2-0-fast-260128 · doubao-seedance-2-0-mini-260615 · glm-5.1 · glm-image · qwen3.6-plus · wan2.7-i2v · wan2.7-r2v · wan2.7-t2v · wan2.7-videoedit · cc-k2.6-code-preview · gemma-4-26b-a4b-it · gemma-4-31b-it · gpt-5.4 · wan2.7-image · wan2.7-image-pro · claude-sonnet-4-6 · coding-xiaomi-mimo-v2.5 · coding-xiaomi-mimo-v2.5-pro · doubao-seed-2-0-lite-260428 · doubao-seed-2-0-mini-260428 · gemini-3.1-flash-image-preview · gemini-3.1-pro-preview · gemini-3.1-pro-preview-customtools · gemini-3.1-pro-preview-search · gpt-5.4-mini · gpt-5.4-nano · gpt-5.5-free · qwen3.5-plus · claude-sonnet-4-6-think · coding-xiaomi-mimo-v2-omni · coding-xiaomi-mimo-v2-pro · gpt-5.3-chat-latest · gpt-5.3-codex · gpt-image-2-free · qwen3.5-122b-a10b · qwen3.5-27b · qwen3.5-35b-a3b · qwen3.5-397b-a17b · qwen3.5-flash · coding-glm-5.1 · doubao-seed-2-0-pro · gpt-5.4-high · gpt-5.4-low · gpt-5.4-pro · qwen3-coder-next · xiaomi-mimo-v2-omni-free · xiaomi-mimo-v2-pro-free · xiaomi-mimo-v2.5-free · xiaomi-mimo-v2.5-pro-free · claude-opus-4-6 · coding-glm-5.1-free · coding-minimax-m2.7-free · glm-5 · glm-5v-turbo · minimax-m2.7 · claude-opus-4-6-think · coding-glm-5-free · coding-glm-5-turbo-free · coding-minimax-m2.5-free · doubao-seed-2-0-code-preview · doubao-seed-2-0-lite-260215 · doubao-seed-2-0-mini · gemini-3-flash-preview · gemini-3-flash-preview-search · glm-5-turbo · cc-glm-5.1 · claude-opus-4-5 · claude-opus-4-5-think · embed-v-4-0 · ernie-image-turbo · gemini-3.1-flash-image-preview-free · mimo-v2-omni · mimo-v2-pro · cohere-command-a · gemini-3-flash-preview-free · cc-minimax-m3 · coding-minimax-m3 · gpt-4.1-free · gpt-4.1-mini-free · gpt-4.1-nano-free · gpt-4o-free · coding-glm-5 · coding-glm-5-turbo · glm-4.7 · veo-3.1-lite-generate-preview · glm-4.7-flash-free · coding-glm-4.7-free · doubao-seedance-1-5-pro-251215 · doubao-seedance-1-0-pro-250528 · doubao-seedance-1-0-pro-fast-251015 · gemini-3-pro-image-preview · gemini-embedding-2 · deepinfra-gemma-4-26b-a4b-it · gpt-5.2-codex · doubao-seedream-5.0-lite · gpt-image-1.5 · gpt-5.2 · gpt-5.2-chat-latest · gpt-5.2-high · gpt-5.2-low · gpt-5.2-pro · gpt-5.1 · gpt-5.1-codex-max · doubao-seed-1-8 · gpt-5.1-chat-latest · gpt-5.1-codex · gpt-5.1-codex-mini · claude-haiku-4-5 · claude-sonnet-4-5 · claude-sonnet-4-5-think · grok-4.20-multi-agent-0309 · mistral-large-3 · cc-glm-5 · cc-glm-5-turbo · cloudflare-glm-5.2 · gemini-2.5-flash-image · grok-4-1-fast-non-reasoning · grok-4-1-fast-reasoning · grok-code-fast-1 · k2.6-code-preview-free · mimo-v2-flash · musesteamer-air-image · qwen3.6-plus-preview-free · zai-glm-5-turbo · gpt-5 · deepseek-v3.2 · deepseek-v3.2-think · gpt-5-codex · DeepSeek-V3.1-Terminus · DeepSeek-V3.1-Think · gpt-5-pro · gpt-5-mini · gpt-5-nano · gpt-5-chat-latest · claude-opus-4-1 · o3-deep-research · kimi-k2.5 · qwen3-max-2026-01-23 · qwen3-vl-flash · qwen3-vl-flash-2026-01-22 · qwen3-vl-plus · cc-minimax-m2.7 · cc-minimax-m2.7-highspeed · minimax-m2.5 · minimax-m2.5-highspeed · mm-minimax-m2.7-highspeed · coding-minimax-m2.7 · coding-minimax-m2.7-highspeed · cc-minimax-m2.5 · cc-minimax-m2.5-highspeed · coding-minimax-m2.5 · coding-minimax-m2.5-highspeed · doubao-seedream-4-5 · sora-2 · sora-2-pro · cc-glm-4.7 · cc-minimax-m2.1 · coding-glm-4.7 · coding-minimax-m2.1 · coding-minimax-m2.1-free · gpt-4o-audio-preview · gpt-4o-mini-audio-preview · minimax-m2.1 · o3 · wan2.6-i2v · wan2.6-t2v · cc-glm-4.6 · coding-glm-4.6 · coding-glm-4.6-free · coding-minimax-m2 · coding-minimax-m2-free · flux-2-flex · flux-2-pro · gemini-2.5-pro · glm-4.6 · glm-4.6v · glm-ocr · kimi-for-coding-free · o3-pro · qianfan-ocr · qianfan-ocr-fast · step-3.5-flash · wan2.2-i2v-plus · wan2.5-i2v-preview · wan2.5-t2v-preview · gemini-2.5-pro-search · kimi-k2-thinking · gemini-2.5-flash · gemini-2.5-flash-preview-09-2025 · glm-4.5v · gemini-2.5-flash-lite · gemini-2.5-flash-lite-nothink · gemini-2.5-flash-lite-preview-09-2025 · gemini-2.5-flash-lite-preview-09-2025-nothink · gemini-2.5-flash-nothink · gemini-2.5-flash-search · gemini-2.5-flash-preview-05-20-nothink · gemini-2.5-flash-preview-05-20-search · DeepSeek-V3-Fast · imagen-4.0 · imagen-4.0-fast-generate-001 · imagen-4.0-generate-001 · imagen-4.0-ultra-generate-001 · imagen-4.0-ultra · gpt-image-1 · gpt-image-1-mini · o4-mini · DeepSeek-OCR · alicloud-kimi-k2-instruct · deepseek-ocr · ernie-5.0-thinking-exp · flux-kontext-max · gemini-2.5-flash-image-preview · glm-4.5 · gpt-4.1 · grok-4 · grok-4-fast-non-reasoning · grok-4-fast-reasoning · kimi-k2-0711 · kimi-k2-instruct · kimi-k2-turbo-preview · paddleocr-vl-0.9b · pp-structurev3 · qwen3-vl-235b-a22b-instruct · qwen3-vl-235b-a22b-thinking · qwen3-vl-30b-a3b-instruct · qwen3-vl-30b-a3b-thinking · veo-3.0-generate-preview · veo-3.1-fast-generate-preview · veo-3.1-generate-preview · aihubmix-router · gpt-4.1-mini · gpt-4.1-nano · gemini-2.5-pro-preview-05-06 · gemini-2.5-pro-preview-03-25 · gemini-2.5-pro-preview-05-06-search · gemini-2.5-pro-preview-03-25-search · qwen3-max-preview · qwen3-max · qwen3-next-80b-a3b-instruct · qwen3-next-80b-a3b-thinking · qwen3-235b-a22b-instruct-2507 · qwen3-235b-a22b-thinking-2507 · qwen3-coder-30b-a3b-instruct · qwen3-coder-480b-a35b-instruct · DeepSeek-V3 · LongCat-Flash-Chat · gemini-2.5-pro-preview-06-05-search · jina-embeddings-v5-text-nano · jina-embeddings-v5-text-small · qwen3-235b-a22b · qwen3-coder-flash · qwen3-coder-plus · qwen3-coder-plus-2025-07-22 · Qwen2.5-VL-72B-Instruct · ernie-5.0-thinking-preview · inclusionAI/Ling-1T · inclusionAI/Ring-1T · bce-reranker-base · codex-mini-latest · doubao-seedream-4-0 · embedding-v1 · ernie-4.5-turbo-latest · glm-4.5-x · gme-qwen2-vl-2b-instruct · gte-rerank-v2 · inclusionAI/Ling-flash-2.0 · inclusionAI/Ling-mini-2.0 · inclusionAI/Ring-flash-2.0 · jina-deepsearch-v1 · jina-embeddings-v4 · jina-reranker-v3 · llama-4-maverick · llama-4-scout · qwen-image · qwen-image-edit · qwen-image-max · qwen-mt-plus · qwen-mt-turbo · qwen3-embedding-0.6b · qwen3-embedding-4b · qwen3-embedding-8b · qwen3-reranker-0.6b · qwen3-reranker-4b · qwen3-reranker-8b · tao-8k · jina-clip-v2 · jina-reranker-m0 · jina-colbert-v2 · DeepSeek-R1 · gpt-4o-search-preview · gpt-4o-mini-search-preview · jina-embeddings-v3 · claude-3-7-sonnet · ernie-4.5 · ernie-4.5-turbo-vl · mimo-v2-flash-free · FLUX-1.1-pro · o3-mini · doubao-seed-1-6 · doubao-seed-1-6-flash · doubao-seed-1-6-lite · doubao-seed-1-6-thinking · qwen3-30b-a3b-instruct-2507 · qwen3-30b-a3b-thinking-2507 · Qwen2-VL-72B-Instruct · Qwen2-VL-7B-Instruct · cc-kimi-for-coding · gemini-embedding-001 · gpt-oss-120b · qwen-3-235b-a22b-thinking-2507 · Qwen/Qwen3-30B-A3B · Qwen/Qwen3-32B · qwen3-32b · Qwen/Qwen3-14B · Qwen/Qwen3-8B · embedding-2 · embedding-3 · gemini-2.5-pro-preview-06-05 · Qwen/Qwen2.5-VL-72B-Instruct · o1 · o1-pro · ByteDance-Seed/Seed-OSS-36B-Instruct · doubao-seed-1-6-250615 · doubao-seed-1-6-flash-250615 · doubao-seed-1-6-thinking-250615 · doubao-seed-1-6-vision-250815 · Doubao-1.5-thinking-pro · cc-minimax-m2 · deepseek-ai/DeepSeek-Prover-V2-671B · gemini-2.5-flash-preview-tts · gemini-2.5-pro-preview-tts · gemma-3-12b-it · gemma-3-27b-it · gemma-3-4b-it · gemma-3n-e4b-it · gemma-3-1b-it · deepseek-r1-distill-llama-70b · gpt-4o-mini-tts · tngtech/DeepSeek-R1T-Chimera · veo-2.0-generate-001 · o1-preview · o1-mini · gpt-4o-2024-11-20 · gpt-4o · gpt-4o-mini · AiHubmix-mistral-medium · ERNIE-X1.1-Preview · Qwen/QwQ-32B · chutesai/Mistral-Small-3.1-24B-Instruct-2503 · ernie-x1.1-preview · minimax-m2 · MiniMaxAI/MiniMax-M1-80k · Qwen/Qwen2.5-VL-32B-Instruct · baidu/ERNIE-4.5-300B-A47B · bge-large-en · bge-large-zh · codestral-latest · ernie-4.5-0.3b · ernie-4.5-turbo-128k-preview · ernie-x1-turbo · kat-dev · llama-3.3-70b · moonshotai/Kimi-Dev-72B · moonshotai/Moonlight-16B-A3B-Instruct · nvidia-nemotron-3-super-120b-a12b · o1-global · qianfan-qi-vl · qwen2.5-vl-72b-instruct · tencent/Hunyuan-A13B-Instruct · unsloth/gemma-3-27b-it · gemini-exp-1206 · gpt-4o-zh · qwen-qwq-32b · unsloth/gemma-3-12b-it · qwen-max-0125 · BAAI/bge-large-en-v1.5 · BAAI/bge-large-zh-v1.5 · BAAI/bge-reranker-v2-m3 · tencent/Hunyuan-MT-7B · V3 · V_2 · V_2_TURBO · V_2A · V_2A_TURBO · V_1 · V_1_TURBO · doubao-embedding-large-text-240915 · kimi-thinking-preview · gpt-4o-2024-08-06 · qwen-plus-2025-07-28 · qwen-plus-latest · sonar · stepfun-ai/step3 · text-embedding-v4 · AiHubmix-Phi-4-mini-reasoning · qwen-turbo-latest · aihub-Phi-4-multimodal-instruct · qwen3-30b-a3b · aihub-Phi-4-mini-instruct · grok-3 · aihub-Phi-4 · claude-3-opus-20240229 · dall-e-3 · doubao-embedding-text-240715 · grok-3-beta · qwen3-14b · grok-3-fast · qwen3-8b · deepseek-ai/DeepSeek-R1-Zero · grok-3-fast-beta · grok-3-mini · qwen3-4b · grok-3-mini-beta · qwen3-1.7b · qwen3-0.6b · alicloud-glm-5 · command-a-03-2025 · grok-3-mini-fast-beta · qwen-3-32b · qwen-turbo-2025-04-28 · qwen-plus-2025-04-28 · THUDM/GLM-Z1-32B-0414 · THUDM/GLM-4.1V-9B-Thinking · text-embedding-004 · THUDM/GLM-4-32B-0414 · THUDM/GLM-Z1-9B-0414 · THUDM/GLM-4-9B-0414 · cc-doubao-seed-code-preview-latest · doubao-seed-code-preview-latest · deepseek-ai/Janus-Pro-7B · glm-zero-preview · qwen-3-235b-a22b-instruct-2507 · coding-glm-4.5-air · deepinfra-nvidia-nemotron-3-nano-30b-a3b2 · glm-4.5-air · gpt-4-32k · nvidia-llama-3.1-nemotron-70b-instruct · nvidia-llama-3.3-nemotron-super-49b-v1.5 · nvidia-nemotron-3-nano-30b-a3b · nvidia-nemotron-nano-12b-v2-vl · nvidia-nemotron-nano-9b-v2 · o1-preview-2024-09-12 · Qwen/QVQ-72B-Preview · Qwen/QwQ-32B-Preview · llama-3.1-sonar-huge-128k-online · aihubmix-Mistral-Large-2411 · llama-3.1-sonar-large-128k-online · aihubmix-Mistral-large-2407 · grok-2-1212 · llama-3.1-70b · wan2.6-t2i · DESCRIBE · UPSCALE · bai-qwen3-vl-235b-a22b-instruct · cc-MiniMax-M2 · cc-deepseek-v3 · cc-deepseek-v3.1 · cc-ernie-4.5-300b-a47b · cc-kimi-dev-72b · cc-kimi-k2-instruct · cc-kimi-k2-instruct-0905 · cc-kimi-k2-thinking · computer-use-preview · gpt-image-test · grok-4.20-beta-0309-non-reasoning · grok-4.20-beta-0309-reasoning · grok-4.20-multi-agent-beta-0309 · jina-reader · jina-search · llama3.1-8b · o1-2024-12-17 · sf-kimi-k2-thinking · Baichuan3-Turbo · Baichuan3-Turbo-128k · Baichuan4 · Baichuan4-Air · Baichuan4-Turbo · DeepSeek-v3 · Doubao-1.5-lite-32k · Doubao-1.5-pro-256k · Doubao-1.5-pro-32k · Doubao-1.5-vision-pro-32k · Doubao-lite-128k · Doubao-lite-32k · Doubao-lite-4k · Doubao-pro-128k · Doubao-pro-256k · Doubao-pro-32k · Doubao-pro-4k · GPT-OSS-20B · Gryphe/MythoMax-L2-13b · MiniMax-Text-01 · Mistral-large-2407 · Qwen/Qwen2-1.5B-Instruct · Qwen/Qwen2-57B-A14B-Instruct · Qwen/Qwen2-72B-Instruct · Qwen/Qwen2-7B-Instruct · Qwen/Qwen2.5-32B-Instruct · Qwen/Qwen2.5-72B-Instruct · Qwen/Qwen2.5-72B-Instruct-128K · Qwen/Qwen2.5-7B-Instruct · Qwen/Qwen2.5-Coder-32B-Instruct · Qwen3-235B-A22B-Thinking-2507 · Stable-Diffusion-3-5-Large · WizardLM/WizardCoder-Python-34B-V1.0 · ahm-Phi-3-5-MoE-instruct · ahm-Phi-3-5-mini-instruct · ahm-Phi-3-5-vision-instruct · ahm-Phi-3-medium-128k · ahm-Phi-3-medium-4k · ahm-Phi-3-small-128k · aihubmix-Codestral-2501 · aihubmix-Cohere-command-r · aihubmix-Jamba-1-5-Large · aihubmix-Llama-3-1-405B-Instruct · aihubmix-Llama-3-1-70B-Instruct · aihubmix-Llama-3-1-8B-Instruct · aihubmix-Llama-3-2-11B-Vision · aihubmix-Llama-3-2-90B-Vision · aihubmix-Llama-3-70B-Instruct · aihubmix-Mistral-large · aihubmix-command-r-08-2024 · aihubmix-command-r-plus · aihubmix-command-r-plus-08-2024 · alicloud-deepseek-v3.2 · alicloud-glm-4.7 · alicloud-kimi-k2-thinking · alicloud-kimi-k2.5 · alicloud-minimax-m2.5 · anthropic-opus-4-6 · azure-deepseek-v3.2 · azure-deepseek-v3.2-speciale · azure-kimi-k2.5 · cbs-glm-4.7 · cerebras-llama-3.3-70b · chatglm_lite · chatglm_pro · chatglm_std · chatglm_turbo · claude-2 · claude-2.0 · claude-2.1 · claude-3-haiku-20240229 · claude-3-haiku-20240307 · claude-3-sonnet-20240229 · claude-instant-1 · claude-instant-1.2 · code-davinci-edit-001 · cogview-3 · cogview-3-plus · command · command-light · command-light-nightly · command-nightly · command-r · command-r-08-2024 · command-r-plus · command-r-plus-08-2024 · dall-e-2 · davinci · davinci-002 · deepinfra-llama-3.1-8b-instant · deepinfra-llama-3.3-70b-instant-turbo · deepinfra-llama-4-maverick-17b-128e-instruct · deepinfra-llama-4-scout-17b-16e-instruct · deepseek-ai/DeepSeek-Coder-V2-Instruct · deepseek-ai/DeepSeek-R1-Distill-Llama-70B · deepseek-ai/DeepSeek-R1-Distill-Llama-8B · deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B · deepseek-ai/DeepSeek-R1-Distill-Qwen-14B · deepseek-ai/DeepSeek-R1-Distill-Qwen-32B · deepseek-ai/DeepSeek-R1-Distill-Qwen-7B · deepseek-ai/DeepSeek-V2-Chat · deepseek-ai/DeepSeek-V2.5 · deepseek-ai/deepseek-llm-67b-chat · deepseek-ai/deepseek-vl2 · deepseek-v3 · distil-whisper-large-v3-en · doubao-1-5-thinking-vision-pro-250428 · fx-flux-2-pro · gemini-2.5-pro-exp-03-25 · gemini-embedding-exp-03-07 · gemini-exp-1114 · gemini-exp-1121 · gemini-pro · gemini-pro-vision · gemma-7b-it · glm-3-turbo · glm-4 · glm-4-flash · glm-4-plus · glm-4.5-airx · glm-4v · glm-4v-plus · google-gemma-3-12b-it · google-gemma-3-27b-it · google-gemma-3-4b-it · google/gemini-exp-1114 · google/gemma-2-27b-it · google/gemma-2-9b-it:free · gpt-3.5-turbo · gpt-3.5-turbo-0301 · gpt-3.5-turbo-0613 · gpt-3.5-turbo-1106 · gpt-3.5-turbo-16k · gpt-3.5-turbo-16k-0613 · gpt-3.5-turbo-instruct · gpt-4 · gpt-4-0125-preview · gpt-4-0314 · gpt-4-0613 · gpt-4-1106-preview · gpt-4-32k-0314 · gpt-4-32k-0613 · gpt-4-turbo · gpt-4-turbo-2024-04-09 · gpt-4-turbo-preview · gpt-4-vision-preview · gpt-4o-2024-05-13 · gpt-4o-mini-2024-07-18 · gpt-oss-20b · grok-2-vision-1212 · grok-vision-beta · groq-llama-3.1-8b-instant · groq-llama-3.3-70b-versatile · groq-llama-4-maverick-17b-128e-instruct · groq-llama-4-scout-17b-16e-instruct · jina-embeddings-v2-base-code · learnlm-1.5-pro-experimental · llama-3.1-405b-instruct · llama-3.1-405b-reasoning · llama-3.1-70b-versatile · llama-3.1-8b-instant · llama-3.1-sonar-small-128k-online · llama-3.2-11b-vision-preview · llama-3.2-1b-preview · llama-3.2-3b-preview · llama-3.2-90b-vision-preview · llama2-70b-4096 · llama2-70b-40960 · llama2-7b-2048 · llama3-70b-8192 · llama3-8b-8192 · llama3-groq-70b-8192-tool-use-preview · llama3-groq-8b-8192-tool-use-preview · meta-llama/Llama-3.2-90B-Vision-Instruct · meta-llama/llama-3.1-405b-instruct:free · meta-llama/llama-3.1-70b-instruct:free · meta-llama/llama-3.1-8b-instruct:free · meta-llama/llama-3.2-11b-vision-instruct:free · meta-llama/llama-3.2-3b-instruct:free · meta/llama-3.1-405b-instruct · meta/llama3-8B-chat · mistralai/mistral-7b-instruct:free · moonshot-kimi-k2.5 · moonshot-v1-128k · moonshot-v1-128k-vision-preview · moonshot-v1-32k · moonshot-v1-32k-vision-preview · moonshot-v1-8k · moonshot-v1-8k-vision-preview · nvidia/Llama-3_1-Nemotron-Ultra-253B-v1 · o1-mini-2024-09-12 · omni-moderation-latest · qwen-flash · qwen-flash-2025-07-28 · qwen-long · qwen-max · qwen-max-longcontext · qwen-plus · qwen-turbo · qwen-turbo-2024-11-01 · qwen2.5-14b-instruct · qwen2.5-32b-instruct · qwen2.5-3b-instruct · qwen2.5-72b-instruct · qwen2.5-7b-instruct · qwen2.5-coder-1.5b-instruct · qwen2.5-coder-7b-instruct · qwen2.5-math-1.5b-instruct · qwen2.5-math-72b-instruct · qwen2.5-math-7b-instruct · step-2-16k · text-ada-001 · text-babbage-001 · text-curie-001 · text-davinci-002 · text-davinci-003 · text-davinci-edit-001 · text-embedding-3-large · text-embedding-3-small · text-embedding-ada-002 · text-embedding-v1 · text-moderation-007 · text-moderation-latest · text-moderation-stable · text-search-ada-doc-001 · tts-1 · tts-1-1106 · tts-1-hd · tts-1-hd-1106 · whisper-1 · whisper-large-v3 · whisper-large-v3-turbo · yi-large · yi-large-rag · yi-large-turbo · yi-lightning · yi-medium · yi-vl-plus · deepseek-r1-distill-qianfan-llama-8b · doubao-1-5-pro-256k-250115 · doubao-1-5-pro-32k-250115 · gpt-4o-2024-08-06-global · gpt-4o-mini-global · meta-llama-3-70b · meta-llama-3-8b · o3-global · o3-mini-global · o3-pro-global · qianfan-chinese-llama-2-13b · qianfan-llama-vl-8b