Models

AIHubMix's in-house automatic model router. Not every request needs the strongest, most expensive model — drawing on internal evals and public leaderboards, we bring together a range of today's mainstream models of varying capability and route each request automatically through a small model we trained. Set model to auto and requests are dispatched to the right model based on their content, lowering cost. Currently in beta; we keep updating and improving it.

GLM-5.3 is Z.AI’s coding and agentic reasoning model, built for complex software engineering, long-running agent tasks, vulnerability analysis, and other demanding workloads. Building on GLM-5.2, it incorporates further post-training improvements to deliver stronger coding performance, better task execution, and greater token efficiency. We currently offer the production-ready GLM-5.3 API with unlimited concurrency, making it well suited for high-throughput workloads, coding agents, and large-scale automation. For a limited time, GLM-5.3 is available at 10% off.

GLM-5.3 is Z.ai’s reasoning model for coding and agentic workflows, designed for complex software engineering, long-running agents, and vulnerability analysis. It uses the same base model as GLM-5.2, with scaled post-training improving coding, task execution, and token efficiency. This model is a limited-time preview version of GLM-5.3, intended for testing and evaluation only. Service stability is not guaranteed, and we do not recommend using it in production environments. We’re waiting for the official commercial API release and will integrate it as soon as official support becomes available.
- Input: $ 0.75 /M
- Output: $ 3.75 /M
- Web Search: $0.014/request
- Cache Storage: $1/h/M tokens
- Input Audio: $1/M tokens
- Input Video: $1/M tokens
Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web development, and knowledge work. It supports a 1M-token context window and adjustable thinking levels. Compared with Gemini 3.6 Flash, it improves coding, tool use, multi-step planning, and instruction following.
Developed by Stealth, Ox Alpha is a reasoning model designed for coding, sustained agentic work, and demanding production workloads. With a massive context length of 1,048,576 tokens, it is ideally suited for long-horizon software engineering and complex reasoning. The model excels at supporting advanced workflows that combine text with other data.

Dots3-Note Preview is an open-weight mixture-of-experts model developed by Dots Studio, featuring 16B active parameters out of 280B total. As the lightest model in the Dots 3 family, it is designed for efficient performance while supporting an expansive context length of 512,000 tokens. This preview version provides an accessible way to experience the capabilities of the Dots 3 architecture.
- Input: $ 0 /M
- Output: $ 0 /M
- Web Search: $0.014/request
- Cache Storage: $1/h/M tokens
- Input Audio: $2/M tokens
- Input Video: $1/M tokens
Gemini 3.7 Flash free version: Free model resources are limited and provided only for trial use; stability cannot be guaranteed, and you may encounter 429 errors during use. If you need to use it in a production environment and require unlimited concurrency with absolute stability, please choose the official version: gemini-3.7-flash

GLM-5.2 is Z.ai’s flagship model for the era of long-horizon tasks. With a truly usable 1M-token context window, it can handle project-level engineering context, execute long-running tasks more reliably, follow engineering standards more consistently, and complete the full development workflow from requirements to multi-platform deployment in a single task.
DeepSeek-V4-Flash-0731(deepseek-v4-flash-0731) is an open-source MoE large language model developed by the Chinese AI company DeepSeek, with support for a million-token context window. It is designed for coding, complex reasoning, tool use, agentic workflows, and long-document processing. Its advantages include strong performance with fewer active parameters and improved efficiency through DSpark speculative decoding. Compared with DeepSeek V4-Flash Preview, it offers significantly stronger coding and agent capabilities, while outperforming DeepSeek V4-Pro Preview on several benchmarks with fewer active parameters.
- Input: $ 0.142 /M
- Output: $ 0.284 /M
- Web Search: $0.00056/request
DeepSeek’s officially released new multimodal visual-understanding model, DeepSeek‑V4‑Flash‑Vision‑Exp, is experimental in nature and supports multimodal inputs. In pure-text capabilities (agents, reasoning, world knowledge, etc.), DeepSeek‑V4‑Flash‑Vision‑Exp is on par with the official DeepSeek‑V4‑Flash release.
DeepSeek V4 Pro 0813 is DeepSeek’s high-performance general-purpose reasoning and agent model, designed for complex reasoning, coding, long-document analysis, and agentic workflows. It supports thinking and non-thinking modes, a 1M-token context window, up to 384K output, tool calling, and the Responses API. Compared with V4 Flash 0731, Pro prioritizes capability on complex tasks, while Flash focuses on speed, cost efficiency, and high concurrency.
- Input: $ 2 /M
- Output: $ 6 /M
Grok 4.6 is xAI’s (SpaceXAI) flagship multimodal reasoning model for coding, long-running agents, knowledge work, and interactive application development. It supports image understanding, a 500K context window, tool calling, and structured outputs. Compared with Grok 4.5, it offers stronger multi-step execution, self-verification, coding, and visual project generation.
DeepSeek V4 Flash 0731 Fast is a high-speed deployment of DeepSeek’s agentic model provided by Wafer, designed for coding, tool use, and high-volume agent workloads. It preserves the capabilities of V4 Flash 0731 while delivering much faster inference. Compared with V4 Pro, it prioritizes latency and execution efficiency.
MAI-Thinking-1 is Microsoft’s first inference model in the MAI series, built for enterprise-scale workloads. With excellent reasoning, mathematical, and general intelligence capabilities, combined with superior cost-effectiveness, it makes high-throughput, 24/7 AI workloads economically viable.
Wan 3.0 (Tongyi Wanxiang 3.0) is an integrated video generation and editing model launched by Alibaba’s Tongyi Lab. It is suitable for scenarios such as advertising, e-commerce, short films, character animation, video editing, and document-to-video. Its advantages include native support for videos up to 30 seconds long and the unified handling of reference content such as text, images, audio, video, documents, spreadsheets, presentations, and web pages, delivering more realistic visuals and audio, stable character consistency, and capabilities for referencing, editing, replication, and driving. Compared with Wan 2.7, which split text-to-video, image-to-video, reference generation, and editing across multiple models, Wan 3.0 integrates these capabilities into a single model, making the creative workflow more unified and flexible.
Wan3.0 Video Prime is Alibaba Cloud’s preview high-speed edition of its All-in-One video model, positioned for faster-turnaround ads, ecommerce, short films, and editing. Aligned with Wan3.0 Video’s capabilities, it supports text-to-video, image-to-video, reference-based generation, document/webpage-to-video, up to 30 seconds at 30fps, 1080P output, adaptive aspect ratios, and native audio. It unifies modes that Wan 2.7 handled through separate task-specific models.
- Input: $ 4$ 2 /M
- Output: $ 20$ 10 /M
- Web Search: $0.01/request
GPT-5.6 Sol (limited-time 50% off) is OpenAI’s frontier reasoning model for complex coding, professional knowledge work, deep research, and long-running agents. It supports a roughly 1.05M-token context window, image understanding, and extensive tool use. Compared with Terra and Luna, Sol prioritizes capability and reliability on demanding tasks.
- Input: $ 2 /M
- Output: $ 0 /M
Seedance 2.5 is ByteDance Seed’s next-generation unified multimodal audio-video generation model, designed for filmmaking, advertising, education, simulation, and long-form content creation. It generates up to 30 seconds of synchronized audio and video in one pass and supports multi-round extensions. Compared with Seedance 2.0, it offers stronger storytelling, smoother transitions, richer reference support, improved realism, and more precise editing.
- Input: $ 0.2 /M
- Output: $ 1.2 /M
- Web Search: $0.01/request
GPT-5.6 Luna is designed for cost-sensitive, high-volume workloads. It roughly corresponds to the nano model tier used in earlier GPT-5 families.
- Input: $ 4 /M
- Output: $ 20 /M
- Web Search: $0.01/request
GPT‑5.6 Sol sets a new standard for both intelligence and efficiency, achieving state-of-the-art results across coding, knowledge work, cybersecurity, and science while outperforming previous and competing frontier models with fewer tokens and at lower estimated cost.
Popular models
Qwen3.8 Max Preview · Kimi K3 · Qwen3.8 Max · Qwen3.7 Flash · GLM 5.2 · Grok 4.5 · Claude Opus 5 · Claude Sonnet 5 · GPT 5.6 Luna · Gemini 3.6 Flash · DeepSeek V4 Flash · GPT 5.5 · Gemini 3.1 Pro Preview
Browse by model author
OpenAI (133) · Anthropic (28) · Google (83) · Grok (26) · Qwen (142) · DeepSeek (39) · Z.AI (65) · ByteDance (46) · Llama (50) · AI21 (2) · Microsoft (14) · Cohere (19) · Mistral (10) · Yi (6) · Moonshot AI (34) · StepFun (4) · Nvidia (17) · Minimax (31) · Perplexity (4) · Baichuan (5) · Ideogram (9) · Jina AI (13) · Stable diffusion (1) · Hunyuan (5) · Baidu (23) · Flux (5) · Meituan (2) · Xiaomi (14) · InclusionAI (7) · BAAI (5) · Agnes (4) · Meta (3) · KLing (2) · Poolside (2) · Dots Studio (1) · Liquid (1) · Stealth (1)
50 free models — no credit card required →
All models (856)
Auto · GLM 5.3 · Coding GLM 5.3 · Gemini 3.7 Flash · Ox Alpha · Dots 3 Note Preview (free) · Gemini 3.7 Flash (free) · GLM 5.2 · DeepSeek V4 Flash 0731 · DeepSeek V4 Flash Vision Exp · DeepSeek V4 Pro 0813 · Grok 4.6 · DeepSeek V4 Flash 0731 Fast · Mai Thinking 1 · Wan3.0 Video · Wan3.0 Video Prime · GPT 5.6 Sol Disc · Doubao Seedance 2.5 260628 · GPT 5.6 Luna · GPT 5.6 Sol · GPT 5.6 Terra · Agnes 2.5 Flash · Agnes 2.5 Pro · Agnes 2.5 Pro Alpha · Agnes Image 2.1 Flash · Grok 4.5 · Lfm 2.5 2.6b (free) · MiniMax H3 · Muse Glimmer 30B · Nemotron Lightning 3.5 30B A3B · Qwen3.8 2.4t A95B · Claude Opus 5 · Gemini 3.6 Flash · Ling 3.0 Tiny (free) · Nemotron 3.5 Lightning (free) · Qwen Image 3.0 · Qwen Image 3.0 Pro · Qwen3.8 Max · Claude Sonnet 5 · Kimi K3 · Ling 3.0 Flash (free) · Muse Spark 1.2 · Qwen3.8 Max Preview · Gemini 3.1 Flash Lite Image · Gemini 3.5 Flash Lite · Gemini 3.5 Flash Lite (free) · Gemini 3.6 Flash (free) · GLM 5.2 Fast Preview · Muse Spark 1.1 · Qwen Audio 3.0 Tts Flash · Qwen Audio 3.0 Tts Plus · Claude Fable 5 · Jina Reranker V3.5 · Claude Opus 4.8 · Hy3 · Doubao Seed 2.1 Pro · Doubao Seed 2.1 Turbo · Mai Image 2.5 Pro · Gemini 3.5 Flash · Grok Build 0.1 · Mai Image 2.5 · Mai Image 2.5 Flash · Coding Kimi K3 · Happyhorse 1.1 I2v · Happyhorse 1.1 R2v · Happyhorse 1.1 T2v · Coding GLM 5.2 (free) · Coding Kimi K3 (free) · Gemini 3.1 Flash Image · GPT Oss 20B (free) · Kimi K2.7 Code · Kimi K2.7 Code Highspeed · Gemini 3 Pro Image · GPT 4o Transcribe Diarize · GPT Audio 1.5 · Hy 3d 3.1 · Kling V3 Omni · Kling Video O1 · Longcat 2.0 · Nemotron Nano 9B V2 (free) · Hy3 Preview · MiniMax M3 · Nemotron Nano 12B V2 VL (free) · Qwen3.7 Flash · Qwen3.7 Plus · Step 3.7 Flash · Claude Opus 4.8 Thinking · Nemotron 3 Super 120B A12B (free) · Nemotron 3 Nano Omni 30B A3B (reasoning) (free) · Nemotron 3 Ultra 550B A55B (free) · Qwen3.7 Max · GPT Image 2 · Nemotron 3.5 Content Safety (free) · Coding GLM 5.2 · ERNIE 5.1 · Gemini 3.1 Flash Lite · gemini-3.1-flash-lite-nothink · Grok 4.3 · Happyhorse 1.0 I2v · Happyhorse 1.0 R2v · Happyhorse 1.0 T2v · Happyhorse 1.0 Video Edit · North Mini Code (free) · GPT 5.5 · GPT 5.5 Pro · Laguna Xs 2.1 (free) · DeepSeek V4 Flash · DeepSeek V4 Pro · Gemma 4 31B It (free) · Command A Plus 05 2026 · Doubao Seedream 5.0 Pro · ERNIE 5.0 · Kimi K2.6 · Laguna S 2.1 (free) · Qwen3.6 Max Preview · Xiaomi Mimo V2.5 · Xiaomi Mimo V2.5 Pro · Claude Opus 4.7 · Claude Opus 4.7 Thinking · GPT Chat · Nemotron 3 Nano 30B A3B (free) · Qwen3.6 27B · Qwen3.6 35B A3B · Qwen3.6 Flash · Cohere Rerank V4.0 Fast · Cohere Rerank V4.0 Pro · Gemma 4 26B A4B It (free) · grok-4-20-non-reasoning · Grok 4 20 (reasoning) · Qwen Image 2.0 · Qwen Image 2.0 Pro · Coding MiniMax M3 (free) · Doubao Seedance 2.0 260128 · Doubao Seedance 2.0 Fast 260128 · Doubao Seedance 2.0 Mini 260615 · GLM 5.1 · GLM Image · Qwen3.6 Plus · Wan2.7 I2v · Wan2.7 R2v · Wan2.7 T2v · Wan2.7 Videoedit · CC K2.6 Code Preview · Gemma 4 26B A4B It · Gemma 4 31B It · GPT 5.4 · Wan2.7 Image · Wan2.7 Image Pro · Claude Sonnet 4.6 · Coding Xiaomi Mimo V2.5 · Coding Xiaomi Mimo V2.5 Pro · Doubao Seed 2.0 Lite 260428 · Doubao Seed 2.0 Mini 260428 · Gemini 3.1 Flash Image Preview · Gemini 3.1 Pro Preview · Gemini 3.1 Pro Preview Customtools · Gemini 3.1 Pro Preview Search · GPT 5.4 Mini · GPT 5.4 Nano · GPT 5.5 (free) · Qwen3.5 Plus · Claude Sonnet 4.6 Thinking · Coding Xiaomi Mimo V2 Omni · Coding Xiaomi Mimo V2 Pro · GPT 5.3 Chat · GPT-5.3-Codex · GPT Image 2 (free) · Qwen3.5 122B A10B · Qwen3.5 27B · Qwen3.5 35B A3B · Qwen3.5 397B A17B · Qwen3.5 Flash · Coding GLM 5.1 · Doubao Seed 2.0 Pro · GPT 5.4 High · GPT 5.4 Low · GPT 5.4 Pro · Qwen3 Coder Next · Xiaomi Mimo V2 Omni (free) · Xiaomi Mimo V2 Pro (free) · Xiaomi Mimo V2.5 (free) · Xiaomi Mimo V2.5 Pro (free) · Claude Opus 4.6 · Coding GLM 5.1 (free) · Coding MiniMax M2.7 (free) · GLM 5 · GLM 5 Vision Turbo · MiniMax M2.7 · Claude Opus 4.6 Thinking · Coding GLM 5 (free) · Coding GLM 5 Turbo (free) · Coding MiniMax M2.5 (free) · Doubao Seed 2.0 Code Preview · Doubao Seed 2.0 Lite 260215 · Doubao Seed 2.0 Mini · Gemini 3 Flash Preview · Gemini 3 Flash Preview Search · GLM 5 Turbo · CC GLM 5.1 · Claude Opus 4.5 · Claude Opus 4.5 Thinking · Embed V 4.0 · ERNIE Image Turbo · Gemini 3.1 Flash Image Preview (free) · MiMo V2 Omni · MiMo V2 Pro · Cohere Command A · Gemini 3 Flash Preview (free) · CC MiniMax M3 · Coding MiniMax M3 · GPT 4.1 (free) · GPT 4.1 Mini (free) · GPT 4.1 Nano (free) · GPT 4o (free) · Coding GLM 5 · Coding GLM 5 Turbo · GLM 4.7 · Veo 3.1 Lite Generate Preview · GLM 4.7 Flash (free) · Coding GLM 4.7 (free) · Doubao Seedance 1.5 Pro 251215 · Doubao Seedance 1.0 Pro 250528 · Doubao Seedance 1.0 Pro Fast 251015 · Gemini 3 Pro Image Preview · Gemini Embedding 2 · Deepinfra Gemma 4 26B A4B It · GPT-5.2-Codex · Doubao Seedream 5.0 Lite · GPT Image 1.5 · Baidu DeepSeek V4 Pro 0813 · GPT 5.2 · GPT 5.2 Chat · GPT 5.2 High · GPT 5.2 Low · GPT 5.2 Pro · GPT 5.1 · GPT-5.1-Codex Max · Doubao Seed 1.8 · GPT 5.1 Chat · GPT-5.1-Codex · GPT-5.1-Codex Mini · Claude Haiku 4.5 · Claude Sonnet 4.5 · Claude Sonnet 4.5 Thinking · Grok 4.20 Multi Agent 0309 · Mistral Large 3 · CC GLM 5 · CC GLM 5 Turbo · Cloudflare Glm 5.2 · Gemini 2.5 Flash Image · grok-4-1-fast-non-reasoning · Grok 4.1 Fast (reasoning) · Grok Code Fast 1 · K2.6 Code Preview (free) · MiMo V2 Flash · Musesteamer Air Image · Qwen3.6 Plus Preview (free) · Zai Glm 5 Turbo · GPT 5 · DeepSeek V3.2 · DeepSeek V3.2 Thinking · GPT-5-Codex · DeepSeek V3.1 Terminus · DeepSeek V3.1 Thinking · GPT 5 Pro · GPT 5 Mini · GPT 5 Nano · GPT 5 Chat · Claude Opus 4.1 · O3 Deep Research · Kimi K2.5 · Qwen3 Max 2026 01-23 · Qwen3 VL Flash · Qwen3 VL Flash 2026 01-22 · Qwen3 VL Plus · CC MiniMax M2.7 · CC MiniMax M2.7 Highspeed · MiniMax M2.5 · MiniMax M2.5 Highspeed · Mm Minimax M2.7 Highspeed · Coding MiniMax M2.7 · Coding MiniMax M2.7 Highspeed · CC MiniMax M2.5 · CC MiniMax M2.5 Highspeed · Coding MiniMax M2.5 · Coding MiniMax M2.5 Highspeed · Doubao Seedream 4.5 · Sora 2 · Sora 2 Pro · CC GLM 4.7 · CC MiniMax M2.1 · Coding GLM 4.7 · Coding MiniMax M2.1 · Coding MiniMax M2.1 (free) · GPT 4o Audio Preview · GPT 4o Mini Audio Preview · MiniMax M2.1 · O3 · Wan2.6 I2v · Wan2.6 T2v · CC GLM 4.6 · Coding GLM 4.6 · Coding GLM 4.6 (free) · Coding MiniMax M2 · Coding MiniMax M2 (free) · Flux 2 Flex · Flux 2 Pro · Gemini 2.5 Pro · GLM 4.6 · GLM 4.6 Vision · GLM Ocr · Kimi For Coding (free) · O3 Pro · Qianfan Ocr · Qianfan Ocr Fast · Step 3.5 Flash · Wan2.2 I2v Plus · Wan2.5 I2v Preview · Wan2.5 T2v Preview · Gemini 2.5 Pro Search · Kimi K2 Thinking · Gemini 2.5 Flash · Gemini 2.5 Flash Preview 09 2025 · GLM 4.5 Vision · Gemini 2.5 Flash Lite · gemini-2.5-flash-lite-nothink · Gemini 2.5 Flash Lite Preview 09 2025 · gemini-2.5-flash-lite-preview-09-2025-nothink · gemini-2.5-flash-nothink · Gemini 2.5 Flash Search · gemini-2.5-flash-preview-05-20-nothink · Gemini 2.5 Flash Preview 05-20 Search · DeepSeek V3 Fast · Imagen 4.0 · Imagen 4.0 Fast Generate 001 · Imagen 4.0 Generate 001 · Imagen 4.0 Ultra Generate 001 · Imagen 4.0 Ultra · GPT Image 1 · GPT Image 1 Mini · O4 Mini · DeepSeek-OCR · Alicloud Kimi K2 Instruct · DeepSeek Ocr · ERNIE 5.0 Thinking Exp · Flux Kontext Max · Gemini 2.5 Flash Image Preview · GLM 4.5 · GPT 4.1 · Grok 4 · grok-4-fast-non-reasoning · Grok 4 Fast (reasoning) · Kimi K2 0711 · Kimi K2 Instruct · Kimi K2 Turbo Preview · Paddleocr VL 0.9b · Pp Structurev3 · Qwen3 VL 235B A22B Instruct · Qwen3 VL 235B A22B Thinking · Qwen3 VL 30B A3B Instruct · Qwen3 VL 30B A3B Thinking · Veo 3.0 Generate Preview · Veo 3.1 Fast Generate Preview · Veo 3.1 Generate Preview · Aihubmix Router · GPT 4.1 Mini · GPT 4.1 Nano · Gemini 2.5 Pro Preview 05-06 · Gemini 2.5 Pro Preview 03-25 · Gemini 2.5 Pro Preview 05-06 Search · Gemini 2.5 Pro Preview 03-25 Search · Qwen3 Max Preview · Qwen3 Max · Qwen3 Next 80B A3B Instruct · Qwen3 Next 80B A3B Thinking · Qwen3 235B A22B Instruct 2507 · Qwen3 235B A22B Thinking 2507 · Qwen3 Coder 30B A3B Instruct · Qwen3 Coder 480B A35B Instruct · DeepSeek V3 · LongCat-Flash-Chat · Gemini 2.5 Pro Preview 06-05 Search · Jina Embeddings V5 Text Nano · Jina Embeddings V5 Text Small · Qwen3 235B A22B · Qwen3 Coder Flash · Qwen3 Coder Plus · Qwen3 Coder Plus 2025 07-22 · ERNIE 5.0 Thinking Preview · inclusionAI/Ling-1T · inclusionAI/Ring-1T · Bce Reranker Base · Codex Mini · Doubao Seedream 4.0 · Embedding V1 · ERNIE 4.5 Turbo · GLM 4.5 X · Gme Qwen2 VL 2B Instruct · Gte Rerank V2 · inclusionAI/Ling-flash-2.0 · inclusionAI/Ling-mini-2.0 · inclusionAI/Ring-flash-2.0 · Jina Deepsearch V1 · Jina Embeddings V4 · Jina Reranker V3 · Llama 4 Maverick · Llama 4 Scout · Qwen Image · Qwen Image Edit · Qwen Image Max · Qwen Mt Plus · Qwen Mt Turbo · Qwen3 Embedding 0.6b · Qwen3 Embedding 4B · Qwen3 Embedding 8B · Qwen3 Reranker 0.6b · Qwen3 Reranker 4B · Qwen3 Reranker 8B · Tao 8K · Jina Clip V2 · Jina Reranker M0 · Jina Colbert V2 · DeepSeek R1 · GPT 4o Search Preview · GPT 4o Mini Search Preview · Jina Embeddings V3 · Claude 3.7 Sonnet · ERNIE 4.5 · ERNIE 4.5 Turbo VL · MiMo V2 Flash (free) · FLUX-1.1-pro · O3 Mini · Doubao Seed 1.6 · Doubao Seed 1.6 Flash · Doubao Seed 1.6 Lite · Doubao Seed 1.6 Thinking · Qwen3 30B A3B Instruct 2507 · Qwen3 30B A3B Thinking 2507 · Qwen2 VL 72B Instruct · Qwen2 VL 7B Instruct · CC Kimi For Coding · Gemini Embedding 001 · gpt-oss-120b · Qwen 3 235B A22B Thinking 2507 · Qwen/Qwen3-30B-A3B · Qwen/Qwen3-32B · Qwen3 32B · Qwen/Qwen3-14B · Qwen/Qwen3-8B · Embedding 2 · Embedding 3 · Gemini 2.5 Pro Preview 06-05 · Qwen/Qwen2.5-VL-72B-Instruct · O1 · O1 Pro · ByteDance-Seed/Seed-OSS-36B-Instruct · Doubao Seed 1.6 250615 · Doubao Seed 1.6 Flash 250615 · Doubao Seed 1.6 Thinking 250615 · Doubao Seed 1.6 Vision 250815 · Doubao 1.5 Thinking Pro · CC MiniMax M2 · deepseek-ai/DeepSeek-Prover-V2-671B · Gemini 2.5 Flash Preview Tts · Gemini 2.5 Pro Preview Tts · Gemma 3 12B It · Gemma 3 27B It · Gemma 3 4B It · Gemma 3n E4B It · Gemma 3 1B It · DeepSeek R1 Distill Llama 70B · GPT 4o Mini Tts · tngtech/DeepSeek-R1T-Chimera · Veo 2.0 Generate 001 · O1 Preview · O1 Mini · GPT 4o 2024 11-20 · GPT 4o · GPT 4o Mini · AiHubmix-mistral-medium · ERNIE X1.1 Preview · Qwen/QwQ-32B · chutesai/Mistral-Small-3.1-24B-Instruct-2503 · ERNIE X1.1 Preview · MiniMax M2 · MiniMaxAI/MiniMax-M1-80k · Qwen/Qwen2.5-VL-32B-Instruct · baidu/ERNIE-4.5-300B-A47B · Bge Large En · Bge Large Zh · Codestral · ERNIE 4.5 0.3b · ERNIE 4.5 Turbo 128K Preview · ERNIE X1 Turbo · Kat Dev · Llama 3.3 70B · moonshotai/Kimi-Dev-72B · moonshotai/Moonlight-16B-A3B-Instruct · Nvidia Nemotron 3 Super 120B A12B · O1 Global · Qianfan Qi VL · Qwen2.5 VL 72B Instruct · tencent/Hunyuan-A13B-Instruct · unsloth/gemma-3-27b-it · gemini-exp-1206 · GPT 4o Zh · Qwen Qwq 32B · unsloth/gemma-3-12b-it · Qwen Max 0125 · BAAI/bge-large-en-v1.5 · BAAI/bge-large-zh-v1.5 · BAAI/bge-reranker-v2-m3 · tencent/Hunyuan-MT-7B · V3 · V_2 · V_2_TURBO · V_2A · V_2A_TURBO · V_1 · V_1_TURBO · Doubao Embedding Large Text 240915 · Kimi Thinking Preview · GPT 4o 2024 08-06 · Qwen Plus 2025 07-28 · Qwen Plus · Sonar · stepfun-ai/step3 · Text Embedding V4 · Aihubmix Phi 4 Mini (reasoning) · Qwen Turbo · Aihub Phi 4 Multimodal Instruct · Qwen3 30B A3B · Aihub Phi 4 Mini Instruct · Grok 3 · Aihub Phi 4 · Claude 3 Opus 20240229 · Dall E 3 · Doubao Embedding Text 240715 · Grok 3 Beta · Qwen3 14B · Grok 3 Fast · Qwen3 8B · deepseek-ai/DeepSeek-R1-Zero · Grok 3 Fast Beta · Grok 3 Mini · Qwen3 4B · Grok 3 Mini Beta · Qwen3 1.7b · Qwen3 0.6b · Alicloud Glm 5 · Command A 03 2025 · Grok 3 Mini Fast Beta · Qwen 3 32B · Qwen Turbo 2025 04-28 · Qwen Plus 2025 04-28 · THUDM/GLM-Z1-32B-0414 · THUDM/GLM-4.1V-9B-Thinking · Text Embedding 004 · THUDM/GLM-4-32B-0414 · THUDM/GLM-Z1-9B-0414 · THUDM/GLM-4-9B-0414 · CC Doubao Seed Code Preview · Doubao Seed Code Preview · deepseek-ai/Janus-Pro-7B · GLM Zero Preview · Qwen 3 235B A22B Instruct 2507 · Coding GLM 4.5 Air · Deepinfra Nvidia Nemotron 3 Nano 30B A3b2 · GLM 4.5 Air · GPT 4 32K · Nvidia Llama 3.1 Nemotron 70B Instruct · Nvidia Llama 3.3 Nemotron Super 49B V1.5 · Nvidia Nemotron 3 Nano 30B A3B · Nvidia Nemotron Nano 12B V2 VL · Nvidia Nemotron Nano 9B V2 · O1 Preview 2024 09-12 · Qwen/QVQ-72B-Preview · Qwen/QwQ-32B-Preview · Llama 3.1 Sonar Huge 128K Online · Aihubmix Mistral Large 2411 · Llama 3.1 Sonar Large 128K Online · Aihubmix Mistral Large 2407 · Grok 2 1212 · Llama 3.1 70B · Wan2.6 T2i · DESCRIBE · UPSCALE · Bai Qwen3 VL 235B A22B Instruct · CC MiniMax M2 · CC DeepSeek V3 · CC DeepSeek V3.1 · CC ERNIE 4.5 300B A47B · CC Kimi Dev 72B · CC Kimi K2 Instruct · CC Kimi K2 Instruct 0905 · CC Kimi K2 Thinking · Computer Use Preview · GPT Image Test · grok-4.20-beta-0309-non-reasoning · Grok 4.20 Beta 0309 (reasoning) · Grok 4.20 Multi Agent Beta 0309 · Jina Reader · Jina Search · Llama3.1 8B · O1 2024 12-17 · Sf Kimi K2 Thinking · Baichuan3 Turbo · Baichuan3 Turbo 128K · Baichuan4 · Baichuan4 Air · Baichuan4 Turbo · DeepSeek V3 · Doubao 1.5 Lite 32K · Doubao 1.5 Pro 256K · Doubao 1.5 Pro 32K · Doubao 1.5 Vision Pro 32K · Doubao Lite 128K · Doubao Lite 32K · Doubao Lite 4K · Doubao Pro 128K · Doubao Pro 256K · Doubao Pro 32K · Doubao Pro 4K · GPT-OSS-20B · Gryphe/MythoMax-L2-13b · MiniMax Text 01 · Mistral Large 2407 · Qwen/Qwen2-1.5B-Instruct · Qwen/Qwen2-57B-A14B-Instruct · Qwen/Qwen2-72B-Instruct · Qwen/Qwen2-7B-Instruct · Qwen/Qwen2.5-32B-Instruct · Qwen/Qwen2.5-72B-Instruct · Qwen/Qwen2.5-72B-Instruct-128K · Qwen/Qwen2.5-7B-Instruct · Qwen/Qwen2.5-Coder-32B-Instruct · Qwen3 235B A22B Thinking 2507 · Stable Diffusion 3.5 Large · WizardLM/WizardCoder-Python-34B-V1.0 · Ahm Phi 3.5 Moe Instruct · Ahm Phi 3.5 Mini Instruct · Ahm Phi 3.5 Vision Instruct · Ahm Phi 3 Medium 128K · Ahm Phi 3 Medium 4K · Ahm Phi 3 Small 128K · Aihubmix Codestral 2501 · Aihubmix Cohere Command R · Aihubmix Jamba 1.5 Large · Aihubmix Llama 3.1 405B Instruct · Aihubmix Llama 3.1 70B Instruct · Aihubmix Llama 3.1 8B Instruct · Aihubmix Llama 3.2 11B Vision · Aihubmix Llama 3.2 90B Vision · Aihubmix Llama 3 70B Instruct · Aihubmix Mistral Large · Aihubmix Command R 08 2024 · Aihubmix Command R Plus · Aihubmix Command R Plus 08 2024 · Alicloud Deepseek V3.2 · Alicloud Glm 4.7 · Alicloud Kimi K2 Thinking · Alicloud Kimi K2.5 · Alicloud Minimax M2.5 · Anthropic Opus 4.6 · Azure Deepseek V3.2 · Azure Deepseek V3.2 Speciale · Azure Kimi K2.5 · Cbs Glm 4.7 · Cerebras Llama 3.3 70B · Chatglm_lite · Chatglm_pro · Chatglm_std · Chatglm_turbo · Claude 2 · Claude 2.0 · Claude 2.1 · Claude 3 Haiku 20240229 · Claude 3 Haiku 20240307 · Claude 3 Sonnet 20240229 · Claude Instant 1 · Claude Instant 1.2 · Code Davinci Edit 001 · Cogview 3 · Cogview 3 Plus · Command · Command Light · Command Light Nightly · Command Nightly · Command R · Command R 08 2024 · Command R Plus · Command R Plus 08 2024 · Dall E 2 · Davinci · Davinci 002 · Deepinfra Llama 3.1 8B Instant · Deepinfra Llama 3.3 70B Instant Turbo · Deepinfra Llama 4 Maverick 17B 128e Instruct · Deepinfra Llama 4 Scout 17B 16e Instruct · deepseek-ai/DeepSeek-Coder-V2-Instruct · deepseek-ai/DeepSeek-R1-Distill-Llama-70B · deepseek-ai/DeepSeek-R1-Distill-Llama-8B · deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B · deepseek-ai/DeepSeek-R1-Distill-Qwen-14B · deepseek-ai/DeepSeek-R1-Distill-Qwen-32B · deepseek-ai/DeepSeek-R1-Distill-Qwen-7B · deepseek-ai/DeepSeek-V2-Chat · deepseek-ai/DeepSeek-V2.5 · deepseek-ai/deepseek-llm-67b-chat · deepseek-ai/deepseek-vl2 · DeepSeek V3 · Distil Whisper Large V3 En · Doubao 1.5 Thinking Vision Pro 250428 · Fx Flux 2 Pro · gemini-2.5-pro-exp-03-25 · gemini-embedding-exp-03-07 · gemini-exp-1114 · gemini-exp-1121 · Gemini Pro · Gemini Pro Vision · Gemma 7B It · GLM 3 Turbo · GLM 4 · GLM 4 Flash · GLM 4 Plus · GLM 4.5 Airx · GLM 4 Vision · GLM 4 Vision Plus · Google Gemma 3 12B It · Google Gemma 3 27B It · Google Gemma 3 4B It · google/gemini-exp-1114 · google/gemma-2-27b-it · google/gemma-2-9b-it:free · GPT 3.5 Turbo · GPT 3.5 Turbo 0301 · GPT 3.5 Turbo 0613 · GPT 3.5 Turbo 1106 · GPT 3.5 Turbo 16K · GPT 3.5 Turbo 16K 0613 · GPT 3.5 Turbo Instruct · GPT 4 · GPT 4 0125 Preview · GPT 4 0314 · GPT 4 0613 · GPT 4 1106 Preview · GPT 4 32K 0314 · GPT 4 32K 0613 · GPT 4 Turbo · GPT 4 Turbo 2024 04-09 · GPT 4 Turbo Preview · GPT 4 Vision Preview · GPT 4o 2024 05-13 · GPT 4o Mini 2024 07-18 · gpt-oss-20b · Grok 2 Vision 1212 · Grok Vision Beta · Groq Llama 3.1 8B Instant · Groq Llama 3.3 70B Versatile · Groq Llama 4 Maverick 17B 128e Instruct · Groq Llama 4 Scout 17B 16e Instruct · Jina Embeddings V2 Base Code · Learnlm 1.5 Pro Experimental · Llama 3.1 405B Instruct · Llama 3.1 405B (reasoning) · Llama 3.1 70B Versatile · Llama 3.1 8B Instant · Llama 3.1 Sonar Small 128K Online · Llama 3.2 11B Vision Preview · Llama 3.2 1B Preview · Llama 3.2 3B Preview · Llama 3.2 90B Vision Preview · Llama2 70B 4096 · Llama2 70B 40960 · Llama2 7B 2048 · Llama3 70B 8192 · Llama3 8B 8192 · Llama3 Groq 70B 8192 Tool Use Preview · Llama3 Groq 8B 8192 Tool Use Preview · meta-llama/Llama-3.2-90B-Vision-Instruct · meta-llama/llama-3.1-405b-instruct:free · meta-llama/llama-3.1-70b-instruct:free · meta-llama/llama-3.1-8b-instruct:free · meta-llama/llama-3.2-11b-vision-instruct:free · meta-llama/llama-3.2-3b-instruct:free · meta/llama-3.1-405b-instruct · meta/llama3-8B-chat · mistralai/mistral-7b-instruct:free · Moonshot Kimi K2.5 · Moonshot V1 128K · Moonshot V1 128K Vision Preview · Moonshot V1 32K · Moonshot V1 32K Vision Preview · Moonshot V1 8K · Moonshot V1 8K Vision Preview · nvidia/Llama-3_1-Nemotron-Ultra-253B-v1 · O1 Mini 2024 09-12 · Omni Moderation · Qwen Flash · Qwen Flash 2025 07-28 · Qwen Long · Qwen Max · Qwen Max Longcontext · Qwen Plus · Qwen Turbo · Qwen Turbo 2024 11-01 · Qwen2.5 14B Instruct · Qwen2.5 32B Instruct · Qwen2.5 3B Instruct · Qwen2.5 72B Instruct · Qwen2.5 7B Instruct · Qwen2.5 Coder 1.5b Instruct · Qwen2.5 Coder 7B Instruct · Qwen2.5 Math 1.5b Instruct · Qwen2.5 Math 72B Instruct · Qwen2.5 Math 7B Instruct · Step 2 16K · Text Ada 001 · Text Babbage 001 · Text Curie 001 · Text Davinci 002 · Text Davinci 003 · Text Davinci Edit 001 · Text Embedding 3 Large · Text Embedding 3 Small · Text Embedding Ada 002 · Text Embedding V1 · Text Moderation 007 · Text Moderation · Text Moderation Stable · Text Search Ada Doc 001 · Tts 1 · Tts 1 1106 · Tts 1 Hd · Tts 1 Hd 1106 · Whisper 1 · Whisper Large V3 · Whisper Large V3 Turbo · Yi Large · Yi Large Rag · Yi Large Turbo · Yi Lightning · Yi Medium · Yi VL Plus · DeepSeek R1 Distill Qianfan Llama 8B · Doubao 1.5 Pro 256K 250115 · Doubao 1.5 Pro 32K 250115 · GPT 4o 2024 08-06 Global · GPT 4o Mini Global · Meta Llama 3 70B · Meta Llama 3 8B · O3 Global · O3 Mini Global · O3 Pro Global · Qianfan Chinese Llama 2 13B · Qianfan Llama VL 8B