Models

Filtre
Total: 857 modèles
API de données de modèle
Types
Tout
Text
Image
Speech
Video
Transcription
Embeddings
Rerank
OCR
Tags
Tout
Featured
Coding
Free
Discount
Développeur
ToutOpenAIAnthropicGoogleGrokQwenDeepSeekZ.AIByteDanceLlamaAI21MicrosoftCohereMistralYiMoonshot AIStepFunNvidiaMinimaxPerplexityBaichuanIdeogramJina AIStable diffusionHunyuanBaiduFluxMeituanInclusionAIBAAIXiaomiKLingMetaPoolsideAgnesLiquidDots Studio
Refine
Newest
Input Modalities
Réinitialiser
autoauto:balancedauto:quality_firstauto:latency_critical

Le routeur automatique de modèles développé par AIHubMix. Toutes les requêtes n'ont pas besoin du modèle le plus puissant et le plus cher — la passerelle combine évaluations internes et classements publics, couvre une gamme de modèles grand public aux capacités variées et route chaque requête automatiquement via un petit modèle que nous avons entraîné. Réglez model sur auto et chaque requête est affectée au modèle adapté selon son contenu, réduisant les coûts. Actuellement en bêta ; nous continuons à le mettre à jour et à l'améliorer.

Facturé au prix normal du modèle réellement utilisé — le routage lui-même est gratuit.Découvrir le routage intelligent
icon
New
jusqu’à 10% de réduction·00:00–23:59 UTCCopy ID
Entrée:$ 1.1268$ 1.01412 /M
Sortie:$ 3.9438$ 3.54942 /M
Contexte:1M
Latence:4.982 s
Débit:29 tok/s

GLM-5.3 is Z.AI’s coding and agentic reasoning model, built for complex software engineering, long-running agent tasks, vulnerability analysis, and other demanding workloads. Building on GLM-5.2, it incorporates further post-training improvements to deliver stronger coding performance, better task execution, and greater token efficiency. We currently offer the production-ready GLM-5.3 API with unlimited concurrency, making it well suited for high-throughput workloads, coding agents, and large-scale automation. For a limited time, GLM-5.3 is available at 10% off.

Entrée:$ 0.06 /M
Sortie:$ 0.21999996 /M
Contexte:-
Latence:-
Débit:-

GLM-5.3 is Z.ai’s reasoning model for coding and agentic workflows, designed for complex software engineering, long-running agents, and vulnerability analysis. It uses the same base model as GLM-5.2, with scaled post-training improving coding, task execution, and token efficiency. This model is a limited-time preview version of GLM-5.3, intended for testing and evaluation only. Service stability is not guaranteed, and we do not recommend using it in production environments. We’re waiting for the official commercial API release and will integrate it as soon as official support becomes available.

  • Entrée: $ 0.75 /M
  • Sortie: $ 3.75 /M
  • Web Search: $0.014/request
  • Cache Storage: $1/h/M tokens
  • Input Audio: $1/M tokens
  • Input Video: $1/M tokens

Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web development, and knowledge work. It supports a 1M-token context window and adjustable thinking levels. Compared with Gemini 3.6 Flash, it improves coding, tool use, multi-step planning, and instruction following.

Entrée:$ 0 /M
Sortie:$ 0 /M
Contexte:128K
Latence:-
Débit:-

GLM 5.2 is a large-scale reasoning model developed by Z Ai that supports text input and output. Featuring a 128,000-token context window, this model is designed to handle complex, data-heavy tasks. It is exceptionally well-suited for long-horizon agent workflows and project-level software engineering.

Entrée:$ 0 /M
Sortie:$ 0 /M
Contexte:512K
Latence:-
Débit:-

Dots3-Note Preview is an open-weight mixture-of-experts model developed by Dots Studio, featuring 16B active parameters out of 280B total. As the lightest model in the Dots 3 family, it is designed for efficient performance while supporting an expansive context length of 512,000 tokens. This preview version provides an accessible way to experience the capabilities of the Dots 3 architecture.

  • Entrée: $ 0 /M
  • Sortie: $ 0 /M
  • Web Search: $0.014/request
  • Cache Storage: $1/h/M tokens
  • Input Audio: $2/M tokens
  • Input Video: $1/M tokens

Gemini 3.7 Flash free version: fFree model resources are limited and provided only for trial use; stability cannot be guaranteed, and you may encounter 429 errors during use. If you need to use it in a production environment and require unlimited concurrency with absolute stability, please choose the official version: gemini-3.7-flash

icon
jusqu’à 30% de réduction·00:00–13:59 UTCCopy ID
Entrée:$ 1.1268$ 0.78876 /M
Sortie:$ 3.9438$ 2.76066 /M
Contexte:1M
Latence:0.948 s
Débit:41 tok/s

GLM-5.2 is Z.ai’s flagship model for the era of long-horizon tasks. With a truly usable 1M-token context window, it can handle project-level engineering context, execute long-running tasks more reliably, follow engineering standards more consistently, and complete the full development workflow from requirements to multi-platform deployment in a single task.

Entrée:$ 0.142 /M
Sortie:$ 0.284 /M
Contexte:1M
Latence:1.011 s
Débit:98 tok/s

DeepSeek-V4-Flash-0731(deepseek-v4-flash-0731) is an open-source MoE large language model developed by the Chinese AI company DeepSeek, with support for a million-token context window. It is designed for coding, complex reasoning, tool use, agentic workflows, and long-document processing. Its advantages include strong performance with fewer active parameters and improved efficiency through DSpark speculative decoding. Compared with DeepSeek V4-Flash Preview, it offers significantly stronger coding and agent capabilities, while outperforming DeepSeek V4-Pro Preview on several benchmarks with fewer active parameters.

Entrée:$ 0.6918 /M
Sortie:$ 2.0754 /M
Contexte:1M
Latence:1.682 s
Débit:60 tok/s

DeepSeek V4 Pro 0813 is DeepSeek’s high-performance general-purpose reasoning and agent model, designed for complex reasoning, coding, long-document analysis, and agentic workflows. It supports thinking and non-thinking modes, a 1M-token context window, up to 384K output, tool calling, and the Responses API. Compared with V4 Flash 0731, Pro prioritizes capability on complex tasks, while Flash focuses on speed, cost efficiency, and high concurrency.

icon
Copy ID
  • Entrée: $ 2 /M
  • Sortie: $ 6 /M

Grok 4.6 is xAI’s (SpaceXAI) flagship multimodal reasoning model for coding, long-running agents, knowledge work, and interactive application development. It supports image understanding, a 500K context window, tool calling, and structured outputs. Compared with Grok 4.5, it offers stronger multi-step execution, self-verification, coding, and visual project generation.

Entrée:$ 2 /M
Sortie:$ 8 /M
Contexte:256K
Latence:6.669 s
Débit:42 tok/s

MAI-Thinking-1 is Microsoft’s first inference model in the MAI series, built for enterprise-scale workloads. With excellent reasoning, mathematical, and general intelligence capabilities, combined with superior cost-effectiveness, it makes high-throughput, 24/7 AI workloads economically viable.

icon
jusqu’à 50% de réduction·00:00–23:59 UTCCopy ID
  • Entrée: $ 5$ 2.5 /M
  • Sortie: $ 30$ 15 /M
  • Web Search: $0.01/request

GPT-5.6 Sol (limited-time 50% off) is OpenAI’s frontier reasoning model for complex coding, professional knowledge work, deep research, and long-running agents. It supports a roughly 1.05M-token context window, image understanding, and extensive tool use. Compared with Terra and Luna, Sol prioritizes capability and reliability on demanding tasks.

  • Entrée: $ 2 /M
  • Sortie: $ 0 /M

Seedance 2.5 is ByteDance Seed’s next-generation unified multimodal audio-video generation model, designed for filmmaking, advertising, education, simulation, and long-form content creation. It generates up to 30 seconds of synchronized audio and video in one pass and supports multi-round extensions. Compared with Seedance 2.0, it offers stronger storytelling, smoother transitions, richer reference support, improved realism, and more precise editing.

  • Entrée: $ 0.2 /M
  • Sortie: $ 1.2 /M
  • Web Search: $0.01/request

GPT-5.6 Luna is designed for cost-sensitive, high-volume workloads. It roughly corresponds to the nano model tier used in earlier GPT-5 families.

  • Entrée: $ 5 /M
  • Sortie: $ 30 /M
  • Web Search: $0.01/request

GPT‑5.6 Sol sets a new standard for both intelligence and efficiency, achieving state-of-the-art results across coding, knowledge work, cybersecurity, and science while outperforming previous and competing frontier models with fewer tokens and at lower estimated cost.

  • Entrée: $ 2 /M
  • Sortie: $ 12 /M
  • Web Search: $0.01/request

GPT-5.6 Terra is designed for workloads that balance intelligence and cost. It roughly corresponds to the mini model tier used in earlier GPT-5 families.

Entrée:$ 0.03 /M
Sortie:$ 0.15 /M
Contexte:512K
Latence:3.931 s
Débit:43 tok/s

Agnes 2.5 Flash is Agnes AI’s fast and efficient language model, an upgraded, fully available model based on Agnes 2.0 Flash. It continues to use an OpenAI-compatible Chat Completions interface and has been optimized for coding tasks, agent workflows, tool calls, multi-turn dialogue, reasoning, and image understanding.

Entrée:$ 0.45 /M
Sortie:$ 0.9 /M
Contexte:1M
Latence:18.56 s
Débit:42 tok/s

Agnes 2.5 Pro is Agnes AI’s paid inference model and the commercially stable version of the Agnes 2.5 Pro Alpha ranking model, suitable for advanced coding, scientific reasoning, long-context analysis, agent workflows, and multimodal understanding. The model is accessed via an OpenAI-compatible Chat Completions API.

Entrée:$ 0.45 /M
Sortie:$ 0.9 /M
Contexte:1M
Latence:4.931 s
Débit:11 tok/s

Agnes 2.5 Pro Alpha is Agnes AI’s paid inference model, suitable for advanced coding, scientific reasoning, long-context analysis, agent workflows, and multimodal understanding. The model is accessed via an OpenAI-compatible Chat Completions API.

Scroll to load more

Popular models

Qwen3.8 Max Preview · Kimi K3 · Qwen3.8 Max · Qwen3.7 Flash · GLM 5.2 · Grok 4.5 · Claude Opus 5 · Claude Sonnet 5 · GPT 5.6 Luna · Gemini 3.6 Flash · DeepSeek V4 Flash · GPT 5.5 · Gemini 3.1 Pro Preview

Browse by provider

OpenAI (133) · Anthropic (28) · Google (83) · Grok (26) · Qwen (141) · DeepSeek (36) · Z.AI (66) · ByteDance (46) · Llama (50) · AI21 (2) · Microsoft (14) · Cohere (19) · Mistral (10) · Yi (6) · Moonshot AI (34) · StepFun (4) · Nvidia (17) · Minimax (31) · Perplexity (4) · Baichuan (5) · Ideogram (9) · Jina AI (13) · Stable diffusion (1) · Hunyuan (5) · Baidu (23) · Flux (5) · Meituan (2) · Xiaomi (14) · InclusionAI (7) · BAAI (5) · Agnes (4) · Meta (3) · KLing (2) · Poolside (2) · Dots Studio (1) · Liquid (1)

51 free models — no credit card required →

All models (852)

auto · glm-5.3 · coding-glm-5.3 · gemini-3.7-flash · glm-5.2-free · dots-3-note-preview-free · gemini-3.7-flash-free · glm-5.2 · deepseek-v4-flash-0731 · deepseek-v4-pro-0813 · grok-4.6 · mai-thinking-1 · gpt-5.6-sol-disc · doubao-seedance-2-5-260628 · gpt-5.6-luna · gpt-5.6-sol · gpt-5.6-terra · agnes-2.5-flash · agnes-2.5-pro · agnes-2.5-pro-alpha · agnes-image-2.1-flash · grok-4.5 · lfm-2.5-2.6b-free · minimax-h3 · muse-glimmer-30b · nemotron-lightning-3.5-30b-a3b · qwen3.8-2.4t-a95b · claude-opus-5 · gemini-3.6-flash · ling-3.0-tiny-free · nemotron-3.5-lightning-free · qwen-image-3.0 · qwen-image-3.0-pro · qwen3.8-max · claude-sonnet-5 · kimi-k3 · ling-3.0-flash-free · muse-spark-1.2 · qwen3.8-max-preview · gemini-3.1-flash-lite-image · gemini-3.5-flash-lite · gemini-3.5-flash-lite-free · gemini-3.6-flash-free · glm-5.2-fast-preview · muse-spark-1.1 · qwen-audio-3.0-tts-flash · qwen-audio-3.0-tts-plus · claude-fable-5 · jina-reranker-v3.5 · claude-opus-4-8 · hy3 · doubao-seed-2-1-pro · doubao-seed-2-1-turbo · mai-image-2.5-pro · gemini-3.5-flash · grok-build-0.1 · mai-image-2.5 · mai-image-2.5-flash · coding-kimi-k3 · happyhorse-1.1-i2v · happyhorse-1.1-r2v · happyhorse-1.1-t2v · coding-glm-5.2-free · coding-kimi-k3-free · gemini-3.1-flash-image · gpt-oss-20b-free · kimi-k2.7-code · kimi-k2.7-code-highspeed · gemini-3-pro-image · gpt-4o-transcribe-diarize · gpt-audio-1.5 · hy-3d-3.1 · kling-v3-omni · kling-video-o1 · longcat-2.0 · nemotron-nano-9b-v2-free · hy3-preview · minimax-m3 · nemotron-nano-12b-v2-vl-free · qwen3.7-flash · qwen3.7-plus · step-3.7-flash · claude-opus-4-8-think · nemotron-3-super-120b-a12b-free · nemotron-3-nano-omni-30b-a3b-reasoning-free · nemotron-3-ultra-550b-a55b-free · qwen3.7-max · gpt-image-2 · nemotron-3.5-content-safety-free · coding-glm-5.2 · ernie-5.1 · gemini-3.1-flash-lite · gemini-3.1-flash-lite-nothink · grok-4.3 · happyhorse-1.0-i2v · happyhorse-1.0-r2v · happyhorse-1.0-t2v · happyhorse-1.0-video-edit · north-mini-code-free · gpt-5.5 · gpt-5.5-pro · laguna-xs-2.1-free · deepseek-v4-flash · deepseek-v4-pro · gemma-4-31b-it-free · command-a-plus-05-2026 · doubao-seedream-5.0-pro · ernie-5.0 · kimi-k2.6 · laguna-s-2.1-free · qwen3.6-max-preview · xiaomi-mimo-v2.5 · xiaomi-mimo-v2.5-pro · claude-opus-4-7 · claude-opus-4-7-think · gpt-chat-latest · nemotron-3-nano-30b-a3b-free · qwen3.6-27b · qwen3.6-35b-a3b · qwen3.6-flash · cohere-rerank-v4.0-fast · cohere-rerank-v4.0-pro · gemma-4-26b-a4b-it-free · grok-4-20-non-reasoning · grok-4-20-reasoning · qwen-image-2.0 · qwen-image-2.0-pro · coding-minimax-m3-free · doubao-seedance-2-0-260128 · doubao-seedance-2-0-fast-260128 · doubao-seedance-2-0-mini-260615 · glm-5.1 · glm-image · qwen3.6-plus · wan2.7-i2v · wan2.7-r2v · wan2.7-t2v · wan2.7-videoedit · cc-k2.6-code-preview · gemma-4-26b-a4b-it · gemma-4-31b-it · gpt-5.4 · wan2.7-image · wan2.7-image-pro · claude-sonnet-4-6 · coding-xiaomi-mimo-v2.5 · coding-xiaomi-mimo-v2.5-pro · doubao-seed-2-0-lite-260428 · doubao-seed-2-0-mini-260428 · gemini-3.1-flash-image-preview · gemini-3.1-pro-preview · gemini-3.1-pro-preview-customtools · gemini-3.1-pro-preview-search · gpt-5.4-mini · gpt-5.4-nano · gpt-5.5-free · qwen3.5-plus · claude-sonnet-4-6-think · coding-xiaomi-mimo-v2-omni · coding-xiaomi-mimo-v2-pro · gpt-5.3-chat-latest · gpt-5.3-codex · gpt-image-2-free · qwen3.5-122b-a10b · qwen3.5-27b · qwen3.5-35b-a3b · qwen3.5-397b-a17b · qwen3.5-flash · coding-glm-5.1 · doubao-seed-2-0-pro · gpt-5.4-high · gpt-5.4-low · gpt-5.4-pro · qwen3-coder-next · xiaomi-mimo-v2-omni-free · xiaomi-mimo-v2-pro-free · xiaomi-mimo-v2.5-free · xiaomi-mimo-v2.5-pro-free · claude-opus-4-6 · coding-glm-5.1-free · coding-minimax-m2.7-free · glm-5 · glm-5v-turbo · minimax-m2.7 · claude-opus-4-6-think · coding-glm-5-free · coding-glm-5-turbo-free · coding-minimax-m2.5-free · doubao-seed-2-0-code-preview · doubao-seed-2-0-lite-260215 · doubao-seed-2-0-mini · gemini-3-flash-preview · gemini-3-flash-preview-search · glm-5-turbo · cc-glm-5.1 · claude-opus-4-5 · claude-opus-4-5-think · embed-v-4-0 · ernie-image-turbo · gemini-3.1-flash-image-preview-free · mimo-v2-omni · mimo-v2-pro · cohere-command-a · gemini-3-flash-preview-free · cc-minimax-m3 · coding-minimax-m3 · gpt-4.1-free · gpt-4.1-mini-free · gpt-4.1-nano-free · gpt-4o-free · coding-glm-5 · coding-glm-5-turbo · glm-4.7 · veo-3.1-lite-generate-preview · glm-4.7-flash-free · coding-glm-4.7-free · doubao-seedance-1-5-pro-251215 · doubao-seedance-1-0-pro-250528 · doubao-seedance-1-0-pro-fast-251015 · gemini-3-pro-image-preview · gemini-embedding-2 · deepinfra-gemma-4-26b-a4b-it · gpt-5.2-codex · doubao-seedream-5.0-lite · gpt-image-1.5 · gpt-5.2 · gpt-5.2-chat-latest · gpt-5.2-high · gpt-5.2-low · gpt-5.2-pro · gpt-5.1 · gpt-5.1-codex-max · doubao-seed-1-8 · gpt-5.1-chat-latest · gpt-5.1-codex · gpt-5.1-codex-mini · claude-haiku-4-5 · claude-sonnet-4-5 · claude-sonnet-4-5-think · grok-4.20-multi-agent-0309 · mistral-large-3 · cc-glm-5 · cc-glm-5-turbo · cloudflare-glm-5.2 · gemini-2.5-flash-image · grok-4-1-fast-non-reasoning · grok-4-1-fast-reasoning · grok-code-fast-1 · k2.6-code-preview-free · mimo-v2-flash · musesteamer-air-image · qwen3.6-plus-preview-free · zai-glm-5-turbo · gpt-5 · deepseek-v3.2 · deepseek-v3.2-think · gpt-5-codex · DeepSeek-V3.1-Terminus · DeepSeek-V3.1-Think · gpt-5-pro · gpt-5-mini · gpt-5-nano · gpt-5-chat-latest · claude-opus-4-1 · o3-deep-research · kimi-k2.5 · qwen3-max-2026-01-23 · qwen3-vl-flash · qwen3-vl-flash-2026-01-22 · qwen3-vl-plus · cc-minimax-m2.7 · cc-minimax-m2.7-highspeed · minimax-m2.5 · minimax-m2.5-highspeed · mm-minimax-m2.7-highspeed · coding-minimax-m2.7 · coding-minimax-m2.7-highspeed · cc-minimax-m2.5 · cc-minimax-m2.5-highspeed · coding-minimax-m2.5 · coding-minimax-m2.5-highspeed · doubao-seedream-4-5 · sora-2 · sora-2-pro · cc-glm-4.7 · cc-minimax-m2.1 · coding-glm-4.7 · coding-minimax-m2.1 · coding-minimax-m2.1-free · gpt-4o-audio-preview · gpt-4o-mini-audio-preview · minimax-m2.1 · o3 · wan2.6-i2v · wan2.6-t2v · cc-glm-4.6 · coding-glm-4.6 · coding-glm-4.6-free · coding-minimax-m2 · coding-minimax-m2-free · flux-2-flex · flux-2-pro · gemini-2.5-pro · glm-4.6 · glm-4.6v · glm-ocr · kimi-for-coding-free · o3-pro · qianfan-ocr · qianfan-ocr-fast · step-3.5-flash · wan2.2-i2v-plus · wan2.5-i2v-preview · wan2.5-t2v-preview · gemini-2.5-pro-search · kimi-k2-thinking · gemini-2.5-flash · gemini-2.5-flash-preview-09-2025 · glm-4.5v · gemini-2.5-flash-lite · gemini-2.5-flash-lite-nothink · gemini-2.5-flash-lite-preview-09-2025 · gemini-2.5-flash-lite-preview-09-2025-nothink · gemini-2.5-flash-nothink · gemini-2.5-flash-search · gemini-2.5-flash-preview-05-20-nothink · gemini-2.5-flash-preview-05-20-search · DeepSeek-V3-Fast · imagen-4.0 · imagen-4.0-fast-generate-001 · imagen-4.0-generate-001 · imagen-4.0-ultra-generate-001 · imagen-4.0-ultra · gpt-image-1 · gpt-image-1-mini · o4-mini · DeepSeek-OCR · alicloud-kimi-k2-instruct · deepseek-ocr · ernie-5.0-thinking-exp · flux-kontext-max · gemini-2.5-flash-image-preview · glm-4.5 · gpt-4.1 · grok-4 · grok-4-fast-non-reasoning · grok-4-fast-reasoning · kimi-k2-0711 · kimi-k2-instruct · kimi-k2-turbo-preview · paddleocr-vl-0.9b · pp-structurev3 · qwen3-vl-235b-a22b-instruct · qwen3-vl-235b-a22b-thinking · qwen3-vl-30b-a3b-instruct · qwen3-vl-30b-a3b-thinking · veo-3.0-generate-preview · veo-3.1-fast-generate-preview · veo-3.1-generate-preview · aihubmix-router · gpt-4.1-mini · gpt-4.1-nano · gemini-2.5-pro-preview-05-06 · gemini-2.5-pro-preview-03-25 · gemini-2.5-pro-preview-05-06-search · gemini-2.5-pro-preview-03-25-search · qwen3-max-preview · qwen3-max · qwen3-next-80b-a3b-instruct · qwen3-next-80b-a3b-thinking · qwen3-235b-a22b-instruct-2507 · qwen3-235b-a22b-thinking-2507 · qwen3-coder-30b-a3b-instruct · qwen3-coder-480b-a35b-instruct · DeepSeek-V3 · LongCat-Flash-Chat · gemini-2.5-pro-preview-06-05-search · jina-embeddings-v5-text-nano · jina-embeddings-v5-text-small · qwen3-235b-a22b · qwen3-coder-flash · qwen3-coder-plus · qwen3-coder-plus-2025-07-22 · Qwen2.5-VL-72B-Instruct · ernie-5.0-thinking-preview · inclusionAI/Ling-1T · inclusionAI/Ring-1T · bce-reranker-base · codex-mini-latest · doubao-seedream-4-0 · embedding-v1 · ernie-4.5-turbo-latest · glm-4.5-x · gme-qwen2-vl-2b-instruct · gte-rerank-v2 · inclusionAI/Ling-flash-2.0 · inclusionAI/Ling-mini-2.0 · inclusionAI/Ring-flash-2.0 · jina-deepsearch-v1 · jina-embeddings-v4 · jina-reranker-v3 · llama-4-maverick · llama-4-scout · qwen-image · qwen-image-edit · qwen-image-max · qwen-mt-plus · qwen-mt-turbo · qwen3-embedding-0.6b · qwen3-embedding-4b · qwen3-embedding-8b · qwen3-reranker-0.6b · qwen3-reranker-4b · qwen3-reranker-8b · tao-8k · jina-clip-v2 · jina-reranker-m0 · jina-colbert-v2 · DeepSeek-R1 · gpt-4o-search-preview · gpt-4o-mini-search-preview · jina-embeddings-v3 · claude-3-7-sonnet · ernie-4.5 · ernie-4.5-turbo-vl · mimo-v2-flash-free · FLUX-1.1-pro · o3-mini · doubao-seed-1-6 · doubao-seed-1-6-flash · doubao-seed-1-6-lite · doubao-seed-1-6-thinking · qwen3-30b-a3b-instruct-2507 · qwen3-30b-a3b-thinking-2507 · Qwen2-VL-72B-Instruct · Qwen2-VL-7B-Instruct · cc-kimi-for-coding · gemini-embedding-001 · gpt-oss-120b · qwen-3-235b-a22b-thinking-2507 · Qwen/Qwen3-30B-A3B · Qwen/Qwen3-32B · qwen3-32b · Qwen/Qwen3-14B · Qwen/Qwen3-8B · embedding-2 · embedding-3 · gemini-2.5-pro-preview-06-05 · Qwen/Qwen2.5-VL-72B-Instruct · o1 · o1-pro · ByteDance-Seed/Seed-OSS-36B-Instruct · doubao-seed-1-6-250615 · doubao-seed-1-6-flash-250615 · doubao-seed-1-6-thinking-250615 · doubao-seed-1-6-vision-250815 · Doubao-1.5-thinking-pro · cc-minimax-m2 · deepseek-ai/DeepSeek-Prover-V2-671B · gemini-2.5-flash-preview-tts · gemini-2.5-pro-preview-tts · gemma-3-12b-it · gemma-3-27b-it · gemma-3-4b-it · gemma-3n-e4b-it · gemma-3-1b-it · deepseek-r1-distill-llama-70b · gpt-4o-mini-tts · tngtech/DeepSeek-R1T-Chimera · veo-2.0-generate-001 · o1-preview · o1-mini · gpt-4o-2024-11-20 · gpt-4o · gpt-4o-mini · AiHubmix-mistral-medium · ERNIE-X1.1-Preview · Qwen/QwQ-32B · chutesai/Mistral-Small-3.1-24B-Instruct-2503 · ernie-x1.1-preview · minimax-m2 · MiniMaxAI/MiniMax-M1-80k · Qwen/Qwen2.5-VL-32B-Instruct · baidu/ERNIE-4.5-300B-A47B · bge-large-en · bge-large-zh · codestral-latest · ernie-4.5-0.3b · ernie-4.5-turbo-128k-preview · ernie-x1-turbo · kat-dev · llama-3.3-70b · moonshotai/Kimi-Dev-72B · moonshotai/Moonlight-16B-A3B-Instruct · nvidia-nemotron-3-super-120b-a12b · o1-global · qianfan-qi-vl · qwen2.5-vl-72b-instruct · tencent/Hunyuan-A13B-Instruct · unsloth/gemma-3-27b-it · gemini-exp-1206 · gpt-4o-zh · qwen-qwq-32b · unsloth/gemma-3-12b-it · qwen-max-0125 · BAAI/bge-large-en-v1.5 · BAAI/bge-large-zh-v1.5 · BAAI/bge-reranker-v2-m3 · tencent/Hunyuan-MT-7B · V3 · V_2 · V_2_TURBO · V_2A · V_2A_TURBO · V_1 · V_1_TURBO · doubao-embedding-large-text-240915 · kimi-thinking-preview · gpt-4o-2024-08-06 · qwen-plus-2025-07-28 · qwen-plus-latest · sonar · stepfun-ai/step3 · text-embedding-v4 · AiHubmix-Phi-4-mini-reasoning · qwen-turbo-latest · aihub-Phi-4-multimodal-instruct · qwen3-30b-a3b · aihub-Phi-4-mini-instruct · grok-3 · aihub-Phi-4 · claude-3-opus-20240229 · dall-e-3 · doubao-embedding-text-240715 · grok-3-beta · qwen3-14b · grok-3-fast · qwen3-8b · deepseek-ai/DeepSeek-R1-Zero · grok-3-fast-beta · grok-3-mini · qwen3-4b · grok-3-mini-beta · qwen3-1.7b · qwen3-0.6b · alicloud-glm-5 · command-a-03-2025 · grok-3-mini-fast-beta · qwen-3-32b · qwen-turbo-2025-04-28 · qwen-plus-2025-04-28 · THUDM/GLM-Z1-32B-0414 · THUDM/GLM-4.1V-9B-Thinking · text-embedding-004 · THUDM/GLM-4-32B-0414 · THUDM/GLM-Z1-9B-0414 · THUDM/GLM-4-9B-0414 · cc-doubao-seed-code-preview-latest · doubao-seed-code-preview-latest · deepseek-ai/Janus-Pro-7B · glm-zero-preview · qwen-3-235b-a22b-instruct-2507 · coding-glm-4.5-air · deepinfra-nvidia-nemotron-3-nano-30b-a3b2 · glm-4.5-air · gpt-4-32k · nvidia-llama-3.1-nemotron-70b-instruct · nvidia-llama-3.3-nemotron-super-49b-v1.5 · nvidia-nemotron-3-nano-30b-a3b · nvidia-nemotron-nano-12b-v2-vl · nvidia-nemotron-nano-9b-v2 · o1-preview-2024-09-12 · Qwen/QVQ-72B-Preview · Qwen/QwQ-32B-Preview · llama-3.1-sonar-huge-128k-online · aihubmix-Mistral-Large-2411 · llama-3.1-sonar-large-128k-online · aihubmix-Mistral-large-2407 · grok-2-1212 · llama-3.1-70b · wan2.6-t2i · DESCRIBE · UPSCALE · bai-qwen3-vl-235b-a22b-instruct · cc-MiniMax-M2 · cc-deepseek-v3 · cc-deepseek-v3.1 · cc-ernie-4.5-300b-a47b · cc-kimi-dev-72b · cc-kimi-k2-instruct · cc-kimi-k2-instruct-0905 · cc-kimi-k2-thinking · computer-use-preview · gpt-image-test · grok-4.20-beta-0309-non-reasoning · grok-4.20-beta-0309-reasoning · grok-4.20-multi-agent-beta-0309 · jina-reader · jina-search · llama3.1-8b · o1-2024-12-17 · sf-kimi-k2-thinking · Baichuan3-Turbo · Baichuan3-Turbo-128k · Baichuan4 · Baichuan4-Air · Baichuan4-Turbo · DeepSeek-v3 · Doubao-1.5-lite-32k · Doubao-1.5-pro-256k · Doubao-1.5-pro-32k · Doubao-1.5-vision-pro-32k · Doubao-lite-128k · Doubao-lite-32k · Doubao-lite-4k · Doubao-pro-128k · Doubao-pro-256k · Doubao-pro-32k · Doubao-pro-4k · GPT-OSS-20B · Gryphe/MythoMax-L2-13b · MiniMax-Text-01 · Mistral-large-2407 · Qwen/Qwen2-1.5B-Instruct · Qwen/Qwen2-57B-A14B-Instruct · Qwen/Qwen2-72B-Instruct · Qwen/Qwen2-7B-Instruct · Qwen/Qwen2.5-32B-Instruct · Qwen/Qwen2.5-72B-Instruct · Qwen/Qwen2.5-72B-Instruct-128K · Qwen/Qwen2.5-7B-Instruct · Qwen/Qwen2.5-Coder-32B-Instruct · Qwen3-235B-A22B-Thinking-2507 · Stable-Diffusion-3-5-Large · WizardLM/WizardCoder-Python-34B-V1.0 · ahm-Phi-3-5-MoE-instruct · ahm-Phi-3-5-mini-instruct · ahm-Phi-3-5-vision-instruct · ahm-Phi-3-medium-128k · ahm-Phi-3-medium-4k · ahm-Phi-3-small-128k · aihubmix-Codestral-2501 · aihubmix-Cohere-command-r · aihubmix-Jamba-1-5-Large · aihubmix-Llama-3-1-405B-Instruct · aihubmix-Llama-3-1-70B-Instruct · aihubmix-Llama-3-1-8B-Instruct · aihubmix-Llama-3-2-11B-Vision · aihubmix-Llama-3-2-90B-Vision · aihubmix-Llama-3-70B-Instruct · aihubmix-Mistral-large · aihubmix-command-r-08-2024 · aihubmix-command-r-plus · aihubmix-command-r-plus-08-2024 · alicloud-deepseek-v3.2 · alicloud-glm-4.7 · alicloud-kimi-k2-thinking · alicloud-kimi-k2.5 · alicloud-minimax-m2.5 · anthropic-opus-4-6 · azure-deepseek-v3.2 · azure-deepseek-v3.2-speciale · azure-kimi-k2.5 · cbs-glm-4.7 · cerebras-llama-3.3-70b · chatglm_lite · chatglm_pro · chatglm_std · chatglm_turbo · claude-2 · claude-2.0 · claude-2.1 · claude-3-haiku-20240229 · claude-3-haiku-20240307 · claude-3-sonnet-20240229 · claude-instant-1 · claude-instant-1.2 · code-davinci-edit-001 · cogview-3 · cogview-3-plus · command · command-light · command-light-nightly · command-nightly · command-r · command-r-08-2024 · command-r-plus · command-r-plus-08-2024 · dall-e-2 · davinci · davinci-002 · deepinfra-llama-3.1-8b-instant · deepinfra-llama-3.3-70b-instant-turbo · deepinfra-llama-4-maverick-17b-128e-instruct · deepinfra-llama-4-scout-17b-16e-instruct · deepseek-ai/DeepSeek-Coder-V2-Instruct · deepseek-ai/DeepSeek-R1-Distill-Llama-70B · deepseek-ai/DeepSeek-R1-Distill-Llama-8B · deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B · deepseek-ai/DeepSeek-R1-Distill-Qwen-14B · deepseek-ai/DeepSeek-R1-Distill-Qwen-32B · deepseek-ai/DeepSeek-R1-Distill-Qwen-7B · deepseek-ai/DeepSeek-V2-Chat · deepseek-ai/DeepSeek-V2.5 · deepseek-ai/deepseek-llm-67b-chat · deepseek-ai/deepseek-vl2 · deepseek-v3 · distil-whisper-large-v3-en · doubao-1-5-thinking-vision-pro-250428 · fx-flux-2-pro · gemini-2.5-pro-exp-03-25 · gemini-embedding-exp-03-07 · gemini-exp-1114 · gemini-exp-1121 · gemini-pro · gemini-pro-vision · gemma-7b-it · glm-3-turbo · glm-4 · glm-4-flash · glm-4-plus · glm-4.5-airx · glm-4v · glm-4v-plus · google-gemma-3-12b-it · google-gemma-3-27b-it · google-gemma-3-4b-it · google/gemini-exp-1114 · google/gemma-2-27b-it · google/gemma-2-9b-it:free · gpt-3.5-turbo · gpt-3.5-turbo-0301 · gpt-3.5-turbo-0613 · gpt-3.5-turbo-1106 · gpt-3.5-turbo-16k · gpt-3.5-turbo-16k-0613 · gpt-3.5-turbo-instruct · gpt-4 · gpt-4-0125-preview · gpt-4-0314 · gpt-4-0613 · gpt-4-1106-preview · gpt-4-32k-0314 · gpt-4-32k-0613 · gpt-4-turbo · gpt-4-turbo-2024-04-09 · gpt-4-turbo-preview · gpt-4-vision-preview · gpt-4o-2024-05-13 · gpt-4o-mini-2024-07-18 · gpt-oss-20b · grok-2-vision-1212 · grok-vision-beta · groq-llama-3.1-8b-instant · groq-llama-3.3-70b-versatile · groq-llama-4-maverick-17b-128e-instruct · groq-llama-4-scout-17b-16e-instruct · jina-embeddings-v2-base-code · learnlm-1.5-pro-experimental · llama-3.1-405b-instruct · llama-3.1-405b-reasoning · llama-3.1-70b-versatile · llama-3.1-8b-instant · llama-3.1-sonar-small-128k-online · llama-3.2-11b-vision-preview · llama-3.2-1b-preview · llama-3.2-3b-preview · llama-3.2-90b-vision-preview · llama2-70b-4096 · llama2-70b-40960 · llama2-7b-2048 · llama3-70b-8192 · llama3-8b-8192 · llama3-groq-70b-8192-tool-use-preview · llama3-groq-8b-8192-tool-use-preview · meta-llama/Llama-3.2-90B-Vision-Instruct · meta-llama/llama-3.1-405b-instruct:free · meta-llama/llama-3.1-70b-instruct:free · meta-llama/llama-3.1-8b-instruct:free · meta-llama/llama-3.2-11b-vision-instruct:free · meta-llama/llama-3.2-3b-instruct:free · meta/llama-3.1-405b-instruct · meta/llama3-8B-chat · mistralai/mistral-7b-instruct:free · moonshot-kimi-k2.5 · moonshot-v1-128k · moonshot-v1-128k-vision-preview · moonshot-v1-32k · moonshot-v1-32k-vision-preview · moonshot-v1-8k · moonshot-v1-8k-vision-preview · nvidia/Llama-3_1-Nemotron-Ultra-253B-v1 · o1-mini-2024-09-12 · omni-moderation-latest · qwen-flash · qwen-flash-2025-07-28 · qwen-long · qwen-max · qwen-max-longcontext · qwen-plus · qwen-turbo · qwen-turbo-2024-11-01 · qwen2.5-14b-instruct · qwen2.5-32b-instruct · qwen2.5-3b-instruct · qwen2.5-72b-instruct · qwen2.5-7b-instruct · qwen2.5-coder-1.5b-instruct · qwen2.5-coder-7b-instruct · qwen2.5-math-1.5b-instruct · qwen2.5-math-72b-instruct · qwen2.5-math-7b-instruct · step-2-16k · text-ada-001 · text-babbage-001 · text-curie-001 · text-davinci-002 · text-davinci-003 · text-davinci-edit-001 · text-embedding-3-large · text-embedding-3-small · text-embedding-ada-002 · text-embedding-v1 · text-moderation-007 · text-moderation-latest · text-moderation-stable · text-search-ada-doc-001 · tts-1 · tts-1-1106 · tts-1-hd · tts-1-hd-1106 · whisper-1 · whisper-large-v3 · whisper-large-v3-turbo · yi-large · yi-large-rag · yi-large-turbo · yi-lightning · yi-medium · yi-vl-plus · deepseek-r1-distill-qianfan-llama-8b · doubao-1-5-pro-256k-250115 · doubao-1-5-pro-32k-250115 · gpt-4o-2024-08-06-global · gpt-4o-mini-global · meta-llama-3-70b · meta-llama-3-8b · o3-global · o3-mini-global · o3-pro-global · qianfan-chinese-llama-2-13b · qianfan-llama-vl-8b