Models

Фильтр
Всего: 857 моделей
API данных модели
Типы
Все
Text
Image
Speech
Video
Transcription
Embeddings
Rerank
OCR
Теги
Все
Featured
Coding
Free
Discount
Разработчик
ВсеOpenAIAnthropicGoogleGrokQwenDeepSeekZ.AIByteDanceLlamaAI21MicrosoftCohereMistralYiMoonshot AIStepFunNvidiaMinimaxPerplexityBaichuanIdeogramJina AIStable diffusionHunyuanBaiduFluxMeituanInclusionAIBAAIXiaomiKLingMetaPoolsideAgnesLiquidDots Studio
Refine
Newest
Input Modalities
Сбросить
autoauto:balancedauto:quality_firstauto:latency_critical

Собственный автоматический маршрутизатор моделей AIHubMix. Не каждому запросу нужна самая мощная и дорогая модель — шлюз объединяет внутренние тесты и публичные рейтинги, охватывает ряд современных моделей разного уровня и автоматически маршрутизирует каждый запрос с помощью обученной нами небольшой модели. Установите model в auto — и запросы будут распределяться по подходящим моделям в зависимости от содержания, снижая затраты. Сейчас в бете; мы продолжаем обновлять и улучшать сервис.

Оплата по обычной цене фактически выбранной модели — сама маршрутизация бесплатна.Подробнее об умной маршрутизации
icon
New
скидка до 10%·00:00–23:59 UTCCopy ID
Ввод:$ 1.1268$ 1.01412 /M
Вывод:$ 3.9438$ 3.54942 /M
Контекст:1M
Задержка:5.006 s
Пропускная способность:29 tok/s

GLM-5.3 is Z.AI’s coding and agentic reasoning model, built for complex software engineering, long-running agent tasks, vulnerability analysis, and other demanding workloads. Building on GLM-5.2, it incorporates further post-training improvements to deliver stronger coding performance, better task execution, and greater token efficiency. We currently offer the production-ready GLM-5.3 API with unlimited concurrency, making it well suited for high-throughput workloads, coding agents, and large-scale automation. For a limited time, GLM-5.3 is available at 10% off.

Ввод:$ 0.06 /M
Вывод:$ 0.21999996 /M
Контекст:-
Задержка:-
Пропускная способность:-

GLM-5.3 is Z.ai’s reasoning model for coding and agentic workflows, designed for complex software engineering, long-running agents, and vulnerability analysis. It uses the same base model as GLM-5.2, with scaled post-training improving coding, task execution, and token efficiency. This model is a limited-time preview version of GLM-5.3, intended for testing and evaluation only. Service stability is not guaranteed, and we do not recommend using it in production environments. We’re waiting for the official commercial API release and will integrate it as soon as official support becomes available.

  • Ввод: $ 0.75 /M
  • Вывод: $ 3.75 /M
  • Web Search: $0.014/request
  • Cache Storage: $1/h/M tokens
  • Input Audio: $1/M tokens
  • Input Video: $1/M tokens

Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web development, and knowledge work. It supports a 1M-token context window and adjustable thinking levels. Compared with Gemini 3.6 Flash, it improves coding, tool use, multi-step planning, and instruction following.

Ввод:$ 0 /M
Вывод:$ 0 /M
Контекст:128K
Задержка:-
Пропускная способность:-

GLM 5.2 is a large-scale reasoning model developed by Z Ai that supports text input and output. Featuring a 128,000-token context window, this model is designed to handle complex, data-heavy tasks. It is exceptionally well-suited for long-horizon agent workflows and project-level software engineering.

Ввод:$ 0 /M
Вывод:$ 0 /M
Контекст:512K
Задержка:-
Пропускная способность:-

Dots3-Note Preview is an open-weight mixture-of-experts model developed by Dots Studio, featuring 16B active parameters out of 280B total. As the lightest model in the Dots 3 family, it is designed for efficient performance while supporting an expansive context length of 512,000 tokens. This preview version provides an accessible way to experience the capabilities of the Dots 3 architecture.

  • Ввод: $ 0 /M
  • Вывод: $ 0 /M
  • Web Search: $0.014/request
  • Cache Storage: $1/h/M tokens
  • Input Audio: $2/M tokens
  • Input Video: $1/M tokens

Gemini 3.7 Flash free version: fFree model resources are limited and provided only for trial use; stability cannot be guaranteed, and you may encounter 429 errors during use. If you need to use it in a production environment and require unlimited concurrency with absolute stability, please choose the official version: gemini-3.7-flash

icon
скидка до 30%·00:00–13:59 UTCCopy ID
Ввод:$ 1.1268$ 0.78876 /M
Вывод:$ 3.9438$ 2.76066 /M
Контекст:1M
Задержка:0.936 s
Пропускная способность:41 tok/s

GLM-5.2 is Z.ai’s flagship model for the era of long-horizon tasks. With a truly usable 1M-token context window, it can handle project-level engineering context, execute long-running tasks more reliably, follow engineering standards more consistently, and complete the full development workflow from requirements to multi-platform deployment in a single task.

Ввод:$ 0.142 /M
Вывод:$ 0.284 /M
Контекст:1M
Задержка:1.011 s
Пропускная способность:98 tok/s

DeepSeek-V4-Flash-0731(deepseek-v4-flash-0731) is an open-source MoE large language model developed by the Chinese AI company DeepSeek, with support for a million-token context window. It is designed for coding, complex reasoning, tool use, agentic workflows, and long-document processing. Its advantages include strong performance with fewer active parameters and improved efficiency through DSpark speculative decoding. Compared with DeepSeek V4-Flash Preview, it offers significantly stronger coding and agent capabilities, while outperforming DeepSeek V4-Pro Preview on several benchmarks with fewer active parameters.

Ввод:$ 0.6918 /M
Вывод:$ 2.0754 /M
Контекст:1M
Задержка:1.682 s
Пропускная способность:60 tok/s

DeepSeek V4 Pro 0813 is DeepSeek’s high-performance general-purpose reasoning and agent model, designed for complex reasoning, coding, long-document analysis, and agentic workflows. It supports thinking and non-thinking modes, a 1M-token context window, up to 384K output, tool calling, and the Responses API. Compared with V4 Flash 0731, Pro prioritizes capability on complex tasks, while Flash focuses on speed, cost efficiency, and high concurrency.

icon
Copy ID
  • Ввод: $ 2 /M
  • Вывод: $ 6 /M

Grok 4.6 is xAI’s (SpaceXAI) flagship multimodal reasoning model for coding, long-running agents, knowledge work, and interactive application development. It supports image understanding, a 500K context window, tool calling, and structured outputs. Compared with Grok 4.5, it offers stronger multi-step execution, self-verification, coding, and visual project generation.

Ввод:$ 2 /M
Вывод:$ 8 /M
Контекст:256K
Задержка:6.669 s
Пропускная способность:42 tok/s

MAI-Thinking-1 is Microsoft’s first inference model in the MAI series, built for enterprise-scale workloads. With excellent reasoning, mathematical, and general intelligence capabilities, combined with superior cost-effectiveness, it makes high-throughput, 24/7 AI workloads economically viable.

icon
скидка до 50%·00:00–23:59 UTCCopy ID
  • Ввод: $ 5$ 2.5 /M
  • Вывод: $ 30$ 15 /M
  • Web Search: $0.01/request

GPT-5.6 Sol (limited-time 50% off) is OpenAI’s frontier reasoning model for complex coding, professional knowledge work, deep research, and long-running agents. It supports a roughly 1.05M-token context window, image understanding, and extensive tool use. Compared with Terra and Luna, Sol prioritizes capability and reliability on demanding tasks.

  • Ввод: $ 2 /M
  • Вывод: $ 0 /M

Seedance 2.5 is ByteDance Seed’s next-generation unified multimodal audio-video generation model, designed for filmmaking, advertising, education, simulation, and long-form content creation. It generates up to 30 seconds of synchronized audio and video in one pass and supports multi-round extensions. Compared with Seedance 2.0, it offers stronger storytelling, smoother transitions, richer reference support, improved realism, and more precise editing.

  • Ввод: $ 0.2 /M
  • Вывод: $ 1.2 /M
  • Web Search: $0.01/request

GPT-5.6 Luna is designed for cost-sensitive, high-volume workloads. It roughly corresponds to the nano model tier used in earlier GPT-5 families.

  • Ввод: $ 5 /M
  • Вывод: $ 30 /M
  • Web Search: $0.01/request

GPT‑5.6 Sol sets a new standard for both intelligence and efficiency, achieving state-of-the-art results across coding, knowledge work, cybersecurity, and science while outperforming previous and competing frontier models with fewer tokens and at lower estimated cost.

  • Ввод: $ 2 /M
  • Вывод: $ 12 /M
  • Web Search: $0.01/request

GPT-5.6 Terra is designed for workloads that balance intelligence and cost. It roughly corresponds to the mini model tier used in earlier GPT-5 families.

Ввод:$ 0.03 /M
Вывод:$ 0.15 /M
Контекст:512K
Задержка:3.931 s
Пропускная способность:43 tok/s

Agnes 2.5 Flash is Agnes AI’s fast and efficient language model, an upgraded, fully available model based on Agnes 2.0 Flash. It continues to use an OpenAI-compatible Chat Completions interface and has been optimized for coding tasks, agent workflows, tool calls, multi-turn dialogue, reasoning, and image understanding.

Ввод:$ 0.45 /M
Вывод:$ 0.9 /M
Контекст:1M
Задержка:18.56 s
Пропускная способность:42 tok/s

Agnes 2.5 Pro is Agnes AI’s paid inference model and the commercially stable version of the Agnes 2.5 Pro Alpha ranking model, suitable for advanced coding, scientific reasoning, long-context analysis, agent workflows, and multimodal understanding. The model is accessed via an OpenAI-compatible Chat Completions API.

Ввод:$ 0.45 /M
Вывод:$ 0.9 /M
Контекст:1M
Задержка:4.931 s
Пропускная способность:11 tok/s

Agnes 2.5 Pro Alpha is Agnes AI’s paid inference model, suitable for advanced coding, scientific reasoning, long-context analysis, agent workflows, and multimodal understanding. The model is accessed via an OpenAI-compatible Chat Completions API.

Scroll to load more

Popular models

Qwen3.8 Max Preview · Kimi K3 · Qwen3.8 Max · Qwen3.7 Flash · GLM 5.2 · Grok 4.5 · Claude Opus 5 · Claude Sonnet 5 · GPT 5.6 Luna · Gemini 3.6 Flash · DeepSeek V4 Flash · GPT 5.5 · Gemini 3.1 Pro Preview

Browse by provider

OpenAI (133) · Anthropic (28) · Google (83) · Grok (26) · Qwen (141) · DeepSeek (36) · Z.AI (66) · ByteDance (46) · Llama (50) · AI21 (2) · Microsoft (14) · Cohere (19) · Mistral (10) · Yi (6) · Moonshot AI (34) · StepFun (4) · Nvidia (17) · Minimax (31) · Perplexity (4) · Baichuan (5) · Ideogram (9) · Jina AI (13) · Stable diffusion (1) · Hunyuan (5) · Baidu (23) · Flux (5) · Meituan (2) · Xiaomi (14) · InclusionAI (7) · BAAI (5) · Agnes (4) · Meta (3) · KLing (2) · Poolside (2) · Dots Studio (1) · Liquid (1)

51 free models — no credit card required →

All models (852)

auto · glm-5.3 · coding-glm-5.3 · gemini-3.7-flash · glm-5.2-free · dots-3-note-preview-free · gemini-3.7-flash-free · glm-5.2 · deepseek-v4-flash-0731 · deepseek-v4-pro-0813 · grok-4.6 · mai-thinking-1 · gpt-5.6-sol-disc · doubao-seedance-2-5-260628 · gpt-5.6-luna · gpt-5.6-sol · gpt-5.6-terra · agnes-2.5-flash · agnes-2.5-pro · agnes-2.5-pro-alpha · agnes-image-2.1-flash · grok-4.5 · lfm-2.5-2.6b-free · minimax-h3 · muse-glimmer-30b · nemotron-lightning-3.5-30b-a3b · qwen3.8-2.4t-a95b · claude-opus-5 · gemini-3.6-flash · ling-3.0-tiny-free · nemotron-3.5-lightning-free · qwen-image-3.0 · qwen-image-3.0-pro · qwen3.8-max · claude-sonnet-5 · kimi-k3 · ling-3.0-flash-free · muse-spark-1.2 · qwen3.8-max-preview · gemini-3.1-flash-lite-image · gemini-3.5-flash-lite · gemini-3.5-flash-lite-free · gemini-3.6-flash-free · glm-5.2-fast-preview · muse-spark-1.1 · qwen-audio-3.0-tts-flash · qwen-audio-3.0-tts-plus · claude-fable-5 · jina-reranker-v3.5 · claude-opus-4-8 · hy3 · doubao-seed-2-1-pro · doubao-seed-2-1-turbo · mai-image-2.5-pro · gemini-3.5-flash · grok-build-0.1 · mai-image-2.5 · mai-image-2.5-flash · coding-kimi-k3 · happyhorse-1.1-i2v · happyhorse-1.1-r2v · happyhorse-1.1-t2v · coding-glm-5.2-free · coding-kimi-k3-free · gemini-3.1-flash-image · gpt-oss-20b-free · kimi-k2.7-code · kimi-k2.7-code-highspeed · gemini-3-pro-image · gpt-4o-transcribe-diarize · gpt-audio-1.5 · hy-3d-3.1 · kling-v3-omni · kling-video-o1 · longcat-2.0 · nemotron-nano-9b-v2-free · hy3-preview · minimax-m3 · nemotron-nano-12b-v2-vl-free · qwen3.7-flash · qwen3.7-plus · step-3.7-flash · claude-opus-4-8-think · nemotron-3-super-120b-a12b-free · nemotron-3-nano-omni-30b-a3b-reasoning-free · nemotron-3-ultra-550b-a55b-free · qwen3.7-max · gpt-image-2 · nemotron-3.5-content-safety-free · coding-glm-5.2 · ernie-5.1 · gemini-3.1-flash-lite · gemini-3.1-flash-lite-nothink · grok-4.3 · happyhorse-1.0-i2v · happyhorse-1.0-r2v · happyhorse-1.0-t2v · happyhorse-1.0-video-edit · north-mini-code-free · gpt-5.5 · gpt-5.5-pro · laguna-xs-2.1-free · deepseek-v4-flash · deepseek-v4-pro · gemma-4-31b-it-free · command-a-plus-05-2026 · doubao-seedream-5.0-pro · ernie-5.0 · kimi-k2.6 · laguna-s-2.1-free · qwen3.6-max-preview · xiaomi-mimo-v2.5 · xiaomi-mimo-v2.5-pro · claude-opus-4-7 · claude-opus-4-7-think · gpt-chat-latest · nemotron-3-nano-30b-a3b-free · qwen3.6-27b · qwen3.6-35b-a3b · qwen3.6-flash · cohere-rerank-v4.0-fast · cohere-rerank-v4.0-pro · gemma-4-26b-a4b-it-free · grok-4-20-non-reasoning · grok-4-20-reasoning · qwen-image-2.0 · qwen-image-2.0-pro · coding-minimax-m3-free · doubao-seedance-2-0-260128 · doubao-seedance-2-0-fast-260128 · doubao-seedance-2-0-mini-260615 · glm-5.1 · glm-image · qwen3.6-plus · wan2.7-i2v · wan2.7-r2v · wan2.7-t2v · wan2.7-videoedit · cc-k2.6-code-preview · gemma-4-26b-a4b-it · gemma-4-31b-it · gpt-5.4 · wan2.7-image · wan2.7-image-pro · claude-sonnet-4-6 · coding-xiaomi-mimo-v2.5 · coding-xiaomi-mimo-v2.5-pro · doubao-seed-2-0-lite-260428 · doubao-seed-2-0-mini-260428 · gemini-3.1-flash-image-preview · gemini-3.1-pro-preview · gemini-3.1-pro-preview-customtools · gemini-3.1-pro-preview-search · gpt-5.4-mini · gpt-5.4-nano · gpt-5.5-free · qwen3.5-plus · claude-sonnet-4-6-think · coding-xiaomi-mimo-v2-omni · coding-xiaomi-mimo-v2-pro · gpt-5.3-chat-latest · gpt-5.3-codex · gpt-image-2-free · qwen3.5-122b-a10b · qwen3.5-27b · qwen3.5-35b-a3b · qwen3.5-397b-a17b · qwen3.5-flash · coding-glm-5.1 · doubao-seed-2-0-pro · gpt-5.4-high · gpt-5.4-low · gpt-5.4-pro · qwen3-coder-next · xiaomi-mimo-v2-omni-free · xiaomi-mimo-v2-pro-free · xiaomi-mimo-v2.5-free · xiaomi-mimo-v2.5-pro-free · claude-opus-4-6 · coding-glm-5.1-free · coding-minimax-m2.7-free · glm-5 · glm-5v-turbo · minimax-m2.7 · claude-opus-4-6-think · coding-glm-5-free · coding-glm-5-turbo-free · coding-minimax-m2.5-free · doubao-seed-2-0-code-preview · doubao-seed-2-0-lite-260215 · doubao-seed-2-0-mini · gemini-3-flash-preview · gemini-3-flash-preview-search · glm-5-turbo · cc-glm-5.1 · claude-opus-4-5 · claude-opus-4-5-think · embed-v-4-0 · ernie-image-turbo · gemini-3.1-flash-image-preview-free · mimo-v2-omni · mimo-v2-pro · cohere-command-a · gemini-3-flash-preview-free · cc-minimax-m3 · coding-minimax-m3 · gpt-4.1-free · gpt-4.1-mini-free · gpt-4.1-nano-free · gpt-4o-free · coding-glm-5 · coding-glm-5-turbo · glm-4.7 · veo-3.1-lite-generate-preview · glm-4.7-flash-free · coding-glm-4.7-free · doubao-seedance-1-5-pro-251215 · doubao-seedance-1-0-pro-250528 · doubao-seedance-1-0-pro-fast-251015 · gemini-3-pro-image-preview · gemini-embedding-2 · deepinfra-gemma-4-26b-a4b-it · gpt-5.2-codex · doubao-seedream-5.0-lite · gpt-image-1.5 · gpt-5.2 · gpt-5.2-chat-latest · gpt-5.2-high · gpt-5.2-low · gpt-5.2-pro · gpt-5.1 · gpt-5.1-codex-max · doubao-seed-1-8 · gpt-5.1-chat-latest · gpt-5.1-codex · gpt-5.1-codex-mini · claude-haiku-4-5 · claude-sonnet-4-5 · claude-sonnet-4-5-think · grok-4.20-multi-agent-0309 · mistral-large-3 · cc-glm-5 · cc-glm-5-turbo · cloudflare-glm-5.2 · gemini-2.5-flash-image · grok-4-1-fast-non-reasoning · grok-4-1-fast-reasoning · grok-code-fast-1 · k2.6-code-preview-free · mimo-v2-flash · musesteamer-air-image · qwen3.6-plus-preview-free · zai-glm-5-turbo · gpt-5 · deepseek-v3.2 · deepseek-v3.2-think · gpt-5-codex · DeepSeek-V3.1-Terminus · DeepSeek-V3.1-Think · gpt-5-pro · gpt-5-mini · gpt-5-nano · gpt-5-chat-latest · claude-opus-4-1 · o3-deep-research · kimi-k2.5 · qwen3-max-2026-01-23 · qwen3-vl-flash · qwen3-vl-flash-2026-01-22 · qwen3-vl-plus · cc-minimax-m2.7 · cc-minimax-m2.7-highspeed · minimax-m2.5 · minimax-m2.5-highspeed · mm-minimax-m2.7-highspeed · coding-minimax-m2.7 · coding-minimax-m2.7-highspeed · cc-minimax-m2.5 · cc-minimax-m2.5-highspeed · coding-minimax-m2.5 · coding-minimax-m2.5-highspeed · doubao-seedream-4-5 · sora-2 · sora-2-pro · cc-glm-4.7 · cc-minimax-m2.1 · coding-glm-4.7 · coding-minimax-m2.1 · coding-minimax-m2.1-free · gpt-4o-audio-preview · gpt-4o-mini-audio-preview · minimax-m2.1 · o3 · wan2.6-i2v · wan2.6-t2v · cc-glm-4.6 · coding-glm-4.6 · coding-glm-4.6-free · coding-minimax-m2 · coding-minimax-m2-free · flux-2-flex · flux-2-pro · gemini-2.5-pro · glm-4.6 · glm-4.6v · glm-ocr · kimi-for-coding-free · o3-pro · qianfan-ocr · qianfan-ocr-fast · step-3.5-flash · wan2.2-i2v-plus · wan2.5-i2v-preview · wan2.5-t2v-preview · gemini-2.5-pro-search · kimi-k2-thinking · gemini-2.5-flash · gemini-2.5-flash-preview-09-2025 · glm-4.5v · gemini-2.5-flash-lite · gemini-2.5-flash-lite-nothink · gemini-2.5-flash-lite-preview-09-2025 · gemini-2.5-flash-lite-preview-09-2025-nothink · gemini-2.5-flash-nothink · gemini-2.5-flash-search · gemini-2.5-flash-preview-05-20-nothink · gemini-2.5-flash-preview-05-20-search · DeepSeek-V3-Fast · imagen-4.0 · imagen-4.0-fast-generate-001 · imagen-4.0-generate-001 · imagen-4.0-ultra-generate-001 · imagen-4.0-ultra · gpt-image-1 · gpt-image-1-mini · o4-mini · DeepSeek-OCR · alicloud-kimi-k2-instruct · deepseek-ocr · ernie-5.0-thinking-exp · flux-kontext-max · gemini-2.5-flash-image-preview · glm-4.5 · gpt-4.1 · grok-4 · grok-4-fast-non-reasoning · grok-4-fast-reasoning · kimi-k2-0711 · kimi-k2-instruct · kimi-k2-turbo-preview · paddleocr-vl-0.9b · pp-structurev3 · qwen3-vl-235b-a22b-instruct · qwen3-vl-235b-a22b-thinking · qwen3-vl-30b-a3b-instruct · qwen3-vl-30b-a3b-thinking · veo-3.0-generate-preview · veo-3.1-fast-generate-preview · veo-3.1-generate-preview · aihubmix-router · gpt-4.1-mini · gpt-4.1-nano · gemini-2.5-pro-preview-05-06 · gemini-2.5-pro-preview-03-25 · gemini-2.5-pro-preview-05-06-search · gemini-2.5-pro-preview-03-25-search · qwen3-max-preview · qwen3-max · qwen3-next-80b-a3b-instruct · qwen3-next-80b-a3b-thinking · qwen3-235b-a22b-instruct-2507 · qwen3-235b-a22b-thinking-2507 · qwen3-coder-30b-a3b-instruct · qwen3-coder-480b-a35b-instruct · DeepSeek-V3 · LongCat-Flash-Chat · gemini-2.5-pro-preview-06-05-search · jina-embeddings-v5-text-nano · jina-embeddings-v5-text-small · qwen3-235b-a22b · qwen3-coder-flash · qwen3-coder-plus · qwen3-coder-plus-2025-07-22 · Qwen2.5-VL-72B-Instruct · ernie-5.0-thinking-preview · inclusionAI/Ling-1T · inclusionAI/Ring-1T · bce-reranker-base · codex-mini-latest · doubao-seedream-4-0 · embedding-v1 · ernie-4.5-turbo-latest · glm-4.5-x · gme-qwen2-vl-2b-instruct · gte-rerank-v2 · inclusionAI/Ling-flash-2.0 · inclusionAI/Ling-mini-2.0 · inclusionAI/Ring-flash-2.0 · jina-deepsearch-v1 · jina-embeddings-v4 · jina-reranker-v3 · llama-4-maverick · llama-4-scout · qwen-image · qwen-image-edit · qwen-image-max · qwen-mt-plus · qwen-mt-turbo · qwen3-embedding-0.6b · qwen3-embedding-4b · qwen3-embedding-8b · qwen3-reranker-0.6b · qwen3-reranker-4b · qwen3-reranker-8b · tao-8k · jina-clip-v2 · jina-reranker-m0 · jina-colbert-v2 · DeepSeek-R1 · gpt-4o-search-preview · gpt-4o-mini-search-preview · jina-embeddings-v3 · claude-3-7-sonnet · ernie-4.5 · ernie-4.5-turbo-vl · mimo-v2-flash-free · FLUX-1.1-pro · o3-mini · doubao-seed-1-6 · doubao-seed-1-6-flash · doubao-seed-1-6-lite · doubao-seed-1-6-thinking · qwen3-30b-a3b-instruct-2507 · qwen3-30b-a3b-thinking-2507 · Qwen2-VL-72B-Instruct · Qwen2-VL-7B-Instruct · cc-kimi-for-coding · gemini-embedding-001 · gpt-oss-120b · qwen-3-235b-a22b-thinking-2507 · Qwen/Qwen3-30B-A3B · Qwen/Qwen3-32B · qwen3-32b · Qwen/Qwen3-14B · Qwen/Qwen3-8B · embedding-2 · embedding-3 · gemini-2.5-pro-preview-06-05 · Qwen/Qwen2.5-VL-72B-Instruct · o1 · o1-pro · ByteDance-Seed/Seed-OSS-36B-Instruct · doubao-seed-1-6-250615 · doubao-seed-1-6-flash-250615 · doubao-seed-1-6-thinking-250615 · doubao-seed-1-6-vision-250815 · Doubao-1.5-thinking-pro · cc-minimax-m2 · deepseek-ai/DeepSeek-Prover-V2-671B · gemini-2.5-flash-preview-tts · gemini-2.5-pro-preview-tts · gemma-3-12b-it · gemma-3-27b-it · gemma-3-4b-it · gemma-3n-e4b-it · gemma-3-1b-it · deepseek-r1-distill-llama-70b · gpt-4o-mini-tts · tngtech/DeepSeek-R1T-Chimera · veo-2.0-generate-001 · o1-preview · o1-mini · gpt-4o-2024-11-20 · gpt-4o · gpt-4o-mini · AiHubmix-mistral-medium · ERNIE-X1.1-Preview · Qwen/QwQ-32B · chutesai/Mistral-Small-3.1-24B-Instruct-2503 · ernie-x1.1-preview · minimax-m2 · MiniMaxAI/MiniMax-M1-80k · Qwen/Qwen2.5-VL-32B-Instruct · baidu/ERNIE-4.5-300B-A47B · bge-large-en · bge-large-zh · codestral-latest · ernie-4.5-0.3b · ernie-4.5-turbo-128k-preview · ernie-x1-turbo · kat-dev · llama-3.3-70b · moonshotai/Kimi-Dev-72B · moonshotai/Moonlight-16B-A3B-Instruct · nvidia-nemotron-3-super-120b-a12b · o1-global · qianfan-qi-vl · qwen2.5-vl-72b-instruct · tencent/Hunyuan-A13B-Instruct · unsloth/gemma-3-27b-it · gemini-exp-1206 · gpt-4o-zh · qwen-qwq-32b · unsloth/gemma-3-12b-it · qwen-max-0125 · BAAI/bge-large-en-v1.5 · BAAI/bge-large-zh-v1.5 · BAAI/bge-reranker-v2-m3 · tencent/Hunyuan-MT-7B · V3 · V_2 · V_2_TURBO · V_2A · V_2A_TURBO · V_1 · V_1_TURBO · doubao-embedding-large-text-240915 · kimi-thinking-preview · gpt-4o-2024-08-06 · qwen-plus-2025-07-28 · qwen-plus-latest · sonar · stepfun-ai/step3 · text-embedding-v4 · AiHubmix-Phi-4-mini-reasoning · qwen-turbo-latest · aihub-Phi-4-multimodal-instruct · qwen3-30b-a3b · aihub-Phi-4-mini-instruct · grok-3 · aihub-Phi-4 · claude-3-opus-20240229 · dall-e-3 · doubao-embedding-text-240715 · grok-3-beta · qwen3-14b · grok-3-fast · qwen3-8b · deepseek-ai/DeepSeek-R1-Zero · grok-3-fast-beta · grok-3-mini · qwen3-4b · grok-3-mini-beta · qwen3-1.7b · qwen3-0.6b · alicloud-glm-5 · command-a-03-2025 · grok-3-mini-fast-beta · qwen-3-32b · qwen-turbo-2025-04-28 · qwen-plus-2025-04-28 · THUDM/GLM-Z1-32B-0414 · THUDM/GLM-4.1V-9B-Thinking · text-embedding-004 · THUDM/GLM-4-32B-0414 · THUDM/GLM-Z1-9B-0414 · THUDM/GLM-4-9B-0414 · cc-doubao-seed-code-preview-latest · doubao-seed-code-preview-latest · deepseek-ai/Janus-Pro-7B · glm-zero-preview · qwen-3-235b-a22b-instruct-2507 · coding-glm-4.5-air · deepinfra-nvidia-nemotron-3-nano-30b-a3b2 · glm-4.5-air · gpt-4-32k · nvidia-llama-3.1-nemotron-70b-instruct · nvidia-llama-3.3-nemotron-super-49b-v1.5 · nvidia-nemotron-3-nano-30b-a3b · nvidia-nemotron-nano-12b-v2-vl · nvidia-nemotron-nano-9b-v2 · o1-preview-2024-09-12 · Qwen/QVQ-72B-Preview · Qwen/QwQ-32B-Preview · llama-3.1-sonar-huge-128k-online · aihubmix-Mistral-Large-2411 · llama-3.1-sonar-large-128k-online · aihubmix-Mistral-large-2407 · grok-2-1212 · llama-3.1-70b · wan2.6-t2i · DESCRIBE · UPSCALE · bai-qwen3-vl-235b-a22b-instruct · cc-MiniMax-M2 · cc-deepseek-v3 · cc-deepseek-v3.1 · cc-ernie-4.5-300b-a47b · cc-kimi-dev-72b · cc-kimi-k2-instruct · cc-kimi-k2-instruct-0905 · cc-kimi-k2-thinking · computer-use-preview · gpt-image-test · grok-4.20-beta-0309-non-reasoning · grok-4.20-beta-0309-reasoning · grok-4.20-multi-agent-beta-0309 · jina-reader · jina-search · llama3.1-8b · o1-2024-12-17 · sf-kimi-k2-thinking · Baichuan3-Turbo · Baichuan3-Turbo-128k · Baichuan4 · Baichuan4-Air · Baichuan4-Turbo · DeepSeek-v3 · Doubao-1.5-lite-32k · Doubao-1.5-pro-256k · Doubao-1.5-pro-32k · Doubao-1.5-vision-pro-32k · Doubao-lite-128k · Doubao-lite-32k · Doubao-lite-4k · Doubao-pro-128k · Doubao-pro-256k · Doubao-pro-32k · Doubao-pro-4k · GPT-OSS-20B · Gryphe/MythoMax-L2-13b · MiniMax-Text-01 · Mistral-large-2407 · Qwen/Qwen2-1.5B-Instruct · Qwen/Qwen2-57B-A14B-Instruct · Qwen/Qwen2-72B-Instruct · Qwen/Qwen2-7B-Instruct · Qwen/Qwen2.5-32B-Instruct · Qwen/Qwen2.5-72B-Instruct · Qwen/Qwen2.5-72B-Instruct-128K · Qwen/Qwen2.5-7B-Instruct · Qwen/Qwen2.5-Coder-32B-Instruct · Qwen3-235B-A22B-Thinking-2507 · Stable-Diffusion-3-5-Large · WizardLM/WizardCoder-Python-34B-V1.0 · ahm-Phi-3-5-MoE-instruct · ahm-Phi-3-5-mini-instruct · ahm-Phi-3-5-vision-instruct · ahm-Phi-3-medium-128k · ahm-Phi-3-medium-4k · ahm-Phi-3-small-128k · aihubmix-Codestral-2501 · aihubmix-Cohere-command-r · aihubmix-Jamba-1-5-Large · aihubmix-Llama-3-1-405B-Instruct · aihubmix-Llama-3-1-70B-Instruct · aihubmix-Llama-3-1-8B-Instruct · aihubmix-Llama-3-2-11B-Vision · aihubmix-Llama-3-2-90B-Vision · aihubmix-Llama-3-70B-Instruct · aihubmix-Mistral-large · aihubmix-command-r-08-2024 · aihubmix-command-r-plus · aihubmix-command-r-plus-08-2024 · alicloud-deepseek-v3.2 · alicloud-glm-4.7 · alicloud-kimi-k2-thinking · alicloud-kimi-k2.5 · alicloud-minimax-m2.5 · anthropic-opus-4-6 · azure-deepseek-v3.2 · azure-deepseek-v3.2-speciale · azure-kimi-k2.5 · cbs-glm-4.7 · cerebras-llama-3.3-70b · chatglm_lite · chatglm_pro · chatglm_std · chatglm_turbo · claude-2 · claude-2.0 · claude-2.1 · claude-3-haiku-20240229 · claude-3-haiku-20240307 · claude-3-sonnet-20240229 · claude-instant-1 · claude-instant-1.2 · code-davinci-edit-001 · cogview-3 · cogview-3-plus · command · command-light · command-light-nightly · command-nightly · command-r · command-r-08-2024 · command-r-plus · command-r-plus-08-2024 · dall-e-2 · davinci · davinci-002 · deepinfra-llama-3.1-8b-instant · deepinfra-llama-3.3-70b-instant-turbo · deepinfra-llama-4-maverick-17b-128e-instruct · deepinfra-llama-4-scout-17b-16e-instruct · deepseek-ai/DeepSeek-Coder-V2-Instruct · deepseek-ai/DeepSeek-R1-Distill-Llama-70B · deepseek-ai/DeepSeek-R1-Distill-Llama-8B · deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B · deepseek-ai/DeepSeek-R1-Distill-Qwen-14B · deepseek-ai/DeepSeek-R1-Distill-Qwen-32B · deepseek-ai/DeepSeek-R1-Distill-Qwen-7B · deepseek-ai/DeepSeek-V2-Chat · deepseek-ai/DeepSeek-V2.5 · deepseek-ai/deepseek-llm-67b-chat · deepseek-ai/deepseek-vl2 · deepseek-v3 · distil-whisper-large-v3-en · doubao-1-5-thinking-vision-pro-250428 · fx-flux-2-pro · gemini-2.5-pro-exp-03-25 · gemini-embedding-exp-03-07 · gemini-exp-1114 · gemini-exp-1121 · gemini-pro · gemini-pro-vision · gemma-7b-it · glm-3-turbo · glm-4 · glm-4-flash · glm-4-plus · glm-4.5-airx · glm-4v · glm-4v-plus · google-gemma-3-12b-it · google-gemma-3-27b-it · google-gemma-3-4b-it · google/gemini-exp-1114 · google/gemma-2-27b-it · google/gemma-2-9b-it:free · gpt-3.5-turbo · gpt-3.5-turbo-0301 · gpt-3.5-turbo-0613 · gpt-3.5-turbo-1106 · gpt-3.5-turbo-16k · gpt-3.5-turbo-16k-0613 · gpt-3.5-turbo-instruct · gpt-4 · gpt-4-0125-preview · gpt-4-0314 · gpt-4-0613 · gpt-4-1106-preview · gpt-4-32k-0314 · gpt-4-32k-0613 · gpt-4-turbo · gpt-4-turbo-2024-04-09 · gpt-4-turbo-preview · gpt-4-vision-preview · gpt-4o-2024-05-13 · gpt-4o-mini-2024-07-18 · gpt-oss-20b · grok-2-vision-1212 · grok-vision-beta · groq-llama-3.1-8b-instant · groq-llama-3.3-70b-versatile · groq-llama-4-maverick-17b-128e-instruct · groq-llama-4-scout-17b-16e-instruct · jina-embeddings-v2-base-code · learnlm-1.5-pro-experimental · llama-3.1-405b-instruct · llama-3.1-405b-reasoning · llama-3.1-70b-versatile · llama-3.1-8b-instant · llama-3.1-sonar-small-128k-online · llama-3.2-11b-vision-preview · llama-3.2-1b-preview · llama-3.2-3b-preview · llama-3.2-90b-vision-preview · llama2-70b-4096 · llama2-70b-40960 · llama2-7b-2048 · llama3-70b-8192 · llama3-8b-8192 · llama3-groq-70b-8192-tool-use-preview · llama3-groq-8b-8192-tool-use-preview · meta-llama/Llama-3.2-90B-Vision-Instruct · meta-llama/llama-3.1-405b-instruct:free · meta-llama/llama-3.1-70b-instruct:free · meta-llama/llama-3.1-8b-instruct:free · meta-llama/llama-3.2-11b-vision-instruct:free · meta-llama/llama-3.2-3b-instruct:free · meta/llama-3.1-405b-instruct · meta/llama3-8B-chat · mistralai/mistral-7b-instruct:free · moonshot-kimi-k2.5 · moonshot-v1-128k · moonshot-v1-128k-vision-preview · moonshot-v1-32k · moonshot-v1-32k-vision-preview · moonshot-v1-8k · moonshot-v1-8k-vision-preview · nvidia/Llama-3_1-Nemotron-Ultra-253B-v1 · o1-mini-2024-09-12 · omni-moderation-latest · qwen-flash · qwen-flash-2025-07-28 · qwen-long · qwen-max · qwen-max-longcontext · qwen-plus · qwen-turbo · qwen-turbo-2024-11-01 · qwen2.5-14b-instruct · qwen2.5-32b-instruct · qwen2.5-3b-instruct · qwen2.5-72b-instruct · qwen2.5-7b-instruct · qwen2.5-coder-1.5b-instruct · qwen2.5-coder-7b-instruct · qwen2.5-math-1.5b-instruct · qwen2.5-math-72b-instruct · qwen2.5-math-7b-instruct · step-2-16k · text-ada-001 · text-babbage-001 · text-curie-001 · text-davinci-002 · text-davinci-003 · text-davinci-edit-001 · text-embedding-3-large · text-embedding-3-small · text-embedding-ada-002 · text-embedding-v1 · text-moderation-007 · text-moderation-latest · text-moderation-stable · text-search-ada-doc-001 · tts-1 · tts-1-1106 · tts-1-hd · tts-1-hd-1106 · whisper-1 · whisper-large-v3 · whisper-large-v3-turbo · yi-large · yi-large-rag · yi-large-turbo · yi-lightning · yi-medium · yi-vl-plus · deepseek-r1-distill-qianfan-llama-8b · doubao-1-5-pro-256k-250115 · doubao-1-5-pro-32k-250115 · gpt-4o-2024-08-06-global · gpt-4o-mini-global · meta-llama-3-70b · meta-llama-3-8b · o3-global · o3-mini-global · o3-pro-global · qianfan-chinese-llama-2-13b · qianfan-llama-vl-8b