Models

フィルター
合計: 870 モデル
モデルデータAPI
種類
すべて
Text
Image
Speech
Video
Transcription
Embeddings
Rerank
OCR
タグ
すべて
Featured
Coding
Free
Discount
開発者
すべてOpenAIAnthropicGoogleGrokQwenDeepSeekZ.AIByteDanceLlamaAI21MicrosoftCohereMistralYiMoonshot AIStepFunNvidiaMinimaxPerplexityBaichuanIdeogramJina AIStable diffusionHunyuanBaiduFluxMeituanInclusionAIBAAIXiaomiKLingMetaPoolsideAgnesLiquidDots StudioStealthThinkingmachinesInceptionUpstageMeta Llama
Refine
Newest
Input Modalities
リセット
autoauto:balancedauto:quality_firstauto:latency_critical

AIHubMix が自社開発したモデル自動ルーターです。すべてのリクエストに最強で最も高価なモデルが必要なわけではありません——ゲートウェイは社内評価と公開ベンチマークを総合し、能力の異なる主流モデル群をカバーし、当社が学習させた小型モデルで自動ルーティングします。リクエスト時に model を auto にするだけで、内容に応じて適切なモデルに自動割り当てされ、コストを削減できます。現在ベータ版で、継続的に更新・改善しています。

実際にルーティングされたモデルの通常価格で課金され、ルーティング自体は無料です。スマートルーティングについて
icon
New
Copy ID
  • 入力: $ 10 /M
  • 出力: $ 50 /M
  • Web Search: $0.01/request

GPT-6 Astra is OpenAI's newest and most intelligent model, with industry-leading performance in computer operations, web browsing, software engineering, scientific research, and professional work. It excels at executing multi-step workflows across code, browsers, and various professional software. Astra can achieve better results with significantly fewer output tokens, making its estimated API cost per task lower.

  • 入力: $ 0.75 /M
  • 出力: $ 3.75 /M
  • Web Search: $0.014/request
  • Cache Storage: $1/h/M tokens
  • Input Audio: $1/M tokens
  • Input Video: $1/M tokens

Gemini 3.8 Flash is Google's most intelligent Flash-series model, designed for long-running software engineering tasks, autonomous agents, and complex enterprise workflows, while retaining the Flash series' fast responsiveness and cost-effectiveness.

  • 入力: $ 11 /M
  • 出力: $ 55 /M
  • Web Search: $0.01/request

Claude-fable-5-1 is Anthropic's most capable newly released model, suitable for the most demanding reasoning and long-running agent work. Claude Fable 5.1, while keeping input and output prices the same as Claude Fable 5, reduces cache read prices to one quarter of the original and is more capable for long-running agent programming, multi-step research, and document, spreadsheet, and slide processing. (The model is extremely expensive and not recommended for casual use.)

icon
New
Standard version50% OFFCopy ID
入力:$ 0.1127$ 0.0563 /M
出力:$ 0.3944$ 0.1972 /M
コンテキスト:1M
レイテンシ:3.257 s
スループット:193 tok/s

GLM-5.3-Flash is a high-efficiency multimodal model from Z.AI. It supports a context window of roughly 1 million tokens, along with text, image, and video inputs, and includes tool-calling capabilities. It is primarily designed for coding agents, complex reasoning, and long-horizon software engineering tasks. Built on the existing GLM technology stack, the model has been further post-trained and optimized to deliver strong performance while placing greater emphasis on inference efficiency, responsiveness, and cost.The model is offered at a limited-time 50% discount; users are welcome to try it.

icon
New
Standard version10% OFFCopy ID
入力:$ 1.1268$ 1.0141 /M
出力:$ 3.9438$ 3.5494 /M
コンテキスト:1M
レイテンシ:2.027 s
スループット:30 tok/s

GLM-5.3 is Z.AI’s coding and agentic reasoning model, built for complex software engineering, long-running agent tasks, vulnerability analysis, and other demanding workloads. Building on GLM-5.2, it incorporates further post-training improvements to deliver stronger coding performance, better task execution, and greater token efficiency. We currently offer the production-ready GLM-5.3 API with unlimited concurrency, making it well suited for high-throughput workloads, coding agents, and large-scale automation. For a limited time, GLM-5.3 is available at 10% off.

icon
New
80% OFFCopy ID
入力:$ 0.2$ 0.04 /M
出力:$ 0.75$ 0.15 /M
コンテキスト:260K
レイテンシ:3.743 s
スループット:110 tok/s

Mercury 2.5 is the latest diffusion-based large language model (dLLM) released by Inception. It is the fastest inference LLM; unlike the sequential token-by-token generation approach, Mercury 2.5 can generate and optimize multiple tokens in parallel, achieving a generation speed of 1,107 tokens per second on standard GPUs. Compared to Mercury 2, its intelligence has increased by more than 10 percentage points, and its quality rivals leading cost-optimized frontier models such as GPT-5.6 Luna (Low), Gemini 3.5 Flash-Lite, and Claude Haiku 4.5.

入力:$ 1.69 /M
出力:$ 5.07 /M
コンテキスト:991K
レイテンシ:4.078 s
スループット:33 tok/s

Qwen3.8-Max-0902 (also known as qwen3.8-max-2026-09-02) is a snapshot version of Alibaba Cloud Tongyi Qianwen's qwen3.8-max. It pushes encoding depth further, enabling it to handle more complex engineering-level projects and long-term autonomous development; collaborative agent capabilities are significantly enhanced, performing more confidently in multi-tool orchestration and end-to-end delivery; visual understanding is comprehensively improved, with more sensitive and accurate chart reasoning, document parsing, and multimodal perception. Continuing the 1-million-context window, reasoning modes, and a complete tool ecosystem, it continues to evolve at a higher level of intelligence.

icon
New
Copy ID
入力:$ 0.845 /M
出力:$ 2.535 /M
コンテキスト:1M
レイテンシ:2.208 s
スループット:13 tok/s

Hy4 Preview is Tencent Hunyuan’s large language model for agents, coding, office automation, and complex tool use.Hy4 preview has a total of 770B parameters and 49B active parameters, and is primarily optimized for agent, coding, and productivity scenarios. It strengthens understanding, planning, tool invocation, and sustained execution capabilities for complex tasks. Compared with the previous generation, Hy4 preview further improves multi-step agents, code development, and productivity tasks, offering better task decomposition, context continuity, instruction following, and long-horizon execution. In coding scenarios, it further enhances code understanding, generation, modification, and the handling of complex engineering tasks; in productivity scenarios, it focuses on improving document processing, information analysis, office automation, game development, webpage generation, and cross-tool collaboration. Hy4 preview is suitable for coding agents, complex tool invocation, and various agent workflows that require multi-step planning and sustained execution, providing more reliable task completion for complex real-world business scenarios.

icon
New
Copy ID
入力:$ 1.375 /M
出力:$ 4.675 /M
コンテキスト:1M
レイテンシ:4.337 s
スループット:87 tok/s

Muse Spark 1.3 is Meta’s multimodal reasoning model, designed for long-running agentic, multi-agent, and complex coding workflows. It can maintain context and task constraints across extended workflows, reconcile conflicting information, and request clarification or confirmation when needed. Compared with Muse Spark 1.2, Muse Spark 1.3 improves efficiency across long-horizon agentic and coding tasks, with fewer unnecessary turns, tool calls, and tokens, while producing more concise outputs.

icon
New
Copy ID
入力:$ 0 /M
出力:$ 0 /M
コンテキスト:256K
レイテンシ:-
スループット:-

The Hy3 official version is honed for real-world business scenarios, using a Mixture-of-Experts (MoE) architecture with 295B total parameters and 21B activated parameters. It natively supports a 256K context window and offers multiple thinking modes: no_think (ultra-fast response), think_low (quick thinking), and think_high (deep reasoning), balancing ultra-fast responses, complex reasoning, and invocation cost. Compared with the Preview version, Hy3—based on real business feedback from Tencent Yuanbao, WorkBuddy, ima, Marvis, and others—focuses on improving the Coding Agent, long-form understanding, multi-turn context continuity, search QA, and complex task execution, performing more stably in reducing hallucinations, improving task completion, and engineering usability. It is better suited to practical scenarios such as frontend tasks, cross-file code development, long-document analysis, office automation, and multi-step Agent workflows.

入力:$ 0 /M
出力:$ 0 /M
コンテキスト:1M
レイテンシ:2.222 s
スループット:72 tok/s

MiniMax-M3 is a versatile multimodal foundation model developed by MiniMax that supports text, image, and video inputs to generate text outputs. With a massive context window of 1,048,576 tokens, it is capable of processing and understanding vast amounts of information. This model is highly optimized for complex tasks, making it exceptionally well-suited for coding and long-horizon agentic workflows.

icon
New
Copy ID
入力:$ 0.1126 /M
出力:$ 0.38 /M
コンテキスト:1M
レイテンシ:2.953 s
スループット:55 tok/s

Qwen3.8 Flash is Alibaba Cloud Qwen’s flagship native vision-language model for coding, office tasks, long-context reasoning, and agent workflows. It supports a 1M-token context, 128K output, web access, and tool calling. Compared with Qwen3.7-Plus, Qwen3.8-Flash significantly reduces training and inference costs—the training overhead is only about one-ninth of the former—while offering stronger capabilities on coding and office tasks.

  • 入力: $ 0.75 /M
  • 出力: $ 3.75 /M
  • Web Search: $0.014/request
  • Cache Storage: $1/h/M tokens
  • Input Audio: $1/M tokens
  • Input Video: $1/M tokens

Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web development, and knowledge work. It supports a 1M-token context window and adjustable thinking levels. Compared with Gemini 3.6 Flash, it improves coding, tool use, multi-step planning, and instruction following.

icon
New
Copy ID
  • 入力: $ 5 /M
  • 出力: $ 5 /M
  • Input Image: $1.6/M tokens

MAI-Image-2.6 is Microsoft's latest image-generation and editing model. It can generate images from text prompts and supports image-guided editing in multiple aspect ratios.

  • 入力: $ 1.75 /M
  • 出力: $ 1.75 /M
  • Input Image: $1.4286/M tokens

MAI-Image-2.6 Flash is Microsoft’s low-latency version of the latest image model MAI-Image-2.6. It supports image generation in multiple aspect ratios and image-guided editing.

入力:$ 0 /M
出力:$ 0 /M
コンテキスト:196K
レイテンシ:-
スループット:-

Developed by Minimax, MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Featuring a generous 196,608-token context window, the model integrates advanced agentic capabilities through multi-agent workflows. It is uniquely built to actively participate in its own evolution, delivering highly adaptable and intelligent performance.

icon
New
Copy ID
入力:$ 0.03 /M
出力:$ 0.15 /M
コンテキスト:512K
レイテンシ:0.767 s
スループット:21 tok/s

Agnes 3.0 Flash is designed for real-world agent tasks and development workflows, covering the full execution chain from task understanding and planning to tool invocation and final delivery. The model focuses on improving stability, instruction-following, factual grounding, and output completeness in complex tasks, helping developers build more reliable agent applications.

icon
New
Copy ID
  • 入力: $ 2 /M
  • 出力: $ 2 /M

MiniMax H3 (minimax-h3) is a general-purpose omni-modal generation model developed by the Chinese AI company MiniMax. It can jointly understand text, images, video, and audio while generating videos of up to 15 seconds in 2K resolution with native stereo sound. It is designed for advertising, branding, e-commerce, product design, UI/UX, and gaming. Compared with Hailuo 01 and Hailuo 02, H3 evolves from specialized video generation into a unified model for multimodal creation, reference-based editing, text rendering, and motion transfer.

icon
New
Copy ID
入力:$ 0 /M
出力:$ 0 /M
コンテキスト:1M
レイテンシ:-
スループット:-

This model actually points to glm-5.3-flash; if you need to use it in production, you can directly call the model name "glm-5.3-flash".

Scroll to load more

Popular models

Qwen3.8 Max Preview · Kimi K3 · Qwen3.8 Max · Qwen3.7 Flash · GLM 5.2 · Grok 4.5 · Claude Opus 5 · Claude Sonnet 5 · GPT 5.6 Luna · Gemini 3.6 Flash · DeepSeek V4 Flash · GPT 5.5 · Gemini 3.1 Pro Preview

Browse by model author

OpenAI (134) · Anthropic (29) · Google (85) · Grok (26) · Qwen (144) · DeepSeek (37) · Z.AI (70) · ByteDance (46) · Llama (51) · AI21 (2) · Microsoft (16) · Cohere (19) · Mistral (10) · Moonshot AI (34) · StepFun (4) · Nvidia (16) · Minimax (34) · Ideogram (9) · Jina AI (13) · Stable diffusion (1) · Hunyuan (7) · Baidu (23) · Flux (5) · Meituan (2) · Xiaomi (14) · InclusionAI (7) · Agnes (5) · BAAI (5) · Meta (3) · KLing (2) · Poolside (2) · Dots Studio (1) · Inception (1) · Liquid (1) · Upstage (1)

56 free models — no credit card required →

40 models deprecating or retired — dates and replacements →

All models (874)

Auto · GPT 6 Astra · Gemini 3.8 Flash · Gemini 3.8 Flash (free) · Claude Fable 5.1 · GLM 5.3 Flash · GLM 5.3 · Mercury 2.5 Preview · Qwen3.8 Max 2026 09-02 · Coding GLM 5.3 · Coding GLM 5.3 Flash (free) · Hy4 Preview · Coding GLM 5.3 Flash · Coding GLM 5.3 (free) · Muse Spark 1.3 · Hy3 (free) · MiniMax M3 (free) · Qwen3.8 Flash · Gemini 3.7 Flash · Mai Image 2.6 · Mai Image 2.6 Flash · MiniMax M2.7 (free) · Agnes 3.0 Flash · MiniMax H3 Max · Ox Alpha · Dots 3 Note Preview (free) · Gemini 3.7 Flash (free) · GLM 5.2 · DeepSeek V4 Flash 0731 · DeepSeek V4 Flash Vision Exp · DeepSeek V4 Pro 0813 · Grok 4.6 · DeepSeek V4 Flash 0731 Fast · Mai Thinking 1 · Wan3.0 Video · Wan3.0 Video Prime · GPT 5.6 Sol Disc · Doubao Seedance 2.5 260628 · GPT 5.6 Luna · GPT 5.6 Sol · GPT 5.6 Terra · Qwen3.8 Max · Agnes 2.5 Flash · Agnes 2.5 Pro · Agnes 2.5 Pro Alpha · Agnes Image 2.1 Flash · Grok 4.5 · Lfm 2.5 2.6b (free) · MiniMax H3 · Qwen3.8 2.4t A95B · Claude Opus 5 · Gemini 3.6 Flash · Ling 3.0 Tiny (free) · Nemotron 3.5 Lightning (free) · Qwen Image 3.0 · Qwen Image 3.0 Pro · Claude Sonnet 5 · Kimi K3 · Ling 3.0 Flash (free) · Muse Spark 1.2 · Qwen3.8 Max Preview · Gemini 3.1 Flash Lite Image · Gemini 3.5 Flash Lite · Gemini 3.5 Flash Lite (free) · Gemini 3.6 Flash (free) · GLM 5.2 Fast Preview · Muse Spark 1.1 · Qwen Audio 3.0 Tts Flash · Qwen Audio 3.0 Tts Plus · Claude Fable 5 · Jina Reranker V3.5 · Claude Opus 4.8 · Hy3 · Doubao Seed 2.1 Pro · Doubao Seed 2.1 Turbo · Mai Image 2.5 Pro · Gemini 3.5 Flash · Grok Build 0.1 · Mai Image 2.5 · Mai Image 2.5 Flash · Coding Kimi K3 · Happyhorse 1.1 I2v · Happyhorse 1.1 R2v · Happyhorse 1.1 T2v · Coding GLM 5.2 (free) · Coding Kimi K3 (free) · Gemini 3.1 Flash Image · GPT Oss 20B (free) · Kimi K2.7 Code · Kimi K2.7 Code Highspeed · Gemini 3 Pro Image · GPT 4o Transcribe Diarize · GPT Audio 1.5 · Hy 3d 3.1 · Kling V3 Omni · Kling Video O1 · Longcat 2.0 · Nemotron Nano 9B V2 (free) · Hy3 Preview · MiniMax M3 · Nemotron Nano 12B V2 VL (free) · Qwen3.7 Flash · Qwen3.7 Plus · Step 3.7 Flash · Claude Opus 4.8 Thinking · Nemotron 3 Super 120B A12B (free) · Nemotron 3 Nano Omni 30B A3B (reasoning) (free) · Nemotron 3 Ultra 550B A55B (free) · Qwen3.7 Max · GPT Image 2 · Nemotron 3.5 Content Safety (free) · Coding GLM 5.2 · ERNIE 5.1 · Gemini 3.1 Flash Lite · gemini-3.1-flash-lite-nothink · Grok 4.3 · Happyhorse 1.0 I2v · Happyhorse 1.0 R2v · Happyhorse 1.0 T2v · Happyhorse 1.0 Video Edit · North Mini Code (free) · GPT 5.5 · GPT 5.5 Pro · Laguna Xs 2.1 (free) · DeepSeek V4 Flash · DeepSeek V4 Pro · Gemma 4 31B It (free) · Command A Plus 05 2026 · Doubao Seedream 5.0 Pro · ERNIE 5.0 · Kimi K2.6 · Laguna S 2.1 (free) · MiMo V2.5 · MiMo V2.5 Pro · Qwen3.6 Max Preview · Claude Opus 4.7 · Claude Opus 4.7 Thinking · GPT Chat · Nemotron 3 Nano 30B A3B (free) · Qwen3.6 27B · Qwen3.6 35B A3B · Qwen3.6 Flash · Cohere Rerank V4.0 Fast · Cohere Rerank V4.0 Pro · Gemma 4 26B A4B It (free) · grok-4-20-non-reasoning · Grok 4 20 (reasoning) · Qwen Image 2.0 · Qwen Image 2.0 Pro · Coding MiniMax M3 (free) · Doubao Seedance 2.0 260128 · Doubao Seedance 2.0 Fast 260128 · Doubao Seedance 2.0 Mini 260615 · GLM 5.1 · GLM Image · Qwen3.6 Plus · Wan2.7 I2v · Wan2.7 R2v · Wan2.7 T2v · Wan2.7 Videoedit · CC K2.6 Code Preview · Gemma 4 26B A4B It · Gemma 4 31B It · GPT 5.4 · Wan2.7 Image · Wan2.7 Image Pro · Claude Sonnet 4.6 · Coding Xiaomi Mimo V2.5 · Coding Xiaomi Mimo V2.5 Pro · Doubao Seed 2.0 Lite 260428 · Doubao Seed 2.0 Mini 260428 · Gemini 3.1 Flash Image Preview · Gemini 3.1 Pro Preview · Gemini 3.1 Pro Preview Customtools · Gemini 3.1 Pro Preview Search · GPT 5.4 Mini · GPT 5.4 Nano · GPT 5.5 (free) · Qwen3.5 Plus · Claude Sonnet 4.6 Thinking · Coding Xiaomi Mimo V2 Omni · Coding Xiaomi Mimo V2 Pro · GPT 5.3 Chat · GPT-5.3-Codex · GPT Image 2 (free) · Qwen3.5 122B A10B · Qwen3.5 27B · Qwen3.5 35B A3B · Qwen3.5 397B A17B · Qwen3.5 Flash · Coding GLM 5.1 · Doubao Seed 2.0 Pro · GPT 5.4 High · GPT 5.4 Low · GPT 5.4 Pro · Qwen3 Coder Next · Xiaomi Mimo V2 Omni (free) · Xiaomi Mimo V2 Pro (free) · Xiaomi Mimo V2.5 (free) · Xiaomi Mimo V2.5 Pro (free) · Claude Opus 4.6 · Coding GLM 5.1 (free) · Coding MiniMax M2.7 (free) · GLM 5 · GLM 5 Vision Turbo · MiniMax M2.7 · Claude Opus 4.6 Thinking · Coding GLM 5 (free) · Coding GLM 5 Turbo (free) · Coding MiniMax M2.5 (free) · Doubao Seed 2.0 Code Preview · Doubao Seed 2.0 Lite 260215 · Doubao Seed 2.0 Mini · Gemini 3 Flash Preview · Gemini 3 Flash Preview Search · GLM 5 Turbo · CC GLM 5.1 · Claude Opus 4.5 · Claude Opus 4.5 Thinking · Embed V 4.0 · ERNIE Image Turbo · Gemini 3.1 Flash Image Preview (free) · MiMo V2 Omni · MiMo V2 Pro · Cohere Command A · Gemini 3 Flash Preview (free) · CC MiniMax M3 · Coding MiniMax M3 · GPT 4.1 (free) · GPT 4.1 Mini (free) · GPT 4.1 Nano (free) · GPT 4o (free) · Coding GLM 5 · Coding GLM 5 Turbo · GLM 4.7 · Veo 3.1 Lite Generate Preview · GLM 4.7 Flash (free) · Coding GLM 4.7 (free) · Doubao Seedance 1.5 Pro 251215 · Doubao Seedance 1.0 Pro 250528 · Doubao Seedance 1.0 Pro Fast 251015 · Gemini 3 Pro Image Preview · Gemini Embedding 2 · Deepinfra Gemma 4 26B A4B It · GPT-5.2-Codex · Doubao Seedream 5.0 Lite · GPT Image 1.5 · GPT 5.2 · GPT 5.2 Chat · GPT 5.2 High · GPT 5.2 Low · GPT 5.2 Pro · GPT 5.1 · GPT-5.1-Codex Max · Doubao Seed 1.8 · GPT 5.1 Chat · GPT-5.1-Codex · GPT-5.1-Codex Mini · Claude Haiku 4.5 · Claude Sonnet 4.5 · Claude Sonnet 4.5 Thinking · Grok 4.20 Multi Agent 0309 · Mistral Large 3 · CC GLM 5 · CC GLM 5 Turbo · Cloudflare Glm 5.2 · Gemini 2.5 Flash Image · grok-4-1-fast-non-reasoning · Grok 4.1 Fast (reasoning) · Grok Code Fast 1 · K2.6 Code Preview (free) · MiMo V2 Flash · Musesteamer Air Image · Qwen3.6 Plus Preview (free) · Zai Glm 5 Turbo · GPT 5 · DeepSeek V3.2 · DeepSeek V3.2 Thinking · GPT-5-Codex · DeepSeek V3.1 Terminus · DeepSeek V3.1 Thinking · GPT 5 Pro · GPT 5 Mini · GPT 5 Nano · GPT 5 Chat · Claude Opus 4.1 · O3 Deep Research · Kimi K2.5 · Qwen3 Max 2026 01-23 · Qwen3 VL Flash · Qwen3 VL Flash 2026 01-22 · Qwen3 VL Plus · CC MiniMax M2.7 · CC MiniMax M2.7 Highspeed · MiniMax M2.5 · MiniMax M2.5 Highspeed · Mm Minimax M2.7 Highspeed · Coding MiniMax M2.7 · Coding MiniMax M2.7 Highspeed · CC MiniMax M2.5 · CC MiniMax M2.5 Highspeed · Coding MiniMax M2.5 · Coding MiniMax M2.5 Highspeed · Doubao Seedream 4.5 · Sora 2 · Sora 2 Pro · CC GLM 4.7 · CC MiniMax M2.1 · Coding GLM 4.7 · Coding MiniMax M2.1 · Coding MiniMax M2.1 (free) · GPT 4o Audio Preview · GPT 4o Mini Audio Preview · MiniMax M2.1 · O3 · Wan2.6 I2v · Wan2.6 T2v · CC GLM 4.6 · Coding GLM 4.6 · Coding GLM 4.6 (free) · Coding MiniMax M2 · Coding MiniMax M2 (free) · Flux 2 Flex · Flux 2 Pro · Gemini 2.5 Pro · GLM 4.6 · GLM 4.6 Vision · GLM Ocr · Kimi For Coding (free) · O3 Pro · Qianfan Ocr · Qianfan Ocr Fast · Step 3.5 Flash · Wan2.2 I2v Plus · Wan2.5 I2v Preview · Wan2.5 T2v Preview · Gemini 2.5 Pro Search · Kimi K2 Thinking · Gemini 2.5 Flash · Gemini 2.5 Flash Preview 09 2025 · GLM 4.5 Vision · Gemini 2.5 Flash Lite · gemini-2.5-flash-lite-nothink · Gemini 2.5 Flash Lite Preview 09 2025 · gemini-2.5-flash-lite-preview-09-2025-nothink · gemini-2.5-flash-nothink · Gemini 2.5 Flash Search · gemini-2.5-flash-preview-05-20-nothink · Gemini 2.5 Flash Preview 05-20 Search · Imagen 4.0 · Imagen 4.0 Fast Generate 001 · Imagen 4.0 Generate 001 · Imagen 4.0 Ultra Generate 001 · Imagen 4.0 Ultra · GPT Image 1 · GPT Image 1 Mini · O4 Mini · DeepSeek-OCR · Alicloud Kimi K2 Instruct · DeepSeek Ocr · ERNIE 5.0 Thinking Exp · Flux Kontext Max · Gemini 2.5 Flash Image Preview · GLM 4.5 · GPT 4.1 · Grok 4 · grok-4-fast-non-reasoning · Grok 4 Fast (reasoning) · Kimi K2 0711 · Kimi K2 Instruct · Kimi K2 Turbo Preview · Paddleocr VL 0.9b · Pp Structurev3 · Qwen3 VL 235B A22B Instruct · Qwen3 VL 235B A22B Thinking · Qwen3 VL 30B A3B Instruct · Qwen3 VL 30B A3B Thinking · Veo 3.0 Generate Preview · Veo 3.1 Fast Generate Preview · Veo 3.1 Generate Preview · AIHubMix Router · GPT 4.1 Mini · GPT 4.1 Nano · Gemini 2.5 Pro Preview 05-06 · Gemini 2.5 Pro Preview 03-25 · Gemini 2.5 Pro Preview 05-06 Search · Gemini 2.5 Pro Preview 03-25 Search · Qwen3 Max Preview · Qwen3 Max · Qwen3 Next 80B A3B Instruct · Qwen3 Next 80B A3B Thinking · Qwen3 235B A22B Instruct 2507 · Qwen3 235B A22B Thinking 2507 · Qwen3 Coder 30B A3B Instruct · Qwen3 Coder 480B A35B Instruct · DeepSeek V3 · LongCat-Flash-Chat · Gemini 2.5 Pro Preview 06-05 Search · Jina Embeddings V5 Text Nano · Jina Embeddings V5 Text Small · Qwen3 235B A22B · Qwen3 Coder Flash · Qwen3 Coder Plus · Qwen3 Coder Plus 2025 07-22 · ERNIE 5.0 Thinking Preview · inclusionAI/Ling-1T · inclusionAI/Ring-1T · Bce Reranker Base · Codex Mini · Doubao Seedream 4.0 · Embedding V1 · ERNIE 4.5 Turbo · GLM 4.5 X · Gme Qwen2 VL 2B Instruct · Gte Rerank V2 · inclusionAI/Ling-flash-2.0 · inclusionAI/Ling-mini-2.0 · inclusionAI/Ring-flash-2.0 · Jina Deepsearch V1 · Jina Embeddings V4 · Jina Reranker V3 · Llama 4 Maverick · Llama 4 Scout · Qwen Image · Qwen Image Edit · Qwen Image Max · Qwen Mt Plus · Qwen Mt Turbo · Qwen3 Embedding 0.6b · Qwen3 Embedding 4B · Qwen3 Embedding 8B · Qwen3 Reranker 0.6b · Qwen3 Reranker 4B · Qwen3 Reranker 8B · Tao 8K · Jina Clip V2 · Jina Reranker M0 · Jina Colbert V2 · GPT 4o Search Preview · GPT 4o Mini Search Preview · Jina Embeddings V3 · Claude 3.7 Sonnet · ERNIE 4.5 · ERNIE 4.5 Turbo VL · MiMo V2 Flash (free) · FLUX-1.1-pro · O3 Mini · Doubao Seed 1.6 · Doubao Seed 1.6 Flash · Doubao Seed 1.6 Lite · Doubao Seed 1.6 Thinking · Qwen3 30B A3B Instruct 2507 · Qwen3 30B A3B Thinking 2507 · Qwen2 VL 72B Instruct · Qwen2 VL 7B Instruct · CC Kimi For Coding · Gemini Embedding 001 · gpt-oss-120b · Qwen 3 235B A22B Thinking 2507 · Qwen/Qwen3-30B-A3B · Qwen/Qwen3-32B · Qwen3 32B · Qwen/Qwen3-14B · Qwen/Qwen3-8B · Embedding 2 · Embedding 3 · Gemini 2.5 Pro Preview 06-05 · Qwen/Qwen2.5-VL-72B-Instruct · O1 · O1 Pro · ByteDance-Seed/Seed-OSS-36B-Instruct · Doubao Seed 1.6 250615 · Doubao Seed 1.6 Flash 250615 · Doubao Seed 1.6 Thinking 250615 · Doubao Seed 1.6 Vision 250815 · Doubao 1.5 Thinking Pro · CC MiniMax M2 · deepseek-ai/DeepSeek-Prover-V2-671B · Gemini 2.5 Flash Preview Tts · Gemini 2.5 Pro Preview Tts · Gemma 3 12B It · Gemma 3 27B It · Gemma 3 4B It · Gemma 3n E4B It · Gemma 3 1B It · DeepSeek R1 Distill Llama 70B · GPT 4o Mini Tts · tngtech/DeepSeek-R1T-Chimera · O1 Preview · O1 Mini · GPT 4o 2024 11-20 · GPT 4o · GPT 4o Mini · AIHubMix Mistral Medium · ERNIE X1.1 Preview · Qwen/QwQ-32B · chutesai/Mistral-Small-3.1-24B-Instruct-2503 · ERNIE X1.1 Preview · MiniMax M2 · MiniMaxAI/MiniMax-M1-80k · Qwen/Qwen2.5-VL-32B-Instruct · baidu/ERNIE-4.5-300B-A47B · Bge Large En · Bge Large Zh · Codestral · ERNIE 4.5 0.3b · ERNIE 4.5 Turbo 128K Preview · ERNIE X1 Turbo · Kat Dev · Llama 3.3 70B · moonshotai/Kimi-Dev-72B · moonshotai/Moonlight-16B-A3B-Instruct · Nvidia Nemotron 3 Super 120B A12B · O1 Global · Qianfan Qi VL · Qwen2.5 VL 72B Instruct · tencent/Hunyuan-A13B-Instruct · unsloth/gemma-3-27b-it · gemini-exp-1206 · GPT 4o Zh · Qwen Qwq 32B · unsloth/gemma-3-12b-it · Qwen Max 0125 · BAAI/bge-large-en-v1.5 · BAAI/bge-large-zh-v1.5 · BAAI/bge-reranker-v2-m3 · tencent/Hunyuan-MT-7B · V3 · V_2 · V_2_TURBO · V_2A · V_2A_TURBO · V_1 · V_1_TURBO · Doubao Embedding Large Text 240915 · Kimi Thinking Preview · GPT 4o 2024 08-06 · Qwen Plus 2025 07-28 · Qwen Plus · stepfun-ai/step3 · Text Embedding V4 · AIHubMix Phi 4 Mini (reasoning) · Qwen Turbo · Aihub Phi 4 Multimodal Instruct · Qwen3 30B A3B · Aihub Phi 4 Mini Instruct · Grok 3 · Aihub Phi 4 · Claude 3 Opus 20240229 · Dall E 3 · Doubao Embedding Text 240715 · Grok 3 Beta · Qwen3 14B · Grok 3 Fast · Qwen3 8B · deepseek-ai/DeepSeek-R1-Zero · Grok 3 Fast Beta · Grok 3 Mini · Qwen3 4B · Grok 3 Mini Beta · Qwen3 1.7b · Qwen3 0.6b · Alicloud Glm 5 · Command A 03 2025 · Grok 3 Mini Fast Beta · Qwen 3 32B · Qwen Turbo 2025 04-28 · Qwen Plus 2025 04-28 · THUDM/GLM-Z1-32B-0414 · THUDM/GLM-4.1V-9B-Thinking · Text Embedding 004 · THUDM/GLM-4-32B-0414 · THUDM/GLM-Z1-9B-0414 · THUDM/GLM-4-9B-0414 · CC Doubao Seed Code Preview · Doubao Seed Code Preview · deepseek-ai/Janus-Pro-7B · GLM Zero Preview · Qwen 3 235B A22B Instruct 2507 · Coding GLM 4.5 Air · Deepinfra Nvidia Nemotron 3 Nano 30B A3b2 · GLM 4.5 Air · Nvidia Llama 3.1 Nemotron 70B Instruct · Nvidia Llama 3.3 Nemotron Super 49B V1.5 · Nvidia Nemotron 3 Nano 30B A3B · Nvidia Nemotron Nano 12B V2 VL · Nvidia Nemotron Nano 9B V2 · O1 Preview 2024 09-12 · Qwen/QVQ-72B-Preview · Qwen/QwQ-32B-Preview · AIHubMix Mistral Large 2411 · AIHubMix Mistral Large 2407 · Grok 2 1212 · Llama 3.1 70B · Wan2.6 T2i · DESCRIBE · UPSCALE · Bai Qwen3 VL 235B A22B Instruct · CC MiniMax M2 · CC DeepSeek V3 · CC DeepSeek V3.1 · CC ERNIE 4.5 300B A47B · CC Kimi Dev 72B · CC Kimi K2 Instruct · CC Kimi K2 Instruct 0905 · CC Kimi K2 Thinking · Computer Use Preview · GPT Image Test · grok-4.20-beta-0309-non-reasoning · Grok 4.20 Beta 0309 (reasoning) · Grok 4.20 Multi Agent Beta 0309 · Jina Reader · Jina Search · Llama3.1 8B · O1 2024 12-17 · Sf Kimi K2 Thinking · DeepSeek V3 · Doubao 1.5 Lite 32K · Doubao 1.5 Pro 256K · Doubao 1.5 Pro 32K · Doubao 1.5 Vision Pro 32K · Doubao Lite 128K · Doubao Lite 32K · Doubao Lite 4K · Doubao Pro 128K · Doubao Pro 256K · Doubao Pro 32K · Doubao Pro 4K · GPT-OSS-20B · Gryphe/MythoMax-L2-13b · MiniMax Text 01 · Mistral Large 2407 · Qwen/Qwen2-1.5B-Instruct · Qwen/Qwen2-57B-A14B-Instruct · Qwen/Qwen2-72B-Instruct · Qwen/Qwen2-7B-Instruct · Qwen/Qwen2.5-32B-Instruct · Qwen/Qwen2.5-72B-Instruct · Qwen/Qwen2.5-72B-Instruct-128K · Qwen/Qwen2.5-7B-Instruct · Qwen/Qwen2.5-Coder-32B-Instruct · Qwen3 235B A22B Thinking 2507 · Stable Diffusion 3.5 Large · WizardLM/WizardCoder-Python-34B-V1.0 · AIHubMix Phi 3.5 MoE Instruct · AIHubMix Phi 3.5 Mini Instruct · AIHubMix Phi 3.5 Vision Instruct · AIHubMix Phi 3 Medium 128K · AIHubMix Phi 3 Medium 4K · AIHubMix Phi 3 Small 128K · AIHubMix Codestral 2501 · AIHubMix Cohere Command R · AIHubMix Jamba 1.5 Large · AIHubMix Llama 3.1 405B Instruct · AIHubMix Llama 3.1 70B Instruct · AIHubMix Llama 3.1 8B Instruct · AIHubMix Llama 3.2 11B Vision · AIHubMix Llama 3.2 90B Vision · AIHubMix Llama 3 70B Instruct · AIHubMix Mistral Large · AIHubMix Command R 08 2024 · AIHubMix Command R Plus · AIHubMix Command R Plus 08 2024 · Alicloud Deepseek V3.2 · Alicloud Glm 4.7 · Alicloud Kimi K2 Thinking · Alicloud Kimi K2.5 · Alicloud Minimax M2.5 · Anthropic Opus 4.6 · Azure Deepseek V3.2 · Azure Deepseek V3.2 Speciale · Azure Kimi K2.5 · Cbs Glm 4.7 · Cerebras Llama 3.3 70B · Chatglm_lite · Chatglm_pro · Chatglm_std · Chatglm_turbo · Claude 2 · Claude 2.0 · Claude 2.1 · Claude 3 Haiku 20240229 · Claude 3 Haiku 20240307 · Claude 3 Sonnet 20240229 · Claude Instant 1 · Claude Instant 1.2 · Code Davinci Edit 001 · Cogview 3 · Cogview 3 Plus · Command · Command Light · Command Light Nightly · Command Nightly · Command R · Command R 08 2024 · Command R Plus · Command R Plus 08 2024 · Dall E 2 · Davinci · Davinci 002 · Deepinfra Llama 3.1 8B Instant · Deepinfra Llama 3.3 70B Instant Turbo · Deepinfra Llama 4 Maverick 17B 128e Instruct · Deepinfra Llama 4 Scout 17B 16e Instruct · deepseek-ai/DeepSeek-Coder-V2-Instruct · deepseek-ai/DeepSeek-R1-Distill-Llama-70B · deepseek-ai/DeepSeek-R1-Distill-Llama-8B · deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B · deepseek-ai/DeepSeek-R1-Distill-Qwen-14B · deepseek-ai/DeepSeek-R1-Distill-Qwen-32B · deepseek-ai/DeepSeek-R1-Distill-Qwen-7B · deepseek-ai/DeepSeek-V2-Chat · deepseek-ai/DeepSeek-V2.5 · deepseek-ai/deepseek-llm-67b-chat · deepseek-ai/deepseek-vl2 · DeepSeek V3 · Distil Whisper Large V3 En · Doubao 1.5 Thinking Vision Pro 250428 · Fx Flux 2 Pro · gemini-2.5-pro-exp-03-25 · gemini-embedding-exp-03-07 · gemini-exp-1114 · gemini-exp-1121 · Gemini Pro · Gemini Pro Vision · Gemma 7B It · GLM 3 Turbo · GLM 4 · GLM 4 Flash · GLM 4 Plus · GLM 4.5 Airx · GLM 4 Vision · GLM 4 Vision Plus · Google Gemma 3 12B It · Google Gemma 3 27B It · Google Gemma 3 4B It · google/gemini-exp-1114 · google/gemma-2-27b-it · google/gemma-2-9b-it:free · GPT 3.5 Turbo · GPT 3.5 Turbo 1106 · GPT 3.5 Turbo 16K · GPT 3.5 Turbo Instruct · GPT 4 · GPT 4 0613 · GPT 4 1106 Preview · GPT 4 Turbo · GPT 4 Turbo 2024 04-09 · GPT 4o 2024 05-13 · GPT 4o Mini 2024 07-18 · gpt-oss-20b · Grok 2 Vision 1212 · Grok Vision Beta · Groq Llama 3.1 8B Instant · Groq Llama 3.3 70B Versatile · Groq Llama 4 Maverick 17B 128e Instruct · Groq Llama 4 Scout 17B 16e Instruct · Jina Embeddings V2 Base Code · Learnlm 1.5 Pro Experimental · Llama 3.1 405B Instruct · Llama 3.1 405B (reasoning) · Llama 3.1 70B Versatile · Llama 3.1 8B Instant · Llama 3.2 11B Vision Preview · Llama 3.2 1B Preview · Llama 3.2 3B Preview · Llama 3.2 90B Vision Preview · Llama 3.3 70B Instruct · Llama2 70B 4096 · Llama2 70B 40960 · Llama2 7B 2048 · Llama3 70B 8192 · Llama3 8B 8192 · Llama3 Groq 70B 8192 Tool Use Preview · Llama3 Groq 8B 8192 Tool Use Preview · meta-llama/Llama-3.2-90B-Vision-Instruct · meta-llama/llama-3.1-405b-instruct:free · meta-llama/llama-3.1-70b-instruct:free · meta-llama/llama-3.1-8b-instruct:free · meta-llama/llama-3.2-11b-vision-instruct:free · meta-llama/llama-3.2-3b-instruct:free · meta/llama-3.1-405b-instruct · meta/llama3-8B-chat · mistralai/mistral-7b-instruct:free · Moonshot Kimi K2.5 · Moonshot V1 128K · Moonshot V1 128K Vision Preview · Moonshot V1 32K · Moonshot V1 32K Vision Preview · Moonshot V1 8K · Moonshot V1 8K Vision Preview · nvidia/Llama-3_1-Nemotron-Ultra-253B-v1 · O1 Mini 2024 09-12 · Omni Moderation · Qwen Flash · Qwen Flash 2025 07-28 · Qwen Long · Qwen Max · Qwen Max Longcontext · Qwen Plus · Qwen Turbo · Qwen Turbo 2024 11-01 · Qwen2.5 14B Instruct · Qwen2.5 32B Instruct · Qwen2.5 3B Instruct · Qwen2.5 72B Instruct · Qwen2.5 7B Instruct · Qwen2.5 Coder 1.5b Instruct · Qwen2.5 Coder 7B Instruct · Qwen2.5 Math 1.5b Instruct · Qwen2.5 Math 72B Instruct · Qwen2.5 Math 7B Instruct · Solar Pro4 · Step 2 16K · Text Ada 001 · Text Babbage 001 · Text Curie 001 · Text Davinci 002 · Text Davinci 003 · Text Davinci Edit 001 · Text Embedding 3 Large · Text Embedding 3 Small · Text Embedding Ada 002 · Text Embedding V1 · Text Moderation 007 · Text Moderation · Text Moderation Stable · Text Search Ada Doc 001 · Tts 1 · Tts 1 1106 · Tts 1 Hd · Tts 1 Hd 1106 · Whisper 1 · Whisper Large V3 · Whisper Large V3 Turbo · DeepSeek R1 Distill Qianfan Llama 8B · Doubao 1.5 Pro 256K 250115 · Doubao 1.5 Pro 32K 250115 · GPT 4o 2024 08-06 Global · GPT 4o Mini Global · Meta Llama 3 70B · Meta Llama 3 8B · O3 Global · O3 Mini Global · O3 Pro Global · Qianfan Chinese Llama 2 13B · Qianfan Llama VL 8B · DeepSeek V3 Fast · Veo 2.0 Generate 001 · Sonar · GPT 4 32K · Llama 3.1 Sonar Huge 128K Online · Llama 3.1 Sonar Large 128K Online · Baichuan3 Turbo · Baichuan3 Turbo 128K · Baichuan4 · Baichuan4 Air · Baichuan4 Turbo · GPT 3.5 Turbo 0301 · GPT 3.5 Turbo 0613 · GPT 3.5 Turbo 16K 0613 · GPT 4 0125 Preview · GPT 4 0314 · GPT 4 32K 0314 · GPT 4 32K 0613 · GPT 4 Turbo Preview · GPT 4 Vision Preview · Llama 3.1 Sonar Small 128K Online · Yi Large · Yi Large Rag · Yi Large Turbo · Yi Lightning · Yi Medium · Yi VL Plus