Models

Filtr
Łącznie: 854 modele
API danych modelu
Typy
Wszystko
Text
Image
Speech
Video
Transcription
Embeddings
Rerank
OCR
Tagi
Wszystko
Featured
Coding
Free
Discount
Deweloper
WszystkoOpenAIAnthropicGoogleGrokQwenDeepSeekZ.AIByteDanceLlamaAI21MicrosoftCohereMistralYiMoonshot AIStepFunNvidiaMinimaxPerplexityBaichuanIdeogramJina AIStable diffusionHunyuanBaiduFluxMeituanInclusionAIBAAIXiaomiKLingMetaPoolsideAgnesLiquidfireworksDots Studio
Refine
Newest
Input Modalities
Resetuj
autoauto:balancedauto:quality_firstauto:latency_critical

Autorski automatyczny router modeli AIHubMix. Nie każde zapytanie wymaga najmocniejszego i najdroższego modelu — brama łączy wewnętrzne testy i publiczne rankingi, obejmuje szereg popularnych modeli o różnych możliwościach i automatycznie kieruje każde zapytanie za pomocą wytrenowanego przez nas małego modelu. Ustaw model na auto, a zapytania trafią do odpowiedniego modelu zależnie od treści, obniżając koszty. Obecnie w wersji beta; stale aktualizujemy i ulepszamy.

Rozliczenie według zwykłej ceny faktycznie użytego modelu — sam routing jest bezpłatny.Poznaj inteligentny routing
Wejście:$ 0.06 /M
Wyjście:$ 0.21999996 /M
Kontekst:-
Opóźnienie:-
Przepustowość:-

GLM-5.3 is Z.ai’s reasoning model for coding and agentic workflows, designed for complex software engineering, long-running agents, and vulnerability analysis. It uses the same base model as GLM-5.2, with scaled post-training improving coding, task execution, and token efficiency. This model is a limited-time preview version of GLM-5.3, intended for testing and evaluation only. Service stability is not guaranteed, and we do not recommend using it in production environments. We’re waiting for the official commercial API release and will integrate it as soon as official support becomes available.

  • Wejście: $ 0.75 /M
  • Wyjście: $ 3.75 /M
  • Web Search: $0.014/request
  • Cache Storage: $1/h/M tokens
  • Input Audio: $1/M tokens
  • Input Video: $1/M tokens

Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web development, and knowledge work. It supports a 1M-token context window and adjustable thinking levels. Compared with Gemini 3.6 Flash, it improves coding, tool use, multi-step planning, and instruction following.

Wejście:$ 0 /M
Wyjście:$ 0 /M
Kontekst:512K
Opóźnienie:-
Przepustowość:-

Dots3-Note Preview is an open-weight mixture-of-experts model developed by Dots Studio, featuring 16B active parameters out of 280B total. As the lightest model in the Dots 3 family, it is designed for efficient performance while supporting an expansive context length of 512,000 tokens. This preview version provides an accessible way to experience the capabilities of the Dots 3 architecture.

  • Wejście: $ 0 /M
  • Wyjście: $ 0 /M
  • Web Search: $0.014/request
  • Cache Storage: $1/h/M tokens
  • Input Audio: $2/M tokens
  • Input Video: $1/M tokens

Gemini 3.7 Flash free version: fFree model resources are limited and provided only for trial use; stability cannot be guaranteed, and you may encounter 429 errors during use. If you need to use it in a production environment and require unlimited concurrency with absolute stability, please choose the official version: gemini-3.7-flash

icon
do 30% zniżki·00:00–23:59 UTCCopy ID
Wejście:$ 1.1268$ 0.78876 /M
Wyjście:$ 3.9438$ 2.76066 /M
Kontekst:1M
Opóźnienie:0.909 s
Przepustowość:45 tok/s

GLM-5.2 is Z.ai’s flagship model for the era of long-horizon tasks. With a truly usable 1M-token context window, it can handle project-level engineering context, execute long-running tasks more reliably, follow engineering standards more consistently, and complete the full development workflow from requirements to multi-platform deployment in a single task.

Wejście:$ 0.142 /M
Wyjście:$ 0.284 /M
Kontekst:1M
Opóźnienie:1.157 s
Przepustowość:117 tok/s

DeepSeek-V4-Flash-0731(deepseek-v4-flash-0731) is an open-source MoE large language model developed by the Chinese AI company DeepSeek, with support for a million-token context window. It is designed for coding, complex reasoning, tool use, agentic workflows, and long-document processing. Its advantages include strong performance with fewer active parameters and improved efficiency through DSpark speculative decoding. Compared with DeepSeek V4-Flash Preview, it offers significantly stronger coding and agent capabilities, while outperforming DeepSeek V4-Pro Preview on several benchmarks with fewer active parameters.

Wejście:$ 0.464 /M
Wyjście:$ 0.928 /M
Kontekst:1M
Opóźnienie:1.021 s
Przepustowość:58 tok/s

DeepSeek V4 Pro 0813 is DeepSeek’s high-performance general-purpose reasoning and agent model, designed for complex reasoning, coding, long-document analysis, and agentic workflows. It supports thinking and non-thinking modes, a 1M-token context window, up to 384K output, tool calling, and the Responses API. Compared with V4 Flash 0731, Pro prioritizes capability on complex tasks, while Flash focuses on speed, cost efficiency, and high concurrency.

icon
Copy ID
  • Wejście: $ 2 /M
  • Wyjście: $ 6 /M

Grok 4.6 is xAI’s (SpaceXAI) flagship multimodal reasoning model for coding, long-running agents, knowledge work, and interactive application development. It supports image understanding, a 500K context window, tool calling, and structured outputs. Compared with Grok 4.5, it offers stronger multi-step execution, self-verification, coding, and visual project generation.

Wejście:$ 2 /M
Wyjście:$ 8 /M
Kontekst:256K
Opóźnienie:6.669 s
Przepustowość:42 tok/s

MAI-Thinking-1 is Microsoft’s first inference model in the MAI series, built for enterprise-scale workloads. With excellent reasoning, mathematical, and general intelligence capabilities, combined with superior cost-effectiveness, it makes high-throughput, 24/7 AI workloads economically viable.

  • Wejście: $ 2 /M
  • Wyjście: $ 0 /M

Seedance 2.5 is ByteDance Seed’s next-generation unified multimodal audio-video generation model, designed for filmmaking, advertising, education, simulation, and long-form content creation. It generates up to 30 seconds of synchronized audio and video in one pass and supports multi-round extensions. Compared with Seedance 2.0, it offers stronger storytelling, smoother transitions, richer reference support, improved realism, and more precise editing.

  • Wejście: $ 0.2 /M
  • Wyjście: $ 1.2 /M
  • Web Search: $0.01/request

GPT-5.6 Luna is designed for cost-sensitive, high-volume workloads. It roughly corresponds to the nano model tier used in earlier GPT-5 families.

  • Wejście: $ 5 /M
  • Wyjście: $ 30 /M
  • Web Search: $0.01/request

GPT‑5.6 Sol sets a new standard for both intelligence and efficiency, achieving state-of-the-art results across coding, knowledge work, cybersecurity, and science while outperforming previous and competing frontier models with fewer tokens and at lower estimated cost.

  • Wejście: $ 2 /M
  • Wyjście: $ 12 /M
  • Web Search: $0.01/request

GPT-5.6 Terra is designed for workloads that balance intelligence and cost. It roughly corresponds to the mini model tier used in earlier GPT-5 families.

Wejście:$ 0.03 /M
Wyjście:$ 0.15 /M
Kontekst:512K
Opóźnienie:0.721 s
Przepustowość:28 tok/s

Agnes 2.5 Flash is Agnes AI’s fast and efficient language model, an upgraded, fully available model based on Agnes 2.0 Flash. It continues to use an OpenAI-compatible Chat Completions interface and has been optimized for coding tasks, agent workflows, tool calls, multi-turn dialogue, reasoning, and image understanding.

Wejście:$ 0.45 /M
Wyjście:$ 0.9 /M
Kontekst:1M
Opóźnienie:3.078 s
Przepustowość:28 tok/s

Agnes 2.5 Pro is Agnes AI’s paid inference model and the commercially stable version of the Agnes 2.5 Pro Alpha ranking model, suitable for advanced coding, scientific reasoning, long-context analysis, agent workflows, and multimodal understanding. The model is accessed via an OpenAI-compatible Chat Completions API.

Wejście:$ 0.45 /M
Wyjście:$ 0.9 /M
Kontekst:1M
Opóźnienie:-
Przepustowość:-

Agnes 2.5 Pro Alpha is Agnes AI’s paid inference model, suitable for advanced coding, scientific reasoning, long-context analysis, agent workflows, and multimodal understanding. The model is accessed via an OpenAI-compatible Chat Completions API.

Quality
standard
standard
$0.003

Agnes Image 2.1 Flash is Agnes AI’s high-performance image generation and image editing model, supporting text-to-image, image-to-image, and multi-image composition. It is suitable for creative design, marketing visuals, e-commerce product images, and social content production.

icon
Copy ID
  • Wejście: $ 2 /M
  • Wyjście: $ 6 /M

Grok 4.5 was trained on datasets spanning knowledge in coding, science, engineering, and math. With both intelligent and efficient reasoning, Grok 4.5 excels at real engineering tasks and exceeds comparable leading models at these tasks.

Wejście:$ 0 /M
Wyjście:$ 0 /M
Kontekst:128K
Opóźnienie:-
Przepustowość:-

LFM-2.5-2.6B is a compact reasoning model developed by Liquid, featuring a generous 128,000 token context length. It is highly suited for agent workflows, data extraction, retrieval-augmented generation (RAG), and long-context processing. However, the developer advises against using this model for agentic coding tasks.

Scroll to load more

50 free models — no credit card required →

All models (849)

auto · coding-glm-5.3 · gemini-3.7-flash · dots-3-note-preview-free · gemini-3.7-flash-free · glm-5.2 · deepseek-v4-flash-0731 · deepseek-v4-pro-0813 · grok-4.6 · mai-thinking-1 · doubao-seedance-2-5-260628 · gpt-5.6-luna · gpt-5.6-sol · gpt-5.6-terra · agnes-2.5-flash · agnes-2.5-pro · agnes-2.5-pro-alpha · agnes-image-2.1-flash · grok-4.5 · lfm-2.5-2.6b-free · minimax-h3 · muse-glimmer-30b · nemotron-lightning-3.5-30b-a3b · qwen3.8-2.4t-a95b · claude-opus-5 · gemini-3.6-flash · ling-3.0-tiny-free · nemotron-3.5-lightning-free · qwen-image-3.0 · qwen-image-3.0-pro · qwen3.8-max · claude-sonnet-5 · kimi-k3 · ling-3.0-flash-free · muse-spark-1.2 · qwen3.8-max-preview · gemini-3.1-flash-lite-image · gemini-3.5-flash-lite · gemini-3.5-flash-lite-free · gemini-3.6-flash-free · glm-5.2-fast-preview · muse-spark-1.1 · qwen-audio-3.0-tts-flash · qwen-audio-3.0-tts-plus · claude-fable-5 · jina-reranker-v3.5 · claude-opus-4-8 · hy3 · doubao-seed-2-1-pro · doubao-seed-2-1-turbo · mai-image-2.5-pro · gemini-3.5-flash · grok-build-0.1 · mai-image-2.5 · mai-image-2.5-flash · coding-kimi-k3 · happyhorse-1.1-i2v · happyhorse-1.1-r2v · happyhorse-1.1-t2v · coding-glm-5.2-free · coding-kimi-k3-free · gemini-3.1-flash-image · gpt-oss-20b-free · kimi-k2.7-code · kimi-k2.7-code-highspeed · gemini-3-pro-image · gpt-4o-transcribe-diarize · gpt-audio-1.5 · hy-3d-3.1 · kling-v3-omni · kling-video-o1 · longcat-2.0 · nemotron-nano-9b-v2-free · hy3-preview · minimax-m3 · nemotron-nano-12b-v2-vl-free · qwen3.7-flash · qwen3.7-plus · step-3.7-flash · claude-opus-4-8-think · nemotron-3-super-120b-a12b-free · nemotron-3-nano-omni-30b-a3b-reasoning-free · nemotron-3-ultra-550b-a55b-free · qwen3.7-max · gpt-image-2 · nemotron-3.5-content-safety-free · coding-glm-5.2 · ernie-5.1 · gemini-3.1-flash-lite · gemini-3.1-flash-lite-nothink · grok-4.3 · happyhorse-1.0-i2v · happyhorse-1.0-r2v · happyhorse-1.0-t2v · happyhorse-1.0-video-edit · north-mini-code-free · gpt-5.5 · gpt-5.5-pro · laguna-xs-2.1-free · deepseek-v4-flash · deepseek-v4-pro · gemma-4-31b-it-free · command-a-plus-05-2026 · doubao-seedream-5.0-pro · ernie-5.0 · kimi-k2.6 · laguna-s-2.1-free · qwen3.6-max-preview · xiaomi-mimo-v2.5 · xiaomi-mimo-v2.5-pro · claude-opus-4-7 · claude-opus-4-7-think · gpt-chat-latest · nemotron-3-nano-30b-a3b-free · qwen3.6-27b · qwen3.6-35b-a3b · qwen3.6-flash · cohere-rerank-v4.0-fast · cohere-rerank-v4.0-pro · gemma-4-26b-a4b-it-free · grok-4-20-non-reasoning · grok-4-20-reasoning · qwen-image-2.0 · qwen-image-2.0-pro · coding-minimax-m3-free · doubao-seedance-2-0-260128 · doubao-seedance-2-0-fast-260128 · doubao-seedance-2-0-mini-260615 · glm-5.1 · glm-image · qwen3.6-plus · wan2.7-i2v · wan2.7-r2v · wan2.7-t2v · wan2.7-videoedit · cc-k2.6-code-preview · gemma-4-26b-a4b-it · gemma-4-31b-it · gpt-5.4 · wan2.7-image · wan2.7-image-pro · claude-sonnet-4-6 · coding-xiaomi-mimo-v2.5 · coding-xiaomi-mimo-v2.5-pro · doubao-seed-2-0-lite-260428 · doubao-seed-2-0-mini-260428 · gemini-3.1-flash-image-preview · gemini-3.1-pro-preview · gemini-3.1-pro-preview-customtools · gemini-3.1-pro-preview-search · gpt-5.4-mini · gpt-5.4-nano · gpt-5.5-free · qwen3.5-plus · claude-sonnet-4-6-think · coding-xiaomi-mimo-v2-omni · coding-xiaomi-mimo-v2-pro · gpt-5.3-chat-latest · gpt-5.3-codex · gpt-image-2-free · qwen3.5-122b-a10b · qwen3.5-27b · qwen3.5-35b-a3b · qwen3.5-397b-a17b · qwen3.5-flash · coding-glm-5.1 · doubao-seed-2-0-pro · gpt-5.4-high · gpt-5.4-low · gpt-5.4-pro · qwen3-coder-next · xiaomi-mimo-v2-omni-free · xiaomi-mimo-v2-pro-free · xiaomi-mimo-v2.5-free · xiaomi-mimo-v2.5-pro-free · claude-opus-4-6 · coding-glm-5.1-free · coding-minimax-m2.7-free · glm-5 · glm-5v-turbo · minimax-m2.7 · claude-opus-4-6-think · coding-glm-5-free · coding-glm-5-turbo-free · coding-minimax-m2.5-free · doubao-seed-2-0-code-preview · doubao-seed-2-0-lite-260215 · doubao-seed-2-0-mini · gemini-3-flash-preview · gemini-3-flash-preview-search · glm-5-turbo · cc-glm-5.1 · claude-opus-4-5 · claude-opus-4-5-think · embed-v-4-0 · ernie-image-turbo · gemini-3.1-flash-image-preview-free · mimo-v2-omni · mimo-v2-pro · cohere-command-a · gemini-3-flash-preview-free · cc-minimax-m3 · coding-minimax-m3 · gpt-4.1-free · gpt-4.1-mini-free · gpt-4.1-nano-free · gpt-4o-free · coding-glm-5 · coding-glm-5-turbo · glm-4.7 · veo-3.1-lite-generate-preview · glm-4.7-flash-free · coding-glm-4.7-free · doubao-seedance-1-5-pro-251215 · doubao-seedance-1-0-pro-250528 · doubao-seedance-1-0-pro-fast-251015 · gemini-3-pro-image-preview · gemini-embedding-2 · deepinfra-gemma-4-26b-a4b-it · gpt-5.2-codex · doubao-seedream-5.0-lite · gpt-image-1.5 · gpt-5.2 · gpt-5.2-chat-latest · gpt-5.2-high · gpt-5.2-low · gpt-5.2-pro · gpt-5.1 · gpt-5.1-codex-max · doubao-seed-1-8 · gpt-5.1-chat-latest · gpt-5.1-codex · gpt-5.1-codex-mini · claude-haiku-4-5 · claude-sonnet-4-5 · claude-sonnet-4-5-think · grok-4.20-multi-agent-0309 · mistral-large-3 · cc-glm-5 · cc-glm-5-turbo · cloudflare-glm-5.2 · gemini-2.5-flash-image · grok-4-1-fast-non-reasoning · grok-4-1-fast-reasoning · grok-code-fast-1 · k2.6-code-preview-free · mimo-v2-flash · musesteamer-air-image · qwen3.6-plus-preview-free · zai-glm-5-turbo · gpt-5 · deepseek-v3.2 · deepseek-v3.2-think · gpt-5-codex · DeepSeek-V3.1-Terminus · DeepSeek-V3.1-Think · gpt-5-pro · gpt-5-mini · gpt-5-nano · gpt-5-chat-latest · claude-opus-4-1 · o3-deep-research · kimi-k2.5 · qwen3-max-2026-01-23 · qwen3-vl-flash · qwen3-vl-flash-2026-01-22 · qwen3-vl-plus · cc-minimax-m2.7 · cc-minimax-m2.7-highspeed · minimax-m2.5 · minimax-m2.5-highspeed · mm-minimax-m2.7-highspeed · coding-minimax-m2.7 · coding-minimax-m2.7-highspeed · cc-minimax-m2.5 · cc-minimax-m2.5-highspeed · coding-minimax-m2.5 · coding-minimax-m2.5-highspeed · doubao-seedream-4-5 · sora-2 · sora-2-pro · cc-glm-4.7 · cc-minimax-m2.1 · coding-glm-4.7 · coding-minimax-m2.1 · coding-minimax-m2.1-free · gpt-4o-audio-preview · gpt-4o-mini-audio-preview · minimax-m2.1 · o3 · wan2.6-i2v · wan2.6-t2v · cc-glm-4.6 · coding-glm-4.6 · coding-glm-4.6-free · coding-minimax-m2 · coding-minimax-m2-free · flux-2-flex · flux-2-pro · gemini-2.5-pro · glm-4.6 · glm-4.6v · glm-ocr · kimi-for-coding-free · o3-pro · qianfan-ocr · qianfan-ocr-fast · step-3.5-flash · wan2.2-i2v-plus · wan2.5-i2v-preview · wan2.5-t2v-preview · gemini-2.5-pro-search · kimi-k2-thinking · gemini-2.5-flash · gemini-2.5-flash-preview-09-2025 · glm-4.5v · gemini-2.5-flash-lite · gemini-2.5-flash-lite-nothink · gemini-2.5-flash-lite-preview-09-2025 · gemini-2.5-flash-lite-preview-09-2025-nothink · gemini-2.5-flash-nothink · gemini-2.5-flash-search · gemini-2.5-flash-preview-05-20-nothink · gemini-2.5-flash-preview-05-20-search · DeepSeek-V3-Fast · imagen-4.0 · imagen-4.0-fast-generate-001 · imagen-4.0-generate-001 · imagen-4.0-ultra-generate-001 · imagen-4.0-ultra · gpt-image-1 · gpt-image-1-mini · o4-mini · DeepSeek-OCR · alicloud-kimi-k2-instruct · deepseek-ocr · ernie-5.0-thinking-exp · flux-kontext-max · gemini-2.5-flash-image-preview · glm-4.5 · gpt-4.1 · grok-4 · grok-4-fast-non-reasoning · grok-4-fast-reasoning · kimi-k2-0711 · kimi-k2-instruct · kimi-k2-turbo-preview · paddleocr-vl-0.9b · pp-structurev3 · qwen3-vl-235b-a22b-instruct · qwen3-vl-235b-a22b-thinking · qwen3-vl-30b-a3b-instruct · qwen3-vl-30b-a3b-thinking · veo-3.0-generate-preview · veo-3.1-fast-generate-preview · veo-3.1-generate-preview · aihubmix-router · gpt-4.1-mini · gpt-4.1-nano · gemini-2.5-pro-preview-05-06 · gemini-2.5-pro-preview-03-25 · gemini-2.5-pro-preview-05-06-search · gemini-2.5-pro-preview-03-25-search · qwen3-max-preview · qwen3-max · qwen3-next-80b-a3b-instruct · qwen3-next-80b-a3b-thinking · qwen3-235b-a22b-instruct-2507 · qwen3-235b-a22b-thinking-2507 · qwen3-coder-30b-a3b-instruct · qwen3-coder-480b-a35b-instruct · DeepSeek-V3 · LongCat-Flash-Chat · gemini-2.5-pro-preview-06-05-search · jina-embeddings-v5-text-nano · jina-embeddings-v5-text-small · qwen3-235b-a22b · qwen3-coder-flash · qwen3-coder-plus · qwen3-coder-plus-2025-07-22 · Qwen2.5-VL-72B-Instruct · ernie-5.0-thinking-preview · inclusionAI/Ling-1T · inclusionAI/Ring-1T · bce-reranker-base · codex-mini-latest · doubao-seedream-4-0 · embedding-v1 · ernie-4.5-turbo-latest · glm-4.5-x · gme-qwen2-vl-2b-instruct · gte-rerank-v2 · inclusionAI/Ling-flash-2.0 · inclusionAI/Ling-mini-2.0 · inclusionAI/Ring-flash-2.0 · jina-deepsearch-v1 · jina-embeddings-v4 · jina-reranker-v3 · llama-4-maverick · llama-4-scout · qwen-image · qwen-image-edit · qwen-image-max · qwen-mt-plus · qwen-mt-turbo · qwen3-embedding-0.6b · qwen3-embedding-4b · qwen3-embedding-8b · qwen3-reranker-0.6b · qwen3-reranker-4b · qwen3-reranker-8b · tao-8k · jina-clip-v2 · jina-reranker-m0 · jina-colbert-v2 · DeepSeek-R1 · gpt-4o-search-preview · gpt-4o-mini-search-preview · jina-embeddings-v3 · claude-3-7-sonnet · ernie-4.5 · ernie-4.5-turbo-vl · mimo-v2-flash-free · FLUX-1.1-pro · o3-mini · doubao-seed-1-6 · doubao-seed-1-6-flash · doubao-seed-1-6-lite · doubao-seed-1-6-thinking · qwen3-30b-a3b-instruct-2507 · qwen3-30b-a3b-thinking-2507 · Qwen2-VL-72B-Instruct · Qwen2-VL-7B-Instruct · cc-kimi-for-coding · gemini-embedding-001 · gpt-oss-120b · qwen-3-235b-a22b-thinking-2507 · Qwen/Qwen3-30B-A3B · Qwen/Qwen3-32B · qwen3-32b · Qwen/Qwen3-14B · Qwen/Qwen3-8B · embedding-2 · embedding-3 · gemini-2.5-pro-preview-06-05 · Qwen/Qwen2.5-VL-72B-Instruct · o1 · o1-pro · ByteDance-Seed/Seed-OSS-36B-Instruct · doubao-seed-1-6-250615 · doubao-seed-1-6-flash-250615 · doubao-seed-1-6-thinking-250615 · doubao-seed-1-6-vision-250815 · Doubao-1.5-thinking-pro · cc-minimax-m2 · deepseek-ai/DeepSeek-Prover-V2-671B · gemini-2.5-flash-preview-tts · gemini-2.5-pro-preview-tts · gemma-3-12b-it · gemma-3-27b-it · gemma-3-4b-it · gemma-3n-e4b-it · gemma-3-1b-it · deepseek-r1-distill-llama-70b · gpt-4o-mini-tts · tngtech/DeepSeek-R1T-Chimera · veo-2.0-generate-001 · o1-preview · o1-mini · gpt-4o-2024-11-20 · gpt-4o · gpt-4o-mini · AiHubmix-mistral-medium · ERNIE-X1.1-Preview · Qwen/QwQ-32B · chutesai/Mistral-Small-3.1-24B-Instruct-2503 · ernie-x1.1-preview · minimax-m2 · MiniMaxAI/MiniMax-M1-80k · Qwen/Qwen2.5-VL-32B-Instruct · baidu/ERNIE-4.5-300B-A47B · bge-large-en · bge-large-zh · codestral-latest · ernie-4.5-0.3b · ernie-4.5-turbo-128k-preview · ernie-x1-turbo · kat-dev · llama-3.3-70b · moonshotai/Kimi-Dev-72B · moonshotai/Moonlight-16B-A3B-Instruct · nvidia-nemotron-3-super-120b-a12b · o1-global · qianfan-qi-vl · qwen2.5-vl-72b-instruct · tencent/Hunyuan-A13B-Instruct · unsloth/gemma-3-27b-it · gemini-exp-1206 · gpt-4o-zh · qwen-qwq-32b · unsloth/gemma-3-12b-it · qwen-max-0125 · BAAI/bge-large-en-v1.5 · BAAI/bge-large-zh-v1.5 · BAAI/bge-reranker-v2-m3 · tencent/Hunyuan-MT-7B · V3 · V_2 · V_2_TURBO · V_2A · V_2A_TURBO · V_1 · V_1_TURBO · doubao-embedding-large-text-240915 · kimi-thinking-preview · gpt-4o-2024-08-06 · qwen-plus-2025-07-28 · qwen-plus-latest · sonar · stepfun-ai/step3 · text-embedding-v4 · AiHubmix-Phi-4-mini-reasoning · qwen-turbo-latest · aihub-Phi-4-multimodal-instruct · qwen3-30b-a3b · aihub-Phi-4-mini-instruct · grok-3 · aihub-Phi-4 · claude-3-opus-20240229 · dall-e-3 · doubao-embedding-text-240715 · grok-3-beta · qwen3-14b · grok-3-fast · qwen3-8b · deepseek-ai/DeepSeek-R1-Zero · grok-3-fast-beta · grok-3-mini · qwen3-4b · grok-3-mini-beta · qwen3-1.7b · qwen3-0.6b · alicloud-glm-5 · command-a-03-2025 · grok-3-mini-fast-beta · qwen-3-32b · qwen-turbo-2025-04-28 · qwen-plus-2025-04-28 · THUDM/GLM-Z1-32B-0414 · THUDM/GLM-4.1V-9B-Thinking · text-embedding-004 · THUDM/GLM-4-32B-0414 · THUDM/GLM-Z1-9B-0414 · THUDM/GLM-4-9B-0414 · cc-doubao-seed-code-preview-latest · doubao-seed-code-preview-latest · deepseek-ai/Janus-Pro-7B · glm-zero-preview · qwen-3-235b-a22b-instruct-2507 · coding-glm-4.5-air · deepinfra-nvidia-nemotron-3-nano-30b-a3b2 · glm-4.5-air · gpt-4-32k · nvidia-llama-3.1-nemotron-70b-instruct · nvidia-llama-3.3-nemotron-super-49b-v1.5 · nvidia-nemotron-3-nano-30b-a3b · nvidia-nemotron-nano-12b-v2-vl · nvidia-nemotron-nano-9b-v2 · o1-preview-2024-09-12 · Qwen/QVQ-72B-Preview · Qwen/QwQ-32B-Preview · llama-3.1-sonar-huge-128k-online · aihubmix-Mistral-Large-2411 · llama-3.1-sonar-large-128k-online · aihubmix-Mistral-large-2407 · grok-2-1212 · llama-3.1-70b · wan2.6-t2i · DESCRIBE · UPSCALE · bai-qwen3-vl-235b-a22b-instruct · cc-MiniMax-M2 · cc-deepseek-v3 · cc-deepseek-v3.1 · cc-ernie-4.5-300b-a47b · cc-kimi-dev-72b · cc-kimi-k2-instruct · cc-kimi-k2-instruct-0905 · cc-kimi-k2-thinking · computer-use-preview · gpt-image-test · grok-4.20-beta-0309-non-reasoning · grok-4.20-beta-0309-reasoning · grok-4.20-multi-agent-beta-0309 · jina-reader · jina-search · llama3.1-8b · o1-2024-12-17 · sf-kimi-k2-thinking · Baichuan3-Turbo · Baichuan3-Turbo-128k · Baichuan4 · Baichuan4-Air · Baichuan4-Turbo · DeepSeek-v3 · Doubao-1.5-lite-32k · Doubao-1.5-pro-256k · Doubao-1.5-pro-32k · Doubao-1.5-vision-pro-32k · Doubao-lite-128k · Doubao-lite-32k · Doubao-lite-4k · Doubao-pro-128k · Doubao-pro-256k · Doubao-pro-32k · Doubao-pro-4k · GPT-OSS-20B · Gryphe/MythoMax-L2-13b · MiniMax-Text-01 · Mistral-large-2407 · Qwen/Qwen2-1.5B-Instruct · Qwen/Qwen2-57B-A14B-Instruct · Qwen/Qwen2-72B-Instruct · Qwen/Qwen2-7B-Instruct · Qwen/Qwen2.5-32B-Instruct · Qwen/Qwen2.5-72B-Instruct · Qwen/Qwen2.5-72B-Instruct-128K · Qwen/Qwen2.5-7B-Instruct · Qwen/Qwen2.5-Coder-32B-Instruct · Qwen3-235B-A22B-Thinking-2507 · Stable-Diffusion-3-5-Large · WizardLM/WizardCoder-Python-34B-V1.0 · ahm-Phi-3-5-MoE-instruct · ahm-Phi-3-5-mini-instruct · ahm-Phi-3-5-vision-instruct · ahm-Phi-3-medium-128k · ahm-Phi-3-medium-4k · ahm-Phi-3-small-128k · aihubmix-Codestral-2501 · aihubmix-Cohere-command-r · aihubmix-Jamba-1-5-Large · aihubmix-Llama-3-1-405B-Instruct · aihubmix-Llama-3-1-70B-Instruct · aihubmix-Llama-3-1-8B-Instruct · aihubmix-Llama-3-2-11B-Vision · aihubmix-Llama-3-2-90B-Vision · aihubmix-Llama-3-70B-Instruct · aihubmix-Mistral-large · aihubmix-command-r-08-2024 · aihubmix-command-r-plus · aihubmix-command-r-plus-08-2024 · alicloud-deepseek-v3.2 · alicloud-glm-4.7 · alicloud-kimi-k2-thinking · alicloud-kimi-k2.5 · alicloud-minimax-m2.5 · anthropic-opus-4-6 · azure-deepseek-v3.2 · azure-deepseek-v3.2-speciale · azure-kimi-k2.5 · cbs-glm-4.7 · cerebras-llama-3.3-70b · chatglm_lite · chatglm_pro · chatglm_std · chatglm_turbo · claude-2 · claude-2.0 · claude-2.1 · claude-3-haiku-20240229 · claude-3-haiku-20240307 · claude-3-sonnet-20240229 · claude-instant-1 · claude-instant-1.2 · code-davinci-edit-001 · cogview-3 · cogview-3-plus · command · command-light · command-light-nightly · command-nightly · command-r · command-r-08-2024 · command-r-plus · command-r-plus-08-2024 · dall-e-2 · davinci · davinci-002 · deepinfra-llama-3.1-8b-instant · deepinfra-llama-3.3-70b-instant-turbo · deepinfra-llama-4-maverick-17b-128e-instruct · deepinfra-llama-4-scout-17b-16e-instruct · deepseek-ai/DeepSeek-Coder-V2-Instruct · deepseek-ai/DeepSeek-R1-Distill-Llama-70B · deepseek-ai/DeepSeek-R1-Distill-Llama-8B · deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B · deepseek-ai/DeepSeek-R1-Distill-Qwen-14B · deepseek-ai/DeepSeek-R1-Distill-Qwen-32B · deepseek-ai/DeepSeek-R1-Distill-Qwen-7B · deepseek-ai/DeepSeek-V2-Chat · deepseek-ai/DeepSeek-V2.5 · deepseek-ai/deepseek-llm-67b-chat · deepseek-ai/deepseek-vl2 · deepseek-v3 · distil-whisper-large-v3-en · doubao-1-5-thinking-vision-pro-250428 · fx-flux-2-pro · gemini-2.5-pro-exp-03-25 · gemini-embedding-exp-03-07 · gemini-exp-1114 · gemini-exp-1121 · gemini-pro · gemini-pro-vision · gemma-7b-it · glm-3-turbo · glm-4 · glm-4-flash · glm-4-plus · glm-4.5-airx · glm-4v · glm-4v-plus · google-gemma-3-12b-it · google-gemma-3-27b-it · google-gemma-3-4b-it · google/gemini-exp-1114 · google/gemma-2-27b-it · google/gemma-2-9b-it:free · gpt-3.5-turbo · gpt-3.5-turbo-0301 · gpt-3.5-turbo-0613 · gpt-3.5-turbo-1106 · gpt-3.5-turbo-16k · gpt-3.5-turbo-16k-0613 · gpt-3.5-turbo-instruct · gpt-4 · gpt-4-0125-preview · gpt-4-0314 · gpt-4-0613 · gpt-4-1106-preview · gpt-4-32k-0314 · gpt-4-32k-0613 · gpt-4-turbo · gpt-4-turbo-2024-04-09 · gpt-4-turbo-preview · gpt-4-vision-preview · gpt-4o-2024-05-13 · gpt-4o-mini-2024-07-18 · gpt-oss-20b · grok-2-vision-1212 · grok-vision-beta · groq-llama-3.1-8b-instant · groq-llama-3.3-70b-versatile · groq-llama-4-maverick-17b-128e-instruct · groq-llama-4-scout-17b-16e-instruct · jina-embeddings-v2-base-code · learnlm-1.5-pro-experimental · llama-3.1-405b-instruct · llama-3.1-405b-reasoning · llama-3.1-70b-versatile · llama-3.1-8b-instant · llama-3.1-sonar-small-128k-online · llama-3.2-11b-vision-preview · llama-3.2-1b-preview · llama-3.2-3b-preview · llama-3.2-90b-vision-preview · llama2-70b-4096 · llama2-70b-40960 · llama2-7b-2048 · llama3-70b-8192 · llama3-8b-8192 · llama3-groq-70b-8192-tool-use-preview · llama3-groq-8b-8192-tool-use-preview · meta-llama/Llama-3.2-90B-Vision-Instruct · meta-llama/llama-3.1-405b-instruct:free · meta-llama/llama-3.1-70b-instruct:free · meta-llama/llama-3.1-8b-instruct:free · meta-llama/llama-3.2-11b-vision-instruct:free · meta-llama/llama-3.2-3b-instruct:free · meta/llama-3.1-405b-instruct · meta/llama3-8B-chat · mistralai/mistral-7b-instruct:free · moonshot-kimi-k2.5 · moonshot-v1-128k · moonshot-v1-128k-vision-preview · moonshot-v1-32k · moonshot-v1-32k-vision-preview · moonshot-v1-8k · moonshot-v1-8k-vision-preview · nvidia/Llama-3_1-Nemotron-Ultra-253B-v1 · o1-mini-2024-09-12 · omni-moderation-latest · qwen-flash · qwen-flash-2025-07-28 · qwen-long · qwen-max · qwen-max-longcontext · qwen-plus · qwen-turbo · qwen-turbo-2024-11-01 · qwen2.5-14b-instruct · qwen2.5-32b-instruct · qwen2.5-3b-instruct · qwen2.5-72b-instruct · qwen2.5-7b-instruct · qwen2.5-coder-1.5b-instruct · qwen2.5-coder-7b-instruct · qwen2.5-math-1.5b-instruct · qwen2.5-math-72b-instruct · qwen2.5-math-7b-instruct · step-2-16k · text-ada-001 · text-babbage-001 · text-curie-001 · text-davinci-002 · text-davinci-003 · text-davinci-edit-001 · text-embedding-3-large · text-embedding-3-small · text-embedding-ada-002 · text-embedding-v1 · text-moderation-007 · text-moderation-latest · text-moderation-stable · text-search-ada-doc-001 · tts-1 · tts-1-1106 · tts-1-hd · tts-1-hd-1106 · whisper-1 · whisper-large-v3 · whisper-large-v3-turbo · yi-large · yi-large-rag · yi-large-turbo · yi-lightning · yi-medium · yi-vl-plus · deepseek-r1-distill-qianfan-llama-8b · doubao-1-5-pro-256k-250115 · doubao-1-5-pro-32k-250115 · gpt-4o-2024-08-06-global · gpt-4o-mini-global · meta-llama-3-70b · meta-llama-3-8b · o3-global · o3-mini-global · o3-pro-global · qianfan-chinese-llama-2-13b · qianfan-llama-vl-8b