Models

Filtro
Total: 854 modelos
API de datos del modelo
Tipos
Todo
Text
Image
Speech
Video
Transcription
Embeddings
Rerank
OCR
Etiquetas
Todo
Featured
Coding
Free
Discount
Desarrollador
TodoOpenAIAnthropicGoogleGrokQwenDeepSeekZ.AIByteDanceLlamaAI21MicrosoftCohereMistralYiMoonshot AIStepFunNvidiaMinimaxPerplexityBaichuanIdeogramJina AIStable diffusionHunyuanBaiduFluxMeituanInclusionAIBAAIXiaomiKLingMetaPoolsideAgnesLiquidfireworksDots Studio
Refine
Price: Low to High
Input Modalities
Reiniciar
autoauto:balancedauto:quality_firstauto:latency_critical

El enrutador automático de modelos desarrollado por AIHubMix. No todas las solicitudes necesitan el modelo más potente y caro: la pasarela combina evaluaciones internas y rankings públicos, cubre una gama de modelos actuales de distintas capacidades y enruta cada solicitud automáticamente mediante un modelo pequeño entrenado por nosotros. Pon model en auto y cada solicitud se asigna al modelo adecuado según su contenido, reduciendo costos. Actualmente en beta; seguimos actualizándolo y mejorándolo.

Se cobra al precio original del modelo realmente utilizado; el enrutamiento en sí es gratuito.Conocer el enrutamiento inteligente
Entrada:$ 0.06 /M
Salida:$ 0.21999996 /M
Contexto:-
Latencia:-
Rendimiento:-

GLM-5.3 is Z.ai’s reasoning model for coding and agentic workflows, designed for complex software engineering, long-running agents, and vulnerability analysis. It uses the same base model as GLM-5.2, with scaled post-training improving coding, task execution, and token efficiency. This model is a limited-time preview version of GLM-5.3, intended for testing and evaluation only. Service stability is not guaranteed, and we do not recommend using it in production environments. We’re waiting for the official commercial API release and will integrate it as soon as official support becomes available.

  • Entrada: $ 0.75 /M
  • Salida: $ 3.75 /M
  • Web Search: $0.014/request
  • Cache Storage: $1/h/M tokens
  • Input Audio: $1/M tokens
  • Input Video: $1/M tokens

Gemini 3.7 Flash is Google’s natively multimodal reasoning model for coding, agents, web development, and knowledge work. It supports a 1M-token context window and adjustable thinking levels. Compared with Gemini 3.6 Flash, it improves coding, tool use, multi-step planning, and instruction following.

Entrada:$ 0 /M
Salida:$ 0 /M
Contexto:512K
Latencia:-
Rendimiento:-

Dots3-Note Preview is an open-weight mixture-of-experts model developed by Dots Studio, featuring 16B active parameters out of 280B total. As the lightest model in the Dots 3 family, it is designed for efficient performance while supporting an expansive context length of 512,000 tokens. This preview version provides an accessible way to experience the capabilities of the Dots 3 architecture.

  • Entrada: $ 0 /M
  • Salida: $ 0 /M
  • Web Search: $0.014/request
  • Cache Storage: $1/h/M tokens
  • Input Audio: $2/M tokens
  • Input Video: $1/M tokens

Gemini 3.7 Flash free version: fFree model resources are limited and provided only for trial use; stability cannot be guaranteed, and you may encounter 429 errors during use. If you need to use it in a production environment and require unlimited concurrency with absolute stability, please choose the official version: gemini-3.7-flash

icon
hasta 30% de descuento·00:00–23:59 UTCCopy ID
Entrada:$ 1.1268$ 0.78876 /M
Salida:$ 3.9438$ 2.76066 /M
Contexto:1M
Latencia:0.909 s
Rendimiento:45 tok/s

GLM-5.2 is Z.ai’s flagship model for the era of long-horizon tasks. With a truly usable 1M-token context window, it can handle project-level engineering context, execute long-running tasks more reliably, follow engineering standards more consistently, and complete the full development workflow from requirements to multi-platform deployment in a single task.

Entrada:$ 0.142 /M
Salida:$ 0.284 /M
Contexto:1M
Latencia:1.157 s
Rendimiento:117 tok/s

DeepSeek-V4-Flash-0731(deepseek-v4-flash-0731) is an open-source MoE large language model developed by the Chinese AI company DeepSeek, with support for a million-token context window. It is designed for coding, complex reasoning, tool use, agentic workflows, and long-document processing. Its advantages include strong performance with fewer active parameters and improved efficiency through DSpark speculative decoding. Compared with DeepSeek V4-Flash Preview, it offers significantly stronger coding and agent capabilities, while outperforming DeepSeek V4-Pro Preview on several benchmarks with fewer active parameters.

Entrada:$ 0.464 /M
Salida:$ 0.928 /M
Contexto:1M
Latencia:1.021 s
Rendimiento:58 tok/s

DeepSeek V4 Pro 0813 is DeepSeek’s high-performance general-purpose reasoning and agent model, designed for complex reasoning, coding, long-document analysis, and agentic workflows. It supports thinking and non-thinking modes, a 1M-token context window, up to 384K output, tool calling, and the Responses API. Compared with V4 Flash 0731, Pro prioritizes capability on complex tasks, while Flash focuses on speed, cost efficiency, and high concurrency.

icon
Copy ID
  • Entrada: $ 2 /M
  • Salida: $ 6 /M

Grok 4.6 is xAI’s (SpaceXAI) flagship multimodal reasoning model for coding, long-running agents, knowledge work, and interactive application development. It supports image understanding, a 500K context window, tool calling, and structured outputs. Compared with Grok 4.5, it offers stronger multi-step execution, self-verification, coding, and visual project generation.

Entrada:$ 2 /M
Salida:$ 8 /M
Contexto:256K
Latencia:6.669 s
Rendimiento:42 tok/s

MAI-Thinking-1 is Microsoft’s first inference model in the MAI series, built for enterprise-scale workloads. With excellent reasoning, mathematical, and general intelligence capabilities, combined with superior cost-effectiveness, it makes high-throughput, 24/7 AI workloads economically viable.

  • Entrada: $ 2 /M
  • Salida: $ 0 /M

Seedance 2.5 is ByteDance Seed’s next-generation unified multimodal audio-video generation model, designed for filmmaking, advertising, education, simulation, and long-form content creation. It generates up to 30 seconds of synchronized audio and video in one pass and supports multi-round extensions. Compared with Seedance 2.0, it offers stronger storytelling, smoother transitions, richer reference support, improved realism, and more precise editing.

  • Entrada: $ 0.2 /M
  • Salida: $ 1.2 /M
  • Web Search: $0.01/request

GPT-5.6 Luna is designed for cost-sensitive, high-volume workloads. It roughly corresponds to the nano model tier used in earlier GPT-5 families.

  • Entrada: $ 5 /M
  • Salida: $ 30 /M
  • Web Search: $0.01/request

GPT‑5.6 Sol sets a new standard for both intelligence and efficiency, achieving state-of-the-art results across coding, knowledge work, cybersecurity, and science while outperforming previous and competing frontier models with fewer tokens and at lower estimated cost.

  • Entrada: $ 2 /M
  • Salida: $ 12 /M
  • Web Search: $0.01/request

GPT-5.6 Terra is designed for workloads that balance intelligence and cost. It roughly corresponds to the mini model tier used in earlier GPT-5 families.

Entrada:$ 0.03 /M
Salida:$ 0.15 /M
Contexto:512K
Latencia:0.721 s
Rendimiento:28 tok/s

Agnes 2.5 Flash is Agnes AI’s fast and efficient language model, an upgraded, fully available model based on Agnes 2.0 Flash. It continues to use an OpenAI-compatible Chat Completions interface and has been optimized for coding tasks, agent workflows, tool calls, multi-turn dialogue, reasoning, and image understanding.

Entrada:$ 0.45 /M
Salida:$ 0.9 /M
Contexto:1M
Latencia:3.078 s
Rendimiento:28 tok/s

Agnes 2.5 Pro is Agnes AI’s paid inference model and the commercially stable version of the Agnes 2.5 Pro Alpha ranking model, suitable for advanced coding, scientific reasoning, long-context analysis, agent workflows, and multimodal understanding. The model is accessed via an OpenAI-compatible Chat Completions API.

Entrada:$ 0.45 /M
Salida:$ 0.9 /M
Contexto:1M
Latencia:-
Rendimiento:-

Agnes 2.5 Pro Alpha is Agnes AI’s paid inference model, suitable for advanced coding, scientific reasoning, long-context analysis, agent workflows, and multimodal understanding. The model is accessed via an OpenAI-compatible Chat Completions API.

Quality
standard
standard
$0.003

Agnes Image 2.1 Flash is Agnes AI’s high-performance image generation and image editing model, supporting text-to-image, image-to-image, and multi-image composition. It is suitable for creative design, marketing visuals, e-commerce product images, and social content production.

icon
Copy ID
  • Entrada: $ 2 /M
  • Salida: $ 6 /M

Grok 4.5 was trained on datasets spanning knowledge in coding, science, engineering, and math. With both intelligent and efficient reasoning, Grok 4.5 excels at real engineering tasks and exceeds comparable leading models at these tasks.

Entrada:$ 0 /M
Salida:$ 0 /M
Contexto:128K
Latencia:-
Rendimiento:-

LFM-2.5-2.6B is a compact reasoning model developed by Liquid, featuring a generous 128,000 token context length. It is highly suited for agent workflows, data extraction, retrieval-augmented generation (RAG), and long-context processing. However, the developer advises against using this model for agentic coding tasks.

Scroll to load more

50 free models — no credit card required →

All models (849)

auto · coding-glm-5.3 · gemini-3.7-flash · dots-3-note-preview-free · gemini-3.7-flash-free · glm-5.2 · deepseek-v4-flash-0731 · deepseek-v4-pro-0813 · grok-4.6 · mai-thinking-1 · doubao-seedance-2-5-260628 · gpt-5.6-luna · gpt-5.6-sol · gpt-5.6-terra · agnes-2.5-flash · agnes-2.5-pro · agnes-2.5-pro-alpha · agnes-image-2.1-flash · grok-4.5 · lfm-2.5-2.6b-free · minimax-h3 · muse-glimmer-30b · nemotron-lightning-3.5-30b-a3b · qwen3.8-2.4t-a95b · claude-opus-5 · gemini-3.6-flash · ling-3.0-tiny-free · nemotron-3.5-lightning-free · qwen-image-3.0 · qwen-image-3.0-pro · qwen3.8-max · claude-sonnet-5 · kimi-k3 · ling-3.0-flash-free · muse-spark-1.2 · qwen3.8-max-preview · gemini-3.1-flash-lite-image · gemini-3.5-flash-lite · gemini-3.5-flash-lite-free · gemini-3.6-flash-free · glm-5.2-fast-preview · muse-spark-1.1 · qwen-audio-3.0-tts-flash · qwen-audio-3.0-tts-plus · claude-fable-5 · jina-reranker-v3.5 · claude-opus-4-8 · hy3 · doubao-seed-2-1-pro · doubao-seed-2-1-turbo · mai-image-2.5-pro · gemini-3.5-flash · grok-build-0.1 · mai-image-2.5 · mai-image-2.5-flash · coding-kimi-k3 · happyhorse-1.1-i2v · happyhorse-1.1-r2v · happyhorse-1.1-t2v · coding-glm-5.2-free · coding-kimi-k3-free · gemini-3.1-flash-image · gpt-oss-20b-free · kimi-k2.7-code · kimi-k2.7-code-highspeed · gemini-3-pro-image · gpt-4o-transcribe-diarize · gpt-audio-1.5 · hy-3d-3.1 · kling-v3-omni · kling-video-o1 · longcat-2.0 · nemotron-nano-9b-v2-free · hy3-preview · minimax-m3 · nemotron-nano-12b-v2-vl-free · qwen3.7-flash · qwen3.7-plus · step-3.7-flash · claude-opus-4-8-think · nemotron-3-super-120b-a12b-free · nemotron-3-nano-omni-30b-a3b-reasoning-free · nemotron-3-ultra-550b-a55b-free · qwen3.7-max · gpt-image-2 · nemotron-3.5-content-safety-free · coding-glm-5.2 · ernie-5.1 · gemini-3.1-flash-lite · gemini-3.1-flash-lite-nothink · grok-4.3 · happyhorse-1.0-i2v · happyhorse-1.0-r2v · happyhorse-1.0-t2v · happyhorse-1.0-video-edit · north-mini-code-free · gpt-5.5 · gpt-5.5-pro · laguna-xs-2.1-free · deepseek-v4-flash · deepseek-v4-pro · gemma-4-31b-it-free · command-a-plus-05-2026 · doubao-seedream-5.0-pro · ernie-5.0 · kimi-k2.6 · laguna-s-2.1-free · qwen3.6-max-preview · xiaomi-mimo-v2.5 · xiaomi-mimo-v2.5-pro · claude-opus-4-7 · claude-opus-4-7-think · gpt-chat-latest · nemotron-3-nano-30b-a3b-free · qwen3.6-27b · qwen3.6-35b-a3b · qwen3.6-flash · cohere-rerank-v4.0-fast · cohere-rerank-v4.0-pro · gemma-4-26b-a4b-it-free · grok-4-20-non-reasoning · grok-4-20-reasoning · qwen-image-2.0 · qwen-image-2.0-pro · coding-minimax-m3-free · doubao-seedance-2-0-260128 · doubao-seedance-2-0-fast-260128 · doubao-seedance-2-0-mini-260615 · glm-5.1 · glm-image · qwen3.6-plus · wan2.7-i2v · wan2.7-r2v · wan2.7-t2v · wan2.7-videoedit · cc-k2.6-code-preview · gemma-4-26b-a4b-it · gemma-4-31b-it · gpt-5.4 · wan2.7-image · wan2.7-image-pro · claude-sonnet-4-6 · coding-xiaomi-mimo-v2.5 · coding-xiaomi-mimo-v2.5-pro · doubao-seed-2-0-lite-260428 · doubao-seed-2-0-mini-260428 · gemini-3.1-flash-image-preview · gemini-3.1-pro-preview · gemini-3.1-pro-preview-customtools · gemini-3.1-pro-preview-search · gpt-5.4-mini · gpt-5.4-nano · gpt-5.5-free · qwen3.5-plus · claude-sonnet-4-6-think · coding-xiaomi-mimo-v2-omni · coding-xiaomi-mimo-v2-pro · gpt-5.3-chat-latest · gpt-5.3-codex · gpt-image-2-free · qwen3.5-122b-a10b · qwen3.5-27b · qwen3.5-35b-a3b · qwen3.5-397b-a17b · qwen3.5-flash · coding-glm-5.1 · doubao-seed-2-0-pro · gpt-5.4-high · gpt-5.4-low · gpt-5.4-pro · qwen3-coder-next · xiaomi-mimo-v2-omni-free · xiaomi-mimo-v2-pro-free · xiaomi-mimo-v2.5-free · xiaomi-mimo-v2.5-pro-free · claude-opus-4-6 · coding-glm-5.1-free · coding-minimax-m2.7-free · glm-5 · glm-5v-turbo · minimax-m2.7 · claude-opus-4-6-think · coding-glm-5-free · coding-glm-5-turbo-free · coding-minimax-m2.5-free · doubao-seed-2-0-code-preview · doubao-seed-2-0-lite-260215 · doubao-seed-2-0-mini · gemini-3-flash-preview · gemini-3-flash-preview-search · glm-5-turbo · cc-glm-5.1 · claude-opus-4-5 · claude-opus-4-5-think · embed-v-4-0 · ernie-image-turbo · gemini-3.1-flash-image-preview-free · mimo-v2-omni · mimo-v2-pro · cohere-command-a · gemini-3-flash-preview-free · cc-minimax-m3 · coding-minimax-m3 · gpt-4.1-free · gpt-4.1-mini-free · gpt-4.1-nano-free · gpt-4o-free · coding-glm-5 · coding-glm-5-turbo · glm-4.7 · veo-3.1-lite-generate-preview · glm-4.7-flash-free · coding-glm-4.7-free · doubao-seedance-1-5-pro-251215 · doubao-seedance-1-0-pro-250528 · doubao-seedance-1-0-pro-fast-251015 · gemini-3-pro-image-preview · gemini-embedding-2 · deepinfra-gemma-4-26b-a4b-it · gpt-5.2-codex · doubao-seedream-5.0-lite · gpt-image-1.5 · gpt-5.2 · gpt-5.2-chat-latest · gpt-5.2-high · gpt-5.2-low · gpt-5.2-pro · gpt-5.1 · gpt-5.1-codex-max · doubao-seed-1-8 · gpt-5.1-chat-latest · gpt-5.1-codex · gpt-5.1-codex-mini · claude-haiku-4-5 · claude-sonnet-4-5 · claude-sonnet-4-5-think · grok-4.20-multi-agent-0309 · mistral-large-3 · cc-glm-5 · cc-glm-5-turbo · cloudflare-glm-5.2 · gemini-2.5-flash-image · grok-4-1-fast-non-reasoning · grok-4-1-fast-reasoning · grok-code-fast-1 · k2.6-code-preview-free · mimo-v2-flash · musesteamer-air-image · qwen3.6-plus-preview-free · zai-glm-5-turbo · gpt-5 · deepseek-v3.2 · deepseek-v3.2-think · gpt-5-codex · DeepSeek-V3.1-Terminus · DeepSeek-V3.1-Think · gpt-5-pro · gpt-5-mini · gpt-5-nano · gpt-5-chat-latest · claude-opus-4-1 · o3-deep-research · kimi-k2.5 · qwen3-max-2026-01-23 · qwen3-vl-flash · qwen3-vl-flash-2026-01-22 · qwen3-vl-plus · cc-minimax-m2.7 · cc-minimax-m2.7-highspeed · minimax-m2.5 · minimax-m2.5-highspeed · mm-minimax-m2.7-highspeed · coding-minimax-m2.7 · coding-minimax-m2.7-highspeed · cc-minimax-m2.5 · cc-minimax-m2.5-highspeed · coding-minimax-m2.5 · coding-minimax-m2.5-highspeed · doubao-seedream-4-5 · sora-2 · sora-2-pro · cc-glm-4.7 · cc-minimax-m2.1 · coding-glm-4.7 · coding-minimax-m2.1 · coding-minimax-m2.1-free · gpt-4o-audio-preview · gpt-4o-mini-audio-preview · minimax-m2.1 · o3 · wan2.6-i2v · wan2.6-t2v · cc-glm-4.6 · coding-glm-4.6 · coding-glm-4.6-free · coding-minimax-m2 · coding-minimax-m2-free · flux-2-flex · flux-2-pro · gemini-2.5-pro · glm-4.6 · glm-4.6v · glm-ocr · kimi-for-coding-free · o3-pro · qianfan-ocr · qianfan-ocr-fast · step-3.5-flash · wan2.2-i2v-plus · wan2.5-i2v-preview · wan2.5-t2v-preview · gemini-2.5-pro-search · kimi-k2-thinking · gemini-2.5-flash · gemini-2.5-flash-preview-09-2025 · glm-4.5v · gemini-2.5-flash-lite · gemini-2.5-flash-lite-nothink · gemini-2.5-flash-lite-preview-09-2025 · gemini-2.5-flash-lite-preview-09-2025-nothink · gemini-2.5-flash-nothink · gemini-2.5-flash-search · gemini-2.5-flash-preview-05-20-nothink · gemini-2.5-flash-preview-05-20-search · DeepSeek-V3-Fast · imagen-4.0 · imagen-4.0-fast-generate-001 · imagen-4.0-generate-001 · imagen-4.0-ultra-generate-001 · imagen-4.0-ultra · gpt-image-1 · gpt-image-1-mini · o4-mini · DeepSeek-OCR · alicloud-kimi-k2-instruct · deepseek-ocr · ernie-5.0-thinking-exp · flux-kontext-max · gemini-2.5-flash-image-preview · glm-4.5 · gpt-4.1 · grok-4 · grok-4-fast-non-reasoning · grok-4-fast-reasoning · kimi-k2-0711 · kimi-k2-instruct · kimi-k2-turbo-preview · paddleocr-vl-0.9b · pp-structurev3 · qwen3-vl-235b-a22b-instruct · qwen3-vl-235b-a22b-thinking · qwen3-vl-30b-a3b-instruct · qwen3-vl-30b-a3b-thinking · veo-3.0-generate-preview · veo-3.1-fast-generate-preview · veo-3.1-generate-preview · aihubmix-router · gpt-4.1-mini · gpt-4.1-nano · gemini-2.5-pro-preview-05-06 · gemini-2.5-pro-preview-03-25 · gemini-2.5-pro-preview-05-06-search · gemini-2.5-pro-preview-03-25-search · qwen3-max-preview · qwen3-max · qwen3-next-80b-a3b-instruct · qwen3-next-80b-a3b-thinking · qwen3-235b-a22b-instruct-2507 · qwen3-235b-a22b-thinking-2507 · qwen3-coder-30b-a3b-instruct · qwen3-coder-480b-a35b-instruct · DeepSeek-V3 · LongCat-Flash-Chat · gemini-2.5-pro-preview-06-05-search · jina-embeddings-v5-text-nano · jina-embeddings-v5-text-small · qwen3-235b-a22b · qwen3-coder-flash · qwen3-coder-plus · qwen3-coder-plus-2025-07-22 · Qwen2.5-VL-72B-Instruct · ernie-5.0-thinking-preview · inclusionAI/Ling-1T · inclusionAI/Ring-1T · bce-reranker-base · codex-mini-latest · doubao-seedream-4-0 · embedding-v1 · ernie-4.5-turbo-latest · glm-4.5-x · gme-qwen2-vl-2b-instruct · gte-rerank-v2 · inclusionAI/Ling-flash-2.0 · inclusionAI/Ling-mini-2.0 · inclusionAI/Ring-flash-2.0 · jina-deepsearch-v1 · jina-embeddings-v4 · jina-reranker-v3 · llama-4-maverick · llama-4-scout · qwen-image · qwen-image-edit · qwen-image-max · qwen-mt-plus · qwen-mt-turbo · qwen3-embedding-0.6b · qwen3-embedding-4b · qwen3-embedding-8b · qwen3-reranker-0.6b · qwen3-reranker-4b · qwen3-reranker-8b · tao-8k · jina-clip-v2 · jina-reranker-m0 · jina-colbert-v2 · DeepSeek-R1 · gpt-4o-search-preview · gpt-4o-mini-search-preview · jina-embeddings-v3 · claude-3-7-sonnet · ernie-4.5 · ernie-4.5-turbo-vl · mimo-v2-flash-free · FLUX-1.1-pro · o3-mini · doubao-seed-1-6 · doubao-seed-1-6-flash · doubao-seed-1-6-lite · doubao-seed-1-6-thinking · qwen3-30b-a3b-instruct-2507 · qwen3-30b-a3b-thinking-2507 · Qwen2-VL-72B-Instruct · Qwen2-VL-7B-Instruct · cc-kimi-for-coding · gemini-embedding-001 · gpt-oss-120b · qwen-3-235b-a22b-thinking-2507 · Qwen/Qwen3-30B-A3B · Qwen/Qwen3-32B · qwen3-32b · Qwen/Qwen3-14B · Qwen/Qwen3-8B · embedding-2 · embedding-3 · gemini-2.5-pro-preview-06-05 · Qwen/Qwen2.5-VL-72B-Instruct · o1 · o1-pro · ByteDance-Seed/Seed-OSS-36B-Instruct · doubao-seed-1-6-250615 · doubao-seed-1-6-flash-250615 · doubao-seed-1-6-thinking-250615 · doubao-seed-1-6-vision-250815 · Doubao-1.5-thinking-pro · cc-minimax-m2 · deepseek-ai/DeepSeek-Prover-V2-671B · gemini-2.5-flash-preview-tts · gemini-2.5-pro-preview-tts · gemma-3-12b-it · gemma-3-27b-it · gemma-3-4b-it · gemma-3n-e4b-it · gemma-3-1b-it · deepseek-r1-distill-llama-70b · gpt-4o-mini-tts · tngtech/DeepSeek-R1T-Chimera · veo-2.0-generate-001 · o1-preview · o1-mini · gpt-4o-2024-11-20 · gpt-4o · gpt-4o-mini · AiHubmix-mistral-medium · ERNIE-X1.1-Preview · Qwen/QwQ-32B · chutesai/Mistral-Small-3.1-24B-Instruct-2503 · ernie-x1.1-preview · minimax-m2 · MiniMaxAI/MiniMax-M1-80k · Qwen/Qwen2.5-VL-32B-Instruct · baidu/ERNIE-4.5-300B-A47B · bge-large-en · bge-large-zh · codestral-latest · ernie-4.5-0.3b · ernie-4.5-turbo-128k-preview · ernie-x1-turbo · kat-dev · llama-3.3-70b · moonshotai/Kimi-Dev-72B · moonshotai/Moonlight-16B-A3B-Instruct · nvidia-nemotron-3-super-120b-a12b · o1-global · qianfan-qi-vl · qwen2.5-vl-72b-instruct · tencent/Hunyuan-A13B-Instruct · unsloth/gemma-3-27b-it · gemini-exp-1206 · gpt-4o-zh · qwen-qwq-32b · unsloth/gemma-3-12b-it · qwen-max-0125 · BAAI/bge-large-en-v1.5 · BAAI/bge-large-zh-v1.5 · BAAI/bge-reranker-v2-m3 · tencent/Hunyuan-MT-7B · V3 · V_2 · V_2_TURBO · V_2A · V_2A_TURBO · V_1 · V_1_TURBO · doubao-embedding-large-text-240915 · kimi-thinking-preview · gpt-4o-2024-08-06 · qwen-plus-2025-07-28 · qwen-plus-latest · sonar · stepfun-ai/step3 · text-embedding-v4 · AiHubmix-Phi-4-mini-reasoning · qwen-turbo-latest · aihub-Phi-4-multimodal-instruct · qwen3-30b-a3b · aihub-Phi-4-mini-instruct · grok-3 · aihub-Phi-4 · claude-3-opus-20240229 · dall-e-3 · doubao-embedding-text-240715 · grok-3-beta · qwen3-14b · grok-3-fast · qwen3-8b · deepseek-ai/DeepSeek-R1-Zero · grok-3-fast-beta · grok-3-mini · qwen3-4b · grok-3-mini-beta · qwen3-1.7b · qwen3-0.6b · alicloud-glm-5 · command-a-03-2025 · grok-3-mini-fast-beta · qwen-3-32b · qwen-turbo-2025-04-28 · qwen-plus-2025-04-28 · THUDM/GLM-Z1-32B-0414 · THUDM/GLM-4.1V-9B-Thinking · text-embedding-004 · THUDM/GLM-4-32B-0414 · THUDM/GLM-Z1-9B-0414 · THUDM/GLM-4-9B-0414 · cc-doubao-seed-code-preview-latest · doubao-seed-code-preview-latest · deepseek-ai/Janus-Pro-7B · glm-zero-preview · qwen-3-235b-a22b-instruct-2507 · coding-glm-4.5-air · deepinfra-nvidia-nemotron-3-nano-30b-a3b2 · glm-4.5-air · gpt-4-32k · nvidia-llama-3.1-nemotron-70b-instruct · nvidia-llama-3.3-nemotron-super-49b-v1.5 · nvidia-nemotron-3-nano-30b-a3b · nvidia-nemotron-nano-12b-v2-vl · nvidia-nemotron-nano-9b-v2 · o1-preview-2024-09-12 · Qwen/QVQ-72B-Preview · Qwen/QwQ-32B-Preview · llama-3.1-sonar-huge-128k-online · aihubmix-Mistral-Large-2411 · llama-3.1-sonar-large-128k-online · aihubmix-Mistral-large-2407 · grok-2-1212 · llama-3.1-70b · wan2.6-t2i · DESCRIBE · UPSCALE · bai-qwen3-vl-235b-a22b-instruct · cc-MiniMax-M2 · cc-deepseek-v3 · cc-deepseek-v3.1 · cc-ernie-4.5-300b-a47b · cc-kimi-dev-72b · cc-kimi-k2-instruct · cc-kimi-k2-instruct-0905 · cc-kimi-k2-thinking · computer-use-preview · gpt-image-test · grok-4.20-beta-0309-non-reasoning · grok-4.20-beta-0309-reasoning · grok-4.20-multi-agent-beta-0309 · jina-reader · jina-search · llama3.1-8b · o1-2024-12-17 · sf-kimi-k2-thinking · Baichuan3-Turbo · Baichuan3-Turbo-128k · Baichuan4 · Baichuan4-Air · Baichuan4-Turbo · DeepSeek-v3 · Doubao-1.5-lite-32k · Doubao-1.5-pro-256k · Doubao-1.5-pro-32k · Doubao-1.5-vision-pro-32k · Doubao-lite-128k · Doubao-lite-32k · Doubao-lite-4k · Doubao-pro-128k · Doubao-pro-256k · Doubao-pro-32k · Doubao-pro-4k · GPT-OSS-20B · Gryphe/MythoMax-L2-13b · MiniMax-Text-01 · Mistral-large-2407 · Qwen/Qwen2-1.5B-Instruct · Qwen/Qwen2-57B-A14B-Instruct · Qwen/Qwen2-72B-Instruct · Qwen/Qwen2-7B-Instruct · Qwen/Qwen2.5-32B-Instruct · Qwen/Qwen2.5-72B-Instruct · Qwen/Qwen2.5-72B-Instruct-128K · Qwen/Qwen2.5-7B-Instruct · Qwen/Qwen2.5-Coder-32B-Instruct · Qwen3-235B-A22B-Thinking-2507 · Stable-Diffusion-3-5-Large · WizardLM/WizardCoder-Python-34B-V1.0 · ahm-Phi-3-5-MoE-instruct · ahm-Phi-3-5-mini-instruct · ahm-Phi-3-5-vision-instruct · ahm-Phi-3-medium-128k · ahm-Phi-3-medium-4k · ahm-Phi-3-small-128k · aihubmix-Codestral-2501 · aihubmix-Cohere-command-r · aihubmix-Jamba-1-5-Large · aihubmix-Llama-3-1-405B-Instruct · aihubmix-Llama-3-1-70B-Instruct · aihubmix-Llama-3-1-8B-Instruct · aihubmix-Llama-3-2-11B-Vision · aihubmix-Llama-3-2-90B-Vision · aihubmix-Llama-3-70B-Instruct · aihubmix-Mistral-large · aihubmix-command-r-08-2024 · aihubmix-command-r-plus · aihubmix-command-r-plus-08-2024 · alicloud-deepseek-v3.2 · alicloud-glm-4.7 · alicloud-kimi-k2-thinking · alicloud-kimi-k2.5 · alicloud-minimax-m2.5 · anthropic-opus-4-6 · azure-deepseek-v3.2 · azure-deepseek-v3.2-speciale · azure-kimi-k2.5 · cbs-glm-4.7 · cerebras-llama-3.3-70b · chatglm_lite · chatglm_pro · chatglm_std · chatglm_turbo · claude-2 · claude-2.0 · claude-2.1 · claude-3-haiku-20240229 · claude-3-haiku-20240307 · claude-3-sonnet-20240229 · claude-instant-1 · claude-instant-1.2 · code-davinci-edit-001 · cogview-3 · cogview-3-plus · command · command-light · command-light-nightly · command-nightly · command-r · command-r-08-2024 · command-r-plus · command-r-plus-08-2024 · dall-e-2 · davinci · davinci-002 · deepinfra-llama-3.1-8b-instant · deepinfra-llama-3.3-70b-instant-turbo · deepinfra-llama-4-maverick-17b-128e-instruct · deepinfra-llama-4-scout-17b-16e-instruct · deepseek-ai/DeepSeek-Coder-V2-Instruct · deepseek-ai/DeepSeek-R1-Distill-Llama-70B · deepseek-ai/DeepSeek-R1-Distill-Llama-8B · deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B · deepseek-ai/DeepSeek-R1-Distill-Qwen-14B · deepseek-ai/DeepSeek-R1-Distill-Qwen-32B · deepseek-ai/DeepSeek-R1-Distill-Qwen-7B · deepseek-ai/DeepSeek-V2-Chat · deepseek-ai/DeepSeek-V2.5 · deepseek-ai/deepseek-llm-67b-chat · deepseek-ai/deepseek-vl2 · deepseek-v3 · distil-whisper-large-v3-en · doubao-1-5-thinking-vision-pro-250428 · fx-flux-2-pro · gemini-2.5-pro-exp-03-25 · gemini-embedding-exp-03-07 · gemini-exp-1114 · gemini-exp-1121 · gemini-pro · gemini-pro-vision · gemma-7b-it · glm-3-turbo · glm-4 · glm-4-flash · glm-4-plus · glm-4.5-airx · glm-4v · glm-4v-plus · google-gemma-3-12b-it · google-gemma-3-27b-it · google-gemma-3-4b-it · google/gemini-exp-1114 · google/gemma-2-27b-it · google/gemma-2-9b-it:free · gpt-3.5-turbo · gpt-3.5-turbo-0301 · gpt-3.5-turbo-0613 · gpt-3.5-turbo-1106 · gpt-3.5-turbo-16k · gpt-3.5-turbo-16k-0613 · gpt-3.5-turbo-instruct · gpt-4 · gpt-4-0125-preview · gpt-4-0314 · gpt-4-0613 · gpt-4-1106-preview · gpt-4-32k-0314 · gpt-4-32k-0613 · gpt-4-turbo · gpt-4-turbo-2024-04-09 · gpt-4-turbo-preview · gpt-4-vision-preview · gpt-4o-2024-05-13 · gpt-4o-mini-2024-07-18 · gpt-oss-20b · grok-2-vision-1212 · grok-vision-beta · groq-llama-3.1-8b-instant · groq-llama-3.3-70b-versatile · groq-llama-4-maverick-17b-128e-instruct · groq-llama-4-scout-17b-16e-instruct · jina-embeddings-v2-base-code · learnlm-1.5-pro-experimental · llama-3.1-405b-instruct · llama-3.1-405b-reasoning · llama-3.1-70b-versatile · llama-3.1-8b-instant · llama-3.1-sonar-small-128k-online · llama-3.2-11b-vision-preview · llama-3.2-1b-preview · llama-3.2-3b-preview · llama-3.2-90b-vision-preview · llama2-70b-4096 · llama2-70b-40960 · llama2-7b-2048 · llama3-70b-8192 · llama3-8b-8192 · llama3-groq-70b-8192-tool-use-preview · llama3-groq-8b-8192-tool-use-preview · meta-llama/Llama-3.2-90B-Vision-Instruct · meta-llama/llama-3.1-405b-instruct:free · meta-llama/llama-3.1-70b-instruct:free · meta-llama/llama-3.1-8b-instruct:free · meta-llama/llama-3.2-11b-vision-instruct:free · meta-llama/llama-3.2-3b-instruct:free · meta/llama-3.1-405b-instruct · meta/llama3-8B-chat · mistralai/mistral-7b-instruct:free · moonshot-kimi-k2.5 · moonshot-v1-128k · moonshot-v1-128k-vision-preview · moonshot-v1-32k · moonshot-v1-32k-vision-preview · moonshot-v1-8k · moonshot-v1-8k-vision-preview · nvidia/Llama-3_1-Nemotron-Ultra-253B-v1 · o1-mini-2024-09-12 · omni-moderation-latest · qwen-flash · qwen-flash-2025-07-28 · qwen-long · qwen-max · qwen-max-longcontext · qwen-plus · qwen-turbo · qwen-turbo-2024-11-01 · qwen2.5-14b-instruct · qwen2.5-32b-instruct · qwen2.5-3b-instruct · qwen2.5-72b-instruct · qwen2.5-7b-instruct · qwen2.5-coder-1.5b-instruct · qwen2.5-coder-7b-instruct · qwen2.5-math-1.5b-instruct · qwen2.5-math-72b-instruct · qwen2.5-math-7b-instruct · step-2-16k · text-ada-001 · text-babbage-001 · text-curie-001 · text-davinci-002 · text-davinci-003 · text-davinci-edit-001 · text-embedding-3-large · text-embedding-3-small · text-embedding-ada-002 · text-embedding-v1 · text-moderation-007 · text-moderation-latest · text-moderation-stable · text-search-ada-doc-001 · tts-1 · tts-1-1106 · tts-1-hd · tts-1-hd-1106 · whisper-1 · whisper-large-v3 · whisper-large-v3-turbo · yi-large · yi-large-rag · yi-large-turbo · yi-lightning · yi-medium · yi-vl-plus · deepseek-r1-distill-qianfan-llama-8b · doubao-1-5-pro-256k-250115 · doubao-1-5-pro-32k-250115 · gpt-4o-2024-08-06-global · gpt-4o-mini-global · meta-llama-3-70b · meta-llama-3-8b · o3-global · o3-mini-global · o3-pro-global · qianfan-chinese-llama-2-13b · qianfan-llama-vl-8b