ling-3.0-flash-free

by Inclusionai

Developed by Inclusionai, ling-3.0-flash-free is a 124B-parameter Mixture-of-Experts (MoE) model with approximately 5.1B parameters activated per token. Featuring an expansive context length of 262,144 tokens, this model is built to handle extensive datasets and long-form content. It is designed with token efficiency and production-scale agentic inference as key priorities to enable seamless developer deployment.

Specifications

Context window262,144 tokens
Modalitiestext
Featuresreasoning, tool_calling, long_context
Endpointschat_completions

Use ling-3.0-flash-free via the AIHubMix unified API — one interface for every major LLM.