gemma-4-26b-a4b-it-free

by Google

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model developed by Google DeepMind. Featuring an expansive context length of 262,144 tokens, it delivers near-31B quality with highly efficient inference. Despite its 25.2B total parameters, only 3.8B are activated per token, making it an incredibly fast and cost-effective solution.

Specifications

Context window262,144 tokens
Modalitiestext, image
Featuresreasoning, tool_calling, long_context
Endpointschat_completions

Paid version: gemma-4-26b-a4b-it

More from Google

Use gemma-4-26b-a4b-it-free via the AIHubMix unified API — one interface for every major LLM.