by Nvidia
NVIDIA Nemotron 3 Nano 30B A3B is a highly efficient small language Mixture of Experts (MoE) model developed by Nvidia. Designed to help developers build specialized agentic AI systems, it delivers exceptional compute efficiency and accuracy. Additionally, it features an impressive context length of 256,000 tokens to support extensive data processing.
| Context window | 256,000 tokens |
| Modalities | text |
| Features | reasoning, tool_calling, long_context |
| Endpoints | chat_completions |
Use nemotron-3-nano-30b-a3b-free via the AIHubMix unified API — one interface for every major LLM.