by Nvidia
NVIDIA Nemotron 3 Nano 30B A3B is a highly efficient small language Mixture of Experts (MoE) model developed by Nvidia. Designed to help developers build specialized agentic AI systems, it delivers exceptional compute efficiency and accuracy. Additionally, it features an impressive context length of 256,000 tokens to support extensive data processing.
NVIDIA Nemotron 3 Nano 30B A3B is a highly efficient small language Mixture of Experts (MoE) model developed by Nvidia. Designed to help developers build specialized agentic AI systems, it delivers exceptional compute efficiency and accuracy. Additionally, it features an impressive context length of 256,000 tokens to support extensive data processing.
nemotron-3-nano-30b-a3b-free has a 256,000 token context window.
nemotron-3-nano-30b-a3b-free accepts text input.
nemotron-3-nano-30b-a3b-free supports reasoning, tool calling and long context. Per-protocol parameter support is listed in the capability table on this page.
nemotron-3-nano-30b-a3b-free is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to nemotron-3-nano-30b-a3b-free — no other code changes needed.
nemotron-3-nano-30b-a3b-free is developed by Nvidia. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
NVIDIA-Nemotron-Nano-9B-v2-free is a large language model trained from scratch by NVIDIA…
Developed by Nvidia, Nemotron-Nano-12B-V2-VL-Free is a 12-billion-parameter open…
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model built on a hybrid…
Developed by Nvidia, NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model…
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model featuring…
Developed by NVIDIA, Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal…
Use nemotron-3-nano-30b-a3b-free via the AIHubMix unified API — one interface for every major LLM.