nvidia-nemotron-3-super-120b-a12b

by Nvidia

An open-source, efficient hybrid Mamba-Transformer MoE model that supports a context length of one million tokens and excels at agent reasoning, programming, planning, and tool invocation.

API Pricing

Input$0.11 / 1M tokens
Output$0.55 / 1M tokens
Cache read$0.03 / 1M tokens

Specifications

Context window1,000,000 tokens
Modalitiestext
Featuresthinking, tool calling, function calling, structured outputs, long context

FAQ

What is nvidia-nemotron-3-super-120b-a12b?

An open-source, efficient hybrid Mamba-Transformer MoE model that supports a context length of one million tokens and excels at agent reasoning, programming, planning, and tool invocation.

What is the context length of nvidia-nemotron-3-super-120b-a12b?

nvidia-nemotron-3-super-120b-a12b has a 1,000,000 token context window.

How much does nvidia-nemotron-3-super-120b-a12b cost?

On AIHubMix, nvidia-nemotron-3-super-120b-a12b costs $0.11 per million input tokens and $0.55 per million output tokens. Cached input reads are billed at $0.03 per million tokens.

What modalities does nvidia-nemotron-3-super-120b-a12b support?

nvidia-nemotron-3-super-120b-a12b accepts text input.

What features does nvidia-nemotron-3-super-120b-a12b support?

nvidia-nemotron-3-super-120b-a12b supports thinking, tool calling, function calling, structured outputs and long context. Per-protocol parameter support is listed in the capability table on this page.

How do I call nvidia-nemotron-3-super-120b-a12b via API?

nvidia-nemotron-3-super-120b-a12b is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to nvidia-nemotron-3-super-120b-a12b — no other code changes needed.

Who develops nvidia-nemotron-3-super-120b-a12b?

nvidia-nemotron-3-super-120b-a12b is developed by Nvidia. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More from Nvidia

nemotron-nano-9b-v2-free

by Nvidia

NVIDIA-Nemotron-Nano-9B-v2-free is a large language model trained from scratch by NVIDIA…

128,000 tokens context

nemotron-nano-12b-v2-vl-free

by Nvidia

Developed by Nvidia, Nemotron-Nano-12B-V2-VL-Free is a 12-billion-parameter open…

128,000 tokens context

nemotron-3-super-120b-a12b-free

by Nvidia

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model built on a hybrid…

262,144 tokens context

nemotron-3-nano-omni-30b-a3b-reasoning-free

by Nvidia

Developed by Nvidia, NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model…

256,000 tokens context

nemotron-3-ultra-550b-a55b-free

by Nvidia

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model featuring…

1,000,000 tokens context

nemotron-3.5-content-safety-free

by Nvidia

Developed by NVIDIA, Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal…

128,000 tokens context

Use nvidia-nemotron-3-super-120b-a12b via the AIHubMix unified API — one interface for every major LLM.