nemotron-nano-12b-v2-vl-free

by Nvidia

Developed by Nvidia, Nemotron-Nano-12B-V2-VL-Free is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It features a powerful 128,000-token context length to handle extensive visual and textual inputs. The model utilizes a hybrid Transformer-Mamba architecture, successfully combining transformer-level accuracy with Mamba's structural advantages.

Specifications

Context window128,000 tokens
Modalitiestext, image
Featuresreasoning, tool calling, long context
Endpointschat_completions

FAQ

What is nemotron-nano-12b-v2-vl-free?

Developed by Nvidia, Nemotron-Nano-12B-V2-VL-Free is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It features a powerful 128,000-token context length to handle extensive visual and textual inputs. The model utilizes a hybrid Transformer-Mamba architecture, successfully combining transformer-level accuracy with Mamba's structural advantages.

What is the context length of nemotron-nano-12b-v2-vl-free?

nemotron-nano-12b-v2-vl-free has a 128,000 token context window.

What modalities does nemotron-nano-12b-v2-vl-free support?

nemotron-nano-12b-v2-vl-free accepts text and image input.

What features does nemotron-nano-12b-v2-vl-free support?

nemotron-nano-12b-v2-vl-free supports reasoning, tool calling and long context. Per-protocol parameter support is listed in the capability table on this page.

How do I call nemotron-nano-12b-v2-vl-free via API?

nemotron-nano-12b-v2-vl-free is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to nemotron-nano-12b-v2-vl-free — no other code changes needed.

Who develops nemotron-nano-12b-v2-vl-free?

nemotron-nano-12b-v2-vl-free is developed by Nvidia. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

More from Nvidia

nemotron-nano-9b-v2-free

by Nvidia

NVIDIA-Nemotron-Nano-9B-v2-free is a large language model trained from scratch by NVIDIA…

128,000 tokens context

nemotron-3-super-120b-a12b-free

by Nvidia

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model built on a hybrid…

262,144 tokens context

nemotron-3-nano-omni-30b-a3b-reasoning-free

by Nvidia

Developed by Nvidia, NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model…

256,000 tokens context

nemotron-3-ultra-550b-a55b-free

by Nvidia

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model featuring…

1,000,000 tokens context

nemotron-3.5-content-safety-free

by Nvidia

Developed by NVIDIA, Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal…

128,000 tokens context

nemotron-3-nano-30b-a3b-free

by Nvidia

NVIDIA Nemotron 3 Nano 30B A3B is a highly efficient small language Mixture of Experts…

256,000 tokens context

Use nemotron-nano-12b-v2-vl-free via the AIHubMix unified API — one interface for every major LLM.