Nemotron 3 Super 120B A12B (free)

nemotron-3-super-120b-a12b-free · Nvidia

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model built on a hybrid Mamba-Transformer architecture. Activating just 12B parameters, it delivers maximum compute efficiency and accuracy for complex multi-agent applications. Additionally, it features an expansive context length of 262,144 tokens to handle large-scale inputs seamlessly.

Specifications

Context1.05M tokens
Modalitiestext
CapabilitiesThinking, Tool calling, Structured outputs
Endpointschat_completions

Frequently asked questions

What is Nemotron 3 Super 120B A12B (free)?

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model built on a hybrid Mamba-Transformer architecture. Activating just 12B parameters, it delivers maximum compute efficiency and accuracy for complex multi-agent applications. Additionally, it features an expansive context length of 262,144 tokens to handle large-scale inputs seamlessly.

What is the context length of Nemotron 3 Super 120B A12B (free)?

Nemotron 3 Super 120B A12B (free) has a 1,048,576 token context window.

What modalities does Nemotron 3 Super 120B A12B (free) support?

Nemotron 3 Super 120B A12B (free) accepts text input.

What capabilities does Nemotron 3 Super 120B A12B (free) support?

Nemotron 3 Super 120B A12B (free) supports reasoning, tool calling and long context. Per-protocol parameter support is listed in the capability table on this page.

How do I call Nemotron 3 Super 120B A12B (free) via API?

Nemotron 3 Super 120B A12B (free) is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to nemotron-3-super-120b-a12b-free — no other code changes needed.

Who created Nemotron 3 Super 120B A12B (free)?

Nemotron 3 Super 120B A12B (free) is developed by Nvidia. AIHubMix aggregates it alongside models from other providers behind one API and one bill.

When was Nemotron 3 Super 120B A12B (free) released?

Nemotron 3 Super 120B A12B (free) was released on March 11, 2026 by Nvidia.

More models from Nvidia

See all Nvidia models →

Nemotron Lightning 3.5 30B A3B

by Nvidia

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, featuring…

$0.05/1M in · $0.2/1M out
262,000 tokens context

Nemotron 3.5 Lightning (free)

by Nvidia

NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, featuring…

1,000,000 tokens context

Nemotron Nano 9B V2 (free)

by Nvidia

NVIDIA-Nemotron-Nano-9B-v2-free is a large language model trained from scratch by NVIDIA…

128,000 tokens context

Nemotron Nano 12B V2 VL (free)

by Nvidia

Developed by Nvidia, Nemotron-Nano-12B-V2-VL-Free is a 12-billion-parameter open…

128,000 tokens context

Nemotron 3 Nano Omni 30B A3B (reasoning) (free)

by Nvidia

Developed by Nvidia, NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model…

256,000 tokens context

Nemotron 3 Ultra 550B A55B (free)

by Nvidia

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model featuring…

1,000,000 tokens context

Use Nemotron 3 Super 120B A12B (free) via the AIHubMix unified API — one interface for every major LLM.