nemotron-lightning-3.5-30b-a3b · Nvidia
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, featuring 3 billion active parameters out of 30 billion total. It supports an extensive context window of up to 1,000,000 tokens. This model is designed to deliver efficient performance for high-throughput agentic workloads and specialized tasks. Compared with Nemotron 3 Super and Ultra, Lightning is smaller and optimized for fast, high-volume execution, while the larger models focus on complex planning and advanced reasoning.
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, featuring 3 billion active parameters out of 30 billion total. It supports an extensive context window of up to 1,000,000 tokens. This model is designed to deliver efficient performance for high-throughput agentic workloads and specialized tasks. Compared with Nemotron 3 Super and Ultra, Lightning is smaller and optimized for fast, high-volume execution, while the larger models focus on complex planning and advanced reasoning.
Nemotron Lightning 3.5 30B A3B has a 1,048,576 token context window.
On AIHubMix, Nemotron Lightning 3.5 30B A3B costs $0.05 per million input tokens and $0.2 per million output tokens. Cached input reads are billed at $0.01 per million tokens.
Nemotron Lightning 3.5 30B A3B accepts text input.
Nemotron Lightning 3.5 30B A3B supports reasoning, tool calling and long context. Per-protocol parameter support is listed in the capability table on this page.
Nemotron Lightning 3.5 30B A3B is available through the AIHubMix unified API. The API is OpenAI-compatible: point your OpenAI SDK at https://aihubmix.com/v1, use your AIHubMix API key, and set the model name to nemotron-lightning-3.5-30b-a3b — no other code changes needed.
Nemotron Lightning 3.5 30B A3B is developed by Nvidia. AIHubMix aggregates it alongside models from other providers behind one API and one bill.
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, featuring…
NVIDIA-Nemotron-Nano-9B-v2-free is a large language model trained from scratch by NVIDIA…
Developed by Nvidia, Nemotron-Nano-12B-V2-VL-Free is a 12-billion-parameter open…
NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model built on a hybrid…
Developed by Nvidia, NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model…
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model featuring…
Use Nemotron Lightning 3.5 30B A3B via the AIHubMix unified API — one interface for every major LLM.