Providers
Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.
The model record exists, but no source-linked provider offer is available yet.
Submit a sourceNVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.
The model record exists, but no source-linked provider offer is available yet.
Submit a sourceRecorded from the source catalog and provider listings.
nvidia/nemotron-3.5-lightning:freeEvery result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.
ModelBench does not infer quality from price, context size, or model name. When a source publishes a comparable result, it appears here with its version and link.
Read the methodologyField-level changes detected between successful source imports.
This record has not changed within the retained import history.
Every figure on this page traces back to one of these records.
Fetch the complete source-linked model record. No key, no account, no rate-limited tier.
GET https://www.modelbench.lol/api/v1/models/nvidia/nemotron-3.5-lightning:freecurl "https://www.modelbench.lol/api/v1/models/nvidia/nemotron-3.5-lightning:free"Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that. It is published by NVIDIA and catalogued here from OpenRouter.
This model is documented as free to use at the listed providers.
NVIDIA: Nemotron 3.5 Lightning (free) accepts up to 1M tokens of context and returns up to 65.536K output tokens.
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
NVIDIA: Nemotron 3 Ultra (free)NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Nemotron 3 Ultra 550B A55BLargest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
NVIDIA: Nemotron 3 Nano Omni (free)NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...