NVIDIA logo

Llama Nemotron Rerank VL 1B v2

Reranking model for improving retrieval quality in search and recommendation systems

Source-linked nemotron Open weights Released 2026-03-31
API data Report
Input modalitiesT
Output modalitiesT
Input priceNot listed
Output priceNot listed
Context window128K
Max output4.096K
Overview

Capabilities and specifications

× Reasoning No
× Tool calling No
? Structured output Unknown
Attachments Yes
Vision input Yes
Open weights Yes
Creator
NVIDIA
Model family
nemotron
Knowledge cutoff
Not documented
License
Not documented
Release date
2026-03-31
Model ID
nvidia/llama-nemotron-rerank-vl-1b-v2
Inference availability

Provider pricing and availability

Provider-specific identifiers, limits, and prices per million tokens.

Report a price
Providers offering Llama Nemotron Rerank VL 1B v2
ProviderProvider model IDContextInputOutputCache readCapabilitiesSource
Nvidia nvidia/llama-nemotron-rerank-vl-1b-v2 128K Not listed Not listed - Not listed Docs

Capability badges appear only when the provider catalog explicitly lists support.

Serving quality

Performance and reliability

Observed serving measurements stay separate from model specifications and are never estimated from price or model size.

Throughput Not reported

Generation speed; higher is faster.

Latency / TTFT Not reported

Time until the first output token arrives.

Uptime Not reported

Successful requests over a measured period.

Performance varies by inference provider, region, and load. Provider-level measurements appear only when a source reports them.

Published evaluations

Benchmarks

Raw results remain attached to their original source and version.

Benchmark registry
No comparable benchmark results

ModelBench does not infer quality from price, context size, or model name.

Read the methodology
Catalog activity

Recent record changes

Detected changes from successful source imports.

Full change log
No recent changes detected

This record has not changed in the retained import history.

Public API

Use this record

Fetch the complete source-linked model record without an API key.

API documentation
Endpoint
GET https://www.modelbench.lol/api/v1/models/nvidia/llama-nemotron-rerank-vl-1b-v2
curl
curl "https://www.modelbench.lol/api/v1/models/nvidia/llama-nemotron-rerank-vl-1b-v2"