Back to nemotron
NVIDIA /

Llama Nemotron Rerank VL 1B v2
Llama Nemotron Rerank VL 1B v2

Open WeightsMultimodaltextimageUpdated March 31, 2026
Some details or benchmark scores on this page are self-reported by developers and unconfirmed.

Llama Nemotron Rerank VL 1B v2

📌 Model Overview

Reranking model for improving retrieval quality in search and recommendation systems

Llama Nemotron Rerank VL 1B v2 is a Open Weights model developed by NVIDIA, released on 2026-03-31. It is designed primarily for Multimodal General workloads. Featuring a 128K tokens context window and 1B parameter count, it offers robust performance for enterprise integration, developers, and researchers.


✨ Key Features & Capabilities

Specification Table
FeatureDescription
Context Window128K tokens capacity for extended prompts and multi-turn workflows
Primary TaskOptimized for Multimodal General
Deploymentself-hostable, api-only
Modalitytext, image
Max OutputMax Output: 4K tokens

⚙️ Technical Specifications

Specification Table
SpecificationDetails
Developer / LabNVIDIA
Release Date2026-03-31
Model TypeOpen Weights
Parameters1B
Context Window128K tokens
Licenseproprietary
Model Familynemotron

💰 Pricing

Specification Table
Tier / UnitRate (USD)
1M input tokens$0
1M output tokens$0

📜 License & Usage

This model is governed by the proprietary license. Please check official developer guidelines before commercial deployment.

Key Features

Max Output: 4K tokens

Feature 01

You might also want to compare

Verified Sources

Tags

open-weights

Model Specs

open-weights

Parameters

1B

Context Window

128K tokens

License

proprietary

Deployment

self-hostableapi-only

Pricing

  • 1M input tokens$0
  • 1M output tokens$0

as of 2026-07-26

Lineage

Curator Notes

Skeleton entry imported from models.dev backfill (2026-07-26). Needs manual enrichment and primary-source verification.

Compare Specs

Compare parameters, context windows, modalities, and benchmark scores of this model side-by-side with others.

Compare Model