# Llama Nemotron Embed VL 1B v2 Usage, Cost & Rank | OpenCode Data

> Llama Nemotron Embed VL 1B v2 costs $0.00 per 1M input tokens and $0.00 per 1M output tokens.

Embedding model for semantic search, retrieval, clustering, and ranking pipelines

- Page: https://opencode.ai/data/nvidia/llama-nemotron-embed-vl-1b-v2
- JSON: https://opencode.ai/data/nvidia/llama-nemotron-embed-vl-1b-v2.json
- Model ID: nvidia/llama-nemotron-embed-vl-1b-v2
- Lab: NVIDIA

## Model facts

| Fact | Value |
| --- | --- |
| Context window | 33K |
| Max output | 2K |
| Knowledge cutoff | - |
| Release date | 2026-02-10 |
| Input modalities | text, image |
| Output modalities | text |
| Reasoning | No |
| Tool calling | No |
| Open weights | Yes |

## Pricing: USD per 1M tokens

| Input | Output | Cached input | Cache write |
| --- | --- | --- | --- |
| $0.00 | $0.00 | - | - |

## Methodology

- Updates: Aggregated every hour. Days and weeks use UTC.
- Tokens: Input, output, reasoning, and cached tokens for each request.
- Users and sessions: Approximate counts of distinct users and OpenCode sessions.
- Cost: Session cost is the average cost per OpenCode session. Token prices are list prices from the OpenCode model catalog.
- Retention: The share of a model's users in one week who use it again the next week.
- Citation: Cite OpenCode Data (opencode.ai/data) with the update time shown at the top of the page.
