# Llama 3.1 Nemotron 70B Instruct Usage, Cost & Rank | OpenCode Data

> Llama 3.1 Nemotron 70B Instruct costs $0.00 per 1M input tokens and $0.00 per 1M output tokens.

Nemotron model for efficient reasoning, coding, and specialized AI agents

- Page: https://opencode.ai/data/nvidia/llama-3-1-nemotron-70b-instruct
- JSON: https://opencode.ai/data/nvidia/llama-3-1-nemotron-70b-instruct.json
- Model ID: nvidia/llama-3.1-nemotron-70b-instruct
- Lab: NVIDIA

## Model facts

| Fact | Value |
| --- | --- |
| Context window | 128K |
| Max output | 8.2K |
| Knowledge cutoff | - |
| Release date | 2025-04-15 |
| Input modalities | text |
| Output modalities | text |
| Reasoning | No |
| Tool calling | Yes |
| Open weights | Yes |

## Pricing: USD per 1M tokens

| Input | Output | Cached input | Cache write |
| --- | --- | --- | --- |
| $0.00 | $0.00 | - | - |

## Methodology

- Updates: Aggregated every hour. Days and weeks use UTC.
- Tokens: Input, output, reasoning, and cached tokens for each request.
- Users and sessions: Approximate counts of distinct users and OpenCode sessions.
- Cost: Session cost is the average cost per OpenCode session. Token prices are list prices from the OpenCode model catalog.
- Retention: The share of a model's users in one week who use it again the next week.
- Citation: Cite OpenCode Data (opencode.ai/data) with the update time shown at the top of the page.
