# Qwen3.8 Flash vs Nemotron 3.5 Lightning 30B A3B - AI Model Comparison

- Page: https://opencode.ai/data/compare/alibaba/qwen3-8-flash/nvidia/nemotron-3-5-lightning
- Updated: 2026-10-02T08:39:37.000Z

|  | Qwen3.8 Flash | Nemotron 3.5 Lightning 30B A3B |
| --- | --- | --- |
| Lab | Alibaba | NVIDIA |
| Context window | 1M | 262K |
| Max output | 131K | 262K |
| Input modalities | text, image, video | text |
| Reasoning | Yes | Yes |
| Tool calling | Yes | Yes |
| Open weights | No | Yes |
| Release date | 2026-08-26 | 2026-08-11 |
| Input price per 1M tokens | $0.15 | $0.05 |
| Output price per 1M tokens | $0.47 | $0.20 |
| Cached input price per 1M tokens | $0.02 | $0.01 |
| OpenCode tokens: past 2 months | 2T | 4.3T |
| Share of all tokens | 0.2% | 0.5% |
| Rank by tokens last week | #17 | #15 |
| Unique users | 151K | 276K |
| Weekly retention | 60.0% | - |

Model pages: [Qwen3.8 Flash](https://opencode.ai/data/alibaba/qwen3-8-flash.md), [Nemotron 3.5 Lightning 30B A3B](https://opencode.ai/data/nvidia/nemotron-3-5-lightning.md)

## Methodology

- Updates: Aggregated every hour. Days and weeks use UTC.
- Tokens: Input, output, reasoning, and cached tokens for each request.
- Users and sessions: Approximate counts of distinct users and OpenCode sessions.
- Cost: Session cost is the average cost per OpenCode session. Token prices are list prices from the OpenCode model catalog.
- Retention: The share of a model's users in one week who use it again the next week.
- Citation: Cite OpenCode Data (opencode.ai/data) with the update time shown at the top of the page.
