# GLM-4.6V-Flash Usage, Cost & Rank | OpenCode Data

> GLM-4.6V-Flash costs $0.00 per 1M input tokens and $0.00 per 1M output tokens.

Lightweight GLM vision model for visual reasoning, documents, and multimodal agents

- Page: https://opencode.ai/data/zhipuai/glm-4-6v-flash
- JSON: https://opencode.ai/data/zhipuai/glm-4-6v-flash.json
- Model ID: zhipuai/glm-4.6v-flash
- Lab: Zhipu

## Model facts

| Fact | Value |
| --- | --- |
| Context window | 128K |
| Max output | 33K |
| Knowledge cutoff | - |
| Release date | 2025-12-08 |
| Input modalities | text, image, video |
| Output modalities | text |
| Reasoning | Yes |
| Tool calling | Yes |
| Open weights | Yes |
| Weights | [Hugging Face](https://huggingface.co/zai-org/GLM-4.6V-Flash) |

## Pricing: USD per 1M tokens

| Input | Output | Cached input | Cache write |
| --- | --- | --- | --- |
| $0.00 | $0.00 | $0.00 | $0.00 |

## Methodology

- Updates: Aggregated every hour. Days and weeks use UTC.
- Tokens: Input, output, reasoning, and cached tokens for each request.
- Users and sessions: Approximate counts of distinct users and OpenCode sessions.
- Cost: Session cost is the average cost per OpenCode session. Token prices are list prices from the OpenCode model catalog.
- Retention: The share of a model's users in one week who use it again the next week.
- Citation: Cite OpenCode Data (opencode.ai/data) with the update time shown at the top of the page.
