Standard & Paid Price
Standard Price
| Meter | Standard Price |
|---|---|
| Input price | $0.2500 / 1M tokens |
| Completion price | $1.5000 / 1M tokens |
| Cache read price | $0.0500 / 1M tokens |
| Cache creation price | $0.3125 / 1M tokens |
Dynamic pricing · 2 tiers
Billing details
Vendor list price comparison
Description
Alibaba Qwen3.6-Flash — fast vision-language model that the vendor highlights for agentic coding and for spatial intelligence, with marked gains in object localisation and detection. 1,000,000-token context window (991,808 max input, 983,616 in thinking mode), up to 65,536 output tokens, and a 131,072-token chain-of-thought budget. Accepts text, image, and video input and returns text. Supports function calling and context caching.
Capabilities
- Input modalities
- text, image, video
- Output modalities
- text
- Tasks
- chat, reasoning, vision, file-video-understanding
API endpoints
POST /v1/messages(anthropic)POST /v1/chat/completions(openai)