qwen3-vl-235b-a22b-thinking
qwen3-vl-235b-a22b-thinking
StreamingTool callingVisionReasoning
Overview
qwen3-vl-235b-a22b-thinking is available on MAX API through the OpenAI-compatible API.
Pricing
Official list prices in USD. Token prices are per 1M tokens.All requests
- Input
- $0.4
- Output
- $4
| Tier | Input | Output |
|---|---|---|
| All requests | $0.4 | $4 |
Cache read applies to prompt tokens served from the prompt cache; cache write applies to tokens written into it.
Example request
curl https://<your-endpoint>/v1/chat/completions \
-H "Authorization: Bearer $API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "qwen3-vl-235b-a22b-thinking",
"messages": [{"role": "user", "content": "Hello!"}]
}'https://<your-endpoint> is a placeholder. Your endpoint is shown in the console after you sign in.