gpt-6-astra
gpt-6-astra
StreamingTool callingVisionReasoningJSON modeWeb search
Overview
gpt-6-astra is available on MAX API through the OpenAI-compatible API.
Pricing
Official list prices in USD. Token prices are per 1M tokens.Input ≤ 272,000 tokens
- Input
- $10
- Output
- $50
- Cache read
- $1
- Cache write
- $12.5
Input > 271,999 tokens
- Input
- $20
- Output
- $75
- Cache read
- $2
- Cache write
- $25
| Tier | Input | Output | Cache read | Cache write |
|---|---|---|---|---|
| Input ≤ 272,000 tokens | $10 | $50 | $1 | $12.5 |
| Input > 271,999 tokens | $20 | $75 | $2 | $25 |
This model uses tiered pricing: the tier is chosen by the total input length of each request.
Cache read applies to prompt tokens served from the prompt cache; cache write applies to tokens written into it.
Example request
curl https://<your-endpoint>/v1/chat/completions \
-H "Authorization: Bearer $API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gpt-6-astra",
"messages": [{"role": "user", "content": "Hello!"}]
}'https://<your-endpoint> is a placeholder. Your endpoint is shown in the console after you sign in.