grok-code-fast
grok-code-fast
StreamingTool callingVisionReasoningJSON mode
Overview
grok-code-fast is available on MAX API through the OpenAI-compatible API.
Pricing
Official list prices in USD. Token prices are per 1M tokens.| Tier | Input | Output | Cache read |
|---|---|---|---|
| Input ≤ 200,000 tokens | $1 | $2 | $0.2 |
| Input > 199,999 tokens | $2 | $4 | $0.4 |
This model uses tiered pricing: the tier is chosen by the total input length of each request.
Cache read applies to prompt tokens served from the prompt cache; cache write applies to tokens written into it.
Example request
curl https://clubs.byte-ai.cn/v1/chat/completions \
-H "Authorization: Bearer $API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "grok-code-fast",
"messages": [{"role": "user", "content": "Hello!"}]
}'