Grok 4.6
text
xai/grok-4.6ZDR supportedxAI's Grok 4.6, a flagship reasoning model for coding, agentic tasks, and visual work. Accepts text and image inputs, and supports function calling and structured outputs.
$6.00list priceper 1M output tokens (prompts up to 200K tokens)
pricing
| unit | list → Elite price |
|---|---|
| per 1M cached input tokens (prompts over 200K tokens) | $1.00 |
| per 1M cached input tokens (prompts up to 200K tokens) | $0.50 |
| per 1M input tokens (prompts over 200K tokens) | $4.00 |
| per 1M input tokens (prompts up to 200K tokens) | $2.00 |
| per 1M output tokens (prompts over 200K tokens) | $12.00 |
| per 1M output tokens (prompts up to 200K tokens) | $6.00 |
No Elite discount on this model right now: everyone pays the list price.
Everyone pays the list price. Reaching Elite (hold or stake 1,337 $XAVA, or hold 1,337,000 $XINF) unlocks the Elite price and a bigger buyback of your token. How it works
openai sdk (python)
from openai import OpenAI
client = OpenAI(
base_url="https://www.zinf.ai/v1",
api_key="xk_live_...", # your Xava Inference key
)
r = client.chat.completions.create(
model="xai/grok-4.6",
messages=[{"role": "user", "content": "Hello!"}],
)
print(r.choices[0].message.content)curl
curl https://www.zinf.ai/v1/chat/completions \
-H "Authorization: Bearer $XINF_API_KEY" \
-H "Content-Type: application/json" \
-d '{"model":"xai/grok-4.6","messages":[{"role":"user","content":"Hello!"}]}'limits
providerxAI
typetext
context window500,000 tokens
max output—
prompt storagenone
provider retentionnone ZDR supported
livesoon
requests 24h—
p50 latency—
uptime 7d—