docs
models / Mistral AI / Mistral Small 3.1 24B Instruct

Mistral Small 3.1 24B Instruct

textmistralai/mistral-small-3.1-24b-instructZDR supported

Building upon Mistral Small 3 (2501), Mistral Small 3.1 (2503) adds state-of-the-art vision understanding and enhances long context capabilities up to 128k tokens without compromising text performance. With 24 billion parameters, this model achieves top-tier capabilities in both text and vision tasks.

$0.555list priceper 1M output tokens
pricing
unitlist price
per 1M input tokens$0.351
per 1M output tokens$0.555

Everyone pays the list price. Reaching Elite (hold or stake 1,337 $XAVA, or hold 1,337,000 $XINF) unlocks the Elite price and a bigger buyback of your token. How it works

openai sdk (python)
from openai import OpenAI

client = OpenAI(
    base_url="https://www.zinf.ai/v1",
    api_key="xk_live_...",  # your Xava Inference key
)

r = client.chat.completions.create(
    model="mistralai/mistral-small-3.1-24b-instruct",
    messages=[{"role": "user", "content": "Hello!"}],
)
print(r.choices[0].message.content)
curl
curl https://www.zinf.ai/v1/chat/completions \
  -H "Authorization: Bearer $XINF_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"mistralai/mistral-small-3.1-24b-instruct","messages":[{"role":"user","content":"Hello!"}]}'
limits
providerMistral AI
typetext
context window128,000 tokens
max output—
prompt storagenone
provider retentionnone ZDR supported
livesoon
requests 24h—
p50 latency—
uptime 7d—