Grok 4
x-ai/grok-4
available
chat
hosted: global
zero_retention
xAI's Grok models focus on reasoning and real-time knowledge, with mini variants for cost-sensitive traffic.
Context window
256,000
tokens
Input
₹313.50
per million tokens
Output
₹1567.50
per million tokens
Licence
varies
verify before production use
Call this model
from openai import OpenAI
client = OpenAI(
base_url="https://thaliai.in/api/v1",
api_key="thali-sk-...",
)
completion = client.chat.completions.create(
model="x-ai/grok-4",
messages=[{"role": "user", "content": "Hello"}],
)
Works with any OpenAI SDK. Streaming, fallback lists and provider preferences are documented in the routing guide.
This model is served via global infrastructure. For workloads
that must stay in India, filter the catalog for
hosted_in: "in".