Thali

← All models

Grok 4

x-ai/grok-4

available chat hosted: global zero_retention

xAI's Grok models focus on reasoning and real-time knowledge, with mini variants for cost-sensitive traffic.

Context window 256,000 tokens
Input ₹313.50 per million tokens
Output ₹1567.50 per million tokens
Licence varies verify before production use

Call this model

from openai import OpenAI client = OpenAI( base_url="https://thaliai.in/api/v1", api_key="thali-sk-...", ) completion = client.chat.completions.create( model="x-ai/grok-4", messages=[{"role": "user", "content": "Hello"}], )

Works with any OpenAI SDK. Streaming, fallback lists and provider preferences are documented in the routing guide.

This model is served via global infrastructure. For workloads that must stay in India, filter the catalog for hosted_in: "in".