Grok 4.7
xAI's Grok 4.7: the newest frontier model - agents, coding and reasoning, 500K context.
First request
from openai import OpenAI
client = OpenAI(
base_url="https://api.altrouter.ai/v1",
api_key="ar-...",
)
resp = client.chat.completions.create(
model="grok-4.7",
messages=[{"role": "user", "content": "Hello!"}],
)
print(resp.choices[0].message.content)import OpenAI from "openai";
const client = new OpenAI({
baseURL: "https://api.altrouter.ai/v1",
apiKey: process.env.ALTROUTER_API_KEY, // ar-...
});
const resp = await client.chat.completions.create({
model: "grok-4.7",
messages: [{ role: "user", content: "Hello!" }],
});
console.log(resp.choices[0].message.content);curl https://api.altrouter.ai/v1/chat/completions \
-H "Authorization: Bearer ar-..." \
-H "Content-Type: application/json" \
-d '{
"model": "grok-4.7",
"messages": [
{
"role": "user",
"content": "Hello!"
}
]
}'The interface is OpenAI-compatible: in existing code only the base_url, the key and the model id change — nothing else.
Grok 4.7 against its price neighbours
| Model | Input / 1M | Output / 1M | Context | |
|---|---|---|---|---|
| Claude Haiku 4.5Anthropic | $0.77 | $4.00 | 200K | Open |
| GPT-5.6 LunaOpenAI | $0.85 | $5.10 | 400K | Open |
| Grok 4.7this pagexAI | $1.70 | $5.10 | 500K | |
| o3OpenAI | $1.70 | $6.80 | 200K | Open |
| Claude Sonnet 5.5Anthropic | $1.70 | $8.50 | 1M | Open |
Prices and parameters come from the same catalog as the price above; this is not a benchmark, it is comparable billing terms.
Limits
Context is 500K tokens per request, and it holds everything you send: the system prompt, the conversation history and the current question together. Whatever does not fit is not seen by the model — trimming the history is your side of the job, and you pay for exactly what fit.
Response length is set by max_tokens, capped at 32,000 tokens per call. A response that hits the cap comes back with finish_reason: "length" — truncated but still billed, so it is worth setting deliberately on long generations.
Supported: a reasoning mode, tool calling, image input.
There is no subscription: every successful request is billed at the price above, and a failed one is not billed at all. That price is already 15% below the vendor's official one — the official price is shown struck through next to it so it can be checked. No xAI account of your own is needed: requests go to api.altrouter.ai and we make the upstream call.
Rate limiting is shared across models and counted per key: 600 requests per minute by default, and a 429 with retry-after: 60 above it. An individual key can get its own limit and a spending cap per day, week, month or lifetime in the dashboard.
When to pick Grok 4.7, and when a neighbour
The cheaper neighbour is GPT-5.6 Luna: $0.85 per 1M input tokens against $1.70 for Grok 4.7. On the same volume the bill comes out 2× smaller. Its context is smaller: 400K against 500K — long documents will need splitting. It is the sensible pick where the task is repetitive and high-volume — labelling, classification, short templated answers — the kind of work where quality is bounded by the prompt rather than the model.
The dearer neighbour is o3: $1.70 per 1M input tokens, or 1× the price. On catalog parameters it offers the same capabilities and the same order of context, so the premium is only justified if Grok 4.7 falls short on your own task.
Guessing is more expensive than checking: the table above is prices and parameters, not a benchmark, and it cannot tell you which model does better on your own prompt. Switching costs one line — the model id in the request changes, the base_url and the key stay — so running both on your own data takes minutes.
OpenAI-compatible chat completions endpoint, with SSE streaming.
/v1/chat/completionsStructured fields (tools, response_format, stop and more) are in the full parameter reference.