Long coding and agent runs
xAI trained Grok 4.7 on problems that take many hours to complete. It checks its own work more carefully than Grok 4.6 and manages long context better.
The Grok 4.7 API for long coding and agent tasks, 65% below xAI's list price
grok-4.7from openai import OpenAIclient = OpenAI( base_url="https://api.zurelay.com/v1", api_key="YOUR_ZURELAY_KEY",)stream = client.chat.completions.create( model="grok-4.7", messages=[{"role": "user", "content": "Explain quicksort in two sentences."}], stream=True,)for chunk in stream: if chunk.choices: print(chunk.choices[0].delta.content or "", end="")Base URL https://api.zurelay.com/v1. Works with any OpenAI SDK.
Pricing
Pay as you go from prepaid credit, with no subscription and no minimum. Requests that fail are never billed.
| Rate | zurelay | xAI list | You save |
|---|---|---|---|
Input per 1M tokens | $0.70 | 65%off | |
Output per 1M tokens | $2.10 | 65%off | |
Cached input per 1M tokens | $0.175 | — | — |
One rate at every prompt length, up to the full 500K context window. Streaming, tool calls and structured outputs cost nothing extra.
Move the sliders to your monthly usage.
$1,248 a year
Overview
Grok 4.7 is xAI's newest model for coding, agentic tasks and knowledge work, with a 500K-token context window and image input. The zurelay Grok 4.7 API gives you the same model through one OpenAI-compatible endpoint, billed per token at 65% below xAI's list price.
Grok 4.7 is xAI's model for coding, agentic tasks and professional knowledge work, released on September 21, 2026. The Grok 4.7 API takes text and image input, returns text, and holds up to 500,000 tokens of context. xAI sets no separate limit on text output. The knowledge cutoff is May 2026.
xAI built Grok 4.7 on a new, larger base model than Grok 4.6. It then trained it with a longer reinforcement learning run on a harder mix of tasks, weighted toward problems that take many hours to finish. xAI says the result checks its own work more carefully, manages long context better and is better at creating documents and presentations. The list price did not change from Grok 4.6.
You control how long Grok 4.7 reasons. xAI's reasoning effort setting takes low, medium, high or xhigh, with high as the default. Use low for short, simple calls and xhigh for the hardest problems. The model also supports function calling and structured outputs, so it drops into agent loops and JSON pipelines without extra parsing.
On zurelay you call Grok 4.7 with the OpenAI SDK you already use: set the base URL to https://api.zurelay.com/v1 and the model to grok-4.7. You pay $0.70 per 1M input tokens and $2.10 per 1M output tokens, against xAI's list price of $2.00 and $6.00. zurelay charges one rate at every prompt length, while xAI doubles its rates once a prompt reaches 200K tokens.
Strengths
xAI trained Grok 4.7 on problems that take many hours to complete. It checks its own work more carefully than Grok 4.6 and manages long context better.
A large codebase, a stack of contracts or a full agent trace fits in one request. Images go in the same prompt as the text.
Set reasoning effort per request, from low to xhigh. Simple calls stay short and cheap, and hard ones get more thinking.
xAI reports that Grok 4.7 is better than Grok 4.6 at creating documents and presentations, which helps report and deck generation.
Use cases
Run multi-step refactors, bug hunts and test-writing loops where the model has to verify its own changes over a long session.
Draft reports, memos and slide outlines from long source material. Ask for structured outputs when the next step needs JSON.
Load hundreds of pages of contracts, specs or logs into the 500K-token window and ask questions across all of it.
Send UI screenshots, charts or scanned pages with your prompt and get text or JSON back.
Get started
No waitlist and no new SDK. If your code already talks to OpenAI, it already talks to zurelay.
Sign up, add credit and create an API key. Set a monthly budget or a rate limit per key if you like.
Change the base URL. Everything else in your code stays the same.
https://api.zurelay.com/v1Send chat completions as usual. Streaming, tool calls and usage reporting work as you expect.
grok-4.7Works with the tools you already use
Compare
$0.18 / $0.36 per 1M tokens
Pick MiMo V2.6 Pro when you need a 1M-token context, audio or video input, or open weights.
MiMo V2.6 Pro API$0.23 / $0.62 per 1M tokens
Pick DeepSeek V4 Pro 0813 for an open-weight agent model with a 1M-token context at a lower list price.
DeepSeek V4 Pro 0813 API$0.037 / $0.15 per 1M tokens
Pick DeepSeek V4.1 Flash for high-volume text and image jobs where the lowest cost per token matters most.
DeepSeek V4.1 Flash APIFAQ
Anything else, write to sales@zurelay.com and an engineer will answer.
Grok 4.7 costs $0.70 per 1M input tokens and $2.10 per 1M output tokens on zurelay, with cached input at $0.175. xAI's list price is $2.00 input and $6.00 output, so you save 65%. zurelay bills one rate at every prompt length, while xAI doubles its rates once a prompt reaches 200K tokens.
Yes. zurelay serves the same grok-4.7 model at 65% below xAI's list price. You pay as you go for the tokens you use, and credit never expires.
Install the official OpenAI SDK, set the base URL to https://api.zurelay.com/v1 and use your zurelay API key. Then pass model "grok-4.7" to chat.completions.create. Your existing prompts, tool definitions and streaming code work unchanged.
Grok 4.7 has a 500,000-token context window. xAI sets no separate limit on text output. The model accepts text and images and returns text.
xAI built Grok 4.7 for coding, agentic tasks and knowledge work. It was trained on problems that take hours to finish, and xAI says it is better than Grok 4.6 at checking its own work, handling long context and creating documents and presentations.
Grok 4.7 uses a new, larger base model and a longer reinforcement learning run on a harder mix of tasks. The context window, input types and list price are the same as Grok 4.6. xAI calls Grok 4.7 the most capable model it has built.
Yes. Grok 4.7 accepts images alongside text in the same request and returns text. It does not generate images or audio.
Streaming works with stream: true, the same as the OpenAI API. You can set your own requests-per-minute cap and monthly budget on each zurelay API key, so one runaway agent can't drain your balance.
Yes. Requests go to grok-4.7 itself, never a smaller or substitute model. Output is sampled, so wording varies from run to run, just as it does on xAI's own API. zurelay is an independent service and is not affiliated with xAI.
Create a key in seconds, point the SDK you already use at zurelay, and every request costs up to 85% less from the first token.