xAI
Operational· 100% uptime, 7 days

Grok 4.7 API

The Grok 4.7 API for long coding and agent tasks, 65% below xAI's list price

Price per 1M tokens65%off
Input
$0.70$2.00
Output
$2.10$6.00
Cached input
$0.175
Model IDgrok-4.7
chat.py
from openai import OpenAI
client = OpenAI(
base_url="https://api.zurelay.com/v1",
api_key="YOUR_ZURELAY_KEY",
)
stream = client.chat.completions.create(
model="grok-4.7",
messages=[{"role": "user", "content": "Explain quicksort in two sentences."}],
stream=True,
)
for chunk in stream:
if chunk.choices:
print(chunk.choices[0].delta.content or "", end="")

Base URL https://api.zurelay.com/v1. Works with any OpenAI SDK.

Context window
500K tokens
Input
Text, Image
Output
Text
Released
Sep 21, 2026
Uptime, 7 days
100%

Pricing

Grok 4.7 API pricing. A fraction of the list price.

Pay as you go from prepaid credit, with no subscription and no minimum. Requests that fail are never billed.

RatezurelayxAI listYou save
Input
per 1M tokens
$0.70$2.0065%off
Output
per 1M tokens
$2.10$6.0065%off
Cached input
per 1M tokens
$0.175——

One rate at every prompt length, up to the full 500K context window. Streaming, tool calls and structured outputs cost nothing extra.

Savings calculator

Move the sliders to your monthly usage.

Input tokens per month50M
Output tokens per month10M
zurelay
$56.00
xAI list price
$160.00
You save every month$104.00

$1,248 a year

Overview

What is Grok 4.7?

Grok 4.7 is xAI's newest model for coding, agentic tasks and knowledge work, with a 500K-token context window and image input. The zurelay Grok 4.7 API gives you the same model through one OpenAI-compatible endpoint, billed per token at 65% below xAI's list price.

Grok 4.7 is xAI's model for coding, agentic tasks and professional knowledge work, released on September 21, 2026. The Grok 4.7 API takes text and image input, returns text, and holds up to 500,000 tokens of context. xAI sets no separate limit on text output. The knowledge cutoff is May 2026.

xAI built Grok 4.7 on a new, larger base model than Grok 4.6. It then trained it with a longer reinforcement learning run on a harder mix of tasks, weighted toward problems that take many hours to finish. xAI says the result checks its own work more carefully, manages long context better and is better at creating documents and presentations. The list price did not change from Grok 4.6.

You control how long Grok 4.7 reasons. xAI's reasoning effort setting takes low, medium, high or xhigh, with high as the default. Use low for short, simple calls and xhigh for the hardest problems. The model also supports function calling and structured outputs, so it drops into agent loops and JSON pipelines without extra parsing.

On zurelay you call Grok 4.7 with the OpenAI SDK you already use: set the base URL to https://api.zurelay.com/v1 and the model to grok-4.7. You pay $0.70 per 1M input tokens and $2.10 per 1M output tokens, against xAI's list price of $2.00 and $6.00. zurelay charges one rate at every prompt length, while xAI doubles its rates once a prompt reaches 200K tokens.

Strengths

Where Grok 4.7 shines. And what teams build with it.

01

Long coding and agent runs

xAI trained Grok 4.7 on problems that take many hours to complete. It checks its own work more carefully than Grok 4.6 and manages long context better.

02

500K tokens of context

A large codebase, a stack of contracts or a full agent trace fits in one request. Images go in the same prompt as the text.

03

Reasoning cost you control

Set reasoning effort per request, from low to xhigh. Simple calls stay short and cheap, and hard ones get more thinking.

04

Documents and presentations

xAI reports that Grok 4.7 is better than Grok 4.6 at creating documents and presentations, which helps report and deck generation.

Use cases

  • Coding agents

    Run multi-step refactors, bug hunts and test-writing loops where the model has to verify its own changes over a long session.

  • Knowledge work

    Draft reports, memos and slide outlines from long source material. Ask for structured outputs when the next step needs JSON.

  • Long-document analysis

    Load hundreds of pages of contracts, specs or logs into the 500K-token window and ask questions across all of it.

  • Screenshots and charts

    Send UI screenshots, charts or scanned pages with your prompt and get text or JSON back.

Get started

Call Grok 4.7 in three steps.

No waitlist and no new SDK. If your code already talks to OpenAI, it already talks to zurelay.

  1. 1

    Create a key

    Sign up, add credit and create an API key. Set a monthly budget or a rate limit per key if you like.

  2. 2

    Point your SDK at zurelay

    Change the base URL. Everything else in your code stays the same.

    https://api.zurelay.com/v1
  3. 3

    Use grok-4.7

    Send chat completions as usual. Streaming, tool calls and usage reporting work as you expect.

    grok-4.7

FAQ

Grok 4.7 API questions.

Anything else, write to sales@zurelay.com and an engineer will answer.

How much does the Grok 4.7 API cost on zurelay?

Grok 4.7 costs $0.70 per 1M input tokens and $2.10 per 1M output tokens on zurelay, with cached input at $0.175. xAI's list price is $2.00 input and $6.00 output, so you save 65%. zurelay bills one rate at every prompt length, while xAI doubles its rates once a prompt reaches 200K tokens.

Is there a cheaper Grok API than xAI's own?

Yes. zurelay serves the same grok-4.7 model at 65% below xAI's list price. You pay as you go for the tokens you use, and credit never expires.

How do I call the Grok 4.7 API with the OpenAI SDK?

Install the official OpenAI SDK, set the base URL to https://api.zurelay.com/v1 and use your zurelay API key. Then pass model "grok-4.7" to chat.completions.create. Your existing prompts, tool definitions and streaming code work unchanged.

What is the Grok 4.7 context window?

Grok 4.7 has a 500,000-token context window. xAI sets no separate limit on text output. The model accepts text and images and returns text.

What is Grok 4.7 best at?

xAI built Grok 4.7 for coding, agentic tasks and knowledge work. It was trained on problems that take hours to finish, and xAI says it is better than Grok 4.6 at checking its own work, handling long context and creating documents and presentations.

How is Grok 4.7 different from Grok 4.6?

Grok 4.7 uses a new, larger base model and a longer reinforcement learning run on a harder mix of tasks. The context window, input types and list price are the same as Grok 4.6. xAI calls Grok 4.7 the most capable model it has built.

Can Grok 4.7 read images?

Yes. Grok 4.7 accepts images alongside text in the same request and returns text. It does not generate images or audio.

Does the Grok 4.7 API support streaming and per-key rate limits?

Streaming works with stream: true, the same as the OpenAI API. You can set your own requests-per-minute cap and monthly budget on each zurelay API key, so one runaway agent can't drain your balance.

Are Grok 4.7 responses on zurelay the same as from xAI?

Yes. Requests go to grok-4.7 itself, never a smaller or substitute model. Output is sampled, so wording varies from run to run, just as it does on xAI's own API. zurelay is an independent service and is not affiliated with xAI.

More models on zurelay

All models

Stop paying list price.
Start saving today.

Create a key in seconds, point the SDK you already use at zurelay, and every request costs up to 85% less from the first token.