Gemini API pricing. Up to 61% below list.
Gemini is Google's family of models. Pro is for complex reasoning and coding, Flash for coding, agents and multi-step work, and Flash-Lite for fast, high-volume jobs. They read text, images, video and audio. Zurelay serves them through one OpenAI-compatible endpoint, at one rate at every prompt length.
- Gemini models
- 5
- Below Google's list
- 58–61%
- From, per 1M input
- $0.10
Pricing
Every Gemini model. Live prices, one key.
Each Gemini model on Zurelay, with Google's list price beside ours.
| Model | Context | Input / 1M | Output / 1M | 10K requests | You save |
|---|---|---|---|---|---|
| Gemini 3.1 Progemini-3.1-pro | 1M | $0.80 | $4.80 | $40.00 | 60%off |
| Gemini 3.8 Flashgemini-3.8-flash | 1M | $0.32 | $1.59 | $14.35 | 58%off |
| Gemini 3.7 Flashgemini-3.7-flash | 1M | $0.32 | $1.59 | $14.35 | 58%off |
| Gemini 3.5 Flash-Litegemini-3.5-flash-lite | 1M | $0.12 | $1.04 | $7.60 | 58%off |
| Gemini 3.1 Flash-Litegemini-3.1-flash-lite | 1M | $0.10 | $0.59 | $4.95 | 61%off |
Prices in US dollars per 1M tokens, live from the catalog, with Google's list price struck through. 10K requests is 10,000 requests of 2,000 tokens in and 500 out, at each price. You save is on output tokens.
Which model
Which Gemini model to use. Same key for all of them.
Complex reasoning
Gemini 3.1 Pro
$0.80 / $4.80 per 1M tokens
Google's Pro model for complex problem-solving and agentic coding, still in preview at Google.
Gemini 3.1 Pro APICoding and agents
Gemini 3.8 Flash
$0.32 / $1.59 per 1M tokens
Google's most intelligent Flash model, built for long-horizon software engineering, autonomous agents and enterprise workflows.
Gemini 3.8 Flash APIFast, high-volume calls
Gemini 3.5 Flash-Lite
$0.12 / $1.04 per 1M tokens
The fastest, most cost-effective model in Google's Gemini 3.5 series, for agent steps, search and document processing.
Gemini 3.5 Flash-Lite APIAlso on Zurelay, for prompts, evals and agents tuned on them: Gemini 3.7 Flash and Gemini 3.1 Flash-Lite.
Quickstart
Change two lines. Keep your code.
Point the OpenAI SDK, or any OpenAI-compatible tool, at https://api.zurelay.com/v1 and use your Zurelay key. Streaming, tool calling and structured output work as documented.
Add credit from $10, with no subscription. The quickstart has the rest.
from openai import OpenAIclient = OpenAI( base_url="https://api.zurelay.com/v1", api_key="YOUR_ZURELAY_KEY",)stream = client.chat.completions.create( model="gemini-3.1-pro", messages=[{"role": "user", "content": "Explain quicksort in two sentences."}], stream=True,)for chunk in stream: if chunk.choices: print(chunk.choices[0].delta.content or "", end="")Verified
The Gemini you asked for. Checked every day.
Model capacity is often bought ahead, through prepaid credits, committed-spend contracts and volume tiers, and not all of it gets used. Sellers offer it below list price. Zurelay tests each route, pools several per model and sends every request to one that is available and verified.
Before a route serves requests, and every day after, it has to pass identity, dated-knowledge and reasoning checks that catch a different or older model sold under a newer name. Failed attempts move to another route, and you aren’t billed for them. How models are verified.
Live health for every model is on the status page.
How much does the Gemini API cost?
On Zurelay, Gemini models cost from $0.10 input and $0.59 output per 1M tokens (Gemini 3.1 Flash-Lite) to $0.80 and $4.80 (Gemini 3.1 Pro), 58–61% below Google's list prices. There's no subscription or minimum spend: you add prepaid credit from $10 and pay for the tokens you use.
Is it the real Gemini?
Yes. Each model is served over several routes run by independent sellers. Every route is tested before it serves requests and every day after, with identity, dated-knowledge and reasoning checks that catch a different or older model sold under the name you asked for. If an answer doesn't look right, send its x-request-id to support and we check the route that served it.
Why is it cheaper than Google's own API?
Sellers offer model capacity they bought ahead, through prepaid credits, committed-spend contracts and volume tiers, below list price. Zurelay tests it, pools several routes per model and sends each request to one that is available and verified. Requests aren't served under Google's own terms for its customers; our Terms explain how they are served.
How do I start calling Gemini through Zurelay?
Create a key, then point the OpenAI SDK, or any OpenAI-compatible tool, at https://api.zurelay.com/v1 with a model ID such as gemini-3.1-pro. Code that already uses the OpenAI SDK only needs the base URL, the key and the model name changed.
Do long prompts cost more?
No. Each Gemini model on Zurelay has one input price and one output price at every prompt length, up to its context window.
Can Gemini read video and audio through Zurelay?
Yes. Send images, PDFs, audio and video in the message, as data or links, in OpenAI's format. They're billed as the input tokens the model counts for them.
Stop paying list price.
Start saving today.
Create a key in seconds, point the SDK you already use at Zurelay, and every request costs up to 90% less from the first token.