Gemini API pricing. Up to 61% below list.

Gemini is Google's family of models. Pro is for complex reasoning and coding, Flash for coding, agents and multi-step work, and Flash-Lite for fast, high-volume jobs. They read text, images, video and audio. Zurelay serves them through one OpenAI-compatible endpoint, at one rate at every prompt length.

Gemini models
5
Below Google's list
58–61%
From, per 1M input
$0.10

Pricing

Every Gemini model. Live prices, one key.

Each Gemini model on Zurelay, with Google's list price beside ours.

Gemini API prices on Zurelay and at Google's list price
ModelContextInput / 1MOutput / 1M10K requestsYou save
Gemini 3.1 Progemini-3.1-pro1M
$0.80
$2.00
$4.80
$12.00
$40.00
$100.00
60%off
Gemini 3.8 Flashgemini-3.8-flash1M
$0.32
$0.75
$1.59
$3.75
$14.35
$33.75
58%off
Gemini 3.7 Flashgemini-3.7-flash1M
$0.32
$0.75
$1.59
$3.75
$14.35
$33.75
58%off
Gemini 3.5 Flash-Litegemini-3.5-flash-lite1M
$0.12
$0.30
$1.04
$2.50
$7.60
$18.50
58%off
Gemini 3.1 Flash-Litegemini-3.1-flash-lite1M
$0.10
$0.25
$0.59
$1.50
$4.95
$12.50
61%off

Prices in US dollars per 1M tokens, live from the catalog, with Google's list price struck through. 10K requests is 10,000 requests of 2,000 tokens in and 500 out, at each price. You save is on output tokens.

Quickstart

Change two lines. Keep your code.

Point the OpenAI SDK, or any OpenAI-compatible tool, at https://api.zurelay.com/v1 and use your Zurelay key. Streaming, tool calling and structured output work as documented.

Add credit from $10, with no subscription. The quickstart has the rest.

chat.py
from openai import OpenAI
client = OpenAI(
base_url="https://api.zurelay.com/v1",
api_key="YOUR_ZURELAY_KEY",
)
stream = client.chat.completions.create(
model="gemini-3.1-pro",
messages=[{"role": "user", "content": "Explain quicksort in two sentences."}],
stream=True,
)
for chunk in stream:
if chunk.choices:
print(chunk.choices[0].delta.content or "", end="")

Verified

The Gemini you asked for. Checked every day.

Model capacity is often bought ahead, through prepaid credits, committed-spend contracts and volume tiers, and not all of it gets used. Sellers offer it below list price. Zurelay tests each route, pools several per model and sends every request to one that is available and verified.

Before a route serves requests, and every day after, it has to pass identity, dated-knowledge and reasoning checks that catch a different or older model sold under a newer name. Failed attempts move to another route, and you aren’t billed for them. How models are verified.

Live health for every model is on the status page.

FAQ

Gemini API questions.

Anything else? Ask our team and we’ll answer by email.

How much does the Gemini API cost?

On Zurelay, Gemini models cost from $0.10 input and $0.59 output per 1M tokens (Gemini 3.1 Flash-Lite) to $0.80 and $4.80 (Gemini 3.1 Pro), 58–61% below Google's list prices. There's no subscription or minimum spend: you add prepaid credit from $10 and pay for the tokens you use.

Is it the real Gemini?

Yes. Each model is served over several routes run by independent sellers. Every route is tested before it serves requests and every day after, with identity, dated-knowledge and reasoning checks that catch a different or older model sold under the name you asked for. If an answer doesn't look right, send its x-request-id to support and we check the route that served it.

Why is it cheaper than Google's own API?

Sellers offer model capacity they bought ahead, through prepaid credits, committed-spend contracts and volume tiers, below list price. Zurelay tests it, pools several routes per model and sends each request to one that is available and verified. Requests aren't served under Google's own terms for its customers; our Terms explain how they are served.

How do I start calling Gemini through Zurelay?

Create a key, then point the OpenAI SDK, or any OpenAI-compatible tool, at https://api.zurelay.com/v1 with a model ID such as gemini-3.1-pro. Code that already uses the OpenAI SDK only needs the base URL, the key and the model name changed.

Do long prompts cost more?

No. Each Gemini model on Zurelay has one input price and one output price at every prompt length, up to its context window.

Can Gemini read video and audio through Zurelay?

Yes. Send images, PDFs, audio and video in the message, as data or links, in OpenAI's format. They're billed as the input tokens the model counts for them.

Stop paying list price.
Start saving today.

Create a key in seconds, point the SDK you already use at Zurelay, and every request costs up to 90% less from the first token.