All models
gcp_vertex

Gemini 3.1 Flash-Lite Image

Our cloud
Try in Playground

Gemini 3.1 Flash-Lite Image, which Google calls Nano Banana 2 Lite, is the efficiency specialist of the image generation family, targeting sub-2 second end-to-end latency. It is built for high-volume interactive developer use cases and real-time consumer applications, and for fast multi-turn local edits such as color swaps and background changes.

The real Google model

Same weights, same API, same answers. We change only how you pay for it: one key for every maker, one bill, no separate account to open. Every figure on this page comes from Google’s own page — check it.

Maker's own page

Price

per 1M tokens

Input

Output

$0.25

$30.00

One image: $0.03 to $0.23

Context

65,536

up to 4,096 out

Modalities

Released

Jun 30, 2026

Knows up to Jan 2025

The prices above are per 1M tokens, the unit every model is billed in. A picture is billed as the tokens it weighs, so a larger one costs more.

What it would cost you

1000 tokens is roughly 750 English words.

Estimated total$15.50

Where this model runs

gcp_vertex
Google Vertex

Similar models

Models built for the same kind of work, at a similar price.

We do not mark up the price of tokens. It is the same price the provider charges. We earn on the top-up fee.

Use it from your code

Model:gemini-3.1-flash-lite-image
main.py
OpenAI Specification v1.0
from openai import OpenAI

# Works with the official openai package (pip install openai)
client = OpenAI(
    base_url="https://api.flintbeam.com/v1",
    api_key="sk-live-your-api-key"
)

response = client.chat.completions.create(
    model="gemini-3.1-flash-lite-image",
    messages=[
        {"role": "system", "content": "You are an experienced software engineer."},
        {"role": "user", "content": "Describe the architecture of a distributed cache."}
    ],
    temperature=0.7,
    max_tokens=1500
)

print(response.choices[0].message.content)