All models
openai

GPT-5 nano

Our cloud
Vendor key$0.05/$0.40per 1M tokens

GPT-5 nano is the fastest, most cost-efficient version of GPT-5. It is good for summarization and classification tasks. For most new speed- and cost-sensitive workloads OpenAI recommends starting with GPT-5.6 Luna.

The real OpenAI model

Same weights, same API, same answers. We change only how you pay for it: one key for every maker, one bill, no separate account to open. Every figure on this page comes from OpenAI’s own page — check it.

Maker's own page

Price

per 1M tokens

Input

Output

$0.05

$0.40

Context

400,000

up to 128,000 out

Modalities

ToolsReasoning

Released

Aug 7, 2025

Knows up to May 31, 2024

What it would cost you

1000 tokens is roughly 750 English words.

Estimated total$0.30

Where this model runs

Our cloud
openai
OpenAI
azure_openai
Microsoft Azure

If one is unavailable the request goes to the next one, and you never see it happen.

Vendor key
openai
OpenAI

Similar models

Models built for the same kind of work, at a similar price.

We do not mark up the price of tokens. It is the same price the provider charges. We earn on the top-up fee.

Use it from your code

Model:gpt-5-nano
main.py
OpenAI Specification v1.0
from openai import OpenAI

# Works with the official openai package (pip install openai)
client = OpenAI(
    base_url="https://api.flintbeam.com/v1",
    api_key="sk-live-your-api-key"
)

response = client.chat.completions.create(
    model="gpt-5-nano",
    messages=[
        {"role": "system", "content": "You are an experienced software engineer."},
        {"role": "user", "content": "Describe the architecture of a distributed cache."}
    ],
    temperature=0.7,
    max_tokens=1500
)

print(response.choices[0].message.content)