GPT-4o mini
Our cloudGPT-4o mini is a fast, affordable small model for focused tasks. It accepts both text and image inputs and produces text outputs, including Structured Outputs. It is ideal for fine-tuning, and outputs from a larger model like GPT-4o can be distilled into it for similar results at lower cost and latency.
The real OpenAI model
Same weights, same API, same answers. We change only how you pay for it: one key for every maker, one bill, no separate account to open. Every figure on this page comes from OpenAI’s own page — check it.
Maker's own pagePrice
per 1M tokens
Input
Output
$0.15
$0.60
Context
128,000
up to 16,384 out
Modalities
Released
Jul 18, 2024
Knows up to Oct 1, 2023
What it would cost you
1000 tokens is roughly 750 English words.
Where this model runs
Similar models
Models built for the same kind of work, at a similar price.
We do not mark up the price of tokens. It is the same price the provider charges. We earn on the top-up fee.
Use it from your code
from openai import OpenAI
# Works with the official openai package (pip install openai)
client = OpenAI(
base_url="https://api.flintbeam.com/v1",
api_key="sk-live-your-api-key"
)
response = client.chat.completions.create(
model="gpt-4o-mini",
messages=[
{"role": "system", "content": "You are an experienced software engineer."},
{"role": "user", "content": "Describe the architecture of a distributed cache."}
],
temperature=0.7,
max_tokens=1500
)
print(response.choices[0].message.content)