GPT-5.4 Mini with Jev
GPT-5.4 Mini is a model from OpenAI, released 2026-03-17. It costs $0.750 per million input tokens and $4.50 per million output tokens, reads up to 400K tokens of context, and writes up to 128,000 tokens in one reply. Through llm11, set model to openai/gpt-5.4-mini.
- Input price
- $0.750 / 1M
- Output price
- $4.50 / 1M
- Context window
- 400K tokens
- Max output
- 128K tokens
- Cached input
- $0.075 / 1M
- Released
- 2026-03-17
- Accepts
- file, image, text
- Intelligence index
- 24.1
Read from the provider's prompt cache
Artificial Analysis, via the catalogue
Read from the live catalogue and refreshed daily. At blended list price, 30% of the 332 priced models we can call cost more.
Using GPT-5.4 Mini with Jev
Jev is TypeSafe AI's System One model. It reads each request first and decides which model in your pool answers, with a calibrated confidence on that call. GPT-5.4 Mini does the answering when Jev sends it work, or when you name it yourself. GPT-5.4 Mini falls in the llm11-balanced band, and the labs and limits it meets make it eligible for that pack. Packs keep a spread of eight models per band, and it is not one of the eight today, so it is only picked when you name it in a custom pool.
Jev is new, so we have not published results for GPT-5.4 Mini with and without it, and this page does not invent any. Every response comes with a receipt that names the model that answered and what it cost against your most expensive model, so you can measure the pairing on your own traffic.
What a request costs
List price arithmetic, the same sum a receipt does. llm11 adds nothing per request; the only fee is on buying credits.
| Request | Tokens in / out | One | A thousand |
|---|---|---|---|
| Short chat turn | 500 / 200 | $0.00128 | $1.28 |
| Question over retrieved documents | 4,000 / 500 | $0.00525 | $5.25 |
| Long document summary | 30,000 / 1,000 | $0.027 | $27.00 |
Cheaper models to try
Nearest in price and below GPT-5.4 Mini, each within a few points of it on the intelligence index.
| Model | Input / 1M | Output / 1M | Context | Index | Released |
|---|---|---|---|---|---|
| SpaceXAI Grok 4.3 | $1.25 | $2.50 | 1M | 24.9 | |
| Google Gemini 3.8 Flash | $0.750 | $3.75 | 1.05M | 40.9 | |
| Google Gemini 3.7 Flash | $0.750 | $3.75 | 1.05M | 39.1 | |
| Google Gemini 3.6 Flash | $0.750 | $3.75 | 1.05M | 34 |
Call it
Same OpenAI request shape. Naming the model pins it, so nothing is routed, and the answer is still checked.
python
from openai import OpenAI
client = OpenAI(base_url="https://www.llm11.com/v1", api_key="llm11_live_...")
res = client.chat.completions.create(
model="openai/gpt-5.4-mini",
messages=[{"role": "user", "content": "Hello"}],
)More from OpenAI
| Model | Input / 1M | Output / 1M | Context | Released |
|---|---|---|---|---|
| GPT-6 Luna | $0.100 | $0.500 | 1.05M | |
| GPT-6 Luna Pro | $0.100 | $0.500 | 1.05M | |
| GPT-6 Sol | $2.00 | $10 | 1.05M | |
| GPT-6 Sol Pro | $2.00 | $10 | 1.05M | |
| GPT-6 Astra | $10 | $50 | 1.05M | |
| GPT-6 Astra Pro | $10 | $50 | 1.05M |
Questions
- What is GPT-5.4 Mini with Jev?
- Jev is TypeSafe AI's System One model. It reads each request first and decides which model in your pool answers, with a calibrated confidence on that call. GPT-5.4 Mini does the answering when Jev sends it work, or when you name it yourself. GPT-5.4 Mini falls in the llm11-balanced band, and the labs and limits it meets make it eligible for that pack. Packs keep a spread of eight models per band, and it is not one of the eight today, so it is only picked when you name it in a custom pool. Routing decisions are made by Jev, TypeSafe AI's System One model.
- Is GPT-5.4 Mini a System One model?
- No. GPT-5.4 Mini answers requests. The System One model is Jev, which decides which model answers each one, so the two do different jobs and llm11 uses both.
- How much does GPT-5.4 Mini cost?
- $0.750 per million input tokens and $4.50 per million output tokens. A 4,000 token prompt with a 500 token answer costs about $0.00525, so a thousand of them is about $5.25. llm11 passes provider list price through and charges 5% when you buy credits.
- What is the GPT-5.4 Mini context window?
- 400,000 tokens, with up to 128,000 tokens in a single reply.
- Does GPT-5.4 Mini support tool calling and structured output?
- The catalogue lists tool calling, structured output and reasoning controls.
- Can I use GPT-5.4 Mini through llm11?
- Yes. Set model to openai/gpt-5.4-mini on the OpenAI-compatible endpoint and the request goes to it directly, with verification still running.
- Does Jev route to GPT-5.4 Mini?
- Routing decisions are made by Jev, TypeSafe AI's System One model. GPT-5.4 Mini falls in the llm11-balanced band, and the labs and limits it meets make it eligible for that pack. Packs keep a spread of eight models per band, and it is not one of the eight today, so it is only picked when you name it in a custom pool.
- What is a cheaper alternative to GPT-5.4 Mini?
- Nearest in price and below it: SpaceXAI Grok 4.3 at $1.25 in and $2.50 out, Google Gemini 3.8 Flash at $0.750 in and $3.75 out and Google Gemini 3.7 Flash at $0.750 in and $3.75 out. Nearest in price and below GPT-5.4 Mini, each within a few points of it on the intelligence index.