Ministral 3 3B 2512 with Jev
Ministral 3 3B 2512 is a model from Mistral, released 2025-12-02. It costs $0.100 per million input tokens and $0.100 per million output tokens, reads up to 131K tokens of context, and writes up to 104,857 tokens in one reply. Through llm11, set model to mistralai/ministral-3b-2512.
- Input price
- $0.100 / 1M
- Output price
- $0.100 / 1M
- Context window
- 131K tokens
- Max output
- 105K tokens
- Cached input
- $0.010 / 1M
- Released
- 2025-12-02
- Accepts
- text, image
- Intelligence index
- 4.8
Read from the provider's prompt cache
Artificial Analysis, via the catalogue
Read from the live catalogue and refreshed daily. At blended list price, 91% of the 332 priced models we can call cost more.
Using Ministral 3 3B 2512 with Jev
Jev is TypeSafe AI's System One model. It reads each request first and decides which model in your pool answers, with a calibrated confidence on that call. Ministral 3 3B 2512 does the answering when Jev sends it work, or when you name it yourself. Ministral 3 3B 2512 falls in the llm11-fast band, and the labs and limits it meets make it eligible for that pack. Packs keep a spread of eight models per band, and it is not one of the eight today, so it is only picked when you name it in a custom pool.
Jev is new, so we have not published results for Ministral 3 3B 2512 with and without it, and this page does not invent any. Every response comes with a receipt that names the model that answered and what it cost against your most expensive model, so you can measure the pairing on your own traffic.
What a request costs
List price arithmetic, the same sum a receipt does. llm11 adds nothing per request; the only fee is on buying credits.
| Request | Tokens in / out | One | A thousand |
|---|---|---|---|
| Short chat turn | 500 / 200 | $0.00007 | $0.070 |
| Question over retrieved documents | 4,000 / 500 | $0.00045 | $0.450 |
| Long document summary | 30,000 / 1,000 | $0.00310 | $3.10 |
Cheaper models to try
Nearest in price and below Ministral 3 3B 2512, each within a few points of it on the intelligence index.
| Model | Input / 1M | Output / 1M | Context | Index | Released |
|---|---|---|---|---|---|
| DeepSeek DeepSeek V4 Flash 0731 | $0.018 | $0.320 | 1.31M | 34.3 | |
| Google Gemma 3 12B | $0.050 | $0.150 | 131K | 3.8 | |
| OpenAI gpt-oss-120b | $0.037 | $0.170 | 131K | 11.6 | |
| OpenAI gpt-oss-20b | $0.018 | $0.090 | 131K | 9 |
Call it
Same OpenAI request shape. Naming the model pins it, so nothing is routed, and the answer is still checked.
python
from openai import OpenAI
client = OpenAI(base_url="https://www.llm11.com/v1", api_key="llm11_live_...")
res = client.chat.completions.create(
model="mistralai/ministral-3b-2512",
messages=[{"role": "user", "content": "Hello"}],
)More from Mistral
| Model | Input / 1M | Output / 1M | Context | Released |
|---|---|---|---|---|
| Mistral Medium 3.5 | $1.50 | $7.50 | 262K | |
| Mistral Small 4 | $0.150 | $0.600 | 262K | |
| Devstral 2 2512 | $0.400 | $2.00 | 262K | |
| Ministral 3 14B 2512 | $0.200 | $0.200 | 262K | |
| Ministral 3 8B 2512 | $0.150 | $0.150 | 262K | |
| Mistral Large 3 2512 | $0.500 | $1.50 | 262K |
Questions
- What is Ministral 3 3B 2512 with Jev?
- Jev is TypeSafe AI's System One model. It reads each request first and decides which model in your pool answers, with a calibrated confidence on that call. Ministral 3 3B 2512 does the answering when Jev sends it work, or when you name it yourself. Ministral 3 3B 2512 falls in the llm11-fast band, and the labs and limits it meets make it eligible for that pack. Packs keep a spread of eight models per band, and it is not one of the eight today, so it is only picked when you name it in a custom pool. Routing decisions are made by Jev, TypeSafe AI's System One model.
- Is Ministral 3 3B 2512 a System One model?
- No. Ministral 3 3B 2512 answers requests. The System One model is Jev, which decides which model answers each one, so the two do different jobs and llm11 uses both.
- How much does Ministral 3 3B 2512 cost?
- $0.100 per million input tokens and $0.100 per million output tokens. A 4,000 token prompt with a 500 token answer costs about $0.00045, so a thousand of them is about $0.450. llm11 passes provider list price through and charges 5% when you buy credits.
- What is the Ministral 3 3B 2512 context window?
- 131,072 tokens, with up to 104,857 tokens in a single reply.
- Does Ministral 3 3B 2512 support tool calling and structured output?
- The catalogue lists tool calling and structured output, but it does not list reasoning controls.
- Can I use Ministral 3 3B 2512 through llm11?
- Yes. Set model to mistralai/ministral-3b-2512 on the OpenAI-compatible endpoint and the request goes to it directly, with verification still running.
- Does Jev route to Ministral 3 3B 2512?
- Routing decisions are made by Jev, TypeSafe AI's System One model. Ministral 3 3B 2512 falls in the llm11-fast band, and the labs and limits it meets make it eligible for that pack. Packs keep a spread of eight models per band, and it is not one of the eight today, so it is only picked when you name it in a custom pool.
- What is a cheaper alternative to Ministral 3 3B 2512?
- Nearest in price and below it: DeepSeek DeepSeek V4 Flash 0731 at $0.018 in and $0.320 out, Google Gemma 3 12B at $0.050 in and $0.150 out and OpenAI gpt-oss-120b at $0.037 in and $0.170 out. Nearest in price and below Ministral 3 3B 2512, each within a few points of it on the intelligence index.