Command R7B (12-2024) with Jev
Command R7B (12-2024) is a model from Cohere, released 2024-12-14. It costs $0.037 per million input tokens and $0.150 per million output tokens, reads up to 128K tokens of context, and writes up to 4,000 tokens in one reply. Through llm11, set model to cohere/command-r7b-12-2024.
- Input price
- $0.037 / 1M
- Output price
- $0.150 / 1M
- Context window
- 128K tokens
- Max output
- 4K tokens
- Released
- 2024-12-14
- Accepts
- text
Read from the live catalogue and refreshed daily. At blended list price, 96% of the 332 priced models we can call cost more.
Using Command R7B (12-2024) with Jev
Jev is TypeSafe AI's System One model. It reads each request first and decides which model in your pool answers, with a calibrated confidence on that call. Command R7B (12-2024) does the answering when Jev sends it work, or when you name it yourself. Command R7B (12-2024) falls in the llm11-fast band, and the labs and limits it meets make it eligible for that pack. Packs keep a spread of eight models per band, and it is not one of the eight today, so it is only picked when you name it in a custom pool.
Jev is new, so we have not published results for Command R7B (12-2024) with and without it, and this page does not invent any. Every response comes with a receipt that names the model that answered and what it cost against your most expensive model, so you can measure the pairing on your own traffic.
What a request costs
List price arithmetic, the same sum a receipt does. llm11 adds nothing per request; the only fee is on buying credits.
| Request | Tokens in / out | One | A thousand |
|---|---|---|---|
| Short chat turn | 500 / 200 | $0.00005 | $0.049 |
| Question over retrieved documents | 4,000 / 500 | $0.00022 | $0.225 |
| Long document summary | 30,000 / 1,000 | $0.00127 | $1.27 |
Cheaper models to try
The catalogue has no benchmark score for Command R7B (12-2024), so this is a price ordering. It does not claim these answer as well.
| Model | Input / 1M | Output / 1M | Context | Released |
|---|---|---|---|---|
| Google Gemma 3 4B | $0.050 | $0.100 | 131K | |
| Mistral Mistral Small 3 | $0.050 | $0.080 | 33K | |
| Meta Llama 3.1 8B Instruct | $0.050 | $0.080 | 131K | |
| Qwen Qwen3.7 Flash | $0.030 | $0.130 | 1M |
Call it
Same OpenAI request shape. Naming the model pins it, so nothing is routed, and the answer is still checked.
python
from openai import OpenAI
client = OpenAI(base_url="https://www.llm11.com/v1", api_key="llm11_live_...")
res = client.chat.completions.create(
model="cohere/command-r7b-12-2024",
messages=[{"role": "user", "content": "Hello"}],
)More from Cohere
| Model | Input / 1M | Output / 1M | Context | Released |
|---|---|---|---|---|
| Command A+ | $0.300 | $1.50 | 192K | |
| Command A | $2.50 | $10 | 256K | |
| Command R (08-2024) | $0.150 | $0.600 | 128K | |
| Command R+ (08-2024) | $2.50 | $10 | 128K |
Questions
- What is Command R7B (12-2024) with Jev?
- Jev is TypeSafe AI's System One model. It reads each request first and decides which model in your pool answers, with a calibrated confidence on that call. Command R7B (12-2024) does the answering when Jev sends it work, or when you name it yourself. Command R7B (12-2024) falls in the llm11-fast band, and the labs and limits it meets make it eligible for that pack. Packs keep a spread of eight models per band, and it is not one of the eight today, so it is only picked when you name it in a custom pool. Routing decisions are made by Jev, TypeSafe AI's System One model.
- Is Command R7B (12-2024) a System One model?
- No. Command R7B (12-2024) answers requests. The System One model is Jev, which decides which model answers each one, so the two do different jobs and llm11 uses both.
- How much does Command R7B (12-2024) cost?
- $0.037 per million input tokens and $0.150 per million output tokens. A 4,000 token prompt with a 500 token answer costs about $0.00022, so a thousand of them is about $0.225. llm11 passes provider list price through and charges 5% when you buy credits.
- What is the Command R7B (12-2024) context window?
- 128,000 tokens, with up to 4,000 tokens in a single reply.
- Does Command R7B (12-2024) support tool calling and structured output?
- The catalogue lists structured output, but it does not list tool calling and reasoning controls.
- Can I use Command R7B (12-2024) through llm11?
- Yes. Set model to cohere/command-r7b-12-2024 on the OpenAI-compatible endpoint and the request goes to it directly, with verification still running.
- Does Jev route to Command R7B (12-2024)?
- Routing decisions are made by Jev, TypeSafe AI's System One model. Command R7B (12-2024) falls in the llm11-fast band, and the labs and limits it meets make it eligible for that pack. Packs keep a spread of eight models per band, and it is not one of the eight today, so it is only picked when you name it in a custom pool.
- What is a cheaper alternative to Command R7B (12-2024)?
- Nearest in price and below it: Google Gemma 3 4B at $0.050 in and $0.100 out, Mistral Mistral Small 3 at $0.050 in and $0.080 out and Meta Llama 3.1 8B Instruct at $0.050 in and $0.080 out. The catalogue has no benchmark score for Command R7B (12-2024), so this is a price ordering. It does not claim these answer as well.