Gemma 2 27B with Jev
Gemma 2 27B is a model from Google, released 2024-07-13. It costs $0.650 per million input tokens and $0.650 per million output tokens, reads up to 8K tokens of context, and writes up to 2,048 tokens in one reply. Through llm11, set model to google/gemma-2-27b-it.
- Input price
- $0.650 / 1M
- Output price
- $0.650 / 1M
- Context window
- 8K tokens
- Max output
- 2K tokens
- Released
- 2024-07-13
- Accepts
- text
Read from the live catalogue and refreshed daily. At blended list price, 53% of the 332 priced models we can call cost more.
Using Gemma 2 27B with Jev
Jev is TypeSafe AI's System One model. It reads each request first and decides which model in your pool answers, with a calibrated confidence on that call. Gemma 2 27B does the answering when Jev sends it work, or when you name it yourself. At list price Gemma 2 27B falls in the llm11-balanced band, but no pack routes to it because its context window is under 32K tokens, the floor for pack members. You can still pin it: pass google/gemma-2-27b-it as model and routing steps aside while verification still runs.
Jev is new, so we have not published results for Gemma 2 27B with and without it, and this page does not invent any. Every response comes with a receipt that names the model that answered and what it cost against your most expensive model, so you can measure the pairing on your own traffic.
What a request costs
List price arithmetic, the same sum a receipt does. llm11 adds nothing per request; the only fee is on buying credits.
| Request | Tokens in / out | One | A thousand |
|---|---|---|---|
| Short chat turn | 500 / 200 | $0.00046 | $0.455 |
| Question over retrieved documents | 4,000 / 500 | $0.00293 | $2.93 |
| Long document summary | 30,000 / 1,000 | $0.020 | $20.15 |
Cheaper models to try
The catalogue has no benchmark score for Gemma 2 27B, so this is a price ordering. It does not claim these answer as well.
| Model | Input / 1M | Output / 1M | Context | Released |
|---|---|---|---|---|
| Qwen Qwen3 VL 235B A22B Instruct | $0.210 | $1.90 | 262K | |
| Cohere Command A+ | $0.300 | $1.50 | 192K | |
| Qwen Qwen3.5 Plus 2026-02-15 | $0.260 | $1.56 | 1M | |
| Google Gemini 3.1 Flash Lite | $0.250 | $1.50 | 1.05M |
Call it
Same OpenAI request shape. Naming the model pins it, so nothing is routed, and the answer is still checked.
python
from openai import OpenAI
client = OpenAI(base_url="https://www.llm11.com/v1", api_key="llm11_live_...")
res = client.chat.completions.create(
model="google/gemma-2-27b-it",
messages=[{"role": "user", "content": "Hello"}],
)More from Google
| Model | Input / 1M | Output / 1M | Context | Released |
|---|---|---|---|---|
| Gemini 3.8 Flash | $0.750 | $3.75 | 1.05M | |
| Gemini 3.7 Flash | $0.750 | $3.75 | 1.05M | |
| Gemini 3.5 Flash Lite | $0.300 | $2.50 | 1.05M | |
| Gemini 3.6 Flash | $0.750 | $3.75 | 1.05M | |
| Gemini 3.5 Flash | $1.50 | $9.00 | 1.05M | |
| Gemini 3.1 Flash Lite | $0.250 | $1.50 | 1.05M |
Questions
- What is Gemma 2 27B with Jev?
- Jev is TypeSafe AI's System One model. It reads each request first and decides which model in your pool answers, with a calibrated confidence on that call. Gemma 2 27B does the answering when Jev sends it work, or when you name it yourself. At list price Gemma 2 27B falls in the llm11-balanced band, but no pack routes to it because its context window is under 32K tokens, the floor for pack members. You can still pin it: pass google/gemma-2-27b-it as model and routing steps aside while verification still runs. Routing decisions are made by Jev, TypeSafe AI's System One model.
- Is Gemma 2 27B a System One model?
- No. Gemma 2 27B answers requests. The System One model is Jev, which decides which model answers each one, so the two do different jobs and llm11 uses both.
- How much does Gemma 2 27B cost?
- $0.650 per million input tokens and $0.650 per million output tokens. A 4,000 token prompt with a 500 token answer costs about $0.00293, so a thousand of them is about $2.93. llm11 passes provider list price through and charges 5% when you buy credits.
- What is the Gemma 2 27B context window?
- 8,192 tokens, with up to 2,048 tokens in a single reply.
- Does Gemma 2 27B support tool calling and structured output?
- The catalogue lists structured output, but it does not list tool calling and reasoning controls.
- Can I use Gemma 2 27B through llm11?
- Yes. Set model to google/gemma-2-27b-it on the OpenAI-compatible endpoint and the request goes to it directly, with verification still running.
- Does Jev route to Gemma 2 27B?
- Routing decisions are made by Jev, TypeSafe AI's System One model. At list price Gemma 2 27B falls in the llm11-balanced band, but no pack routes to it because its context window is under 32K tokens, the floor for pack members. You can still pin it: pass google/gemma-2-27b-it as model and routing steps aside while verification still runs.
- What is a cheaper alternative to Gemma 2 27B?
- Nearest in price and below it: Qwen Qwen3 VL 235B A22B Instruct at $0.210 in and $1.90 out, Cohere Command A+ at $0.300 in and $1.50 out and Qwen Qwen3.5 Plus 2026-02-15 at $0.260 in and $1.56 out. The catalogue has no benchmark score for Gemma 2 27B, so this is a price ordering. It does not claim these answer as well.