Inference.net models
llm11 can call 2 Inference.net models through one API, with Jev, TypeSafe AI's System One model, deciding which one answers each request. Blended prices run from $0.060 to $0.095 per million tokens. The newest is Schematron V2 Small, released 2026-09-12.
| Model | Input / 1M | Output / 1M | Context | Released |
|---|---|---|---|---|
| Schematron V2 Small | $0.050 | $0.230 | 128K | |
| Schematron V2 Turbo | $0.030 | $0.150 | 128K |
Read from the live catalogue and refreshed daily. Prices are per million tokens.
Questions
- Which Inference.net models can I call through llm11?
- 2 today, listed with live prices on this page. The list follows the upstream catalogue, so a new Inference.net model appears here without anyone editing the page.
- What is the cheapest and the most expensive Inference.net model?
- By blended price, weighting three parts input to one part output: Schematron V2 Turbo is cheapest at $0.060 per million tokens, and Schematron V2 Small is the most expensive at $0.095.
- What is the newest Inference.net model?
- Schematron V2 Small, released 2026-09-12. Call it with model set to inference-net/schematron-v2-small.