- Schematron V2 Turbo — $0.1 sample workload
- Schematron V2 Small — $0.17 sample workload
Inference.net LLM API Models & Pricing
Browse Inference.net LLM API models with normalized prompt pricing, output pricing, context windows, release dates, and popularity signals.
2Models tracked
2Priced models
0Popular models
2Latest releases
Provider shortlist
Use Inference.net as a model-level shortlist, not a single default choice.
Start from the cost, context, and popularity picks below, then open the model pages or head-to-head comparisons before choosing an API.
Cost pick
Schematron V2 Turbo
$0.1 sample workload
Context pick
Schematron V2 Turbo
128K context window
Popular pick
No current pick
No popular models are available.
When to avoid
Verify model fit
Do not choose by provider name alone. Compare model-level input price, output price, context window, release timing, and current availability before production use.
- Schematron V2 Turbo — 128K
- Schematron V2 Small — 128K
Inference.net comparisons
Search Inference.net comparisons| Comparison | Sample Cost Winner | Larger Context | Newest Release |
|---|---|---|---|
| 🔥DeepSeek V4 Flash 0423 vs NewSchematron V2 Turbo | Schematron V2 Turbo | DeepSeek V4 Flash 0423 | |
| 🔥DeepSeek V4 Flash 0731 vs NewSchematron V2 Turbo | Schematron V2 Turbo | DeepSeek V4 Flash 0731 | |
| 🔥Gemini 3.8 Flash vs NewSchematron V2 Turbo | Schematron V2 Turbo | Gemini 3.8 Flash | |
| 🔥GLM 5.3 Flash vs NewSchematron V2 Turbo | Schematron V2 Turbo | GLM 5.3 Flash | |
| 🔥GLM 5.3 vs NewSchematron V2 Turbo | Schematron V2 Turbo | GLM 5.3 | |
| 🔥GPT-5.6 Luna vs NewSchematron V2 Turbo | Schematron V2 Turbo | GPT-5.6 Luna | |
| 🔥Hy3 vs NewSchematron V2 Turbo | Schematron V2 Turbo | Hy3 | |
| 🔥Hy4 preview vs NewSchematron V2 Turbo | Schematron V2 Turbo | Hy4 preview |
Cross-Provider Alternatives
Open alternatives hubUse this shortlist when Inference.net is not a hard requirement and your first constraint is workload cost.
| Alternative | Provider | Sample Cost | Input / 1M | Output / 1M | Context |
|---|---|---|---|---|---|
| 🔥Nemotron 3 Ultra (free) | NVIDIA | $0 | $0 | $0 | 1M |
| NewLing 3.0 Flash VL (free) | inclusionAI | $0 | $0 | $0 | 262.14K |
| NewNex-N2.5-Mini (free) | Nex AGI | $0 | $0 | $0 | 262.14K |
| NewNex-N2.5-Pro (free) | Nex AGI | $0 | $0 | $0 | 262.14K |
| NewLing 3.0 Flash Sante (free) | inclusionAI | $0 | $0 | $0 | 262.14K |
| Ling 3.0 Flash Fin (free) | inclusionAI | $0 | $0 | $0 | 262.14K |
Inference.net model catalog
Browse all models| Model | Prompt | Output | Context | Popularity |
|---|---|---|---|---|
| NewSchematron V2 Turbo | $0.03 | $0.15 | 128K | |
| NewSchematron V2 Small | $0.05 | $0.23 | 128K |