Inference.net LLM API Models & Pricing

Browse Inference.net LLM API models with normalized prompt pricing, output pricing, context windows, release dates, and popularity signals.

2Models tracked
2Priced models
0Popular models
2Latest releases

Provider shortlist

Use Inference.net as a model-level shortlist, not a single default choice.

Start from the cost, context, and popularity picks below, then open the model pages or head-to-head comparisons before choosing an API.

Cost pick Schematron V2 Turbo $0.1 sample workload
Context pick Schematron V2 Turbo 128K context window
Popular pick No current pick No popular models are available.
When to avoid Verify model fit Do not choose by provider name alone. Compare model-level input price, output price, context window, release timing, and current availability before production use.

Inference.net comparisons

Search Inference.net comparisons
ComparisonSample Cost WinnerLarger ContextNewest Release
🔥DeepSeek V4 Flash 0423 vs NewSchematron V2 TurboSchematron V2 TurboDeepSeek V4 Flash 0423
🔥DeepSeek V4 Flash 0731 vs NewSchematron V2 TurboSchematron V2 TurboDeepSeek V4 Flash 0731
🔥Gemini 3.8 Flash vs NewSchematron V2 TurboSchematron V2 TurboGemini 3.8 Flash
🔥GLM 5.3 Flash vs NewSchematron V2 TurboSchematron V2 TurboGLM 5.3 Flash
🔥GLM 5.3 vs NewSchematron V2 TurboSchematron V2 TurboGLM 5.3
🔥GPT-5.6 Luna vs NewSchematron V2 TurboSchematron V2 TurboGPT-5.6 Luna
🔥Hy3 vs NewSchematron V2 TurboSchematron V2 TurboHy3
🔥Hy4 preview vs NewSchematron V2 TurboSchematron V2 TurboHy4 preview

Cross-Provider Alternatives

Open alternatives hub

Use this shortlist when Inference.net is not a hard requirement and your first constraint is workload cost.

AlternativeProviderSample CostInput / 1MOutput / 1MContext
🔥Nemotron 3 Ultra (free)NVIDIA$0$0$01M
NewLing 3.0 Flash VL (free)inclusionAI$0$0$0262.14K
NewNex-N2.5-Mini (free)Nex AGI$0$0$0262.14K
NewNex-N2.5-Pro (free)Nex AGI$0$0$0262.14K
NewLing 3.0 Flash Sante (free)inclusionAI$0$0$0262.14K
Ling 3.0 Flash Fin (free)inclusionAI$0$0$0262.14K

Inference.net model catalog

Browse all models
ModelPromptOutputContextPopularity
NewSchematron V2 Turbo$0.03$0.15128K
NewSchematron V2 Small$0.05$0.23128K