Inference.net models
Inference.net builds the Schematron models, small models trained for one job: turning HTML into structured JSON against a caller-supplied schema. Zeplik carries both V2 servings — Turbo, tuned for throughput on high-volume extraction, and Small, tuned for extraction quality on complex schemas and long pages.
Ask a Inference.net model now
All 2 Inference.net models
Schematron V2 TurboFast and light
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads.
128K ctx1.8 cr/M outSeptember 2026
Schematron V2 SmallFast and light
Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages.
128K ctx2.8 cr/M outSeptember 2026
Pricing at a glance
| Model | Context | Input cr/M | Output cr/M | Released |
|---|---|---|---|---|
| Schematron V2 Turbo | 128K | 0.36 | 1.8 | September 2026 |
| Schematron V2 Small | 128K | 0.60 | 2.8 | September 2026 |
Frequently asked questions
- Which Inference.net models can I use on Zeplik?
- Zeplik carries 2 Inference.net models: Schematron V2 Turbo and Schematron V2 Small. They all run in the same chat, so you do not need a separate Inference.net account or API key.
- How much do Inference.net models cost on Zeplik?
- Output tokens run from 1.8 to 2.8 credits per million depending on which Inference.net model you pick (raw provider rates $0.15 to $0.23 per million). 1 credit = $0.10, computed from the raw provider rate with a 1.20x margin.
- Which Inference.net model has the largest context window?
- Schematron V2 Turbo, at 128K tokens (128,000). The context window covers the conversation plus any documents you attach.
- Do Inference.net models support images and tool calling?
- None of the 2 accept image input, and none support tool calling.
- Do I need a paid plan to use Inference.net models on Zeplik?
- Both need a paid plan (Plus or above). Usage is then billed in credits from your monthly allowance.
- What is the newest Inference.net model on Zeplik?
- Schematron V2 Turbo, released September 2026. Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads.
Other providers
- OpenAI (100)
- Qwen (53)
- Google (42)
- Anthropic (29)
- Mistral (26)
- Z.ai (18)
- DeepSeek (15)
- NVIDIA (11)
- All models
Explore Zeplik
- 467 AI skillsReady-to-run expert methods that run on any Inference.net model, or any other model you pick.
- 997 integrationsConnect Gmail, Slack, GitHub, Notion and hundreds more so the assistant can work in your apps.
- Simple pricingPlus, Pro and Max plans with a monthly credit allowance that works across every model.