Inference Net
Every Inference Net model UltraGPT can reach, kept in sync with the live catalog — pricing, context window and capabilities for each one.
2
Models tracked
1
Providers
0
Reasoning models
0
See images (vision)
128K
Largest context window
0
Free to run
All Inference Net models.
Search and filter this provider's lineup — pricing per million tokens, context window, modalities, reasoning effort and independent benchmark scores where available.
2 of 2 models
Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
