Provider

Inference Net

Every Inference Net model UltraGPT can reach, kept in sync with the live catalog — pricing, context window and capabilities for each one.

2 models from Inference Net Up to 128K tokens of context
Open the appBrowse every provider

2

Models tracked

1

Providers

0

Reasoning models

0

See images (vision)

128K

Largest context window

0

Free to run

Directory

All Inference Net models.

Search and filter this provider's lineup — pricing per million tokens, context window, modalities, reasoning effort and independent benchmark scores where available.

2 of 2 models

Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...

StreamingJSON mode
Context window128K tok
Max output4.1K tok
ReleasedSep 12, 2026
Knowledge cutoff
TextText
Input$0.05 / 1M tok
Output$0.23 / 1M tok
Cache read$0.05 / 1M tok

Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...

StreamingJSON mode
Context window128K tok
Max output8.2K tok
ReleasedSep 12, 2026
Knowledge cutoff
TextText
Input$0.03 / 1M tok
Output$0.15 / 1M tok
Cache read$0.03 / 1M tok

Every Inference Net model above, one subscription away.

UltraGPT puts this whole directory behind a single login and one flat monthly price — no separate bill per provider.