NVIDIA: Nemotron 3.5 Content Safety (free)
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
Intelligence overview
128K tok
Context window
8.2K tok
Max output
What NVIDIA: Nemotron 3.5 Content Safety (free) can do.
Reasoning
Extended reasoning
Works through problems step by step before answering, trading latency for accuracy on harder tasks.
Multimodal
Image understanding
Reads and reasons over images alongside text in the same prompt.
Development
Streaming
Streams tokens as they're generated instead of waiting on the full response.
Pay for what you use.
Every model in UltraGPT runs behind one flat subscription — this is the underlying per-token cost the catalog tracks.
Free
Free
Built for serious work.
Think
Works through multi-step problems with extended reasoning before it answers.
Research
Holds up to 128K tokens of source material in context without losing the thread.
Create
Reads images alongside a prompt to ground creative and design work in what's actually on screen.
Ask NVIDIA: Nemotron 3.5 Content Safety (free) anything.
A preview of the real UltraGPT composer — the model you see here is the model you get.
Continues in the UltraGPT app — one login, every model.
How NVIDIA: Nemotron 3.5 Content Safety (free) stacks up.
The closest models in the catalog by provider, capability and score — real entries, pulled from the same directory.
Technical details.
Specification
- Context window
- 128K tokens
- Max output
- 8.2K tokens
- Input modalities
- Text + Image
- Output modalities
- Text
- Release date
- Jun 4, 2026
- Streaming
- Supported
- Tool calling
- Not supported
- JSON mode
- Not supported
More from NVIDIA, and everything else.
This model is one entry in the full UltraGPT catalog — browse the rest, compare pricing and context windows, or filter by capability.
