Developed by NVIDIA, Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model fine-tuned from Google Gemma-3-4B. Supporting an expansive context length of 128,000 tokens, this model moderates both user inputs and generated responses for LLMs and VLMs. It provides robust safety filtering to ensure aligned and secure interactions across multiple modalities.
Pricing
- Input Tokens: $0 /M tokens
- Output Tokens: $0 /M tokens
- Cache Read: $0 /M tokens
Input Modalities
- Text
- Vision
Output Modalities
- Text
Context length
- 131K tokens
Capabilities
- Thinking
- Streaming
- Tool calling
- Web search
- URL context
- Code interpreter
- Computer use
- File search
- Memory tool
- Structured outputs
- Citations
- Prompt caching
- Background mode
- Server-side sessions
Try this model
Python
