Nvidia model page

NVIDIA: Nemotron 3.5 Content Safety

nvidia/nemotron-3.5-content-safety

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

TextVisionStructured outputLong context
Input / 1M
$0.26
Output / 1M
$0.26
Context
131K
Markup
30%

Use this model through UAI

A dedicated public landing page for this synced catalog row.

Provider
Nvidia

Synced from the active catalog

Context window
131K tokens

Validated for gateway requests

UAI input
$0.26

Charged from your UAI balance

UAI output
$0.26

Same API key and billing flow

This page exists for direct linking, onboarding, and SEO. The same model row powers /v1/models, the public catalog, the app pricing screen, and request validation for /v1/chat/completions.

Why use this model through UAI

The same model route, but with one account, one balance, and one API shape.

Long-context inputs

NVIDIA: Nemotron 3.5 Content Safety exposes 131K tokens, which makes it a good fit for large documents, transcripts, and retrieval-heavy chat flows.

Multimodal prompts

The synced capabilities for this row indicate image or vision support, so you can route multimodal requests through the same UAI account and billing flow.

OpenAI-compatible access

Call nvidia/nemotron-3.5-content-safety with the same UAI API key, the same /v1/chat/completions shape, and the same request validation used across the rest of the catalog.

Pricing and access

Current synced numbers for this exact model row.

UAI input / 1M tokens
$0.26
UAI output / 1M tokens
$0.26
Nvidia input / 1M
$0.20
Nvidia output / 1M
$0.20
Context window
131,072
Current markup
30%

Need another option? Browse the full models catalog or compare rows in the signed-in pricing table.

Quickstart

Use this model with your UAI key over the standard OpenAI-compatible route.

Create a key in API Keys, then send requests to https://uai.sh/v1/chat/completions.

curl
curl https://uai.sh/v1/chat/completions \
  -H "Authorization: Bearer uai-YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "nvidia/nemotron-3.5-content-safety",
    "messages": [{"role": "user", "content": "Say hi in one short sentence."}],
    "stream": false
  }'
Python (OpenAI SDK)
from openai import OpenAI

client = OpenAI(base_url="https://uai.sh/v1", api_key="uai-YOUR_API_KEY")
response = client.chat.completions.create(
    model="nvidia/nemotron-3.5-content-safety",
    messages=[{"role": "user", "content": "Say hi in one short sentence."}],
    stream=False
)
print(response.choices[0].message.content)
Node.js (OpenAI SDK)
import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://uai.sh/v1",
  apiKey: "uai-YOUR_API_KEY"
});

const response = await client.chat.completions.create({
  model: "nvidia/nemotron-3.5-content-safety",
  messages: [{ role: "user", content: "Say hi in one short sentence." }],
  stream: false
});

console.log(response.choices[0].message.content);

Related models

Other active rows to compare without leaving the public catalog.

nvidia
NVIDIA: Nemotron 3.5 Lightning
nvidia/nemotron-3.5-lightning
Input / 1M$0.10
Context262K tokens
nvidia
NVIDIA: Nemotron 3.5 Content Safety (free)
nvidia/nemotron-3.5-content-safety:free
Input / 1M$0.00
Context128K tokens
nvidia
NVIDIA: Nemotron 3.5 Lightning (free)
nvidia/nemotron-3.5-lightning:free
Input / 1M$0.00
Context1M tokens