Nemotron 3.5 Content Safety API

nvidia/nemotron-3.5-content-safety
A compact multimodal safety moderator, Nemotron 3.5 Content Safety classifies prompts, images, and optional responses as safe or unsafe, especially for AI input/output moderation.
Context
128K tokens
Input
$0.27508 / 1M tokens
Output
$0.27508 / 1M tokens
Released
Jun 4, 2026

How to use Nemotron 3.5 Content Safety API

Install any OpenAI-compatible SDK, point it at api.aimlapi.com/v1, and set the model to nvidia/nemotron-3.5-content-safety.
import requests

r = requests.post(
    "https://api.aimlapi.com/v1/chat/completions",
    headers={"Authorization": "Bearer " + AIMLAPI_KEY},
    json={
      "model": "nvidia/nemotron-3.5-content-safety",
      "messages": [
        {
          "role": "user",
          "content": "Hello!"
        }
      ]
    },
)
print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", {
  method: "POST",
  headers: {
    Authorization: `Bearer ${process.env.AIMLAPI_KEY}`,
    "Content-Type": "application/json",
  },
  body: JSON.stringify({
    "model": "nvidia/nemotron-3.5-content-safety",
    "messages": [
      {
        "role": "user",
        "content": "Hello!"
      }
    ]
  }),
});
console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \
  -H "Authorization: Bearer $AIMLAPI_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"nvidia/nemotron-3.5-content-safety","messages":[{"role":"user","content":"Hello!"}]}'

OpenAI-compatible — swap the base URL and it works with your existing SDK.

Nemotron 3.5 Content Safety API Pricing

TypePrice
Input
$0.27508 / 1M tokens
Output
$0.27508 / 1M tokens

Nemotron 3.5 Content Safety Benchmarks

BenchmarkScoreWhat it measuresSourceRetrieved
MMMU
0.023
College-level multimodal understanding + reasoningSourceSeptember 21, 2026
DocVQA
0.058
Question answering over document imagesSourceSeptember 21, 2026
AI2D
0.001
Question answering over science diagramsSourceSeptember 21, 2026

Nemotron 3.5 Content Safety vs other models

ModelInputOutputContextBest for
$0.27508 / 1M tokens
$0.27508 / 1M tokens
128K tokens
Vision and document understanding
$0.52 / 1M
$0.52 / 1M
131K tokens
Reasoning + agents
$0.123786 / 1M tokens
$0.61893 / 1M tokens
256K tokens
Coding + agents
$0.06877 / 1M tokens
$0.27508 / 1M tokens
256K tokens
Coding + agents

Frequently asked questions

The model has a 131,072-token context window and can return up to 117,964 output tokens in a single response.

$0.27508 per 1M input tokens and $0.27508 per 1M output tokens on AI/ML API. Input and output are priced the same.

It accepts image and text input and returns text. That lets a single request carry a prompt, an optional image, and an optional response to be judged together.

Yes. Streaming and vision are supported.

June 4, 2026.

Use nvidia/nemotron-3.5-content-safety as the model id.

Send a request to the chat completions endpoint with the model id nvidia/nemotron-3.5-content-safety. The request shape is the same as for any other chat model on AI/ML API.

Start building with Nemotron 3.5 Content Safety

Get API Key
1000+ models, one API.