import requests r = requests.post( "https://api.aimlapi.com/v1/chat/completions", headers={"Authorization": "Bearer " + AIMLAPI_KEY}, json={ "model": "nvidia/nemotron-3.5-content-safety", "messages": [ { "role": "user", "content": "Hello!" } ] }, ) print(r.json())
const r = await fetch("https://api.aimlapi.com/v1/chat/completions", { method: "POST", headers: { Authorization: `Bearer ${process.env.AIMLAPI_KEY}`, "Content-Type": "application/json", }, body: JSON.stringify({ "model": "nvidia/nemotron-3.5-content-safety", "messages": [ { "role": "user", "content": "Hello!" } ] }), }); console.log(await r.json());
curl -X POST https://api.aimlapi.com/v1/chat/completions \ -H "Authorization: Bearer $AIMLAPI_KEY" \ -H "Content-Type: application/json" \ -d '{"model":"nvidia/nemotron-3.5-content-safety","messages":[{"role":"user","content":"Hello!"}]}'
OpenAI-compatible — swap the base URL and it works with your existing SDK.
| Type | Price |
|---|---|
| Input | |
| Output | |
| Model | Input | Output | Context | Best for |
|---|---|---|---|---|
Nemotron 3.5 Content Safety This page | Vision and document understanding | |||
| Reasoning + agents | ||||
| Coding + agents | ||||
| Coding + agents |
The model has a 131,072-token context window and can return up to 117,964 output tokens in a single response.
$0.27508 per 1M input tokens and $0.27508 per 1M output tokens on AI/ML API. Input and output are priced the same.
It accepts image and text input and returns text. That lets a single request carry a prompt, an optional image, and an optional response to be judged together.
Yes. Streaming and vision are supported.
June 4, 2026.
Use nvidia/nemotron-3.5-content-safety as the model id.
Send a request to the chat completions endpoint with the model id nvidia/nemotron-3.5-content-safety. The request shape is the same as for any other chat model on AI/ML API.