nemotron-content-safety-reasoning-4b
Summary
nemotron-content-safety-reasoning-4b is an open-weight language model from NVIDIA, released on 22 Jan 2026. It accepts text and generates text, with a 128K-token context window and up to 4K output tokens per response. NVIDIA has deprecated this model.
Safety model for policy screening, moderation, and risk-aware routing workflows
Last verified 4 Oct 2026. Every figure on this page links to its source.
Specifications
| Developer | NVIDIA |
|---|---|
| Release date | 22 Jan 2026 [models.dev] |
| Status | Deprecated [models.dev] |
| Weights | Open weights [models.dev] |
| Context window | 128K (128,000 tokens) [models.dev] |
| Max output | 4K tokens [models.dev] |
| Input | Text [models.dev] |
| Output | Text [models.dev] |
| Reasoning mode | Yes [models.dev] |
| Tool calling | No [models.dev] |
nemotron-content-safety-reasoning-4b API pricing by provider
USD per million tokens. Sorted by input price.
We do not currently track a paid API listing for this model.
Frequently asked questions
What is the context window of nemotron-content-safety-reasoning-4b?
nemotron-content-safety-reasoning-4b has a context window of 128,000 tokens (128K) and can generate up to 4,096 tokens in a single response, according to models.dev.
Is nemotron-content-safety-reasoning-4b open source?
nemotron-content-safety-reasoning-4b is an open-weight model. Check the license terms before commercial use.
When was nemotron-content-safety-reasoning-4b released?
NVIDIA released nemotron-content-safety-reasoning-4b on 22 Jan 2026, according to models.dev.
Change history
No changes detected since we started tracking this model. We re-check every source daily.