The best uncensored NVIDIA model is Nemotron 3 Nano Omni (free) (14.9 intelligence) — served without a provider-level moderation layer. Nemotron 3 Nano 30B A3B (7.4) is next.
Large language models served without a provider-level moderation layer, ranked by intelligence. "Uncensored" here means the API endpoint doesn't add its own filtering on top of the model — each model's own alignment training still applies, and open-weight entries can be self-hosted for full control. Relevant for fiction, roleplay, red-teaming and safety research, where an extra moderation layer causes over-refusal on legitimate prompts.
The best uncensored NVIDIA model is Nemotron 3 Nano Omni (free) (14.9 intelligence) — served without a provider-level moderation layer. Nemotron 3 Nano 30B A3B (7.4) is next.
Nemotron 3 Nano 30B A3B (7.4) is the closest alternative on this metric. See the full ranking above for the tradeoffs.
modelgrep tracks 10 NVIDIA models with live benchmarks, speed, latency and per-provider pricing, led on intelligence by Nemotron 3 Nano Omni (free). 2 of them qualify for this ranking.