Model profile
Nemotron 3 Super 100B
Evidence coverage
1 of 321 tracked benchmarks is published. 1 is verified and 0 provisional. 1 of 8 categories are measured.
- Published / tracked
- 1 / 321
- Verified
- 1
- Provisional
- 0
- Categories with evidence
- 1 / 8
Evidence by category
- Agentic1 benchmarkVerified
- Coding0 benchmarksNot measured
- Reasoning0 benchmarksNot measured
- Knowledge0 benchmarksNot measured
- Math0 benchmarksNot measured
- Multilingual0 benchmarksNot measured
- Multimodal0 benchmarksNot measured
- Inst. Following0 benchmarksNot measured
Nemotron 3 Super 100B ranks #117 out of 200 models on the public leaderboard with an overall score of 50.08/100. It does not yet have enough sourced coverage for BenchLM's verified leaderboard. While not a frontier model, it offers specific advantages depending on the use case.
Nemotron 3 Super 100B is a open weight model with a 1M token context window. It processes queries without explicit chain-of-thought reasoning, offering faster response times and lower token usage.
This profile currently has 1 of 321 tracked benchmarks. BenchLM only exposes non-generated benchmark rows publicly, so missing categories stay blank until a sourced evaluation is available.
Peer position
Exact provisional scores and ranks for the closest listed peers. A score can appear before a model clears the evidence threshold for a rank, so equal scores can have different rank states.
Range 49.89–50.28
- DeepSeek Coder 2.0DeepSeekCompare#11450.28DeepSeek Coder 2.0 is #114 with a score of 50.28.
- Seed 1.6ByteDanceCompare#11550.23Seed 1.6 is #115 with a score of 50.23.
- GPT-OSS 120BOpenAICompare#11650.08GPT-OSS 120B is #116 with a score of 50.08.
- Nemotron 3 Super 100BCurrent modelNVIDIA#11750.08Nemotron 3 Super 100B is #117 with a score of 50.08.
- o4-mini (high)OpenAICompare#11849.98o4-mini (high) is #118 with a score of 49.98.
- DeepSeekMath V2DeepSeekCompare#11949.89DeepSeekMath V2 is #119 with a score of 49.89.
- Qwen2.5-1MAlibabaCompare#12049.89Qwen2.5-1M is #120 with a score of 49.89.
Category percentile
More
Relative position among models eligible for each sourced category. A higher percentile means a stronger position within that category's ranked cohort; 100 is highest.
Category evidence
Scores and ranks appear only where this model has published benchmark evidence. Categories without displayable source records remain not measured.
| Category | Score | Rank | Percentile | Weight | Benchmarks | Evidence |
|---|---|---|---|---|---|---|
| AgenticRank Not rankedWeight 22%1 benchmarkVerified | 56.7 | Not ranked | Not available | 22% | 1 benchmark | Verified |
| CodingWeight 20%0 benchmarksNot measured | Not measured | Not ranked | Not available | 20% | 0 benchmarks | Not measured |
| ReasoningWeight 17%0 benchmarksNot measured | Not measured | Not ranked | Not available | 17% | 0 benchmarks | Not measured |
| KnowledgeWeight 12%0 benchmarksNot measured | Not measured | Not ranked | Not available | 12% | 0 benchmarks | Not measured |
| MathWeight 5%0 benchmarksNot measured | Not measured | Not ranked | Not available | 5% | 0 benchmarks | Not measured |
| MultilingualWeight 7%0 benchmarksNot measured | Not measured | Not ranked | Not available | 7% | 0 benchmarks | Not measured |
| MultimodalWeight 12%0 benchmarksNot measured | Not measured | Not ranked | Not available | 12% | 0 benchmarks | Not measured |
| Inst. FollowingWeight 5%0 benchmarksNot measured | Not measured | Not ranked | Not available | 5% | 0 benchmarks | Not measured |
Chatbot Arena performance
Scroll horizontally to inspect confidence intervals and vote counts.
| View | Elo | Confidence interval | Votes |
|---|---|---|---|
| Text Overall | 1260 | Not available | Not available |
Benchmark Details
Rows below have a displayable published verification record. Each source link and provenance note remains in the page HTML while its category is closed. Source-unverified manual rows and generated rows stay hidden.
Agentic1 benchmark
Frequently Asked Questions
How does Nemotron 3 Super 100B perform overall in AI benchmarks?
Nemotron 3 Super 100B has 1 published benchmark scores on BenchLM, but it does not yet have enough non-generated coverage to receive a global overall rank.
Is Nemotron 3 Super 100B good for agentic tool use and computer tasks?
Nemotron 3 Super 100B has visible benchmark coverage in agentic tool use and computer tasks, but BenchLM does not currently assign it a global category rank there.
Is Nemotron 3 Super 100B open source?
Yes, Nemotron 3 Super 100B is an open weight model created by NVIDIA, meaning it can be downloaded and run locally or fine-tuned for specific use cases.
Does Nemotron 3 Super 100B have full benchmark coverage on BenchLM?
Not yet. Nemotron 3 Super 100B currently has 1 published benchmark scores out of the 321 benchmarks BenchLM tracks. BenchLM only exposes non-generated public benchmark rows, so missing categories stay blank until a sourced evaluation is available.
What is the context window size of Nemotron 3 Super 100B?
Nemotron 3 Super 100B has a published context window of 1M, which determines how much text it can process in a single interaction.
Related Resources
Choose with this week’s evidence
Join 2,000+ readers for ranking moves, new releases, pricing changes, and the evidence behind them.
Free. One email per week.