RankingOmniScience Accuracy · source release snapshot; version not specified

OmniScience Accuracy · source release snapshot; version not specified

Data updated 12 Sept 2026

Bucket
Supporting evidence
Unit
percent
Direction
Higher is better
Version
source release snapshot; version not specified
Display harness
nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16
Board
https://huggingface.co/nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16

The available records have no admitted matched comparison in the capability core. Raw results remain available below.

Compare published benchmark results with category weights →

Models

ModelScoreHarnessEvidenceSource-recorded date
Qwen3.5-397B-A17BQwen35.9%
Reported settings & source

NVIDIA evaluation harness/settings per benchmark in model card; comparator scores are NVIDIA-reported under that protocol.

First-party reported result. Original cell: 35.9

nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 · Performance table: OmniScience Accuracy · reviewed 2026-09-06

nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16lab self-report2026-09-06
Kimi-K2.6Moonshot35.5%
Reported settings & source

NVIDIA evaluation harness/settings per benchmark in model card; comparator scores are NVIDIA-reported under that protocol.

First-party reported result. Original cell: 35.5

nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 · Performance table: OmniScience Accuracy · reviewed 2026-09-06

nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16lab self-report2026-09-06
GLM-5.1Z.ai31.3%
Reported settings & source

NVIDIA evaluation harness/settings per benchmark in model card; comparator scores are NVIDIA-reported under that protocol.

First-party reported result. Original cell: 31.3

nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 · Performance table: OmniScience Accuracy · reviewed 2026-09-06

nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16lab self-report2026-09-06
NVIDIA Nemotron 3 UltraNVIDIA24.1%
Reported settings & source

NVIDIA evaluation harness/settings per benchmark in model card; comparator scores are NVIDIA-reported under that protocol.

First-party reported result. Original cell: 24.1 The linked NVIDIA reproduction recipe explicitly enables thinking for this model and benchmark. It does not establish Low/Medium/High/Max or a fixed reasoning-token budget; the request adapter removes max_tokens and max_completion_tokens. Comparator settings are not specified by this recipe.

nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 · Performance table: OmniScience Accuracy · reviewed 2026-09-06

nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16lab self-report2026-09-06
MiniMax-M2.7MiniMax20.5%
Reported settings & source

NVIDIA evaluation harness/settings per benchmark in model card; comparator scores are NVIDIA-reported under that protocol.

First-party reported result. Original cell: 20.5

nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 · Performance table: OmniScience Accuracy · reviewed 2026-09-06

nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16lab self-report2026-09-06