More Models from NVIDIA
Compare other engines available in this laboratory.
SWE-Bench Verified: 70.7–71.9%, PinchBench: 90.0%, RULER @1M: 94.7%, LiveCodeBench v6: 89.0%, IOI 2025: 570, GPQA (no tools): 87.0%, IFBench: 81.7%, IMOAnswerBench: 88.6–92.3%, Terminal-Bench 2.1: 56.4%, Artificial Analysis Intelligence Index: 48 (highest US open model) | NVIDIA's most capable model (released Jun 4, 2026).
SWE-Bench Verified: 51.56% (BF16) / 52.80% (NVFP4), PinchBench: 85.37%, GPQA Diamond: 75.44%, MMLU Pro: 81.94%, Terminal-Bench 2.1: 24.58%, HLE: 11.72%, IFBench (loose): 71.88%, BrowseComp: 36.97%, AA Intelligence Index: 23.6, AA Coding Index: 26.8, AA Agentic Index: 13.8, AA Non-Hallucination Rate: 62.4% | NVIDIA's highest-efficiency model, purpose-built for the execution layer of always-on agents (released Aug 11, 2026).