More Models from DeepSeek
Compare other engines available in this laboratory.
Terminal-Bench 2.1: 87.9%, DeepSWE: 62.7%, CyberGym: 83.3%, AutomationBench: 31.8%, Toolathlon-Verified: 74.1%, HLE w/ tools: 60.0%, GPQA Diamond: 90.1%, Artificial Analysis Intelligence Index: 53.2, AA Coding Index: 68.8, AA Agentic Index: 49.6 | The official GA release of V4 Pro, superseding the April preview (released Aug 13, 2026).
Terminal-Bench 2.1: 82.7%, CyberGym: 76.7%, Toolathlon-Verified: 70.3%, DSBench-FullStack: 68.7%, DSBench-Hard: 59.6%, DeepSWE: 54.4%, NL2Repo: 54.2%, Agents' Last Exam: 25.2%, AutomationBench Public: 25.1%, Artificial Analysis Intelligence Index: 50 | The official GA release of V4 Flash, superseding the April preview (released Jul 31, 2026).
Strong on long-document and routine coding | Cost-effective high-volume tasks, long-document processing, routine coding, local deployment on consumer hardware.