ObviousBench

PUBLIC MODEL RELIABILITY · LAB VIEW

DeepSeek

8 public configurations across 8 model families in the 2026-07-25 release. The global cost frontier stays available as context.

GLOBAL REFERENCE RETAINED

Where DeepSeek sits in the complete public field

The lab is highlighted without changing the frontier calculation or zoom. This preserves the comparison frame that a provider-only chart would otherwise hide.

DeepSeek in the global fieldAll 432 public configurations on the same global log-cost and pass3 scale as the homepage frontier. Public field DeepSeekGlobal Pareto
DeepSeek configurations highlighted in the global ObviousBench cost and reliability field 100% 95% 90% 80% 60% 40% 20% 10% More expensive Cheaper

This is a fixed global reference, not a recomputed DeepSeek frontier.

CURATED DETAIL

Decision pages for DeepSeek

These highlighted cards have an indexable evidence page. The rest remain in the full explorer rather than being turned into thin pages.

FULL LAB VIEW

All observed DeepSeek model families

Sorted by highest observed answer pass3, then the lowest cost at that score. Release dates are not inferred because this release data does not encode a consistent vendor release-date field.

DeepSeek v4 ProBest observed 79.9% · $0.04693
DeepSeek v4 FlashBest observed 73.6% · $0.008068
DeepSeek v3.2 ExpBest observed 61.1% · $0.005534
DeepSeek v3.1 TerminusBest observed 58.3% · $0.01981
DeepSeek Chat v3.1Best observed 54.9% · $0.00504
DeepSeek v3.2Best observed 50.7% · $0.003792
DeepSeek Chat v3 0324Best observed 48.6% · $0.006335
DeepSeek ChatBest observed 42.4% · $0.007478