ObviousBench

PUBLIC MODEL RELIABILITY · LAB VIEW

xAI

19 public configurations across 5 model families in the 2026-07-25 release. The global cost frontier stays available as context.

GLOBAL REFERENCE RETAINED

Where xAI sits in the complete public field

The lab is highlighted without changing the frontier calculation or zoom. This preserves the comparison frame that a provider-only chart would otherwise hide.

xAI in the global fieldAll 432 public configurations on the same global log-cost and pass3 scale as the homepage frontier. Public field xAIGlobal Pareto
xAI configurations highlighted in the global ObviousBench cost and reliability field 100% 95% 90% 80% 60% 40% 20% 10% More expensive Cheaper

This is a fixed global reference, not a recomputed xAI frontier.

CURATED DETAIL

Decision pages for xAI

These highlighted cards have an indexable evidence page. The rest remain in the full explorer rather than being turned into thin pages.

FULL LAB VIEW

All observed xAI model families

Sorted by highest observed answer pass3, then the lowest cost at that score. Release dates are not inferred because this release data does not encode a consistent vendor release-date field.

Grok 4.5Best observed 100% · $0.5465
Grok Build 0.1Best observed 99.3% · $0.6161
Grok 4.3Best observed 98.6% · $0.4246
Grok 4.20Best observed 98.6% · $0.8269
Grok 4.1 FastBest observed 93.8% · $0.07883