ObviousBench

PUBLIC MODEL RELIABILITY · CONFIGURATION DETAIL

GPT 5.6 Terra

Observed settings from the 2026-08-13 public release. This is a configuration readout, not a claim about general model quality or a vendor guarantee.

Highest observed answer pass³77.1%

None / disabled · $0.07145

Settings observed6

ordered deepest to lightest

Benchmark sample144

items · 3 answers each

Observed answer pass3 by reasoning settingSettings are ordered from deeper to lighter / no reported reasoning.
GPT 5.6 Terra observed pass³ and full-run cost by setting 100% 95% 90% 70% 96.5%Max$0.2091 95.8%Extra high$0.2154 95.8%High$0.1947 94.4%Medium$0.1747 94.4%Low$0.1674 77.1%None / disabled$0.07145

Point labels show answer pass3 and estimated full-run cost. The table below retains the exact released observations.

RELEASED OBSERVATIONS

All visible settings

Answer pass3 means all three sampled answers were correct. Cost is the estimated public API cost for a full benchmark run.

SettingAnswer pass3Full-run costReasoning evidence
Max96.5%$0.2091Max
Extra high95.8%$0.2154Extra high
High95.8%$0.1947High
Medium94.4%$0.1747Medium
Low94.4%$0.1674Low
None / disabled77.1%$0.07145None / disabled

How to read this page

These values are a release-bound observation set, not an estimate of every task or production workflow. Compare the full public field and global Pareto reference before making a cost or reliability decision.

Compare every public configuration on the cost frontier →