Excellent formatting, no regulatory credibility
The presentation-substance gap in frontier Claude models — a Panel essay following the three-way and Fable-5 briefings
The three frontier Claude models — Sonnet 4.6, Opus 4.7, Fable 5 — land within a three-point band for materially-safe answers on live regulatory questions: 11–13%. That is the reliability floor of the configuration, not a rounding error. This essay argues that bigger, newer, or search-heavier does not close the gap — and that the presentation optimisation is on the wrong side of the tension for regulated use.
Read essay →