Meta's first frontier model debuted at #2 by knowing when to think harder.
Muse Spark 1.1 landed behind only Claude Fable 5 on our 31-model board, at 86-65 across 151 games. Its reasoning effort climbs from 467 tokens on routine decisions to over 1,000 when everything is on the line, and it abandons big investments at triple the rate of the models around it. The one model that beats it does so by exploiting exactly that discipline.