Zuckerberg’s Muse Code Loses to Anthropic on Meta’s Own Benchmark Charts
Mark Zuckerberg launched Muse Code in beta on Wednesday, Meta's first artificial intelligence (AI) coding agent. Anthropic's Claude Opus 5 beats it in all four comparisons Meta published at launch.
Meta released those charts anyway. The company is selling a cheaper tool rather than a better one. Independent test data suggests the gap is wider than Meta showed.
Follow us on X to get the latest news as it happens
Meta's Own Charts Hand Anthropic Every Round
Muse Spark 1.2 is the model inside Muse Code. It scored 82.9% on Terminal-Bench 2.1.
Claude Opus 5 scored 86.7% on the same test. Terminal-Bench comes from the Laude Institute and Stanford researchers. It sets 89 real jobs spanning system repair, data work, and security.
… Continue reading the full article at the original source below.



