TBPN

← Full issue

September 3, 2026

Meta’s MuseSpark 1.3 posts strong benchmark results but does not sweep evaluations

Meta’s MuseSpark 1.3 was reported to score 75.4% higher than Gemini 3.8 Flash on DeepSuite, surpassing Opus 5 and GPT-5.6 SOL in that comparison. It scored 62 on the Artificial Intelligence Index, reportedly behind only the newest Claude models.

The model was not a universal leader: Opus 5 still outperformed it in several professional-work and computer-use evaluations. The comparability of DeepSuite with the other assessments, as well as independent reproducibility of the results, was not established.

Privacy ·