TBPN

September 3, 2026

52 stories / новостей

Meta’s MuseSpark 1.3 posts strong benchmark results but does not sweep evaluations

Meta’s MuseSpark 1.3 was reported to score 75.4% higher than Gemini 3.8 Flash on DeepSuite, surpassing Opus 5 and GPT-5.6 SOL in that comparison. It scored 62 on the Artificial Intelligence Index, reportedly behind only the newest Claude models.

The model was not a universal leader: Opus 5 still outperformed it in several professional-work and computer-use evaluations. The comparability of DeepSuite with the other assessments, as well as independent reproducibility of the results, was not established.

MuseSpark 1.3 от Meta показала сильные результаты в бенчмарках, но не стала лидером во всех оценках

По приведённым результатам, MuseSpark 1.3 от Meta набрала на 75,4% больше Gemini 3.8 Flash в DeepSuite, опередив в этом сравнении Opus 5 и GPT-5.6 SOL. В индексе искусственного интеллекта модель получила 62 балла — предположительно уступив только новейшим моделям Claude.

Модель не стала безусловным лидером: Opus 5 по-прежнему превосходил её в ряде оценок профессиональной работы и использования компьютера. Сопоставимость DeepSuite с другими оценками и независимая воспроизводимость результатов не были установлены.

Read the full issue / Читать весь выпуск
Privacy ·