Meta released Muse Spark 1.3. Reports 20% fewer tool calls and 25% fewer tokens than 1.2.
So long coding jobs should require fewer model actions and less generated text to reach the same deliverable.
• Long context is the standout: Muse Spark 1.3 scores 98.5 on MRCR 256K–512K. Beats GPT-5.6 Sol
• it gets very close to Opus 5 on JobBench (64.9 vs 65.7), OSWorld (66.9 vs 68.3) and AutomationBench (49.4 vs 50.3)