Descriptions:
Fahd Mirza delivers a first look at Meta’s Muse Spark 1.1, the latest multimodal reasoning model from Meta Super Intelligence Labs. Built as a major upgrade over the original Muse Spark, the model features a 1-million token context window with active context management, agentic orchestration capabilities (functioning as both a main orchestrator and sub-agent), and strong performance on professional tool-use and MCP benchmarks — where it leads ahead of Claude Opus and GPT-5.5 on Job Bench and MCP Atlas respectively.
The video walks through several live tests run via meta.ai’s thinking mode: a highly complex physics simulation prompt requiring hundreds of independently simulated balls forming a rainbow Mandela pattern (judged better than Opus 4.8’s output), a multimodal physics-reasoning challenge identifying which truck in an image is braking (which Claude Sonnet 5 reportedly failed), and multilingual translation across nearly 80 languages. Mirza also scrolls through Meta’s official benchmarks covering coding, deep research Q&A, and multimodal tasks.
On coding benchmarks, Muse Spark 1.1 shows a large jump over the original Spark but does not top long-horizon tests against Opus or GPT-5.5. On Deep Search QA, it ties roughly with Opus while trailing GPT-5.5. The model is currently available through meta.ai in thinking mode and via a new public preview Meta Model API for developers.
📺 Source: Fahd Mirza · Published July 09, 2026
🏷️ Format: Review







