Meta Unveils Muse Spark 1.2: Vision-to-Code, Robot Navigation, Audio-Visual Understanding
AIatMeta · x · 2026-08-21
Meta officially announced Muse Spark 1.2, supporting a broad range of multimodal tasks: turning visuals into working code, translating perception into physical action, and robust audio-visual understanding for video-heavy enterprise workflows.
Alongside new evals, Meta shared demos of the model's visual understanding and reasoning — starting with one where Muse Spark parses multimodal observations and calls tools to guide a robot through an unstructured environment to find a rubber duck.
Related event: Meta Unveils Muse Spark 1.2 with Visual-to-Code and Robot Orchestration(7 posts)→
More from Models
- Liquid AI releases DSpark draft models, speeding up decoding by up to 4x — JosephJacks_ · 2026-08-21
- Liquid AI Releases DSpark: Speculative Decoding Up to 3.18x Faster — helloiamleonie · 2026-08-21
- Grok Bot releases version 0.23.0 update with improvements — XFreeze · 2026-08-21
- User Warns Muse Spark is "Arrogant" and Dangerous, Advising Against Public Use — saibharadwaj · 2026-08-21
- Harvey unveils Tenet, a model post-trained for long-horizon legal work — rhythmrg · 2026-08-21
- Gemma hits 1 billion downloads as community explores use cases from underwater to space — osanseviero · 2026-08-21