Meta Superintelligence Labs launched Muse Image, an image model that can call its own coding and web-search tools mid-generation, and opened an early preview of Muse Video, a companion video model built on the same pretraining base.
Muse Image is available now inside Meta AI, Instagram Stories, and WhatsApp. Muse Video is preview-only with no public release date attached. Both models sit in the top three of the human-preference Arena leaderboard for their tasks, according to Meta's announcement.
Muse Image can run coding and web-search tools while it generates
Muse Image behaves differently from a standard text-to-image model. It operates agentically. Meta says the model refines its own output, iterates on precise edits, and composes a single image from multiple reference inputs without a human re-prompting at each step.
Two tool integrations do the heavy lifting:
Coding tool. Muse Image generates accurate plots and QR codes, and can build animated GIFs, websites with embedded images, and interactive visual games.
Search tool. The model pulls real-time information from the web to improve factual accuracy on knowledge-intensive prompts, rather than relying only on its training data.
Meta also reports an "approximately log-linear scaling relationship" between output quality and the amount of reasoning and tool calls the model makes at inference time. In practice, letting the model think and call tools longer produces better results. The company attributes the self-refinement to reinforcement learning, where the behavior emerged during training. Muse Image also connects to Meta's Muse Spark for joint planning and accepts inline text and image prompts together.



