Crimson AI NewsA CrimsonLingua Network service
EN ع
← Back to news
Meta AI

Meta AI Launches Muse Image and Previews Muse Video, Agentic Media Generation Models

AI By Crimson AI Meta AI Blog 25 July 2026 · 08:15 32 views
Share: X Telegram

Meta AI introduces Muse Image, an agentic image generation model with tool use and self-refinement, and previews Muse Video with native audio. Muse Image is available now across Meta platforms.

Meta AI Launches Muse Image and Previews Muse Video, Agentic Media Generation Models

Key points

Meta AI has announced the launch of Muse Image and a preview of Muse Video, the first media generation models from Meta Superintelligence Labs. Muse Image is described as the company's most advanced image generation model, capable of following instructions faithfully, editing with precision, and composing from multiple reference images. It also features agentic tool use and integrates with Muse Spark.

Muse Image is available today across the Meta AI app, on meta.ai, Instagram Stories in the US, and WhatsApp in select countries, with Facebook access coming soon. Muse Video is expected to be released soon for creators and within Meta AI.

Unlike traditional prompt-to-image models, Muse Image operates as an agent: it can invoke search and coding tools to improve accuracy, self-refine its outputs, and scale performance with test-time compute. The model learns to write and execute code for accurate plots and QR codes, and to search the web for factual grounding. Self-refinement emerges during reinforcement learning, allowing the model to make local edits or regenerate entirely when needed.

Muse Image also excels at multi-reference image composition, combining elements from several input images. It holds the No. 2 spot on the Arena leaderboard for text-to-image, single-image editing, and multi-image editing based on human preference Elo rankings as of July 5, 2026.

Muse Video, built on the same pretraining base, offers competitive performance in prompt adherence, visual fidelity, and temporal consistency, with native audio support. It ranks No. 3 in text-to-video on Arena. Meta is investing in improving audio-video synchronization and physically accurate fast motion.

To promote transparency, Muse Image includes Content Seal, an invisible watermarking system that persists through cropping, compression, and screenshots. Meta is previewing a detection tool to verify Content Seal watermarks, with plans to extend the system to video.

ModelCategoryArena Rank (Elo)Availability
Muse ImageText-to-Image, Image Editing#2Available now
Muse VideoText-to-Video#3Preview, coming soon
Source
Meta AI · Meta AI Blog
Related news
Meta AI
Meta AI 27 Jul 2026

Meta’s AI Models Power Assistive Robotics Platform for Wheelchair Users

Meta’s open-source AI vision models DINO and SAM are being integrated into a new assistive robotics platform called RAMMP, led by...

36
Research paper
Meta AI 25 Jul 2026

Meta AI Launches Muse Spark: A Multimodal Reasoning Model Toward Personal Superintelligence

Meta AI introduces Muse Spark, the first model from Meta Superintelligence Labs, featuring native multimodal reasoning, tool use,...

40
Meta AI
Meta AI 25 Jul 2026

Meta Unveils Advanced AI Scaling Framework and Safety Report for Muse Spark

Meta details its updated Advanced AI Scaling Framework, a new Safety & Preparedness Report for its Muse Spark model, and advances...

42