Muse Realtime Avatar: Real-Time Embodiment for Expressive AI Characters

Loading video
Loading videoMuse Realtime Avatar is Meta AI Research's embodiment technology that turns Muse Realtime Voice into expressive, interactive avatars. Conditioned on reference media, it brings any character into a live conversation: a photographic portrait responds through subtle expressions, a full-body illustration gestures and shifts posture as it speaks, and animals or everyday objects become expressive without losing what makes them distinctive. Voice and avatar form a single streaming system, with the voice model producing speech tokens that the avatar consumes, and identity holds from one conversational turn to the next. Real-time conversation is the first interaction it enables in Muse.
Interactive AvatarDigital HumanReal-time generationVoice InteractionEmbodimentWorld ModelsMetaHuman-robot interactionEmbodied AI
Category: research
Author: @AIatMeta
Date: 2026-09-24T00:00:00
Duration: 121.588s





