cs.CVSep 30, 2026

MegaAvatar: Controllable Talking Avatar Generation

Authors: Junyao Gao, Sibo Liu, Weidong Zhang, Cairong Zhao, Jun Zhang

Organizations: Tencent · Tongji University

Abstract

This report presents \textbf{MegaAvatar}, a controllable talking avatar generation framework built on top of the Wan2.2-TI2V-5B model. Compared with previous talking-avatar methods that mainly rely on audio or reference-image conditioning, we introduce additional SMPL-X-derived 3D guidance, enabling global control over body pose and head motion. Specifically, we render the driving SMPL-X sequence into dense mesh frames and encode them with a lightweight 3D convolutional encoder, whose outputs are injected into the latent tokens to provide overall motion control. Furthermore, we extend Wan2.2-TI2V-5B with additional audio and face cross-attention modules to enable fine-grained expression control and preserve the input identity, respectively. In addition, we implement an audio-to-SMPL-X model to predict an SMPL-X sequence conditioned on the reference image and input audio, allowing MegaAvatar to support audio-driven inference without user-provided SMPL-X frames. Experiments show that MegaAvatar achieves high-quality talking avatar generation with controllable body and head motion, speech-synchronized facial expressions, and consistent identity preservation. MegaAvatar also supports inference with flexible resolutions and video lengths. Codes, dataset, models will be avaliable in https://github.com/Jeoyal/MegaAvatar

Figures & tables

Explore similar work

CardsList
  1. Avatar V: Scaling Video-Reference Avatar Video Generation

    Jun 11, 2026Benjamin Liang, Ce Chen, Desmond Lin +20Video GenerationHuman Motion Generation

  2. AptAvatar: Fast and Vivid Long-Form Audio-Driven Video Generation for Production-Ready Avatars

    Jul 27, 2026Hengyuan Zhang, Jingna Sun, Meiguang Jin +1Audio-Video GenerationVideo Generation

  3. EmbodiedHead: Real-Time Listening and Speaking Avatar for Conversational Agents

    Apr 19, 2026Yu Zhang, Kaiyuan Shen, Yang LiAvatarsTurn-Taking