Zero-shot talking avatar generation aims at synthesizing natural talking videos from speech and a single portrait image. Previous methods have relied on domain-specific heuristics such as warping-based motion representation and 3D Morphable Models, which limit the naturalness and diversity of the...
![[ICLR 2024] GAIA: Zero-shot Talking Avatar Generation | Yuchi Wang (王宇驰)](https://wangyuchi369.github.io/publication/2023-gaia/featured.png)