Akapulu Labs logo Akapulu Labs Research

Expressive Avatars, Emotional Voice, and Smarter ASR

Today's digest covers speech-driven 3D facial animation, relightable avatar reconstruction from a single image, expressive human motion disentanglement, hierarchical reward optimization for emotional TTS, and edit-flow refinement for non-autoregressive ASR — a broad sweep of lifelike digital human and voice research.

Expressive Avatars, Emotional Voice, and Smarter ASR

MindFlow teaser image illustrating harmonized cognitive semantics and acoustic dynamics in facial animation of dyadic conversations. From MindFlow.

Talking Avatars & Facial Animation

Digital Humans & Avatar Reconstruction

TTS & Voice Synthesis

ASR Architectures