arXiv · 2606.10753
Deploying Speech-Driven 3D Facial Animation in Unreal Engine for Production-Ready Digital Humans
Abstract
Speech-driven 3D facial animation research has shown promising results, but most methods rely on representations that are not compatible with production pipelines. In this work, we present a deployable system that bridges this gap by enabling speech-driven 3D facial animation directly in Unreal Engine (UE) using ARKit-compatible representations. We construct 3DMEAD-ARKit dataset by converting the MEAD corpus into blendshape sequences using MediaPipe, and retrain FaceDiffuser and ProbTalk3D-X to generate stochastic and emotion controllable animations. We further develop a modular UE plugin with a Python backend that supports model selection, and parameter control. We compare the results to two existing commercial tools: Epic Games' MetaHuman speech-driven animator and Nvidia Audio2Face with a perceptual user study. The results highlight the importance of comparisons among academic and commercial pipelines. We recommend watching the supplementary video. We also plan to do live demonstrations of our work at Siggraph 2026 conference.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Alessandro Busacchi, Kazi Injamamul Haque, Zerrin Yumak. 2026-06-09. Deploying Speech-Driven 3D Facial Animation in Unreal Engine for Production-Ready Digital Humans. https://doi.org/10.1145/3799825.3818695
Cite the original work for its findings. Save a collection to share your selection of sources.