Skip to main content
GameDev.net gamedev.net
Research Paper

This is an academic paper or technical research. Key findings may require technical background to fully understand.

Explore Research Radar

PRO Tired of ads? Read GameDev.net ad-free and help keep the community independent with GameDev Pro — $3/month.

arXiv cs.GR
arXiv cs.GR Research
· 1 week, 4 days ago • Hyung Kyu Kim, Byungchan Hwang, Hak Gu Kim

Seeing Speech: Learning Visible Articulatory Dynamics for Speech-Driven 3D Facial Animation

Briefing

arXiv cs.GR details a new framework for speech-driven 3D facial animation that focuses on visible articulation rather than only vertex-level fit. The goal is to make mouth motion track speech in a way that looks more anatomically plausible, especially around the lips where current systems still struggle with one-to-many audio-to-motion mapping.

The method breaks visible speech into three directional articulatory motions: spreading, opening, and protrusion. A Speech--Articulatory Memory (SAM) module uses a key-value memory to retrieve and decode those motions from phonetic context, while a Topology-aware Articulatory Composition (TAC) stage combines them with mesh topology so the final facial motion stays surface-consistent.

For teams building character animation pipelines, the practical angle is better lip sync without relying purely on black-box regression from audio to vertices. That could matter for games, cinematics, virtual avatars, and any real-time or offline facial rig where believable mouth shapes are a visible quality marker.

The authors say experiments on VOCASET and TFHP beat standard reconstruction metrics and also reduce visible articulatory distance and velocity errors for lip motion. A user study reportedly preferred the results for both lip sync and realism, which is the kind of signal animation teams tend to care about when choosing between technically strong and actually convincing facial motion.

“speech-consistent visible articulation remains difficult”

— Authors · Motivation for the new method
Original source
Read on arXiv cs.GR
At a glance
what
A new speech-driven 3D facial animation framework models visible articulation with directional mouth motions.
who
Hyung Kyu Kim, Byungchan Hwang, and Hak Gu Kim; published on arXiv cs.GR.
when
Submitted 24 Sep 2026; arXiv:2609.30517.
impact
Could improve lip sync realism for game characters, avatars, and cinematic facial animation.
Signal Positive

Promising quality gains for facial animation workflows

Discuss

Follow facial-animation updates

See relevant stories in your personalized news feed.

Sign in to follow

Continue on GameDev.net

Useful next steps related to this story.

Game development news without the noise

One useful weekly briefing. No daily flood.

Sending your confirmation email…

Discussion

Loading comments...

Recommended resources

Graphics Programming Resources

See full guide
Real-Time Rendering, Fourth Edition cover
Editor pick Community pick

Real-Time Rendering, Fourth Edition

Amazon · Book

Real-Time Rendering combines fundamental principles with guidance on the latest techniques to provide a complete reference on three-dimensional interactive computer graphics. It will help you increase speed and improve image quality and learn the features and limitations of acceleration algorithms and graphics APIs. This latest fourth edition has been updated to include a chapter on virtual reality and augmented reality and covers new topics such as visual appearance, global illumination, and curves and curved surfaces. It is for anyone serious about computer graphics who wants to learn about algorithms that create synthetic images fast enough that the viewer can interact with a virtual environment.

GameDev.net may earn a commission if you purchase through these links. This helps fund the site at no extra cost to you.

Programming with wgpu in Rust cover
Editor pick

Programming with wgpu in Rust

Amazon · Book

Unlock the full power of modern graphics programming with wgpu and Rust. This comprehensive guide takes you from foundational GPU concepts to advanced real-time rendering and compute techniques—equipping you to build fast, safe, and cross-platform graphics applications. Written for intermediate to advanced Rust developers, this book provides clear explanations, hands-on examples, and detailed insights into how GPUs process and render data. You’ll explore everything from the fundamentals of buffers, shaders, and pipelines to advanced topics like deferred rendering, shadow mapping, and GPU compute workloads.

GameDev.net may earn a commission if you purchase through these links. This helps fund the site at no extra cost to you.

GameDev.net may earn a commission if you purchase through these links. This helps fund the site at no extra cost to you.