Back to search

Article

RAE-NeRF: Residual-Based Audio-Video Encoder with Denoising in Talking Head Synchronization

2025-09-26

Abstract excerpt

In recent years, speech-driven facial synthesis has attracted significant attention due to its wide applications in virtual humans, remote conferencing, and digital human generation. However, existing methods still face limitations in terms of realism, synchronization, and robustness, primarily due to noise interference in speech signals and insufficient precision in audio-visual feature fusion. To address these c...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
5cf39b87-f231-5f78-95d2-7fadea7ad6a4
DOI
10.20944/preprints202509.2231.v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
RAE-NeRF: Residual-Based Audio-Video Encoder with Denoising in Talking Head SynchronizationDOI 10.20944/preprints202509.2231.v1
Select a neighboring publication to make it the new centre.