Aller au contenu principal
Accès ouvert déclaré 2025 article

EGENEFACE: HIGH QUALITY GENERALIZED TALKING FACE SYNTHESIZING ENABLED BY AUDIO

0Citations signalées, ce qui n’est pas une note de qualité
0Institutions déclarées
0Pays d’affiliation déclarés

Le résumé fourni par la source

A critical issue in virtual reality and filmmaking is the generation of photorealistic video portraits from random spoken recordings.The use of neural radiance fields is to enhance 3D realism and image authenticity has been the subject of several recent studies.Nevertheless, the limited training sample scale of earlier NeRF-based techniques restricts their applicability to out-ofdomain sounds.The previous model builds upon GeneFace, which introduced variational auto encoder-based framework with Neural Radiance Field (NeRF) rendering for lip-syncing tasks.While GeneFace demonstrated strong results in terms of synchronization accuracy and visual quality, several limitations hindered its performance in challenging scenarios such as noisy audio, out-of-domain speech, and temporal in stability.Our proposed model, EGeneface, addresses these gaps though.In terms of Generator of variational motion, we reduced Landmark Distance by 5.9% and improved lip synchronization accuracy.For advanced renderer for NeRF, we lowered Frechet Inception Distance to 20.15, improving the visible realism of created videos.In training with noise to increase robustness, we increased Sync Confidence in OOD scenarios by 12%, achieving more reliable lip-syncing across diverse audio environments.For Real-Time deployment optimization, we improved frame processing rate from 30 FPS to 35 FPS, making suitable for real-time applications.

Ce résumé expose les affirmations des auteurs. BNTIC ne l’interprète pas comme une validation indépendante des résultats.

Le contrôle bibliographique ouvert

DOI retrouvé dans Crossref DOI retrouvé ; titre concordant.

Titre Crossref
EGENEFACE: HIGH QUALITY GENERALIZED TALKING FACE SYNTHESIZING ENABLED BY AUDIO
Date Crossref
28/09/2025
Éditeur
Faculty of Engineering, University of Kragujevac
Type
journal-article

Ce recoupement confirme des métadonnées liées au DOI. Il ne confirme ni la méthode ni les conclusions de l’étude, et il ne compte pas comme une seconde source scientifique indépendante.

Les sujets associés

Face recognition and analysisSpeech and Audio ProcessingFace and Expression Recognition

BNTIC News n’est pas le producteur de ces données. Les publications sont interrogées à la demande dans Crossref, OpenAIRE, DOAJ, Europe PMC, HAL, DataCite, AfricArXiv, ROR et la Banque mondiale, sans clé d’accès. OpenAlex reste optionnel. Aucun service payant n’est nécessaire et aucune donnée externe n’est enregistrée en base. Consulter les sources et leurs limites.