One-shot Motion Personalization for Audio-driven Portrait Video Generation
Finetuning audio-driven portrait video diffusion model using a single viewo for personalized motion
Finetuning audio-driven portrait video diffusion model using a single viewo for personalized motion
Robust deepfake detection method and cross-generator dataset by leveraging cross-attention features to capture audio-visual correlations.
An end-to-end method for animating a 3D face mesh with arbitrary shape and triangulation from a given speech audio.
A method that enables direct retargeting between two facial meshes with different shapes and mesh structures.
Retargeting facial expression from a source human performance video to a target stylized 3D character using local patches.
One-Shot Audio-driven 3D talking head generation with enhanced 3D consistency using NeRF and generative knowledge from single image input.
Creating an animatable stylized 3D face mesh with one example pair.
Generating 3D human texture from a single image using sampling and refinement process by utilizing geometry information.
Extracting a sketch from an image in the style of a given reference sketch while preserving the visual content of the image.
A method for generating 3D human texture from a single image based on SMPL model, using sampling and refinement process.
A two-player conversational defense game that uses voice conversation as an input.