Article
Identity-Preserving Text-to-Video Generation by Frequency Decomposition
2024-12-02
Abstract excerpt
Identity-preserving text-to-video (IPT2V) generation aims to create high-fidelity videos with consistent human identity. It is an important task in video generation but remains an open problem for generative models. This paper pushes the technical frontier of IPT2V in two directions that have not been resolved in the literature: (1) A tuning-free pipeline without tedious case-by-case finetuning, and (2) A frequenc...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 5073ad26-25c2-50e1-94ca-3d80f26c380c
- DOI
- 10.32388/tziid6
