Back to search

Article

Identity-Preserving Text-to-Video Generation by Frequency Decomposition

2024-12-02

Abstract excerpt

Identity-preserving text-to-video (IPT2V) generation aims to create high-fidelity videos with consistent human identity. It is an important task in video generation but remains an open problem for generative models. This paper pushes the technical frontier of IPT2V in two directions that have not been resolved in the literature: (1) A tuning-free pipeline without tedious case-by-case finetuning, and (2) A frequenc...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
5073ad26-25c2-50e1-94ca-3d80f26c380c
DOI
10.32388/tziid6
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Identity-Preserving Text-to-Video Generation by Frequency DecompositionDOI 10.32388/tziid6
Select a neighboring publication to make it the new centre.