Article
DiscoverNet: Unsupervised Video Object Discovery through Discrete Visual Tokenization and Language Model-Inspired Reconstruction
2024-11-05
Abstract excerpt
<title>Abstract</title> <p>Object discovery, the task of separating objects from the background without human annotations, continues to be an unsolved challenge in the field of computer vision. Existing methods face difficulties due to object-background ambiguity and slot drift, as they rely on clustering based on only low-level features. In this work, we introduce DiscoverNet, a comprehensive framework that enha...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 84817b2a-d41f-5ed8-95b9-29b10236e73c
- DOI
- 10.21203/rs.3.rs-5310579/v1
