Back to search

Article

DiscoverNet: Unsupervised Video Object Discovery through Discrete Visual Tokenization and Language Model-Inspired Reconstruction

2024-11-05

Abstract excerpt

<title>Abstract</title> <p>Object discovery, the task of separating objects from the background without human annotations, continues to be an unsolved challenge in the field of computer vision. Existing methods face difficulties due to object-background ambiguity and slot drift, as they rely on clustering based on only low-level features. In this work, we introduce DiscoverNet, a comprehensive framework that enha...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
84817b2a-d41f-5ed8-95b9-29b10236e73c
DOI
10.21203/rs.3.rs-5310579/v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
DiscoverNet: Unsupervised Video Object Discovery through Discrete Visual Tokenization and Language Model-Inspired ReconstructionDOI 10.21203/rs.3.rs-5310579/v1
Select a neighboring publication to make it the new centre.