conference-paper
Object-Centric Representation Learning with Attention Mechanism
Research footprint
At a glance
- Citations
- 0
- References
- 23
- Comments
- 0
Paper overview
Abstract
For object-centric representation learning, several slot-based methods, that separate objects using masks and learn the objects separately, are proposed. While these methods are proved to be useful on various downstream tasks, it is known that they require a significant amount of computation for training. We propose the introduction of attention mechanisms into slot-based method to simplify and speed up the computation. We pick ViMON as the base structure and propose two methods, named AttnViMON and SFA. We evaluate them in terms of reconstruction error and computation time, and a downstream task. The proposed methods demonstrate that they achieve significant speed-up while showing even better performance.
Record transparency
Publication details
- DOI
- 10.1109/imcom60618.2024.10418364
- OpenAlex
- W4391742829
- Document type
- conference-paper
- Language
- EN
- Last metadata update
Comments
Log in to join the discussion.