Motion-Refined DINOSAUR for Unsupervised Multi-Object Discovery
Fuente:
arXiv
Saved in:
| Main Authors: | Gong, Xinrui, Hahn, Oliver, Reich, Christoph, Singh, Krishnakant, Schaub-Meyer, Simone, Cremers, Daniel, Roth, Stefan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating Object-Centric Models beyond Object Discovery
by: Singh, Krishnakant, et al.
Published: (2026)
by: Singh, Krishnakant, et al.
Published: (2026)
GLASS: Guided Latent Slot Diffusion for Object-Centric Learning
by: Singh, Krishnakant, et al.
Published: (2024)
by: Singh, Krishnakant, et al.
Published: (2024)
Boosting Unsupervised Semantic Segmentation with Principal Mask Proposals
by: Hahn, Oliver, et al.
Published: (2024)
by: Hahn, Oliver, et al.
Published: (2024)
MUFASA: A Multi-Layer Framework for Slot Attention
by: Bock, Sebastian, et al.
Published: (2026)
by: Bock, Sebastian, et al.
Published: (2026)
Scene-Centric Unsupervised Panoptic Segmentation
by: Hahn, Oliver, et al.
Published: (2025)
by: Hahn, Oliver, et al.
Published: (2025)
Is Synthetic Data all We Need? Benchmarking the Robustness of Models Trained with Synthetic Images
by: Singh, Krishnakant, et al.
Published: (2024)
by: Singh, Krishnakant, et al.
Published: (2024)
Feed-Forward SceneDINO for Unsupervised Semantic Scene Completion
by: Jevtić, Aleksandar, et al.
Published: (2025)
by: Jevtić, Aleksandar, et al.
Published: (2025)
A Perspective on Deep Vision Performance with Standard Image and Video Codecs
by: Reich, Christoph, et al.
Published: (2024)
by: Reich, Christoph, et al.
Published: (2024)
Benchmarking the Attribution Quality of Vision Models
by: Hesse, Robin, et al.
Published: (2024)
by: Hesse, Robin, et al.
Published: (2024)
INSID3: Training-Free In-Context Segmentation with DINOv3
by: Cuttano, Claudia, et al.
Published: (2026)
by: Cuttano, Claudia, et al.
Published: (2026)
Removing Cost Volumes from Optical Flow Estimators
by: Kiefhaber, Simon, et al.
Published: (2025)
by: Kiefhaber, Simon, et al.
Published: (2025)
Semantic Self-adaptation: Enhancing Generalization with a Single Sample
by: Bahmani, Sherwin, et al.
Published: (2022)
by: Bahmani, Sherwin, et al.
Published: (2022)
Efficient Masked Attention Transformer for Few-Shot Classification and Segmentation
by: Carrión-Ojeda, Dustin, et al.
Published: (2025)
by: Carrión-Ojeda, Dustin, et al.
Published: (2025)
Boosting Omnidirectional Stereo Matching with a Pre-trained Depth Foundation Model
by: Endres, Jannik, et al.
Published: (2025)
by: Endres, Jannik, et al.
Published: (2025)
Disentangling Polysemantic Channels in Convolutional Neural Networks
by: Hesse, Robin, et al.
Published: (2025)
by: Hesse, Robin, et al.
Published: (2025)
What is Missing? Explaining Neurons Activated by Absent Concepts
by: Hesse, Robin, et al.
Published: (2026)
by: Hesse, Robin, et al.
Published: (2026)
Beyond Accuracy: What Matters in Designing Well-Behaved Image Classification Models?
by: Hesse, Robin, et al.
Published: (2025)
by: Hesse, Robin, et al.
Published: (2025)
DIAGen: Semantically Diverse Image Augmentation with Generative Models for Few-Shot Learning
by: Lingenberg, Tobias, et al.
Published: (2024)
by: Lingenberg, Tobias, et al.
Published: (2024)
ART: Adaptive Relation Tuning for Generalized Relation Prediction
by: Sudhakaran, Gopika, et al.
Published: (2025)
by: Sudhakaran, Gopika, et al.
Published: (2025)
EgoFlow: Gradient-Guided Flow Matching for Egocentric 6DoF Object Motion Generation
by: Saroha, Abhishek, et al.
Published: (2026)
by: Saroha, Abhishek, et al.
Published: (2026)
LeAD-M3D: Leveraging Asymmetric Distillation for Real-Time Monocular 3D Detection
by: Meier, Johannes, et al.
Published: (2025)
by: Meier, Johannes, et al.
Published: (2025)
CoRe-GS: Coarse-to-Refined Gaussian Splatting with Semantic Object Focus
by: Schieber, Hannah, et al.
Published: (2025)
by: Schieber, Hannah, et al.
Published: (2025)
FlowFeat: Pixel-Dense Embedding of Motion Profiles
by: Araslanov, Nikita, et al.
Published: (2025)
by: Araslanov, Nikita, et al.
Published: (2025)
Benchmarking Single-Step Inpainting Methods for Multi-Object 3D Gaussian Splatting Scenes
by: Dröge, Finn, et al.
Published: (2026)
by: Dröge, Finn, et al.
Published: (2026)
Future Slot Prediction for Unsupervised Object Discovery in Surgical Video
by: Liao, Guiqiu, et al.
Published: (2025)
by: Liao, Guiqiu, et al.
Published: (2025)
Enhancing Object Discovery for Unsupervised Instance Segmentation and Object Detection
by: Feng, Xingyu, et al.
Published: (2025)
by: Feng, Xingyu, et al.
Published: (2025)
Ensemble Foreground Management for Unsupervised Object Discovery
by: Wu, Ziling, et al.
Published: (2025)
by: Wu, Ziling, et al.
Published: (2025)
Unsupervised Discovery of Object-Centric Neural Fields
by: Luo, Rundong, et al.
Published: (2024)
by: Luo, Rundong, et al.
Published: (2024)
HEAP: Unsupervised Object Discovery and Localization with Contrastive Grouping
by: Zhang, Xin, et al.
Published: (2023)
by: Zhang, Xin, et al.
Published: (2023)
The TYC Dataset for Understanding Instance-Level Semantics and Motions of Cells in Microstructures
by: Reich, Christoph, et al.
Published: (2023)
by: Reich, Christoph, et al.
Published: (2023)
Benchmarking Video Frame Interpolation
by: Kiefhaber, Simon, et al.
Published: (2024)
by: Kiefhaber, Simon, et al.
Published: (2024)
Appearance-Based Refinement for Object-Centric Motion Segmentation
by: Xie, Junyu, et al.
Published: (2023)
by: Xie, Junyu, et al.
Published: (2023)
Shape Your Ground: Refining Road Surfaces Beyond Planar Representations
by: Dhaouadi, Oussema, et al.
Published: (2025)
by: Dhaouadi, Oussema, et al.
Published: (2025)
Unsupervised Object Discovery: A Comprehensive Survey and Unified Taxonomy
by: Villa-Vásquez, José-Fabian, et al.
Published: (2024)
by: Villa-Vásquez, José-Fabian, et al.
Published: (2024)
Implicit Motion-Compensated Network for Unsupervised Video Object Segmentation
by: Xi, Lin, et al.
Published: (2022)
by: Xi, Lin, et al.
Published: (2022)
Learning Motion and Temporal Cues for Unsupervised Video Object Segmentation
by: Zhuge, Yunzhi, et al.
Published: (2025)
by: Zhuge, Yunzhi, et al.
Published: (2025)
SparseAlign: A Fully Sparse Framework for Cooperative Object Detection
by: Yuan, Yunshuang, et al.
Published: (2025)
by: Yuan, Yunshuang, et al.
Published: (2025)
Let Your Image Move with Your Motion! -- Implicit Multi-Object Multi-Motion Transfer
by: Li, Yuze, et al.
Published: (2026)
by: Li, Yuze, et al.
Published: (2026)
Masked Multi-Query Slot Attention for Unsupervised Object Discovery
by: Pramanik, Rishav, et al.
Published: (2024)
by: Pramanik, Rishav, et al.
Published: (2024)
Treating Motion as Option with Output Selection for Unsupervised Video Object Segmentation
by: Cho, Suhwan, et al.
Published: (2023)
by: Cho, Suhwan, et al.
Published: (2023)
Similar Items
-
Evaluating Object-Centric Models beyond Object Discovery
by: Singh, Krishnakant, et al.
Published: (2026) -
GLASS: Guided Latent Slot Diffusion for Object-Centric Learning
by: Singh, Krishnakant, et al.
Published: (2024) -
Boosting Unsupervised Semantic Segmentation with Principal Mask Proposals
by: Hahn, Oliver, et al.
Published: (2024) -
MUFASA: A Multi-Layer Framework for Slot Attention
by: Bock, Sebastian, et al.
Published: (2026) -
Scene-Centric Unsupervised Panoptic Segmentation
by: Hahn, Oliver, et al.
Published: (2025)