Object-Centric Pretraining via Target Encoder Bootstrapping
Fuente:
arXiv
Saved in:
| Main Authors: | Đukić, Nikola, Lebailly, Tim, Tuytelaars, Tinne |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Simple Framework for Open-Vocabulary Zero-Shot Segmentation
by: Stegmüller, Thomas, et al.
Published: (2024)
by: Stegmüller, Thomas, et al.
Published: (2024)
CrIBo: Self-Supervised Learning via Cross-Image Object-Level Bootstrapping
by: Lebailly, Tim, et al.
Published: (2023)
by: Lebailly, Tim, et al.
Published: (2023)
Implicit Gaussian Splatting with Efficient Multi-Level Tri-Plane Representation
by: Wu, Minye, et al.
Published: (2024)
by: Wu, Minye, et al.
Published: (2024)
Analysis of Spatial augmentation in Self-supervised models in the purview of training and test distributions
by: Jha, Abhishek, et al.
Published: (2024)
by: Jha, Abhishek, et al.
Published: (2024)
Spec-Gloss Surfels and Normal-Diffuse Priors for Relightable Glossy Objects
by: Kouros, Georgios, et al.
Published: (2025)
by: Kouros, Georgios, et al.
Published: (2025)
Remembering by Reconstructing: Domain Incremental Learning With Test-Time Training on Video Streams
by: Swinnen, Jonathan, et al.
Published: (2026)
by: Swinnen, Jonathan, et al.
Published: (2026)
Charm: The Missing Piece in ViT fine-tuning for Image Aesthetic Assessment
by: Behrad, Fatemeh, et al.
Published: (2025)
by: Behrad, Fatemeh, et al.
Published: (2025)
PEO: Training-Free Aesthetic Quality Enhancement in Pre-Trained Text-to-Image Diffusion Models with Prompt Embedding Optimization
by: Margaryan, Hovhannes, et al.
Published: (2025)
by: Margaryan, Hovhannes, et al.
Published: (2025)
RGS-DR: Deferred Reflections and Residual Shading in 2D Gaussian Splatting
by: Kouros, Georgios, et al.
Published: (2025)
by: Kouros, Georgios, et al.
Published: (2025)
Visually-Aware Context Modeling for News Image Captioning
by: Qu, Tingyu, et al.
Published: (2023)
by: Qu, Tingyu, et al.
Published: (2023)
Introducing Routing Functions to Vision-Language Parameter-Efficient Fine-Tuning with Low-Rank Bottlenecks
by: Qu, Tingyu, et al.
Published: (2024)
by: Qu, Tingyu, et al.
Published: (2024)
Same accuracy, twice as fast: continuous training surpasses retraining from scratch
by: Verwimp, Eli, et al.
Published: (2025)
by: Verwimp, Eli, et al.
Published: (2025)
Eff-GRot: Efficient and Generalizable Rotation Estimation with Transformers
by: Mathioulakis, Fanis, et al.
Published: (2025)
by: Mathioulakis, Fanis, et al.
Published: (2025)
Contrastive Learning for Multi-Object Tracking with Transformers
by: De Plaen, Pierre-François, et al.
Published: (2023)
by: De Plaen, Pierre-François, et al.
Published: (2023)
DM-Align: Leveraging the Power of Natural Language Instructions to Make Changes to Images
by: Trusca, Maria Mihaela, et al.
Published: (2024)
by: Trusca, Maria Mihaela, et al.
Published: (2024)
Recasting Generic Pretrained Vision Transformers As Object-Centric Scene Encoders For Manipulation Policies
by: Qian, Jianing, et al.
Published: (2024)
by: Qian, Jianing, et al.
Published: (2024)
Unsupervised Parameter Efficient Source-free Post-pretraining
by: Jha, Abhishek, et al.
Published: (2025)
by: Jha, Abhishek, et al.
Published: (2025)
Forgetting of task-specific knowledge in model merging-based continual learning
by: Hess, Timm, et al.
Published: (2025)
by: Hess, Timm, et al.
Published: (2025)
Towards More Accurate Personalized Image Generation: Addressing Overfitting and Evaluation Bias
by: Li, Mingxiao, et al.
Published: (2025)
by: Li, Mingxiao, et al.
Published: (2025)
Animate Your Motion: Turning Still Images into Dynamic Videos
by: Li, Mingxiao, et al.
Published: (2024)
by: Li, Mingxiao, et al.
Published: (2024)
TS-LLaVA: Constructing Visual Tokens through Thumbnail-and-Sampling for Training-Free Video Large Language Models
by: Qu, Tingyu, et al.
Published: (2024)
by: Qu, Tingyu, et al.
Published: (2024)
Is this chart lying to me? Automating the detection of misleading visualizations
by: Tonglet, Jonathan, et al.
Published: (2025)
by: Tonglet, Jonathan, et al.
Published: (2025)
Object-Attribute Binding in Text-to-Image Generation: Evaluation and Control
by: Trusca, Maria Mihaela, et al.
Published: (2024)
by: Trusca, Maria Mihaela, et al.
Published: (2024)
OASIS: Online Sample Selection for Continual Visual Instruction Tuning
by: Lee, Minjae, et al.
Published: (2025)
by: Lee, Minjae, et al.
Published: (2025)
Unveiling the Ambiguity in Neural Inverse Rendering: A Parameter Compensation Analysis
by: Kouros, Georgios, et al.
Published: (2024)
by: Kouros, Georgios, et al.
Published: (2024)
Classifying Novel 3D-Printed Objects without Retraining: Towards Post-Production Automation in Additive Manufacturing
by: Mathioulakis, Fanis, et al.
Published: (2026)
by: Mathioulakis, Fanis, et al.
Published: (2026)
Bootstrap Segmentation Foundation Model under Distribution Shift via Object-Centric Learning
by: Tang, Luyao, et al.
Published: (2024)
by: Tang, Luyao, et al.
Published: (2024)
BG-Triangle: Bézier Gaussian Triangle for 3D Vectorization and Rendering
by: Wu, Minye, et al.
Published: (2025)
by: Wu, Minye, et al.
Published: (2025)
On the Role of Individual Differences in Current Approaches to Computational Image Aesthetics
by: Chen, Li-Wei, et al.
Published: (2025)
by: Chen, Li-Wei, et al.
Published: (2025)
The Common Stability Mechanism behind most Self-Supervised Learning Approaches
by: Jha, Abhishek, et al.
Published: (2024)
by: Jha, Abhishek, et al.
Published: (2024)
Self-Supervised Learning with a Multi-Task Latent Space Objective
by: De Plaen, Pierre-François, et al.
Published: (2026)
by: De Plaen, Pierre-François, et al.
Published: (2026)
ContextFusion and Bootstrap: An Effective Approach to Improve Slot Attention-Based Object-Centric Learning
by: Tian, Pinzhuo, et al.
Published: (2025)
by: Tian, Pinzhuo, et al.
Published: (2025)
DAVE: Diagnostic benchmark for Audio Visual Evaluation
by: Radevski, Gorjan, et al.
Published: (2025)
by: Radevski, Gorjan, et al.
Published: (2025)
Two Complementary Perspectives to Continual Learning: Ask Not Only What to Optimize, But Also How
by: Hess, Timm, et al.
Published: (2023)
by: Hess, Timm, et al.
Published: (2023)
Prediction Error-based Classification for Class-Incremental Learning
by: Zając, Michał, et al.
Published: (2023)
by: Zając, Michał, et al.
Published: (2023)
Adversarial Dependence Minimization
by: De Plaen, Pierre-François, et al.
Published: (2025)
by: De Plaen, Pierre-François, et al.
Published: (2025)
Diversity-Driven View Subset Selection for Indoor Novel View Synthesis
by: Wang, Zehao, et al.
Published: (2024)
by: Wang, Zehao, et al.
Published: (2024)
Knowledge Accumulation in Continually Learned Representations and the Issue of Feature Forgetting
by: Hess, Timm, et al.
Published: (2023)
by: Hess, Timm, et al.
Published: (2023)
Synthetic Captions for Open-Vocabulary Zero-Shot Segmentation
by: Lebailly, Tim, et al.
Published: (2025)
by: Lebailly, Tim, et al.
Published: (2025)
Navigating the Nuances: A Fine-grained Evaluation of Vision-Language Navigation
by: Wang, Zehao, et al.
Published: (2024)
by: Wang, Zehao, et al.
Published: (2024)
Similar Items
-
A Simple Framework for Open-Vocabulary Zero-Shot Segmentation
by: Stegmüller, Thomas, et al.
Published: (2024) -
CrIBo: Self-Supervised Learning via Cross-Image Object-Level Bootstrapping
by: Lebailly, Tim, et al.
Published: (2023) -
Implicit Gaussian Splatting with Efficient Multi-Level Tri-Plane Representation
by: Wu, Minye, et al.
Published: (2024) -
Analysis of Spatial augmentation in Self-supervised models in the purview of training and test distributions
by: Jha, Abhishek, et al.
Published: (2024) -
Spec-Gloss Surfels and Normal-Diffuse Priors for Relightable Glossy Objects
by: Kouros, Georgios, et al.
Published: (2025)