Human Gaze Boosts Object-Centered Representation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Schaumlöffel, Timothy, Aubret, Arthur, Roig, Gemma, Triesch, Jochen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Object Semantic Similarity with Self-Supervision
by: Aubret, Arthur, et al.
Published: (2024)
by: Aubret, Arthur, et al.
Published: (2024)
Caregiver Talk Shapes Toddler Vision: A Computational Study of Dyadic Play
by: Schaumlöffel, Timothy, et al.
Published: (2023)
by: Schaumlöffel, Timothy, et al.
Published: (2023)
Temporal Slowness in Central Vision Drives Semantic Object Learning
by: Schaumlöffel, Timothy, et al.
Published: (2026)
by: Schaumlöffel, Timothy, et al.
Published: (2026)
Seeing the Whole in the Parts in Self-Supervised Representation Learning
by: Aubret, Arthur, et al.
Published: (2025)
by: Aubret, Arthur, et al.
Published: (2025)
Self-supervised visual learning from interactions with objects
by: Aubret, Arthur, et al.
Published: (2024)
by: Aubret, Arthur, et al.
Published: (2024)
Simulated Cortical Magnification Supports Self-Supervised Object Learning
by: Yu, Zhengyang, et al.
Published: (2025)
by: Yu, Zhengyang, et al.
Published: (2025)
Evaluation of Randomization through Style Transfer for Enhanced Domain Generalization
by: Eisenhardt, Dustin, et al.
Published: (2026)
by: Eisenhardt, Dustin, et al.
Published: (2026)
Toddlers' Active Gaze Behavior Supports Self-Supervised Object Learning
by: Yu, Zhengyang, et al.
Published: (2024)
by: Yu, Zhengyang, et al.
Published: (2024)
Mechanisms of Object Localization in Vision-Language Models
by: Schaumlöffel, Timothy, et al.
Published: (2026)
by: Schaumlöffel, Timothy, et al.
Published: (2026)
Self-Supervised Learning of Color Constancy
by: Ernst, Markus R., et al.
Published: (2024)
by: Ernst, Markus R., et al.
Published: (2024)
Contextual inference from single objects in Vision-Language models
by: Vilas, Martina G., et al.
Published: (2026)
by: Vilas, Martina G., et al.
Published: (2026)
Boosting Object Representation Learning via Motion and Object Continuity
by: Delfosse, Quentin, et al.
Published: (2022)
by: Delfosse, Quentin, et al.
Published: (2022)
FovEx: Human-Inspired Explanations for Vision Transformers and Convolutional Neural Networks
by: Panda, Mahadev Prasad, et al.
Published: (2024)
by: Panda, Mahadev Prasad, et al.
Published: (2024)
Efficient Unsupervised Shortcut Learning Detection and Mitigation in Transformers
by: Kuhn, Lukas, et al.
Published: (2025)
by: Kuhn, Lukas, et al.
Published: (2025)
Grouped Discrete Representation for Object-Centric Learning
by: Zhao, Rongzhen, et al.
Published: (2024)
by: Zhao, Rongzhen, et al.
Published: (2024)
Zero-Shot Object-Centric Representation Learning
by: Didolkar, Aniket, et al.
Published: (2024)
by: Didolkar, Aniket, et al.
Published: (2024)
Boosting Open Set Recognition Performance through Modulated Representation Learning
by: Kundu, Amit Kumar, et al.
Published: (2025)
by: Kundu, Amit Kumar, et al.
Published: (2025)
Classification of freshwater snails of the genus Radomaniola with multimodal triplet networks
by: Vetter, Dennis, et al.
Published: (2024)
by: Vetter, Dennis, et al.
Published: (2024)
FORLA: Federated Object-centric Representation Learning with Slot Attention
by: Liao, Guiqiu, et al.
Published: (2025)
by: Liao, Guiqiu, et al.
Published: (2025)
Suppressing Uncertainty in Gaze Estimation
by: Wang, Shijing, et al.
Published: (2024)
by: Wang, Shijing, et al.
Published: (2024)
From Vicious to Virtuous Cycles: Synergistic Representation Learning for Unsupervised Video Object-Centric Learning
by: Seong, Hyun Seok, et al.
Published: (2026)
by: Seong, Hyun Seok, et al.
Published: (2026)
Multistream Gaze Estimation with Anatomical Eye Region Isolation by Synthetic to Real Transfer Learning
by: Mahmud, Zunayed, et al.
Published: (2022)
by: Mahmud, Zunayed, et al.
Published: (2022)
CSGaze: Context-aware Social Gaze Prediction
by: Madan, Surbhi, et al.
Published: (2025)
by: Madan, Surbhi, et al.
Published: (2025)
BioHuman: Learning Biomechanical Human Representations from Video
by: Huo, Yujun, et al.
Published: (2026)
by: Huo, Yujun, et al.
Published: (2026)
Identifiable Object Representations under Spatial Ambiguities
by: Kori, Avinash, et al.
Published: (2025)
by: Kori, Avinash, et al.
Published: (2025)
Are Object-Centric Representations Better At Compositional Generalization?
by: Kapl, Ferdinand, et al.
Published: (2026)
by: Kapl, Ferdinand, et al.
Published: (2026)
Object-Centric Relational Representations for Image Generation
by: Butera, Luca, et al.
Published: (2023)
by: Butera, Luca, et al.
Published: (2023)
Toward Real-World IoT Security: Concept Drift-Resilient IoT Botnet Detection via Latent Space Representation Learning and Alignment
by: Wasswa, Hassan, et al.
Published: (2025)
by: Wasswa, Hassan, et al.
Published: (2025)
CauSkelNet: Causal Representation Learning for Human Behaviour Analysis
by: Gu, Xingrui, et al.
Published: (2024)
by: Gu, Xingrui, et al.
Published: (2024)
On Explaining Knowledge Distillation: Measuring and Visualising the Knowledge Transfer Process
by: Adhane, Gereziher, et al.
Published: (2024)
by: Adhane, Gereziher, et al.
Published: (2024)
Enhancing Pre-trained Representation Classifiability can Boost its Interpretability
by: Shen, Shufan, et al.
Published: (2025)
by: Shen, Shufan, et al.
Published: (2025)
OCRT: Boosting Foundation Models in the Open World with Object-Concept-Relation Triad
by: Tang, Luyao, et al.
Published: (2025)
by: Tang, Luyao, et al.
Published: (2025)
Gaze4HRI: Zero-shot Benchmarking Gaze Estimation Neural-Networks for Human-Robot Interaction
by: Sezer, Berk, et al.
Published: (2026)
by: Sezer, Berk, et al.
Published: (2026)
Eyes on the Image: Gaze Supervised Multimodal Learning for Chest X-ray Diagnosis and Report Generation
by: Riju, Tanjim Islam, et al.
Published: (2025)
by: Riju, Tanjim Islam, et al.
Published: (2025)
A Transformer-Based Model for the Prediction of Human Gaze Behavior on Videos
by: Ozdel, Suleyman, et al.
Published: (2024)
by: Ozdel, Suleyman, et al.
Published: (2024)
EM-Net: Gaze Estimation with Expectation Maximization Algorithm
by: Cheng, Zhang, et al.
Published: (2024)
by: Cheng, Zhang, et al.
Published: (2024)
Measuring the Impact of Scene Level Objects on Object Detection: Towards Quantitative Explanations of Detection Decisions
by: Haar, Lynn Vonder, et al.
Published: (2024)
by: Haar, Lynn Vonder, et al.
Published: (2024)
Multiple Object Stitching for Unsupervised Representation Learning
by: Shen, Chengchao, et al.
Published: (2025)
by: Shen, Chengchao, et al.
Published: (2025)
Explicitly Disentangled Representations in Object-Centric Learning
by: Majellaro, Riccardo, et al.
Published: (2024)
by: Majellaro, Riccardo, et al.
Published: (2024)
Cross-Dataset Gaze Estimation by Evidential Inter-intra Fusion
by: Wang, Shijing, et al.
Published: (2024)
by: Wang, Shijing, et al.
Published: (2024)
Similar Items
-
Learning Object Semantic Similarity with Self-Supervision
by: Aubret, Arthur, et al.
Published: (2024) -
Caregiver Talk Shapes Toddler Vision: A Computational Study of Dyadic Play
by: Schaumlöffel, Timothy, et al.
Published: (2023) -
Temporal Slowness in Central Vision Drives Semantic Object Learning
by: Schaumlöffel, Timothy, et al.
Published: (2026) -
Seeing the Whole in the Parts in Self-Supervised Representation Learning
by: Aubret, Arthur, et al.
Published: (2025) -
Self-supervised visual learning from interactions with objects
by: Aubret, Arthur, et al.
Published: (2024)