Temporal Slowness in Central Vision Drives Semantic Object Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Schaumlöffel, Timothy, Aubret, Arthur, Roig, Gemma, Triesch, Jochen |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Object Semantic Similarity with Self-Supervision
by: Aubret, Arthur, et al.
Published: (2024)
by: Aubret, Arthur, et al.
Published: (2024)
Human Gaze Boosts Object-Centered Representation Learning
by: Schaumlöffel, Timothy, et al.
Published: (2025)
by: Schaumlöffel, Timothy, et al.
Published: (2025)
Caregiver Talk Shapes Toddler Vision: A Computational Study of Dyadic Play
by: Schaumlöffel, Timothy, et al.
Published: (2023)
by: Schaumlöffel, Timothy, et al.
Published: (2023)
Mechanisms of Object Localization in Vision-Language Models
by: Schaumlöffel, Timothy, et al.
Published: (2026)
by: Schaumlöffel, Timothy, et al.
Published: (2026)
Simulated Cortical Magnification Supports Self-Supervised Object Learning
by: Yu, Zhengyang, et al.
Published: (2025)
by: Yu, Zhengyang, et al.
Published: (2025)
Seeing the Whole in the Parts in Self-Supervised Representation Learning
by: Aubret, Arthur, et al.
Published: (2025)
by: Aubret, Arthur, et al.
Published: (2025)
Self-supervised visual learning from interactions with objects
by: Aubret, Arthur, et al.
Published: (2024)
by: Aubret, Arthur, et al.
Published: (2024)
Contextual inference from single objects in Vision-Language models
by: Vilas, Martina G., et al.
Published: (2026)
by: Vilas, Martina G., et al.
Published: (2026)
Toddlers' Active Gaze Behavior Supports Self-Supervised Object Learning
by: Yu, Zhengyang, et al.
Published: (2024)
by: Yu, Zhengyang, et al.
Published: (2024)
Evaluation of Randomization through Style Transfer for Enhanced Domain Generalization
by: Eisenhardt, Dustin, et al.
Published: (2026)
by: Eisenhardt, Dustin, et al.
Published: (2026)
Self-Supervised Learning of Color Constancy
by: Ernst, Markus R., et al.
Published: (2024)
by: Ernst, Markus R., et al.
Published: (2024)
Semantics Meets Temporal Correspondence: Self-supervised Object-centric Learning in Videos
by: Qian, Rui, et al.
Published: (2023)
by: Qian, Rui, et al.
Published: (2023)
Spatio-Temporal Attention for Consistent Video Semantic Segmentation in Automated Driving
by: Varghese, Serin, et al.
Published: (2026)
by: Varghese, Serin, et al.
Published: (2026)
Re-assessing the evidence for mental rotation abilities in children using computational models
by: Aubret, Arthur, et al.
Published: (2025)
by: Aubret, Arthur, et al.
Published: (2025)
AdaDrive: Self-Adaptive Slow-Fast System for Language-Grounded Autonomous Driving
by: Zhang, Ruifei, et al.
Published: (2025)
by: Zhang, Ruifei, et al.
Published: (2025)
DoppDrive: Doppler-Driven Temporal Aggregation for Improved Radar Object Detection
by: Haitman, Yuval, et al.
Published: (2025)
by: Haitman, Yuval, et al.
Published: (2025)
Temporal-Spatial Object Relations Modeling for Vision-and-Language Navigation
by: Huang, Bowen, et al.
Published: (2024)
by: Huang, Bowen, et al.
Published: (2024)
Evaluating Vision Foundation Models for Pixel and Object Classification in Microscopy
by: Teuber, Carolin, et al.
Published: (2026)
by: Teuber, Carolin, et al.
Published: (2026)
SOCO: Benchmarking Semantic Object Correspondence in Vision Foundation Models
by: Dünkel, Olaf, et al.
Published: (2026)
by: Dünkel, Olaf, et al.
Published: (2026)
Mask-RadarNet: Enhancing Transformer With Spatial-Temporal Semantic Context for Radar Object Detection in Autonomous Driving
by: Wu, Yuzhi, et al.
Published: (2024)
by: Wu, Yuzhi, et al.
Published: (2024)
A Survey of Deep Learning Based Radar and Vision Fusion for 3D Object Detection in Autonomous Driving
by: Wu, Di, et al.
Published: (2024)
by: Wu, Di, et al.
Published: (2024)
Training-Free Semantic Multi-Object Tracking with Vision-Language Models
by: Bonat, Laurence, et al.
Published: (2026)
by: Bonat, Laurence, et al.
Published: (2026)
FovEx: Human-Inspired Explanations for Vision Transformers and Convolutional Neural Networks
by: Panda, Mahadev Prasad, et al.
Published: (2024)
by: Panda, Mahadev Prasad, et al.
Published: (2024)
Memory-Augmented Vision-Language Agents for Persistent and Semantically Consistent Object Captioning
by: Galliena, Tommaso, et al.
Published: (2026)
by: Galliena, Tommaso, et al.
Published: (2026)
SlowFocus: Enhancing Fine-grained Temporal Understanding in Video LLM
by: Nie, Ming, et al.
Published: (2026)
by: Nie, Ming, et al.
Published: (2026)
Learning with Unmasked Tokens Drives Stronger Vision Learners
by: Kim, Taekyung, et al.
Published: (2023)
by: Kim, Taekyung, et al.
Published: (2023)
Semantic-Supervised Spatial-Temporal Fusion for LiDAR-based 3D Object Detection
by: Wang, Chaoqun, et al.
Published: (2025)
by: Wang, Chaoqun, et al.
Published: (2025)
Compositional Video Synthesis by Temporal Object-Centric Learning
by: Akan, Adil Kaan, et al.
Published: (2025)
by: Akan, Adil Kaan, et al.
Published: (2025)
Efficient Unsupervised Shortcut Learning Detection and Mitigation in Transformers
by: Kuhn, Lukas, et al.
Published: (2025)
by: Kuhn, Lukas, et al.
Published: (2025)
Independently Keypoint Learning for Small Object Semantic Correspondence
by: Jin, Hailong, et al.
Published: (2024)
by: Jin, Hailong, et al.
Published: (2024)
Temporal Object-Aware Vision Transformer for Few-Shot Video Object Detection
by: Kumar, Yogesh, et al.
Published: (2025)
by: Kumar, Yogesh, et al.
Published: (2025)
SlowFast-SCI: Slow-Fast Deep Unfolding Learning for Spectral Compressive Imaging
by: Zeng, Haijin, et al.
Published: (2025)
by: Zeng, Haijin, et al.
Published: (2025)
LaST-VLA: Thinking in Latent Spatio-Temporal Space for Vision-Language-Action in Autonomous Driving
by: Luo, Yuechen, et al.
Published: (2026)
by: Luo, Yuechen, et al.
Published: (2026)
Fast-Slow Test-Time Adaptation for Online Vision-and-Language Navigation
by: Gao, Junyu, et al.
Published: (2023)
by: Gao, Junyu, et al.
Published: (2023)
DriveFix: Spatio-Temporally Coherent Driving Scene Restoration
by: Si, Heyu, et al.
Published: (2026)
by: Si, Heyu, et al.
Published: (2026)
Temporal Contrastive Learning for Video Temporal Reasoning in Large Vision-Language Models
by: Souza, Rafael, et al.
Published: (2024)
by: Souza, Rafael, et al.
Published: (2024)
Harnessing Vision-Language Pretrained Models with Temporal-Aware Adaptation for Referring Video Object Segmentation
by: Zhou, Zikun, et al.
Published: (2024)
by: Zhou, Zikun, et al.
Published: (2024)
Divide and Merge: Motion and Semantic Learning in End-to-End Autonomous Driving
by: Shen, Yinzhe, et al.
Published: (2025)
by: Shen, Yinzhe, et al.
Published: (2025)
Learning Motion and Temporal Cues for Unsupervised Video Object Segmentation
by: Zhuge, Yunzhi, et al.
Published: (2025)
by: Zhuge, Yunzhi, et al.
Published: (2025)
PhysMamba: Efficient Remote Physiological Measurement with SlowFast Temporal Difference Mamba
by: Luo, Chaoqi, et al.
Published: (2024)
by: Luo, Chaoqi, et al.
Published: (2024)
Similar Items
-
Learning Object Semantic Similarity with Self-Supervision
by: Aubret, Arthur, et al.
Published: (2024) -
Human Gaze Boosts Object-Centered Representation Learning
by: Schaumlöffel, Timothy, et al.
Published: (2025) -
Caregiver Talk Shapes Toddler Vision: A Computational Study of Dyadic Play
by: Schaumlöffel, Timothy, et al.
Published: (2023) -
Mechanisms of Object Localization in Vision-Language Models
by: Schaumlöffel, Timothy, et al.
Published: (2026) -
Simulated Cortical Magnification Supports Self-Supervised Object Learning
by: Yu, Zhengyang, et al.
Published: (2025)