Saved in:
| Main Authors: | Ebouky, Brown, Chhatkuli, Ajad, Malossi, Cristiano, Studer, Christoph, Assaf, Roy, Bartezzaghi, Andrea |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2509.17816 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GazeVLM: Active Vision via Internal Attention Control for Multimodal Reasoning
by: Ebouky, Brown, et al.
Published: (2026)
by: Ebouky, Brown, et al.
Published: (2026)
VP Lab: a PEFT-Enabled Visual Prompting Laboratory for Semantic Segmentation
by: Avogaro, Niccolo, et al.
Published: (2025)
by: Avogaro, Niccolo, et al.
Published: (2025)
Show or Tell? Effectively prompting Vision-Language Models for semantic segmentation
by: Avogaro, Niccolo, et al.
Published: (2025)
by: Avogaro, Niccolo, et al.
Published: (2025)
Continuous Pose for Monocular Cameras in Neural Implicit Representation
by: Ma, Qi, et al.
Published: (2023)
by: Ma, Qi, et al.
Published: (2023)
Self-supervised Shape Completion via Involution and Implicit Correspondences
by: Liu, Mengya, et al.
Published: (2024)
by: Liu, Mengya, et al.
Published: (2024)
Inferring Compositional 4D Scenes without Ever Seeing One
by: Gokmen, Ahmet Berke, et al.
Published: (2025)
by: Gokmen, Ahmet Berke, et al.
Published: (2025)
iHuman: Instant Animatable Digital Humans From Monocular Videos
by: Paudel, Pramish, et al.
Published: (2024)
by: Paudel, Pramish, et al.
Published: (2024)
One2Any: One-Reference 6D Pose Estimation for Any Object
by: Liu, Mengya, et al.
Published: (2025)
by: Liu, Mengya, et al.
Published: (2025)
VF-NeRF: Learning Neural Vector Fields for Indoor Scene Reconstruction
by: Puigjaner, Albert Gassol, et al.
Published: (2024)
by: Puigjaner, Albert Gassol, et al.
Published: (2024)
Vision Transformers with Hierarchical Attention
by: Liu, Yun, et al.
Published: (2021)
by: Liu, Yun, et al.
Published: (2021)
Revisiting Continual Semantic Segmentation with Pre-trained Vision Models
by: Zhang, Duzhen, et al.
Published: (2025)
by: Zhang, Duzhen, et al.
Published: (2025)
Q-SAM2: Accurate Quantization for Segment Anything Model 2
by: Farronato, Nicola, et al.
Published: (2025)
by: Farronato, Nicola, et al.
Published: (2025)
GASP: Unifying Geometric and Semantic Self-Supervised Pre-training for Autonomous Driving
by: Ljungbergh, William, et al.
Published: (2025)
by: Ljungbergh, William, et al.
Published: (2025)
Eliciting Reasoning in Language Models with Cognitive Tools
by: Ebouky, Brown, et al.
Published: (2025)
by: Ebouky, Brown, et al.
Published: (2025)
On the Viability of Monocular Depth Pre-training for Semantic Segmentation
by: Lao, Dong, et al.
Published: (2022)
by: Lao, Dong, et al.
Published: (2022)
Enhancing SAR Object Detection with Self-Supervised Pre-training on Masked Auto-Encoders
by: Pu, Xinyang, et al.
Published: (2025)
by: Pu, Xinyang, et al.
Published: (2025)
Enhancing Vision-Language Pre-training with Rich Supervisions
by: Gao, Yuan, et al.
Published: (2024)
by: Gao, Yuan, et al.
Published: (2024)
Weakly Supervised Co-training with Swapping Assignments for Semantic Segmentation
by: Yang, Xinyu, et al.
Published: (2024)
by: Yang, Xinyu, et al.
Published: (2024)
Outline-Guided Object Inpainting with Diffusion Models
by: Pobitzer, Markus, et al.
Published: (2024)
by: Pobitzer, Markus, et al.
Published: (2024)
UNIP: Rethinking Pre-trained Attention Patterns for Infrared Semantic Segmentation
by: Zhang, Tao, et al.
Published: (2025)
by: Zhang, Tao, et al.
Published: (2025)
MaskDiffusion: Exploiting Pre-trained Diffusion Models for Semantic Segmentation
by: Kawano, Yasufumi, et al.
Published: (2024)
by: Kawano, Yasufumi, et al.
Published: (2024)
Effect of Rotation Angle in Self-Supervised Pre-training is Dataset-Dependent
by: Saranchuk, Amy, et al.
Published: (2024)
by: Saranchuk, Amy, et al.
Published: (2024)
Exploiting the Semantic Knowledge of Pre-trained Text-Encoders for Continual Learning
by: Yu, Lu, et al.
Published: (2024)
by: Yu, Lu, et al.
Published: (2024)
Fake It Right: Injecting Anatomical Logic into Synthetic Supervised Pre-training for Medical Segmentation
by: Tang, Jiaqi, et al.
Published: (2026)
by: Tang, Jiaqi, et al.
Published: (2026)
Self-Supervised Pre-training with Symmetric Superimposition Modeling for Scene Text Recognition
by: Gao, Zuan, et al.
Published: (2024)
by: Gao, Zuan, et al.
Published: (2024)
PatchContrast: Self-Supervised Pre-training for 3D Object Detection
by: Shrout, Oren, et al.
Published: (2023)
by: Shrout, Oren, et al.
Published: (2023)
BRAVEn: Improving Self-Supervised Pre-training for Visual and Auditory Speech Recognition
by: Haliassos, Alexandros, et al.
Published: (2024)
by: Haliassos, Alexandros, et al.
Published: (2024)
A Closer Look at Benchmarking Self-Supervised Pre-training with Image Classification
by: Marks, Markus, et al.
Published: (2024)
by: Marks, Markus, et al.
Published: (2024)
In Pursuit of Pixel Supervision for Visual Pre-training
by: Yang, Lihe, et al.
Published: (2025)
by: Yang, Lihe, et al.
Published: (2025)
Formula-Supervised Visual-Geometric Pre-training
by: Yamada, Ryosuke, et al.
Published: (2024)
by: Yamada, Ryosuke, et al.
Published: (2024)
Industrial Synthetic Segment Pre-training
by: Mae, Shinichi, et al.
Published: (2025)
by: Mae, Shinichi, et al.
Published: (2025)
Self-Supervised Pre-training with Combined Datasets for 3D Perception in Autonomous Driving
by: Wang, Shumin, et al.
Published: (2025)
by: Wang, Shumin, et al.
Published: (2025)
Cross-Scale Pretraining: Enhancing Self-Supervised Learning for Low-Resolution Satellite Imagery for Semantic Segmentation
by: Waithaka, John, et al.
Published: (2026)
by: Waithaka, John, et al.
Published: (2026)
There is No VAE: End-to-End Pixel-Space Generative Modeling via Self-Supervised Pre-training
by: Lei, Jiachen, et al.
Published: (2025)
by: Lei, Jiachen, et al.
Published: (2025)
Self-Supervised Pre-training Tasks for an fMRI Time-series Transformer in Autism Detection
by: Zhou, Yinchi, et al.
Published: (2024)
by: Zhou, Yinchi, et al.
Published: (2024)
DINO-Tracker: Taming DINO for Self-Supervised Point Tracking in a Single Video
by: Tumanyan, Narek, et al.
Published: (2024)
by: Tumanyan, Narek, et al.
Published: (2024)
A Curriculum-style Self-training Approach for Source-Free Semantic Segmentation
by: Wang, Yuxi, et al.
Published: (2021)
by: Wang, Yuxi, et al.
Published: (2021)
Endo-CLIP: Progressive Self-Supervised Pre-training on Raw Colonoscopy Records
by: He, Yili, et al.
Published: (2025)
by: He, Yili, et al.
Published: (2025)
SelfMedHPM: Self Pre-training With Hard Patches Mining Masked Autoencoders For Medical Image Segmentation
by: Lv, Yunhao, et al.
Published: (2025)
by: Lv, Yunhao, et al.
Published: (2025)
SGTC: Semantic-Guided Triplet Co-training for Sparsely Annotated Semi-Supervised Medical Image Segmentation
by: Yan, Ke, et al.
Published: (2024)
by: Yan, Ke, et al.
Published: (2024)
Similar Items
-
GazeVLM: Active Vision via Internal Attention Control for Multimodal Reasoning
by: Ebouky, Brown, et al.
Published: (2026) -
VP Lab: a PEFT-Enabled Visual Prompting Laboratory for Semantic Segmentation
by: Avogaro, Niccolo, et al.
Published: (2025) -
Show or Tell? Effectively prompting Vision-Language Models for semantic segmentation
by: Avogaro, Niccolo, et al.
Published: (2025) -
Continuous Pose for Monocular Cameras in Neural Implicit Representation
by: Ma, Qi, et al.
Published: (2023) -
Self-supervised Shape Completion via Involution and Implicit Correspondences
by: Liu, Mengya, et al.
Published: (2024)