Gespeichert in:
| Hauptverfasser: | Marrie, Juliette, Arbel, Michael, Mairal, Julien, Larlus, Diane |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2402.11305 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LUDVIG: Learning-Free Uplifting of 2D Visual Features to Gaussian Splatting Scenes
von: Marrie, Juliette, et al.
Veröffentlicht: (2024)
von: Marrie, Juliette, et al.
Veröffentlicht: (2024)
MicroFlow: Domain-Specific Optical Flow for Ground Deformation Estimation in Seismic Events
von: Bertrand, Juliette, et al.
Veröffentlicht: (2025)
von: Bertrand, Juliette, et al.
Veröffentlicht: (2025)
Unsupervised Imaging Inverse Problems with Diffusion Distribution Matching
von: Meanti, Giacomo, et al.
Veröffentlicht: (2025)
von: Meanti, Giacomo, et al.
Veröffentlicht: (2025)
Beyond MMSE: Enhancing PnP Restoration with ProxiMAP
von: Vert, Kenta, et al.
Veröffentlicht: (2026)
von: Vert, Kenta, et al.
Veröffentlicht: (2026)
UNIC: Universal Classification Models via Multi-teacher Distillation
von: Sariyildiz, Mert Bulent, et al.
Veröffentlicht: (2024)
von: Sariyildiz, Mert Bulent, et al.
Veröffentlicht: (2024)
Layered Motion Fusion: Lifting Motion Segmentation to 3D in Egocentric Videos
von: Tschernezki, Vadim, et al.
Veröffentlicht: (2025)
von: Tschernezki, Vadim, et al.
Veröffentlicht: (2025)
What could go wrong? Discovering and describing failure modes in computer vision
von: Csurka, Gabriela, et al.
Veröffentlicht: (2024)
von: Csurka, Gabriela, et al.
Veröffentlicht: (2024)
CASA: Cross-Attention over Self-Attention for Efficient Vision-Language Fusion
von: Böhle, Moritz, et al.
Veröffentlicht: (2025)
von: Böhle, Moritz, et al.
Veröffentlicht: (2025)
Mesh4D: 4D Mesh Reconstruction and Tracking from Monocular Video
von: Jiang, Zeren, et al.
Veröffentlicht: (2026)
von: Jiang, Zeren, et al.
Veröffentlicht: (2026)
DUNE: Distilling a Universal Encoder from Heterogeneous 2D and 3D Teachers
von: Sariyildiz, Mert Bulent, et al.
Veröffentlicht: (2025)
von: Sariyildiz, Mert Bulent, et al.
Veröffentlicht: (2025)
PanSt3R: Multi-view Consistent Panoptic Segmentation
von: Zust, Lojze, et al.
Veröffentlicht: (2025)
von: Zust, Lojze, et al.
Veröffentlicht: (2025)
Visual Instruction Pretraining for Domain-Specific Foundation Models
von: Li, Yuxuan, et al.
Veröffentlicht: (2025)
von: Li, Yuxuan, et al.
Veröffentlicht: (2025)
Vision Transformers Need Registers
von: Darcet, Timothée, et al.
Veröffentlicht: (2023)
von: Darcet, Timothée, et al.
Veröffentlicht: (2023)
Task Alignment: A simple and effective proxy for model merging in computer vision
von: de Jorge, Pau, et al.
Veröffentlicht: (2026)
von: de Jorge, Pau, et al.
Veröffentlicht: (2026)
Weatherproofing Retrieval for Localization with Generative AI and Geometric Consistency
von: Kalantidis, Yannis, et al.
Veröffentlicht: (2024)
von: Kalantidis, Yannis, et al.
Veröffentlicht: (2024)
Geo4D: Leveraging Video Generators for Geometric 4D Scene Reconstruction
von: Jiang, Zeren, et al.
Veröffentlicht: (2025)
von: Jiang, Zeren, et al.
Veröffentlicht: (2025)
SpectralEarth-FM: Bringing Hyperspectral Imagery into Multimodal Earth Observation Pretraining
von: Braham, Nassim Ait Ali, et al.
Veröffentlicht: (2026)
von: Braham, Nassim Ait Ali, et al.
Veröffentlicht: (2026)
Cross-Modal Knowledge Distillation from Spatial Transcriptomics to Histology
von: Hizmi, Arbel, et al.
Veröffentlicht: (2026)
von: Hizmi, Arbel, et al.
Veröffentlicht: (2026)
Cluster and Predict Latent Patches for Improved Masked Image Modeling
von: Darcet, Timothée, et al.
Veröffentlicht: (2025)
von: Darcet, Timothée, et al.
Veröffentlicht: (2025)
Task Specific Pretraining with Noisy Labels for Remote Sensing Image Segmentation
von: Liu, Chenying, et al.
Veröffentlicht: (2024)
von: Liu, Chenying, et al.
Veröffentlicht: (2024)
EPIC Fields: Marrying 3D Geometry and Video Understanding
von: Tschernezki, Vadim, et al.
Veröffentlicht: (2023)
von: Tschernezki, Vadim, et al.
Veröffentlicht: (2023)
Is Large-Scale Pretraining the Secret to Good Domain Generalization?
von: Teterwak, Piotr, et al.
Veröffentlicht: (2024)
von: Teterwak, Piotr, et al.
Veröffentlicht: (2024)
Optimal transport unlocks end-to-end learning for single-molecule localization
von: Seailles, Romain, et al.
Veröffentlicht: (2025)
von: Seailles, Romain, et al.
Veröffentlicht: (2025)
SLAD : Shared LoRA Adapters for Task Specific Distillation
von: Bensaid, Reda, et al.
Veröffentlicht: (2026)
von: Bensaid, Reda, et al.
Veröffentlicht: (2026)
PANDAS: Prototype-based Novel Class Discovery and Detection
von: Hayes, Tyler L., et al.
Veröffentlicht: (2024)
von: Hayes, Tyler L., et al.
Veröffentlicht: (2024)
Task-Specific Knowledge Distillation from the Vision Foundation Model for Enhanced Medical Image Segmentation
von: Liang, Pengchen, et al.
Veröffentlicht: (2025)
von: Liang, Pengchen, et al.
Veröffentlicht: (2025)
Fast Semisupervised Unmixing Using Nonconvex Optimization
von: Rasti, Behnood, et al.
Veröffentlicht: (2024)
von: Rasti, Behnood, et al.
Veröffentlicht: (2024)
Image Processing and Machine Learning for Hyperspectral Unmixing: An Overview and the HySUPP Python Package
von: Rasti, Behnood, et al.
Veröffentlicht: (2023)
von: Rasti, Behnood, et al.
Veröffentlicht: (2023)
Pretrained Visual Uncertainties
von: Kirchhof, Michael, et al.
Veröffentlicht: (2024)
von: Kirchhof, Michael, et al.
Veröffentlicht: (2024)
SSL-AD: Spatiotemporal Self-Supervised Learning for Generalizability and Adaptability Across Alzheimer's Prediction Tasks and Datasets
von: Kaczmarek, Emily, et al.
Veröffentlicht: (2025)
von: Kaczmarek, Emily, et al.
Veröffentlicht: (2025)
Are Pretrained Image Matchers Good Enough for SAR-Optical Satellite Registration?
von: Corley, Isaac, et al.
Veröffentlicht: (2026)
von: Corley, Isaac, et al.
Veröffentlicht: (2026)
Syn4D: A Multiview Synthetic 4D Dataset
von: Jiang, Zeren, et al.
Veröffentlicht: (2026)
von: Jiang, Zeren, et al.
Veröffentlicht: (2026)
Parameter-Efficient Fine-Tuning of Large Pretrained Models for Instance Segmentation Tasks
von: Baker, Nermeen Abou, et al.
Veröffentlicht: (2026)
von: Baker, Nermeen Abou, et al.
Veröffentlicht: (2026)
CogVLM: Visual Expert for Pretrained Language Models
von: Wang, Weihan, et al.
Veröffentlicht: (2023)
von: Wang, Weihan, et al.
Veröffentlicht: (2023)
SpectralEarth: Training Hyperspectral Foundation Models at Scale
von: Braham, Nassim Ait Ali, et al.
Veröffentlicht: (2024)
von: Braham, Nassim Ait Ali, et al.
Veröffentlicht: (2024)
What Makes a Good Dataset for Knowledge Distillation?
von: Frank, Logan, et al.
Veröffentlicht: (2024)
von: Frank, Logan, et al.
Veröffentlicht: (2024)
Leveraging Vision-Language Foundation Models to Reveal Hidden Image-Attribute Relationships in Medical Imaging
von: Kumar, Amar, et al.
Veröffentlicht: (2025)
von: Kumar, Amar, et al.
Veröffentlicht: (2025)
Distilling Vision-Language Pretraining for Efficient Cross-Modal Retrieval
von: Jang, Young Kyun, et al.
Veröffentlicht: (2024)
von: Jang, Young Kyun, et al.
Veröffentlicht: (2024)
Just rotate it! Uncertainty estimation in closed-source models via multiple queries
von: Pitas, Konstantinos, et al.
Veröffentlicht: (2024)
von: Pitas, Konstantinos, et al.
Veröffentlicht: (2024)
From Generalist to Specialist: Adapting Vision Language Models via Task-Specific Visual Instruction Tuning
von: Bai, Yang, et al.
Veröffentlicht: (2024)
von: Bai, Yang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
LUDVIG: Learning-Free Uplifting of 2D Visual Features to Gaussian Splatting Scenes
von: Marrie, Juliette, et al.
Veröffentlicht: (2024) -
MicroFlow: Domain-Specific Optical Flow for Ground Deformation Estimation in Seismic Events
von: Bertrand, Juliette, et al.
Veröffentlicht: (2025) -
Unsupervised Imaging Inverse Problems with Diffusion Distribution Matching
von: Meanti, Giacomo, et al.
Veröffentlicht: (2025) -
Beyond MMSE: Enhancing PnP Restoration with ProxiMAP
von: Vert, Kenta, et al.
Veröffentlicht: (2026) -
UNIC: Universal Classification Models via Multi-teacher Distillation
von: Sariyildiz, Mert Bulent, et al.
Veröffentlicht: (2024)