UDON: Universal Dynamic Online distillatioN for generic image representations
Fuente:
arXiv
Saved in:
| Main Authors: | Ypsilantis, Nikolaos-Antonios, Chen, Kaifeng, Araujo, André, Chum, Ondřej |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Infusing fine-grained visual knowledge to Vision-Language Models
by: Ypsilantis, Nikolaos-Antonios, et al.
Published: (2025)
by: Ypsilantis, Nikolaos-Antonios, et al.
Published: (2025)
Co-Segmentation without any Pixel-level Supervision with Application to Large-Scale Sketch Classification
by: Ypsilantis, Nikolaos-Antonios, et al.
Published: (2024)
by: Ypsilantis, Nikolaos-Antonios, et al.
Published: (2024)
ILIAS: Instance-Level Image retrieval At Scale
by: Kordopatis-Zilos, Giorgos, et al.
Published: (2025)
by: Kordopatis-Zilos, Giorgos, et al.
Published: (2025)
Crafting Distribution Shifts for Validation and Training in Single Source Domain Generalization
by: Efthymiadis, Nikos, et al.
Published: (2024)
by: Efthymiadis, Nikos, et al.
Published: (2024)
Dark Side Augmentation: Generating Diverse Night Examples for Metric Learning
by: Mohwald, Albert, et al.
Published: (2023)
by: Mohwald, Albert, et al.
Published: (2023)
Composed Image Retrieval for Training-Free Domain Conversion
by: Efthymiadis, Nikos, et al.
Published: (2024)
by: Efthymiadis, Nikos, et al.
Published: (2024)
Composed Image Retrieval for Remote Sensing
by: Psomas, Bill, et al.
Published: (2024)
by: Psomas, Bill, et al.
Published: (2024)
CrossFlowDG: Bridging the Modality Gap with Cross-modal Flow Matching for Domain Generalization
by: Kritikos, Antonios, et al.
Published: (2026)
by: Kritikos, Antonios, et al.
Published: (2026)
Global-to-Local or Local-to-Global? Enhancing Image Retrieval with Efficient Local Search and Effective Global Re-ranking
by: Aiger, Dror, et al.
Published: (2025)
by: Aiger, Dror, et al.
Published: (2025)
Instance-Level Composed Image Retrieval
by: Psomas, Bill, et al.
Published: (2025)
by: Psomas, Bill, et al.
Published: (2025)
Benchmarking Composed Image Retrieval for Applied Earth Observation
by: Psomas, Bill, et al.
Published: (2026)
by: Psomas, Bill, et al.
Published: (2026)
Universal Neural Architecture Space: Covering ConvNets, Transformers and Everything in Between
by: Týbl, Ondřej, et al.
Published: (2025)
by: Týbl, Ondřej, et al.
Published: (2025)
HAMMR: HierArchical MultiModal React agents for generic VQA
by: Castrejon, Lluis, et al.
Published: (2024)
by: Castrejon, Lluis, et al.
Published: (2024)
The Age-specific Alzheimer 's Disease Prediction with Characteristic Constraints in Nonuniform Time Span
by: Hong, Xin, et al.
Published: (2025)
by: Hong, Xin, et al.
Published: (2025)
Universal dimensions of visual representation
by: Chen, Zirui, et al.
Published: (2024)
by: Chen, Zirui, et al.
Published: (2024)
3D Engine-ready Photorealistic Avatars via Dynamic Textures
by: Wang, Yifan, et al.
Published: (2025)
by: Wang, Yifan, et al.
Published: (2025)
GeoDTR+: Toward generic cross-view geolocalization via geometric disentanglement
by: Zhang, Xiaohan, et al.
Published: (2023)
by: Zhang, Xiaohan, et al.
Published: (2023)
Learning predictable and robust neural representations by straightening image sequences
by: Niu, Xueyan, et al.
Published: (2024)
by: Niu, Xueyan, et al.
Published: (2024)
Learned representation-guided diffusion models for large-image generation
by: Graikos, Alexandros, et al.
Published: (2023)
by: Graikos, Alexandros, et al.
Published: (2023)
Unconditional CNN denoisers contain sparse semantic representation of images
by: Kadkhodaie, Zahra, et al.
Published: (2025)
by: Kadkhodaie, Zahra, et al.
Published: (2025)
Predicting 3D representations for Dynamic Scenes
by: Qi, Di, et al.
Published: (2025)
by: Qi, Di, et al.
Published: (2025)
Downscaling climate projections to 1 km with single-image super resolution
by: Košťál, Petr, et al.
Published: (2025)
by: Košťál, Petr, et al.
Published: (2025)
Exploring PCA-based feature representations of image pixels via CNN to enhance food image segmentation
by: Dai, Ying
Published: (2024)
by: Dai, Ying
Published: (2024)
Training-free Neural Architecture Search through Variance of Knowledge of Deep Network Weights
by: Týbl, Ondřej, et al.
Published: (2025)
by: Týbl, Ondřej, et al.
Published: (2025)
DORA: Dynamic Online Reinforcement Agent for Token Merging in Vision Transformers
by: He, Kaixuan, et al.
Published: (2026)
by: He, Kaixuan, et al.
Published: (2026)
Dynamic watermarks in images generated by diffusion models
by: Chen, Yunzhuo, et al.
Published: (2025)
by: Chen, Yunzhuo, et al.
Published: (2025)
Dynamically enhanced static handwriting representation for Parkinson's disease detection
by: Diaz, Moises, et al.
Published: (2024)
by: Diaz, Moises, et al.
Published: (2024)
Deep kernel representations of latent space features for low-dose PET-MR imaging robust to variable dose reduction
by: Pain, Cameron Dennis, et al.
Published: (2024)
by: Pain, Cameron Dennis, et al.
Published: (2024)
Adaptive Begin-of-Video Tokens for Autoregressive Video Diffusion Models
by: Cheng, Tianle, et al.
Published: (2025)
by: Cheng, Tianle, et al.
Published: (2025)
2D Triangle Splatting for Direct Differentiable Mesh Training
by: Sheng, Kaifeng, et al.
Published: (2025)
by: Sheng, Kaifeng, et al.
Published: (2025)
Second-order Gaussian directional derivative representations for image high-resolution corner detection
by: Xie, Dongbo, et al.
Published: (2026)
by: Xie, Dongbo, et al.
Published: (2026)
DynOMo: Online Point Tracking by Dynamic Online Monocular Gaussian Reconstruction
by: Seidenschwarz, Jenny, et al.
Published: (2024)
by: Seidenschwarz, Jenny, et al.
Published: (2024)
DartControl: A Diffusion-Based Autoregressive Motion Model for Real-Time Text-Driven Motion Control
by: Zhao, Kaifeng, et al.
Published: (2024)
by: Zhao, Kaifeng, et al.
Published: (2024)
TIPS: Text-Image Pretraining with Spatial awareness
by: Maninis, Kevis-Kokitsi, et al.
Published: (2024)
by: Maninis, Kevis-Kokitsi, et al.
Published: (2024)
Y-MAP-Net: Real-time depth, normals, segmentation, multi-label captioning and 2D human pose in RGB images
by: Qammaz, Ammar, et al.
Published: (2024)
by: Qammaz, Ammar, et al.
Published: (2024)
Robust image representations with counterfactual contrastive learning
by: Roschewitz, Mélanie, et al.
Published: (2024)
by: Roschewitz, Mélanie, et al.
Published: (2024)
Disentangling representations of retinal images with generative models
by: Müller, Sarah, et al.
Published: (2024)
by: Müller, Sarah, et al.
Published: (2024)
Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing
by: Gao, Kaifeng, et al.
Published: (2024)
by: Gao, Kaifeng, et al.
Published: (2024)
Learning Vision from Models Rivals Learning Vision from Data
by: Tian, Yonglong, et al.
Published: (2023)
by: Tian, Yonglong, et al.
Published: (2023)
EgoM2P: Egocentric Multimodal Multitask Pretraining
by: Li, Gen, et al.
Published: (2025)
by: Li, Gen, et al.
Published: (2025)
Similar Items
-
Infusing fine-grained visual knowledge to Vision-Language Models
by: Ypsilantis, Nikolaos-Antonios, et al.
Published: (2025) -
Co-Segmentation without any Pixel-level Supervision with Application to Large-Scale Sketch Classification
by: Ypsilantis, Nikolaos-Antonios, et al.
Published: (2024) -
ILIAS: Instance-Level Image retrieval At Scale
by: Kordopatis-Zilos, Giorgos, et al.
Published: (2025) -
Crafting Distribution Shifts for Validation and Training in Single Source Domain Generalization
by: Efthymiadis, Nikos, et al.
Published: (2024) -
Dark Side Augmentation: Generating Diverse Night Examples for Metric Learning
by: Mohwald, Albert, et al.
Published: (2023)