UDON: Universal Dynamic Online distillatioN for generic image representations
Fuente:
arXiv
Guardado en:
| Autores principales: | Ypsilantis, Nikolaos-Antonios, Chen, Kaifeng, Araujo, André, Chum, Ondřej |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Infusing fine-grained visual knowledge to Vision-Language Models
por: Ypsilantis, Nikolaos-Antonios, et al.
Publicado: (2025)
por: Ypsilantis, Nikolaos-Antonios, et al.
Publicado: (2025)
Co-Segmentation without any Pixel-level Supervision with Application to Large-Scale Sketch Classification
por: Ypsilantis, Nikolaos-Antonios, et al.
Publicado: (2024)
por: Ypsilantis, Nikolaos-Antonios, et al.
Publicado: (2024)
ILIAS: Instance-Level Image retrieval At Scale
por: Kordopatis-Zilos, Giorgos, et al.
Publicado: (2025)
por: Kordopatis-Zilos, Giorgos, et al.
Publicado: (2025)
Crafting Distribution Shifts for Validation and Training in Single Source Domain Generalization
por: Efthymiadis, Nikos, et al.
Publicado: (2024)
por: Efthymiadis, Nikos, et al.
Publicado: (2024)
Dark Side Augmentation: Generating Diverse Night Examples for Metric Learning
por: Mohwald, Albert, et al.
Publicado: (2023)
por: Mohwald, Albert, et al.
Publicado: (2023)
Composed Image Retrieval for Training-Free Domain Conversion
por: Efthymiadis, Nikos, et al.
Publicado: (2024)
por: Efthymiadis, Nikos, et al.
Publicado: (2024)
Composed Image Retrieval for Remote Sensing
por: Psomas, Bill, et al.
Publicado: (2024)
por: Psomas, Bill, et al.
Publicado: (2024)
CrossFlowDG: Bridging the Modality Gap with Cross-modal Flow Matching for Domain Generalization
por: Kritikos, Antonios, et al.
Publicado: (2026)
por: Kritikos, Antonios, et al.
Publicado: (2026)
Global-to-Local or Local-to-Global? Enhancing Image Retrieval with Efficient Local Search and Effective Global Re-ranking
por: Aiger, Dror, et al.
Publicado: (2025)
por: Aiger, Dror, et al.
Publicado: (2025)
Instance-Level Composed Image Retrieval
por: Psomas, Bill, et al.
Publicado: (2025)
por: Psomas, Bill, et al.
Publicado: (2025)
Benchmarking Composed Image Retrieval for Applied Earth Observation
por: Psomas, Bill, et al.
Publicado: (2026)
por: Psomas, Bill, et al.
Publicado: (2026)
Universal Neural Architecture Space: Covering ConvNets, Transformers and Everything in Between
por: Týbl, Ondřej, et al.
Publicado: (2025)
por: Týbl, Ondřej, et al.
Publicado: (2025)
HAMMR: HierArchical MultiModal React agents for generic VQA
por: Castrejon, Lluis, et al.
Publicado: (2024)
por: Castrejon, Lluis, et al.
Publicado: (2024)
The Age-specific Alzheimer 's Disease Prediction with Characteristic Constraints in Nonuniform Time Span
por: Hong, Xin, et al.
Publicado: (2025)
por: Hong, Xin, et al.
Publicado: (2025)
Universal dimensions of visual representation
por: Chen, Zirui, et al.
Publicado: (2024)
por: Chen, Zirui, et al.
Publicado: (2024)
3D Engine-ready Photorealistic Avatars via Dynamic Textures
por: Wang, Yifan, et al.
Publicado: (2025)
por: Wang, Yifan, et al.
Publicado: (2025)
GeoDTR+: Toward generic cross-view geolocalization via geometric disentanglement
por: Zhang, Xiaohan, et al.
Publicado: (2023)
por: Zhang, Xiaohan, et al.
Publicado: (2023)
Learning predictable and robust neural representations by straightening image sequences
por: Niu, Xueyan, et al.
Publicado: (2024)
por: Niu, Xueyan, et al.
Publicado: (2024)
Learned representation-guided diffusion models for large-image generation
por: Graikos, Alexandros, et al.
Publicado: (2023)
por: Graikos, Alexandros, et al.
Publicado: (2023)
Unconditional CNN denoisers contain sparse semantic representation of images
por: Kadkhodaie, Zahra, et al.
Publicado: (2025)
por: Kadkhodaie, Zahra, et al.
Publicado: (2025)
Predicting 3D representations for Dynamic Scenes
por: Qi, Di, et al.
Publicado: (2025)
por: Qi, Di, et al.
Publicado: (2025)
Downscaling climate projections to 1 km with single-image super resolution
por: Košťál, Petr, et al.
Publicado: (2025)
por: Košťál, Petr, et al.
Publicado: (2025)
Exploring PCA-based feature representations of image pixels via CNN to enhance food image segmentation
por: Dai, Ying
Publicado: (2024)
por: Dai, Ying
Publicado: (2024)
Training-free Neural Architecture Search through Variance of Knowledge of Deep Network Weights
por: Týbl, Ondřej, et al.
Publicado: (2025)
por: Týbl, Ondřej, et al.
Publicado: (2025)
DORA: Dynamic Online Reinforcement Agent for Token Merging in Vision Transformers
por: He, Kaixuan, et al.
Publicado: (2026)
por: He, Kaixuan, et al.
Publicado: (2026)
Dynamic watermarks in images generated by diffusion models
por: Chen, Yunzhuo, et al.
Publicado: (2025)
por: Chen, Yunzhuo, et al.
Publicado: (2025)
Dynamically enhanced static handwriting representation for Parkinson's disease detection
por: Diaz, Moises, et al.
Publicado: (2024)
por: Diaz, Moises, et al.
Publicado: (2024)
Deep kernel representations of latent space features for low-dose PET-MR imaging robust to variable dose reduction
por: Pain, Cameron Dennis, et al.
Publicado: (2024)
por: Pain, Cameron Dennis, et al.
Publicado: (2024)
Adaptive Begin-of-Video Tokens for Autoregressive Video Diffusion Models
por: Cheng, Tianle, et al.
Publicado: (2025)
por: Cheng, Tianle, et al.
Publicado: (2025)
2D Triangle Splatting for Direct Differentiable Mesh Training
por: Sheng, Kaifeng, et al.
Publicado: (2025)
por: Sheng, Kaifeng, et al.
Publicado: (2025)
Second-order Gaussian directional derivative representations for image high-resolution corner detection
por: Xie, Dongbo, et al.
Publicado: (2026)
por: Xie, Dongbo, et al.
Publicado: (2026)
DynOMo: Online Point Tracking by Dynamic Online Monocular Gaussian Reconstruction
por: Seidenschwarz, Jenny, et al.
Publicado: (2024)
por: Seidenschwarz, Jenny, et al.
Publicado: (2024)
DartControl: A Diffusion-Based Autoregressive Motion Model for Real-Time Text-Driven Motion Control
por: Zhao, Kaifeng, et al.
Publicado: (2024)
por: Zhao, Kaifeng, et al.
Publicado: (2024)
TIPS: Text-Image Pretraining with Spatial awareness
por: Maninis, Kevis-Kokitsi, et al.
Publicado: (2024)
por: Maninis, Kevis-Kokitsi, et al.
Publicado: (2024)
Y-MAP-Net: Real-time depth, normals, segmentation, multi-label captioning and 2D human pose in RGB images
por: Qammaz, Ammar, et al.
Publicado: (2024)
por: Qammaz, Ammar, et al.
Publicado: (2024)
Robust image representations with counterfactual contrastive learning
por: Roschewitz, Mélanie, et al.
Publicado: (2024)
por: Roschewitz, Mélanie, et al.
Publicado: (2024)
Disentangling representations of retinal images with generative models
por: Müller, Sarah, et al.
Publicado: (2024)
por: Müller, Sarah, et al.
Publicado: (2024)
Ca2-VDM: Efficient Autoregressive Video Diffusion Model with Causal Generation and Cache Sharing
por: Gao, Kaifeng, et al.
Publicado: (2024)
por: Gao, Kaifeng, et al.
Publicado: (2024)
Learning Vision from Models Rivals Learning Vision from Data
por: Tian, Yonglong, et al.
Publicado: (2023)
por: Tian, Yonglong, et al.
Publicado: (2023)
EgoM2P: Egocentric Multimodal Multitask Pretraining
por: Li, Gen, et al.
Publicado: (2025)
por: Li, Gen, et al.
Publicado: (2025)
Ejemplares similares
-
Infusing fine-grained visual knowledge to Vision-Language Models
por: Ypsilantis, Nikolaos-Antonios, et al.
Publicado: (2025) -
Co-Segmentation without any Pixel-level Supervision with Application to Large-Scale Sketch Classification
por: Ypsilantis, Nikolaos-Antonios, et al.
Publicado: (2024) -
ILIAS: Instance-Level Image retrieval At Scale
por: Kordopatis-Zilos, Giorgos, et al.
Publicado: (2025) -
Crafting Distribution Shifts for Validation and Training in Single Source Domain Generalization
por: Efthymiadis, Nikos, et al.
Publicado: (2024) -
Dark Side Augmentation: Generating Diverse Night Examples for Metric Learning
por: Mohwald, Albert, et al.
Publicado: (2023)