RADAR: Relative Angular Divergence Across Representations
Fuente:
arXiv
Salvato in:
| Autori principali: | Cadet, Xavier, Nowak, Mateusz, Chin, Peter |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ABCD: All Biases Come Disguised
di: Nowak, Mateusz, et al.
Pubblicazione: (2026)
di: Nowak, Mateusz, et al.
Pubblicazione: (2026)
VoD-3DGS: View-opacity-Dependent 3D Gaussian Splatting
di: Nowak, Mateusz, et al.
Pubblicazione: (2025)
di: Nowak, Mateusz, et al.
Pubblicazione: (2025)
DocAtlas: Multilingual Document Understanding Across 80+ Languages
di: Heakl, Ahmed, et al.
Pubblicazione: (2026)
di: Heakl, Ahmed, et al.
Pubblicazione: (2026)
Ferret-UI 2: Mastering Universal User Interface Understanding Across Platforms
di: Li, Zhangheng, et al.
Pubblicazione: (2024)
di: Li, Zhangheng, et al.
Pubblicazione: (2024)
Hyperbolic Multimodal Representation Learning for Biological Taxonomies
di: Gong, ZeMing, et al.
Pubblicazione: (2025)
di: Gong, ZeMing, et al.
Pubblicazione: (2025)
Mixture of Group Experts for Learning Invariant Representations
di: Kang, Lei, et al.
Pubblicazione: (2025)
di: Kang, Lei, et al.
Pubblicazione: (2025)
Verbalized Representation Learning for Interpretable Few-Shot Generalization
di: Yang, Cheng-Fu, et al.
Pubblicazione: (2024)
di: Yang, Cheng-Fu, et al.
Pubblicazione: (2024)
T-MARS: Improving Visual Representations by Circumventing Text Feature Learning
di: Maini, Pratyush, et al.
Pubblicazione: (2023)
di: Maini, Pratyush, et al.
Pubblicazione: (2023)
Vision-Language Models Create Cross-Modal Task Representations
di: Luo, Grace, et al.
Pubblicazione: (2024)
di: Luo, Grace, et al.
Pubblicazione: (2024)
JPEG-LM: LLMs as Image Generators with Canonical Codec Representations
di: Han, Xiaochuang, et al.
Pubblicazione: (2024)
di: Han, Xiaochuang, et al.
Pubblicazione: (2024)
REBEL: Reinforcement Learning via Regressing Relative Rewards
di: Gao, Zhaolin, et al.
Pubblicazione: (2024)
di: Gao, Zhaolin, et al.
Pubblicazione: (2024)
Debiasing Large Vision-Language Models by Ablating Protected Attribute Representations
di: Ratzlaff, Neale, et al.
Pubblicazione: (2024)
di: Ratzlaff, Neale, et al.
Pubblicazione: (2024)
General Transform: A Unified Framework for Adaptive Transform to Enhance Representations
di: Budiutama, Gekko, et al.
Pubblicazione: (2025)
di: Budiutama, Gekko, et al.
Pubblicazione: (2025)
MOFI: Learning Image Representations from Noisy Entity Annotated Images
di: Wu, Wentao, et al.
Pubblicazione: (2023)
di: Wu, Wentao, et al.
Pubblicazione: (2023)
Cross-modal Causal Relation Alignment for Video Question Grounding
di: Chen, Weixing, et al.
Pubblicazione: (2025)
di: Chen, Weixing, et al.
Pubblicazione: (2025)
Modeling Multimodal Social Interactions: New Challenges and Baselines with Densely Aligned Representations
di: Lee, Sangmin, et al.
Pubblicazione: (2024)
di: Lee, Sangmin, et al.
Pubblicazione: (2024)
URRL-IMVC: Unified and Robust Representation Learning for Incomplete Multi-View Clustering
di: Teng, Ge, et al.
Pubblicazione: (2024)
di: Teng, Ge, et al.
Pubblicazione: (2024)
No Train, all Gain: Self-Supervised Gradients Improve Deep Frozen Representations
di: Simoncini, Walter, et al.
Pubblicazione: (2024)
di: Simoncini, Walter, et al.
Pubblicazione: (2024)
BridgeTower: Building Bridges Between Encoders in Vision-Language Representation Learning
di: Xu, Xiao, et al.
Pubblicazione: (2022)
di: Xu, Xiao, et al.
Pubblicazione: (2022)
FocalLens: Instruction Tuning Enables Zero-Shot Conditional Image Representations
di: Hsieh, Cheng-Yu, et al.
Pubblicazione: (2025)
di: Hsieh, Cheng-Yu, et al.
Pubblicazione: (2025)
Prioritizing Image-Related Tokens Enhances Vision-Language Pre-Training
di: Chen, Yangyi, et al.
Pubblicazione: (2025)
di: Chen, Yangyi, et al.
Pubblicazione: (2025)
Diffusion-RPO: Aligning Diffusion Models through Relative Preference Optimization
di: Gu, Yi, et al.
Pubblicazione: (2024)
di: Gu, Yi, et al.
Pubblicazione: (2024)
LatentExplainer: Explaining Latent Representations in Deep Generative Models with Multimodal Large Language Models
di: Zhu, Mengdan, et al.
Pubblicazione: (2024)
di: Zhu, Mengdan, et al.
Pubblicazione: (2024)
If CLIP Could Talk: Understanding Vision-Language Model Representations Through Their Preferred Concept Descriptions
di: Esfandiarpoor, Reza, et al.
Pubblicazione: (2024)
di: Esfandiarpoor, Reza, et al.
Pubblicazione: (2024)
Reefknot: A Comprehensive Benchmark for Relation Hallucination Evaluation, Analysis and Mitigation in Multimodal Large Language Models
di: Zheng, Kening, et al.
Pubblicazione: (2024)
di: Zheng, Kening, et al.
Pubblicazione: (2024)
Bridging Neural and Symbolic Representations with Transitional Dictionary Learning
di: Cheng, Junyan, et al.
Pubblicazione: (2023)
di: Cheng, Junyan, et al.
Pubblicazione: (2023)
RADAR: Enhancing Radiology Report Generation with Supplementary Knowledge Injection
di: Hou, Wenjun, et al.
Pubblicazione: (2025)
di: Hou, Wenjun, et al.
Pubblicazione: (2025)
Traveling Across Languages: Benchmarking Cross-Lingual Consistency in Multimodal LLMs
di: Wang, Hao, et al.
Pubblicazione: (2025)
di: Wang, Hao, et al.
Pubblicazione: (2025)
Learning Representations on the Unit Sphere: Investigating Angular Gaussian and von Mises-Fisher Distributions for Online Continual Learning
di: Michel, Nicolas, et al.
Pubblicazione: (2023)
di: Michel, Nicolas, et al.
Pubblicazione: (2023)
How Do Vision-Language Models Process Conflicting Information Across Modalities?
di: Hua, Tianze, et al.
Pubblicazione: (2025)
di: Hua, Tianze, et al.
Pubblicazione: (2025)
RAVEN: Query-Guided Representation Alignment for Question Answering over Audio, Video, Embedded Sensors, and Natural Language
di: Biswas, Subrata, et al.
Pubblicazione: (2025)
di: Biswas, Subrata, et al.
Pubblicazione: (2025)
Roboflow100-VL: A Multi-Domain Object Detection Benchmark for Vision-Language Models
di: Robicheaux, Peter, et al.
Pubblicazione: (2025)
di: Robicheaux, Peter, et al.
Pubblicazione: (2025)
Red-Teaming Text-to-Image Models via In-Context Experience Replay and Semantic-Preserving Prompt Rewriting
di: Chin, Zhi-Yi, et al.
Pubblicazione: (2024)
di: Chin, Zhi-Yi, et al.
Pubblicazione: (2024)
MLLMs-Augmented Visual-Language Representation Learning
di: Liu, Yanqing, et al.
Pubblicazione: (2023)
di: Liu, Yanqing, et al.
Pubblicazione: (2023)
Guiding Vision-Language Model Selection for Visual Question-Answering Across Tasks, Domains, and Knowledge Types
di: Sinha, Neelabh, et al.
Pubblicazione: (2024)
di: Sinha, Neelabh, et al.
Pubblicazione: (2024)
MM-Spatial: Exploring 3D Spatial Understanding in Multimodal LLMs
di: Daxberger, Erik, et al.
Pubblicazione: (2025)
di: Daxberger, Erik, et al.
Pubblicazione: (2025)
Unified Lexical Representation for Interpretable Visual-Language Alignment
di: Li, Yifan, et al.
Pubblicazione: (2024)
di: Li, Yifan, et al.
Pubblicazione: (2024)
Tracing Representation Progression: Analyzing and Enhancing Layer-Wise Similarity
di: Jiang, Jiachen, et al.
Pubblicazione: (2024)
di: Jiang, Jiachen, et al.
Pubblicazione: (2024)
OSCaR: Object State Captioning and State Change Representation
di: Nguyen, Nguyen, et al.
Pubblicazione: (2024)
di: Nguyen, Nguyen, et al.
Pubblicazione: (2024)
MAMA: Meta-optimized Angular Margin Contrastive Framework for Video-Language Representation Learning
di: Nguyen, Thong, et al.
Pubblicazione: (2024)
di: Nguyen, Thong, et al.
Pubblicazione: (2024)
Documenti analoghi
-
ABCD: All Biases Come Disguised
di: Nowak, Mateusz, et al.
Pubblicazione: (2026) -
VoD-3DGS: View-opacity-Dependent 3D Gaussian Splatting
di: Nowak, Mateusz, et al.
Pubblicazione: (2025) -
DocAtlas: Multilingual Document Understanding Across 80+ Languages
di: Heakl, Ahmed, et al.
Pubblicazione: (2026) -
Ferret-UI 2: Mastering Universal User Interface Understanding Across Platforms
di: Li, Zhangheng, et al.
Pubblicazione: (2024) -
Hyperbolic Multimodal Representation Learning for Biological Taxonomies
di: Gong, ZeMing, et al.
Pubblicazione: (2025)