Saved in:
| Main Author: | Chun, Sanghyuk |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2505.19614 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Improved Probabilistic Image-Text Representations
by: Chun, Sanghyuk
Published: (2023)
by: Chun, Sanghyuk
Published: (2023)
LongProLIP: A Probabilistic Vision-Language Model with Long Context Text
by: Chun, Sanghyuk, et al.
Published: (2025)
by: Chun, Sanghyuk, et al.
Published: (2025)
DNNs May Determine Major Properties of Their Outputs Early, with Timing Possibly Driven by Bias
by: Park, Song, et al.
Published: (2025)
by: Park, Song, et al.
Published: (2025)
Probabilistic Language-Image Pre-Training
by: Chun, Sanghyuk, et al.
Published: (2024)
by: Chun, Sanghyuk, et al.
Published: (2024)
Advances in Multiple Instance Learning for Whole Slide Image Analysis: Techniques, Challenges, and Future Directions
by: Wang, Jun, et al.
Published: (2024)
by: Wang, Jun, et al.
Published: (2024)
Similarity of Neural Architectures using Adversarial Attack Transferability
by: Hwang, Jaehui, et al.
Published: (2022)
by: Hwang, Jaehui, et al.
Published: (2022)
Pretrained Diffusion Models Are Inherently Skipped-Step Samplers
by: Xu, Wenju
Published: (2025)
by: Xu, Wenju
Published: (2025)
Toward Inherently Robust VLMs Against Visual Perception Attacks
by: MohajerAnsari, Pedram, et al.
Published: (2025)
by: MohajerAnsari, Pedram, et al.
Published: (2025)
Discriminative Feature Attributions: Bridging Post Hoc Explainability and Inherent Interpretability
by: Bhalla, Usha, et al.
Published: (2023)
by: Bhalla, Usha, et al.
Published: (2023)
Zero-Shot Defense Against Toxic Images via Inherent Multimodal Alignment in LVLMs
by: Zhao, Wei, et al.
Published: (2025)
by: Zhao, Wei, et al.
Published: (2025)
X-SiT: Inherently Interpretable Surface Vision Transformers for Dementia Diagnosis
by: Bongratz, Fabian, et al.
Published: (2025)
by: Bongratz, Fabian, et al.
Published: (2025)
ProtoEFNet: Dynamic Prototype Learning for Inherently Interpretable Ejection Fraction Estimation in Echocardiography
by: Ghamary, Yeganeh, et al.
Published: (2025)
by: Ghamary, Yeganeh, et al.
Published: (2025)
B-cosification: Transforming Deep Neural Networks to be Inherently Interpretable
by: Arya, Shreyash, et al.
Published: (2024)
by: Arya, Shreyash, et al.
Published: (2024)
Attri-Net: A Globally and Locally Inherently Interpretable Model for Multi-Label Classification Using Class-Specific Counterfactuals
by: Sun, Susu, et al.
Published: (2024)
by: Sun, Susu, et al.
Published: (2024)
Multimodal Learning Without Labeled Multimodal Data: Guarantees and Applications
by: Liang, Paul Pu, et al.
Published: (2023)
by: Liang, Paul Pu, et al.
Published: (2023)
Multi-view Disentanglement for Reinforcement Learning with Multiple Cameras
by: Dunion, Mhairi, et al.
Published: (2024)
by: Dunion, Mhairi, et al.
Published: (2024)
Generalized Contrastive Learning for Universal Multimodal Retrieval
by: Lee, Jungsoo, et al.
Published: (2025)
by: Lee, Jungsoo, et al.
Published: (2025)
Multimodal Representation Learning by Alternating Unimodal Adaptation
by: Zhang, Xiaohui, et al.
Published: (2023)
by: Zhang, Xiaohui, et al.
Published: (2023)
Unsupervised Object-Centric Learning from Multiple Unspecified Viewpoints
by: Yuan, Jinyang, et al.
Published: (2024)
by: Yuan, Jinyang, et al.
Published: (2024)
xMIL: Insightful Explanations for Multiple Instance Learning in Histopathology
by: Hense, Julius, et al.
Published: (2024)
by: Hense, Julius, et al.
Published: (2024)
Modeling Multimodal Social Interactions: New Challenges and Baselines with Densely Aligned Representations
by: Lee, Sangmin, et al.
Published: (2024)
by: Lee, Sangmin, et al.
Published: (2024)
MILES: Modality-Informed Learning Rate Scheduler for Balancing Multimodal Learning
by: Guerra-Manzanares, Alejandro, et al.
Published: (2025)
by: Guerra-Manzanares, Alejandro, et al.
Published: (2025)
Learning with Unmasked Tokens Drives Stronger Vision Learners
by: Kim, Taekyung, et al.
Published: (2023)
by: Kim, Taekyung, et al.
Published: (2023)
On the Value of Cross-Modal Misalignment in Multimodal Representation Learning
by: Cai, Yichao, et al.
Published: (2025)
by: Cai, Yichao, et al.
Published: (2025)
Hierarchy-Guided Multimodal Representation Learning for Taxonomic Inference
by: Ahmed, Sk Miraj, et al.
Published: (2026)
by: Ahmed, Sk Miraj, et al.
Published: (2026)
Toward Unified Multimodal Representation Learning for Autonomous Driving
by: Tao, Ximeng, et al.
Published: (2026)
by: Tao, Ximeng, et al.
Published: (2026)
Multimodal Learning for Embryo Viability Prediction in Clinical IVF
by: Kim, Junsik, et al.
Published: (2024)
by: Kim, Junsik, et al.
Published: (2024)
Collaborative Learning with Multiple Foundation Models for Source-Free Domain Adaptation
by: Lee, Huisoo, et al.
Published: (2025)
by: Lee, Huisoo, et al.
Published: (2025)
Evaluating Multiple Instance Learning Strategies for Automated Sebocyte Droplet Counting
by: Adelipour, Maryam, et al.
Published: (2025)
by: Adelipour, Maryam, et al.
Published: (2025)
Sm: enhanced localization in Multiple Instance Learning for medical imaging classification
by: Castro-Macías, Francisco M., et al.
Published: (2024)
by: Castro-Macías, Francisco M., et al.
Published: (2024)
Contrastive Multiple Instance Learning for Weakly Supervised Person ReID
by: Tyo, Jacob, et al.
Published: (2024)
by: Tyo, Jacob, et al.
Published: (2024)
A Wander Through the Multimodal Landscape: Efficient Transfer Learning via Low-rank Sequence Multimodal Adapter
by: Guo, Zirun, et al.
Published: (2024)
by: Guo, Zirun, et al.
Published: (2024)
Deep Learning Meets OBIA: Tasks, Challenges, Strategies, and Perspectives
by: Ma, Lei, et al.
Published: (2024)
by: Ma, Lei, et al.
Published: (2024)
Doubly Perturbed Task Free Continual Learning
by: Lee, Byung Hyun, et al.
Published: (2023)
by: Lee, Byung Hyun, et al.
Published: (2023)
Principled Multimodal Representation Learning
by: Liu, Xiaohao, et al.
Published: (2025)
by: Liu, Xiaohao, et al.
Published: (2025)
Fine-Scale Soil Mapping in Alaska with Multimodal Machine Learning
by: Lin, Yijun, et al.
Published: (2025)
by: Lin, Yijun, et al.
Published: (2025)
Differential-informed Sample Selection Accelerates Multimodal Contrastive Learning
by: Zhao, Zihua, et al.
Published: (2025)
by: Zhao, Zihua, et al.
Published: (2025)
Learning Multimodal Latent Space with EBM Prior and MCMC Inference
by: Yuan, Shiyu, et al.
Published: (2024)
by: Yuan, Shiyu, et al.
Published: (2024)
Resilient Vision-Tabular Multimodal Learning under Modality Missingness
by: Caruso, Camillo Maria, et al.
Published: (2026)
by: Caruso, Camillo Maria, et al.
Published: (2026)
TerraFlow: Multimodal, Multitemporal Representation Learning for Earth Observation
by: Puriy, Nazar, et al.
Published: (2026)
by: Puriy, Nazar, et al.
Published: (2026)
Similar Items
-
Improved Probabilistic Image-Text Representations
by: Chun, Sanghyuk
Published: (2023) -
LongProLIP: A Probabilistic Vision-Language Model with Long Context Text
by: Chun, Sanghyuk, et al.
Published: (2025) -
DNNs May Determine Major Properties of Their Outputs Early, with Timing Possibly Driven by Bias
by: Park, Song, et al.
Published: (2025) -
Probabilistic Language-Image Pre-Training
by: Chun, Sanghyuk, et al.
Published: (2024) -
Advances in Multiple Instance Learning for Whole Slide Image Analysis: Techniques, Challenges, and Future Directions
by: Wang, Jun, et al.
Published: (2024)