Self-supervised vision-langage alignment of deep learning representations for bone X-rays analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Englebert, Alexandre, Collin, Anne-Sophie, Cornu, Olivier, De Vleeschouwer, Christophe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
X-InstructBLIP: A Framework for aligning X-Modal instruction-aware representations to LLMs and Emergent Cross-modal Reasoning
von: Panagopoulou, Artemis, et al.
Veröffentlicht: (2023)
von: Panagopoulou, Artemis, et al.
Veröffentlicht: (2023)
Self-supervised video pretraining yields robust and more human-aligned visual representations
von: Parthasarathy, Nikhil, et al.
Veröffentlicht: (2022)
von: Parthasarathy, Nikhil, et al.
Veröffentlicht: (2022)
DeepSet SimCLR: Self-supervised deep sets for improved pathology representation learning
von: Torpey, David, et al.
Veröffentlicht: (2024)
von: Torpey, David, et al.
Veröffentlicht: (2024)
Self-supervised structured object representation learning
von: Hadjerci, Oussama, et al.
Veröffentlicht: (2025)
von: Hadjerci, Oussama, et al.
Veröffentlicht: (2025)
Body Part-Based Representation Learning for Occluded Person Re-Identification
von: Somers, Vladimir, et al.
Veröffentlicht: (2022)
von: Somers, Vladimir, et al.
Veröffentlicht: (2022)
Keypoint Promptable Re-Identification
von: Somers, Vladimir, et al.
Veröffentlicht: (2024)
von: Somers, Vladimir, et al.
Veröffentlicht: (2024)
Interpreting the structure of multi-object representations in vision encoders
von: Khajuria, Tarun, et al.
Veröffentlicht: (2024)
von: Khajuria, Tarun, et al.
Veröffentlicht: (2024)
Sustainable self-supervised learning for speech representations
von: Lugo, Luis, et al.
Veröffentlicht: (2024)
von: Lugo, Luis, et al.
Veröffentlicht: (2024)
What to align in multimodal contrastive learning?
von: Dufumier, Benoit, et al.
Veröffentlicht: (2024)
von: Dufumier, Benoit, et al.
Veröffentlicht: (2024)
RUMPL: Ray-Based Transformers for Universal Multi-View 2D to 3D Human Pose Lifting
von: Ghasemzadeh, Seyed Abolfazl, et al.
Veröffentlicht: (2025)
von: Ghasemzadeh, Seyed Abolfazl, et al.
Veröffentlicht: (2025)
MPL: Lifting 3D Human Pose from Multi-view 2D Poses
von: Ghasemzadeh, Seyed Abolfazl, et al.
Veröffentlicht: (2024)
von: Ghasemzadeh, Seyed Abolfazl, et al.
Veröffentlicht: (2024)
CheXalign: Preference fine-tuning in chest X-ray interpretation models without human feedback
von: Hein, Dennis, et al.
Veröffentlicht: (2024)
von: Hein, Dennis, et al.
Veröffentlicht: (2024)
Explore and Explain: Self-supervised Navigation and Recounting
von: Bigazzi, Roberto, et al.
Veröffentlicht: (2020)
von: Bigazzi, Roberto, et al.
Veröffentlicht: (2020)
Towards aligned body representations in vision models
von: Gizdov, Andrey, et al.
Veröffentlicht: (2025)
von: Gizdov, Andrey, et al.
Veröffentlicht: (2025)
Vocabulary-free few-shot learning for Vision-Language Models
von: Zanella, Maxime, et al.
Veröffentlicht: (2025)
von: Zanella, Maxime, et al.
Veröffentlicht: (2025)
Shared and Private Information Learning in Multimodal Sentiment Analysis with Deep Modal Alignment and Self-supervised Multi-Task Learning
von: Lai, Songning, et al.
Veröffentlicht: (2023)
von: Lai, Songning, et al.
Veröffentlicht: (2023)
Self-supervised learning of speech representations with Dutch archival data
von: Vaessen, Nik, et al.
Veröffentlicht: (2025)
von: Vaessen, Nik, et al.
Veröffentlicht: (2025)
Multi-object event graph representation learning for Video Question Answering
von: Wang, Yanan, et al.
Veröffentlicht: (2024)
von: Wang, Yanan, et al.
Veröffentlicht: (2024)
Altogether: Image Captioning via Re-aligning Alt-text
von: Xu, Hu, et al.
Veröffentlicht: (2024)
von: Xu, Hu, et al.
Veröffentlicht: (2024)
Self-supervised learning of video representations from a child's perspective
von: Orhan, A. Emin, et al.
Veröffentlicht: (2024)
von: Orhan, A. Emin, et al.
Veröffentlicht: (2024)
From Vision To Language through Graph of Events in Space and Time: An Explainable Self-supervised Approach
von: Masala, Mihai, et al.
Veröffentlicht: (2025)
von: Masala, Mihai, et al.
Veröffentlicht: (2025)
On the Geometric Accuracy of Implicit and Primitive-based Representations Derived from View Rendering Constraints
von: De Smijter, Elias, et al.
Veröffentlicht: (2025)
von: De Smijter, Elias, et al.
Veröffentlicht: (2025)
A Vision-Language Foundation Model to Enhance Efficiency of Chest X-ray Interpretation
von: Chen, Zhihong, et al.
Veröffentlicht: (2024)
von: Chen, Zhihong, et al.
Veröffentlicht: (2024)
Enhancing Chest X-ray Classification through Knowledge Injection in Cross-Modality Learning
von: Yan, Yang, et al.
Veröffentlicht: (2025)
von: Yan, Yang, et al.
Veröffentlicht: (2025)
NeRF-based Spacecraft Reconstruction from Monocular Imagery Under Illumination Variability and Pose Uncertainty
von: Legrand, Antoine, et al.
Veröffentlicht: (2026)
von: Legrand, Antoine, et al.
Veröffentlicht: (2026)
CAD-Free Learning of Spacecraft Pose Estimators via NeRF-Based Augmentations
von: Legrand, Antoine, et al.
Veröffentlicht: (2026)
von: Legrand, Antoine, et al.
Veröffentlicht: (2026)
Online Gaussian Test-Time Adaptation of Vision-Language Models
von: Fuchs, Clément, et al.
Veröffentlicht: (2025)
von: Fuchs, Clément, et al.
Veröffentlicht: (2025)
NeRF-based Visualization of 3D Cues Supporting Data-Driven Spacecraft Pose Estimation
von: Legrand, Antoine, et al.
Veröffentlicht: (2025)
von: Legrand, Antoine, et al.
Veröffentlicht: (2025)
Pretraining Vision-Language Model for Difference Visual Question Answering in Longitudinal Chest X-rays
von: Cho, Yeongjae, et al.
Veröffentlicht: (2024)
von: Cho, Yeongjae, et al.
Veröffentlicht: (2024)
Assessing the alignment between infants' visual and linguistic experience using multimodal language models
von: Tan, Alvin Wei Ming, et al.
Veröffentlicht: (2025)
von: Tan, Alvin Wei Ming, et al.
Veröffentlicht: (2025)
VICCA: Visual Interpretation and Comprehension of Chest X-ray Anomalies in Generated Report Without Human Feedback
von: Picha, Sayeh Gholipour, et al.
Veröffentlicht: (2025)
von: Picha, Sayeh Gholipour, et al.
Veröffentlicht: (2025)
Uni-Mlip: Unified Self-supervision for Medical Vision Language Pre-training
von: Bawazir, Ameera, et al.
Veröffentlicht: (2024)
von: Bawazir, Ameera, et al.
Veröffentlicht: (2024)
Vascular anatomy-aware self-supervised pre-training for X-ray angiogram analysis
von: Huang, De-Xing, et al.
Veröffentlicht: (2026)
von: Huang, De-Xing, et al.
Veröffentlicht: (2026)
Semantic Textual Similarity Assessment in Chest X-ray Reports Using a Domain-Specific Cosine-Based Metric
von: Picha, Sayeh Gholipour, et al.
Veröffentlicht: (2024)
von: Picha, Sayeh Gholipour, et al.
Veröffentlicht: (2024)
ChEX: Interactive Localization and Region Description in Chest X-rays
von: Müller, Philip, et al.
Veröffentlicht: (2024)
von: Müller, Philip, et al.
Veröffentlicht: (2024)
Advancing human-centric AI for robust X-ray analysis through holistic self-supervised learning
von: Moutakanni, Théo, et al.
Veröffentlicht: (2024)
von: Moutakanni, Théo, et al.
Veröffentlicht: (2024)
EVA-X: A Foundation Model for General Chest X-ray Analysis with Self-supervised Learning
von: Yao, Jingfeng, et al.
Veröffentlicht: (2024)
von: Yao, Jingfeng, et al.
Veröffentlicht: (2024)
CAMELTrack: Context-Aware Multi-cue ExpLoitation for Online Multi-Object Tracking
von: Somers, Vladimir, et al.
Veröffentlicht: (2025)
von: Somers, Vladimir, et al.
Veröffentlicht: (2025)
Look & Mark: Leveraging Radiologist Eye Fixations and Bounding boxes in Multimodal Large Language Models for Chest X-ray Report Generation
von: Kim, Yunsoo, et al.
Veröffentlicht: (2025)
von: Kim, Yunsoo, et al.
Veröffentlicht: (2025)
Extending global-local view alignment for self-supervised learning with remote sensing imagery
von: Wanyan, Xinye, et al.
Veröffentlicht: (2023)
von: Wanyan, Xinye, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
X-InstructBLIP: A Framework for aligning X-Modal instruction-aware representations to LLMs and Emergent Cross-modal Reasoning
von: Panagopoulou, Artemis, et al.
Veröffentlicht: (2023) -
Self-supervised video pretraining yields robust and more human-aligned visual representations
von: Parthasarathy, Nikhil, et al.
Veröffentlicht: (2022) -
DeepSet SimCLR: Self-supervised deep sets for improved pathology representation learning
von: Torpey, David, et al.
Veröffentlicht: (2024) -
Self-supervised structured object representation learning
von: Hadjerci, Oussama, et al.
Veröffentlicht: (2025) -
Body Part-Based Representation Learning for Occluded Person Re-Identification
von: Somers, Vladimir, et al.
Veröffentlicht: (2022)