Self-supervised vision-langage alignment of deep learning representations for bone X-rays analysis
Fuente:
arXiv
Guardado en:
| Autores principales: | Englebert, Alexandre, Collin, Anne-Sophie, Cornu, Olivier, De Vleeschouwer, Christophe |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
X-InstructBLIP: A Framework for aligning X-Modal instruction-aware representations to LLMs and Emergent Cross-modal Reasoning
por: Panagopoulou, Artemis, et al.
Publicado: (2023)
por: Panagopoulou, Artemis, et al.
Publicado: (2023)
Self-supervised video pretraining yields robust and more human-aligned visual representations
por: Parthasarathy, Nikhil, et al.
Publicado: (2022)
por: Parthasarathy, Nikhil, et al.
Publicado: (2022)
DeepSet SimCLR: Self-supervised deep sets for improved pathology representation learning
por: Torpey, David, et al.
Publicado: (2024)
por: Torpey, David, et al.
Publicado: (2024)
Self-supervised structured object representation learning
por: Hadjerci, Oussama, et al.
Publicado: (2025)
por: Hadjerci, Oussama, et al.
Publicado: (2025)
Body Part-Based Representation Learning for Occluded Person Re-Identification
por: Somers, Vladimir, et al.
Publicado: (2022)
por: Somers, Vladimir, et al.
Publicado: (2022)
Keypoint Promptable Re-Identification
por: Somers, Vladimir, et al.
Publicado: (2024)
por: Somers, Vladimir, et al.
Publicado: (2024)
Interpreting the structure of multi-object representations in vision encoders
por: Khajuria, Tarun, et al.
Publicado: (2024)
por: Khajuria, Tarun, et al.
Publicado: (2024)
Sustainable self-supervised learning for speech representations
por: Lugo, Luis, et al.
Publicado: (2024)
por: Lugo, Luis, et al.
Publicado: (2024)
What to align in multimodal contrastive learning?
por: Dufumier, Benoit, et al.
Publicado: (2024)
por: Dufumier, Benoit, et al.
Publicado: (2024)
RUMPL: Ray-Based Transformers for Universal Multi-View 2D to 3D Human Pose Lifting
por: Ghasemzadeh, Seyed Abolfazl, et al.
Publicado: (2025)
por: Ghasemzadeh, Seyed Abolfazl, et al.
Publicado: (2025)
MPL: Lifting 3D Human Pose from Multi-view 2D Poses
por: Ghasemzadeh, Seyed Abolfazl, et al.
Publicado: (2024)
por: Ghasemzadeh, Seyed Abolfazl, et al.
Publicado: (2024)
CheXalign: Preference fine-tuning in chest X-ray interpretation models without human feedback
por: Hein, Dennis, et al.
Publicado: (2024)
por: Hein, Dennis, et al.
Publicado: (2024)
Explore and Explain: Self-supervised Navigation and Recounting
por: Bigazzi, Roberto, et al.
Publicado: (2020)
por: Bigazzi, Roberto, et al.
Publicado: (2020)
Towards aligned body representations in vision models
por: Gizdov, Andrey, et al.
Publicado: (2025)
por: Gizdov, Andrey, et al.
Publicado: (2025)
Vocabulary-free few-shot learning for Vision-Language Models
por: Zanella, Maxime, et al.
Publicado: (2025)
por: Zanella, Maxime, et al.
Publicado: (2025)
Shared and Private Information Learning in Multimodal Sentiment Analysis with Deep Modal Alignment and Self-supervised Multi-Task Learning
por: Lai, Songning, et al.
Publicado: (2023)
por: Lai, Songning, et al.
Publicado: (2023)
Self-supervised learning of speech representations with Dutch archival data
por: Vaessen, Nik, et al.
Publicado: (2025)
por: Vaessen, Nik, et al.
Publicado: (2025)
Multi-object event graph representation learning for Video Question Answering
por: Wang, Yanan, et al.
Publicado: (2024)
por: Wang, Yanan, et al.
Publicado: (2024)
Altogether: Image Captioning via Re-aligning Alt-text
por: Xu, Hu, et al.
Publicado: (2024)
por: Xu, Hu, et al.
Publicado: (2024)
Self-supervised learning of video representations from a child's perspective
por: Orhan, A. Emin, et al.
Publicado: (2024)
por: Orhan, A. Emin, et al.
Publicado: (2024)
From Vision To Language through Graph of Events in Space and Time: An Explainable Self-supervised Approach
por: Masala, Mihai, et al.
Publicado: (2025)
por: Masala, Mihai, et al.
Publicado: (2025)
On the Geometric Accuracy of Implicit and Primitive-based Representations Derived from View Rendering Constraints
por: De Smijter, Elias, et al.
Publicado: (2025)
por: De Smijter, Elias, et al.
Publicado: (2025)
A Vision-Language Foundation Model to Enhance Efficiency of Chest X-ray Interpretation
por: Chen, Zhihong, et al.
Publicado: (2024)
por: Chen, Zhihong, et al.
Publicado: (2024)
Enhancing Chest X-ray Classification through Knowledge Injection in Cross-Modality Learning
por: Yan, Yang, et al.
Publicado: (2025)
por: Yan, Yang, et al.
Publicado: (2025)
NeRF-based Spacecraft Reconstruction from Monocular Imagery Under Illumination Variability and Pose Uncertainty
por: Legrand, Antoine, et al.
Publicado: (2026)
por: Legrand, Antoine, et al.
Publicado: (2026)
CAD-Free Learning of Spacecraft Pose Estimators via NeRF-Based Augmentations
por: Legrand, Antoine, et al.
Publicado: (2026)
por: Legrand, Antoine, et al.
Publicado: (2026)
Online Gaussian Test-Time Adaptation of Vision-Language Models
por: Fuchs, Clément, et al.
Publicado: (2025)
por: Fuchs, Clément, et al.
Publicado: (2025)
NeRF-based Visualization of 3D Cues Supporting Data-Driven Spacecraft Pose Estimation
por: Legrand, Antoine, et al.
Publicado: (2025)
por: Legrand, Antoine, et al.
Publicado: (2025)
Pretraining Vision-Language Model for Difference Visual Question Answering in Longitudinal Chest X-rays
por: Cho, Yeongjae, et al.
Publicado: (2024)
por: Cho, Yeongjae, et al.
Publicado: (2024)
Assessing the alignment between infants' visual and linguistic experience using multimodal language models
por: Tan, Alvin Wei Ming, et al.
Publicado: (2025)
por: Tan, Alvin Wei Ming, et al.
Publicado: (2025)
VICCA: Visual Interpretation and Comprehension of Chest X-ray Anomalies in Generated Report Without Human Feedback
por: Picha, Sayeh Gholipour, et al.
Publicado: (2025)
por: Picha, Sayeh Gholipour, et al.
Publicado: (2025)
Uni-Mlip: Unified Self-supervision for Medical Vision Language Pre-training
por: Bawazir, Ameera, et al.
Publicado: (2024)
por: Bawazir, Ameera, et al.
Publicado: (2024)
Vascular anatomy-aware self-supervised pre-training for X-ray angiogram analysis
por: Huang, De-Xing, et al.
Publicado: (2026)
por: Huang, De-Xing, et al.
Publicado: (2026)
Semantic Textual Similarity Assessment in Chest X-ray Reports Using a Domain-Specific Cosine-Based Metric
por: Picha, Sayeh Gholipour, et al.
Publicado: (2024)
por: Picha, Sayeh Gholipour, et al.
Publicado: (2024)
ChEX: Interactive Localization and Region Description in Chest X-rays
por: Müller, Philip, et al.
Publicado: (2024)
por: Müller, Philip, et al.
Publicado: (2024)
Advancing human-centric AI for robust X-ray analysis through holistic self-supervised learning
por: Moutakanni, Théo, et al.
Publicado: (2024)
por: Moutakanni, Théo, et al.
Publicado: (2024)
EVA-X: A Foundation Model for General Chest X-ray Analysis with Self-supervised Learning
por: Yao, Jingfeng, et al.
Publicado: (2024)
por: Yao, Jingfeng, et al.
Publicado: (2024)
CAMELTrack: Context-Aware Multi-cue ExpLoitation for Online Multi-Object Tracking
por: Somers, Vladimir, et al.
Publicado: (2025)
por: Somers, Vladimir, et al.
Publicado: (2025)
Look & Mark: Leveraging Radiologist Eye Fixations and Bounding boxes in Multimodal Large Language Models for Chest X-ray Report Generation
por: Kim, Yunsoo, et al.
Publicado: (2025)
por: Kim, Yunsoo, et al.
Publicado: (2025)
Extending global-local view alignment for self-supervised learning with remote sensing imagery
por: Wanyan, Xinye, et al.
Publicado: (2023)
por: Wanyan, Xinye, et al.
Publicado: (2023)
Ejemplares similares
-
X-InstructBLIP: A Framework for aligning X-Modal instruction-aware representations to LLMs and Emergent Cross-modal Reasoning
por: Panagopoulou, Artemis, et al.
Publicado: (2023) -
Self-supervised video pretraining yields robust and more human-aligned visual representations
por: Parthasarathy, Nikhil, et al.
Publicado: (2022) -
DeepSet SimCLR: Self-supervised deep sets for improved pathology representation learning
por: Torpey, David, et al.
Publicado: (2024) -
Self-supervised structured object representation learning
por: Hadjerci, Oussama, et al.
Publicado: (2025) -
Body Part-Based Representation Learning for Occluded Person Re-Identification
por: Somers, Vladimir, et al.
Publicado: (2022)