CURE: Curriculum-guided Multi-task Training for Reliable Anatomy Grounded Report Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Messina, Pablo, Villa, Andrés, Alcázar, Juan León, Sánchez, Karen, Hinojosa, Carlos, Parra, Denis, Soto, Álvaro, Ghanem, Bernard |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Behind the Magic, MERLIM: Multi-modal Evaluation Benchmark for Large Image-Language Models
by: Villa, Andrés, et al.
Published: (2023)
by: Villa, Andrés, et al.
Published: (2023)
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models
by: Villa, Andrés, et al.
Published: (2025)
by: Villa, Andrés, et al.
Published: (2025)
MoDA: Modulation Adapter for Fine-Grained Visual Grounding in Instructional MLLMs
by: Barrios, Wayner, et al.
Published: (2025)
by: Barrios, Wayner, et al.
Published: (2025)
Extracting and Encoding: Leveraging Large Language Models and Medical Knowledge to Enhance Radiological Text Representation
by: Messina, Pablo, et al.
Published: (2024)
by: Messina, Pablo, et al.
Published: (2024)
SAVeS: Steering Safety Judgments in Vision-Language Models via Semantic Cues
by: Hinojosa, Carlos, et al.
Published: (2026)
by: Hinojosa, Carlos, et al.
Published: (2026)
Multimodal Safety Evaluation in Generative Agent Social Simulations
by: Vera, Alhim, et al.
Published: (2025)
by: Vera, Alhim, et al.
Published: (2025)
ProjGuard: Safety Monitoring for Computer-Use Agents via Low-Dimensional Projections
by: Contreras, Kebin, et al.
Published: (2026)
by: Contreras, Kebin, et al.
Published: (2026)
MoCA-Video: Motion-Aware Concept Alignment for Consistent Video Editing
by: Zhang, Tong, et al.
Published: (2025)
by: Zhang, Tong, et al.
Published: (2025)
CO2Wounds-V2: Extended Chronic Wounds Dataset From Leprosy Patients
by: Sanchez, Karen, et al.
Published: (2024)
by: Sanchez, Karen, et al.
Published: (2024)
ColorMAE: Exploring data-independent masking strategies in Masked AutoEncoders
by: Hinojosa, Carlos, et al.
Published: (2024)
by: Hinojosa, Carlos, et al.
Published: (2024)
Skill-Aligned Annotation for Reliable Evaluation in Text-to-Image Generation
by: Eldesokey, Abdelrahman, et al.
Published: (2026)
by: Eldesokey, Abdelrahman, et al.
Published: (2026)
CURE-Med: Curriculum-Informed Reinforcement Learning for Multilingual Medical Reasoning
by: Onyame, Eric, et al.
Published: (2026)
by: Onyame, Eric, et al.
Published: (2026)
Learning Semantic Segmentation with Query Points Supervision on Aerial Images
by: Rivier, Santiago, et al.
Published: (2023)
by: Rivier, Santiago, et al.
Published: (2023)
MambaStyle: Efficient StyleGAN Inversion for Real Image Editing with State-Space Models
by: Lopez, Jhon, et al.
Published: (2025)
by: Lopez, Jhon, et al.
Published: (2025)
CURE: Centroid-guided Unsupervised Representation Erasure for Facial Recognition Systems
by: Shivam, Fnu, et al.
Published: (2025)
by: Shivam, Fnu, et al.
Published: (2025)
Train Long, Think Short: Curriculum Learning for Efficient Reasoning
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
Compressed-Language Models for Understanding Compressed File Formats: a JPEG Exploration
by: Pérez, Juan C., et al.
Published: (2024)
by: Pérez, Juan C., et al.
Published: (2024)
Privacy-preserving Optics for Enhancing Protection in Face De-identification
by: Lopez, Jhon, et al.
Published: (2024)
by: Lopez, Jhon, et al.
Published: (2024)
Forget Less, Retain More: A Lightweight Regularizer for Rehearsal-Based Continual Learning
by: Alssum, Lama, et al.
Published: (2025)
by: Alssum, Lama, et al.
Published: (2025)
Anatomy-guided Pathology Segmentation
by: Jaus, Alexander, et al.
Published: (2024)
by: Jaus, Alexander, et al.
Published: (2024)
BOLT: Boost Large Vision-Language Model Without Training for Long-form Video Understanding
by: Liu, Shuming, et al.
Published: (2025)
by: Liu, Shuming, et al.
Published: (2025)
OpenTAD: A Unified Framework and Comprehensive Study of Temporal Action Detection
by: Liu, Shuming, et al.
Published: (2025)
by: Liu, Shuming, et al.
Published: (2025)
UnMix-NeRF: Spectral Unmixing Meets Neural Radiance Fields
by: Perez, Fabian, et al.
Published: (2025)
by: Perez, Fabian, et al.
Published: (2025)
Hala Technical Report: Building Arabic-Centric Instruction & Translation Models at Scale
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
by: Hammoud, Hasan Abed Al Kader, et al.
Published: (2025)
SoccerLens: Grounded Soccer Video Understanding Beyond Accuracy
by: Elsharkawi, Ismael, et al.
Published: (2026)
by: Elsharkawi, Ismael, et al.
Published: (2026)
CURE: Controlled Unlearning for Robust Embeddings -- Mitigating Conceptual Shortcuts in Pre-Trained Language Models
by: Kocak, Aysenur, et al.
Published: (2025)
by: Kocak, Aysenur, et al.
Published: (2025)
MeXtract: Light-Weight Metadata Extraction from Scientific Papers
by: Alyafeai, Zaid, et al.
Published: (2025)
by: Alyafeai, Zaid, et al.
Published: (2025)
MOLE: Metadata Extraction and Validation in Scientific Papers Using LLMs
by: Alyafeai, Zaid, et al.
Published: (2025)
by: Alyafeai, Zaid, et al.
Published: (2025)
Video Self-Stitching Graph Network for Temporal Action Localization
by: Zhao, Chen, et al.
Published: (2020)
by: Zhao, Chen, et al.
Published: (2020)
CURE-OOD: Benchmarking Out-of-Distribution Detection for Survival Prediction
by: Zhao, Wenjie, et al.
Published: (2026)
by: Zhao, Wenjie, et al.
Published: (2026)
CURE: A Multimodal Benchmark for Clinical Understanding and Retrieval Evaluation
by: Gu, Yannian, et al.
Published: (2026)
by: Gu, Yannian, et al.
Published: (2026)
CURE: Cultural Understanding and Reasoning Evaluation - A Framework for "Thick" Culture Alignment Evaluation in LLMs
by: Vo, Truong, et al.
Published: (2025)
by: Vo, Truong, et al.
Published: (2025)
IP-CRR: Information Pursuit for Interpretable Classification of Chest Radiology Reports
by: Ge, Yuyan, et al.
Published: (2025)
by: Ge, Yuyan, et al.
Published: (2025)
CURE: Concept Unlearning via Orthogonal Representation Editing in Diffusion Models
by: Biswas, Shristi Das, et al.
Published: (2025)
by: Biswas, Shristi Das, et al.
Published: (2025)
CURE:Circuit-Aware Unlearning for LLM-based Recommendation
by: Chen, Ziheng, et al.
Published: (2026)
by: Chen, Ziheng, et al.
Published: (2026)
SPARF: Large-Scale Learning of 3D Sparse Radiance Fields from Few Input Images
by: Hamdi, Abdullah, et al.
Published: (2022)
by: Hamdi, Abdullah, et al.
Published: (2022)
Reasoning Vectors: Transferring Chain-of-Thought Capabilities via Task Arithmetic
by: Zbeeb, Mohammad, et al.
Published: (2025)
by: Zbeeb, Mohammad, et al.
Published: (2025)
Mind-the-Glitch: Visual Correspondence for Detecting Inconsistencies in Subject-Driven Generation
by: Eldesokey, Abdelrahman, et al.
Published: (2025)
by: Eldesokey, Abdelrahman, et al.
Published: (2025)
Exploring Missing Modality in Multimodal Egocentric Datasets
by: Ramazanova, Merey, et al.
Published: (2024)
by: Ramazanova, Merey, et al.
Published: (2024)
Pix4Point: Image Pretrained Standard Transformers for 3D Point Cloud Understanding
by: Qian, Guocheng, et al.
Published: (2022)
by: Qian, Guocheng, et al.
Published: (2022)
Similar Items
-
Behind the Magic, MERLIM: Multi-modal Evaluation Benchmark for Large Image-Language Models
by: Villa, Andrés, et al.
Published: (2023) -
EAGLE: Enhanced Visual Grounding Minimizes Hallucinations in Instructional Multimodal Models
by: Villa, Andrés, et al.
Published: (2025) -
MoDA: Modulation Adapter for Fine-Grained Visual Grounding in Instructional MLLMs
by: Barrios, Wayner, et al.
Published: (2025) -
Extracting and Encoding: Leveraging Large Language Models and Medical Knowledge to Enhance Radiological Text Representation
by: Messina, Pablo, et al.
Published: (2024) -
SAVeS: Steering Safety Judgments in Vision-Language Models via Semantic Cues
by: Hinojosa, Carlos, et al.
Published: (2026)