EVA-X: A Foundation Model for General Chest X-ray Analysis with Self-supervised Learning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yao, Jingfeng, Wang, Xinggang, Song, Yuehao, Zhao, Huangxuan, Ma, Jun, Chen, Yajie, Liu, Wenyu, Wang, Bo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
UniX: Unifying Autoregression and Diffusion for Chest X-Ray Understanding and Generation
von: Zhang, Ruiheng, et al.
Veröffentlicht: (2026)
von: Zhang, Ruiheng, et al.
Veröffentlicht: (2026)
ViTGaze: Gaze Following with Interaction Features in Vision Transformers
von: Song, Yuehao, et al.
Veröffentlicht: (2024)
von: Song, Yuehao, et al.
Veröffentlicht: (2024)
GaraMoSt: Parallel Multi-Granularity Motion and Structural Modeling for Efficient Multi-Frame Interpolation in DSA Images
von: Xu, Ziyang, et al.
Veröffentlicht: (2024)
von: Xu, Ziyang, et al.
Veröffentlicht: (2024)
DeltaMIL: Gated Memory Integration for Efficient and Discriminative Whole Slide Image Analysis
von: Zhu, Yueting, et al.
Veröffentlicht: (2025)
von: Zhu, Yueting, et al.
Veröffentlicht: (2025)
FasterDiT: Towards Faster Diffusion Transformers Training without Architecture Modification
von: Yao, Jingfeng, et al.
Veröffentlicht: (2024)
von: Yao, Jingfeng, et al.
Veröffentlicht: (2024)
PersonViT: Large-scale Self-supervised Vision Transformer for Person Re-Identification
von: Hu, Bin, et al.
Veröffentlicht: (2024)
von: Hu, Bin, et al.
Veröffentlicht: (2024)
Matte Anything: Interactive Natural Image Matting with Segment Anything Models
von: Yao, Jingfeng, et al.
Veröffentlicht: (2023)
von: Yao, Jingfeng, et al.
Veröffentlicht: (2023)
RAD-2: Scaling Reinforcement Learning in a Generator-Discriminator Framework
von: Gao, Hao, et al.
Veröffentlicht: (2026)
von: Gao, Hao, et al.
Veröffentlicht: (2026)
Towards Scalable Pre-training of Visual Tokenizers for Generation
von: Yao, Jingfeng, et al.
Veröffentlicht: (2025)
von: Yao, Jingfeng, et al.
Veröffentlicht: (2025)
MoSt-DSA: Modeling Motion and Structural Interactions for Direct Multi-Frame Interpolation in DSA Images
von: Xu, Ziyang, et al.
Veröffentlicht: (2024)
von: Xu, Ziyang, et al.
Veröffentlicht: (2024)
Reconstruction vs. Generation: Taming Optimization Dilemma in Latent Diffusion Models
von: Yao, Jingfeng, et al.
Veröffentlicht: (2025)
von: Yao, Jingfeng, et al.
Veröffentlicht: (2025)
OTCXR: Rethinking Self-supervised Alignment using Optimal Transport for Chest X-ray Analysis
von: Gorade, Vandan, et al.
Veröffentlicht: (2024)
von: Gorade, Vandan, et al.
Veröffentlicht: (2024)
Foundation X: Integrating Classification, Localization, and Segmentation through Lock-Release Pretraining Strategy for Chest X-ray Analysis
von: Islam, Nahid Ul, et al.
Veröffentlicht: (2025)
von: Islam, Nahid Ul, et al.
Veröffentlicht: (2025)
TOGS: Gaussian Splatting with Temporal Opacity Offset for Real-Time 4D DSA Rendering
von: Zhang, Shuai, et al.
Veröffentlicht: (2024)
von: Zhang, Shuai, et al.
Veröffentlicht: (2024)
Turbo-VAED: Fast and Stable Transfer of Video-VAEs to Mobile Devices
von: Zou, Ya, et al.
Veröffentlicht: (2025)
von: Zou, Ya, et al.
Veröffentlicht: (2025)
DiffusionVL: Translating Any Autoregressive Models into Diffusion Vision Language Models
von: Zeng, Lunbin, et al.
Veröffentlicht: (2025)
von: Zeng, Lunbin, et al.
Veröffentlicht: (2025)
CheXPO: Preference Optimization for Chest X-ray VLMs with Counterfactual Rationale
von: Liang, Xiao, et al.
Veröffentlicht: (2025)
von: Liang, Xiao, et al.
Veröffentlicht: (2025)
Chest X-ray Foundation Model with Global and Local Representations Integration
von: Yang, Zefan, et al.
Veröffentlicht: (2025)
von: Yang, Zefan, et al.
Veröffentlicht: (2025)
Visual Generation Tuning
von: Guo, Jiahao, et al.
Veröffentlicht: (2025)
von: Guo, Jiahao, et al.
Veröffentlicht: (2025)
LKCell: Efficient Cell Nuclei Instance Segmentation with Large Convolution Kernels
von: Cui, Ziwei, et al.
Veröffentlicht: (2024)
von: Cui, Ziwei, et al.
Veröffentlicht: (2024)
Self-Supervised Learning for Building Robust Pediatric Chest X-ray Classification Models
von: Cheng, Sheng, et al.
Veröffentlicht: (2024)
von: Cheng, Sheng, et al.
Veröffentlicht: (2024)
DiffusionDriveV2: Reinforcement Learning-Constrained Truncated Diffusion Modeling in End-to-End Autonomous Driving
von: Zou, Jialv, et al.
Veröffentlicht: (2025)
von: Zou, Jialv, et al.
Veröffentlicht: (2025)
Thought Graph Traversal for Test-time Scaling in Chest X-ray VLLMs
von: Yao, Yue, et al.
Veröffentlicht: (2025)
von: Yao, Yue, et al.
Veröffentlicht: (2025)
Enhanced Contrastive Learning with Multi-view Longitudinal Data for Chest X-ray Report Generation
von: Liu, Kang, et al.
Veröffentlicht: (2025)
von: Liu, Kang, et al.
Veröffentlicht: (2025)
MolSight: Optical Chemical Structure Recognition with SMILES Pretraining, Multi-Granularity Learning and Reinforcement Learning
von: Zhang, Wenrui, et al.
Veröffentlicht: (2025)
von: Zhang, Wenrui, et al.
Veröffentlicht: (2025)
CGF-DETR: Cross-Gated Fusion DETR for Enhanced Pneumonia Detection in Chest X-rays
von: Wu, Yefeng, et al.
Veröffentlicht: (2025)
von: Wu, Yefeng, et al.
Veröffentlicht: (2025)
AlphaDrive: Unleashing the Power of VLMs in Autonomous Driving via Reinforcement Learning and Reasoning
von: Jiang, Bo, et al.
Veröffentlicht: (2025)
von: Jiang, Bo, et al.
Veröffentlicht: (2025)
GaussTR: Foundation Model-Aligned Gaussian Transformer for Self-Supervised 3D Spatial Understanding
von: Jiang, Haoyi, et al.
Veröffentlicht: (2024)
von: Jiang, Haoyi, et al.
Veröffentlicht: (2024)
Causality-inspired Discriminative Feature Learning in Triple Domains for Gait Recognition
von: Xiong, Haijun, et al.
Veröffentlicht: (2024)
von: Xiong, Haijun, et al.
Veröffentlicht: (2024)
EVA-02: A Visual Representation for Neon Genesis
von: Fang, Yuxin, et al.
Veröffentlicht: (2023)
von: Fang, Yuxin, et al.
Veröffentlicht: (2023)
A Vision-Language Foundation Model to Enhance Efficiency of Chest X-ray Interpretation
von: Chen, Zhihong, et al.
Veröffentlicht: (2024)
von: Chen, Zhihong, et al.
Veröffentlicht: (2024)
Phrase-grounded Fact-checking for Automatically Generated Chest X-ray Reports
von: Mahmood, Razi, et al.
Veröffentlicht: (2025)
von: Mahmood, Razi, et al.
Veröffentlicht: (2025)
A Reasoning-Enabled Vision-Language Foundation Model for Chest X-ray Interpretation
von: Zhang, Yabin, et al.
Veröffentlicht: (2026)
von: Zhang, Yabin, et al.
Veröffentlicht: (2026)
WeakSAM: Segment Anything Meets Weakly-supervised Instance-level Recognition
von: Zhu, Lianghui, et al.
Veröffentlicht: (2024)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2024)
WeakTr: Exploring Plain Vision Transformer for Weakly-supervised Semantic Segmentation
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
Gait Recognition via Collaborating Discriminative and Generative Diffusion Models
von: Xiong, Haijun, et al.
Veröffentlicht: (2025)
von: Xiong, Haijun, et al.
Veröffentlicht: (2025)
DriveLaW:Unifying Planning and Video Generation in a Latent Driving World
von: Xia, Tianze, et al.
Veröffentlicht: (2025)
von: Xia, Tianze, et al.
Veröffentlicht: (2025)
Efficient Chest X-ray Representation Learning via Semantic-Partitioned Contrastive Learning
von: Feng, Wangyu, et al.
Veröffentlicht: (2026)
von: Feng, Wangyu, et al.
Veröffentlicht: (2026)
Addressing Asynchronicity in Clinical Multimodal Fusion via Individualized Chest X-ray Generation
von: Yao, Wenfang, et al.
Veröffentlicht: (2024)
von: Yao, Wenfang, et al.
Veröffentlicht: (2024)
MMRad-22K: A Structured Multimodal Evidence Dataset for Chest X-ray Report Generation
von: Zhao, Yichen, et al.
Veröffentlicht: (2026)
von: Zhao, Yichen, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
UniX: Unifying Autoregression and Diffusion for Chest X-Ray Understanding and Generation
von: Zhang, Ruiheng, et al.
Veröffentlicht: (2026) -
ViTGaze: Gaze Following with Interaction Features in Vision Transformers
von: Song, Yuehao, et al.
Veröffentlicht: (2024) -
GaraMoSt: Parallel Multi-Granularity Motion and Structural Modeling for Efficient Multi-Frame Interpolation in DSA Images
von: Xu, Ziyang, et al.
Veröffentlicht: (2024) -
DeltaMIL: Gated Memory Integration for Efficient and Discriminative Whole Slide Image Analysis
von: Zhu, Yueting, et al.
Veröffentlicht: (2025) -
FasterDiT: Towards Faster Diffusion Transformers Training without Architecture Modification
von: Yao, Jingfeng, et al.
Veröffentlicht: (2024)