LoC-Path: Learning to Compress for Pathology Multimodal Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Qingqiao, Lyu, Weimin, Xu, Meilong, Qi, Kehan, Hu, Xiaoling, Gupta, Saumya, Zhou, Jiawei, Chen, Chao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unrolled Networks are Conditional Probability Flows in MRI Reconstruction
by: Qi, Kehan, et al.
Published: (2025)
by: Qi, Kehan, et al.
Published: (2025)
Efficient Whole Slide Pathology VQA via Token Compression
by: Lyu, Weimin, et al.
Published: (2025)
by: Lyu, Weimin, et al.
Published: (2025)
Topo-R1: Detecting Topological Anomalies via Vision-Language Models
by: Xu, Meilong, et al.
Published: (2026)
by: Xu, Meilong, et al.
Published: (2026)
Semi-supervised Segmentation of Histopathology Images with Noise-Aware Topological Consistency
by: Xu, Meilong, et al.
Published: (2023)
by: Xu, Meilong, et al.
Published: (2023)
Bézier Meets Diffusion: Robust Generation Across Domains for Medical Image Segmentation
by: Li, Chen, et al.
Published: (2025)
by: Li, Chen, et al.
Published: (2025)
TopoCellGen: Generating Histopathology Cell Topology with a Diffusion Model
by: Xu, Meilong, et al.
Published: (2024)
by: Xu, Meilong, et al.
Published: (2024)
Act Like a Pathologist: Tissue-Aware Whole Slide Image Reasoning
by: Huang, Wentao, et al.
Published: (2026)
by: Huang, Wentao, et al.
Published: (2026)
MATCH: Multi-faceted Adaptive Topo-Consistency for Semi-Supervised Histopathology Segmentation
by: Xu, Meilong, et al.
Published: (2025)
by: Xu, Meilong, et al.
Published: (2025)
Spatial Diffusion for Cell Layout Generation
by: Li, Chen, et al.
Published: (2024)
by: Li, Chen, et al.
Published: (2024)
LoC-LIC: Low Complexity Learned Image Coding Using Hierarchical Feature Transforms
by: Ameen, Ayman A., et al.
Published: (2025)
by: Ameen, Ayman A., et al.
Published: (2025)
Backdooring Vision-Language Models with Out-Of-Distribution Data
by: Lyu, Weimin, et al.
Published: (2024)
by: Lyu, Weimin, et al.
Published: (2024)
ReinPath: A Multimodal Reinforcement Learning Approach for Pathology
by: Zhou, Kangcheng, et al.
Published: (2026)
by: Zhou, Kangcheng, et al.
Published: (2026)
Multimodal Model for Computational Pathology:Representation Learning and Image Compression
by: Wu, Peihang, et al.
Published: (2026)
by: Wu, Peihang, et al.
Published: (2026)
Walking Your Frog Fast in 4 LoC
by: Meinert, Nis
Published: (2024)
by: Meinert, Nis
Published: (2024)
PathMR: Multimodal Visual Reasoning for Interpretable Pathology Diagnosis
by: Zhang, Ye, et al.
Published: (2025)
by: Zhang, Ye, et al.
Published: (2025)
HiPath: Hierarchical Vision-Language Alignment for Structured Pathology Report Prediction
by: Yuan, Ruicheng, et al.
Published: (2026)
by: Yuan, Ruicheng, et al.
Published: (2026)
Learning Topological Representations for Deep Image Understanding
by: Hu, Xiaoling
Published: (2024)
by: Hu, Xiaoling
Published: (2024)
TopoDiffusionNet: A Topology-aware Diffusion Model
by: Gupta, Saumya, et al.
Published: (2024)
by: Gupta, Saumya, et al.
Published: (2024)
RB-FT: Rationale-Bootstrapped Fine-Tuning for Video Classification
by: Xu, Meilong, et al.
Published: (2025)
by: Xu, Meilong, et al.
Published: (2025)
Non-Termination Proving: 100 Million LoC and Beyond
by: Vanegue, Julien, et al.
Published: (2025)
by: Vanegue, Julien, et al.
Published: (2025)
Semi-Supervised Contrastive VAE for Disentanglement of Digital Pathology Images
by: Hasan, Mahmudul, et al.
Published: (2024)
by: Hasan, Mahmudul, et al.
Published: (2024)
PPE: Positional Preservation Embedding for Token Compression in Multimodal Large Language Models
by: Huang, Mouxiao, et al.
Published: (2025)
by: Huang, Mouxiao, et al.
Published: (2025)
PathFound: An Agentic Multimodal Model Activating Evidence-seeking Pathological Diagnosis
by: Hua, Shengyi, et al.
Published: (2025)
by: Hua, Shengyi, et al.
Published: (2025)
EvoCut: Multi-Layer Evolution-Aware Visual Token Compression for Efficient Large Vision-Language Models
by: Lu, Hongyu, et al.
Published: (2026)
by: Lu, Hongyu, et al.
Published: (2026)
PathMMU: A Massive Multimodal Expert-Level Benchmark for Understanding and Reasoning in Pathology
by: Sun, Yuxuan, et al.
Published: (2024)
by: Sun, Yuxuan, et al.
Published: (2024)
Visual Attention Drifts,but Anchors Hold:Mitigating Hallucination in Multimodal Large Language Models via Cross-Layer Visual Anchors
by: Yang, Chengxu, et al.
Published: (2026)
by: Yang, Chengxu, et al.
Published: (2026)
Are Multimodal Large Language Models Ready for Omnidirectional Spatial Reasoning?
by: Dongfang, Zihao, et al.
Published: (2025)
by: Dongfang, Zihao, et al.
Published: (2025)
PathFL: Multi-Alignment Federated Learning for Pathology Image Segmentation
by: Zhang, Yuan, et al.
Published: (2025)
by: Zhang, Yuan, et al.
Published: (2025)
RankByGene: Gene-Guided Histopathology Representation Learning Through Cross-Modal Ranking Consistency
by: Huang, Wentao, et al.
Published: (2024)
by: Huang, Wentao, et al.
Published: (2024)
ORCE: Order-Aware Alignment of Verbalized Confidence in Large Language Models
by: Li, Chen, et al.
Published: (2026)
by: Li, Chen, et al.
Published: (2026)
TrojVLM: Backdoor Attack Against Vision Language Models
by: Lyu, Weimin, et al.
Published: (2024)
by: Lyu, Weimin, et al.
Published: (2024)
Navigating Gigapixel Pathology Images with Large Multimodal Models
by: Buckley, Thomas A., et al.
Published: (2025)
by: Buckley, Thomas A., et al.
Published: (2025)
Multi-Modal Proxy Learning Towards Personalized Visual Multiple Clustering
by: Yao, Jiawei, et al.
Published: (2024)
by: Yao, Jiawei, et al.
Published: (2024)
PathAR: Structure-First Autoregressive Synthesis of Multimodal Pathology Images
by: Zhang, Yuan, et al.
Published: (2026)
by: Zhang, Yuan, et al.
Published: (2026)
GAMBIT: A Gamified Jailbreak Framework for Multimodal Large Language Models
by: Hu, Xiangdong, et al.
Published: (2026)
by: Hu, Xiangdong, et al.
Published: (2026)
Content-Aware Mamba for Learned Image Compression
by: Chen, Yunuo, et al.
Published: (2025)
by: Chen, Yunuo, et al.
Published: (2025)
PolyPath: Adapting a Large Multimodal Model for Multi-slide Pathology Report Generation
by: Ahmed, Faruk, et al.
Published: (2025)
by: Ahmed, Faruk, et al.
Published: (2025)
Learning Brain Tumor Representation in 3D High-Resolution MR Images via Interpretable State Space Models
by: Hu, Qingqiao, et al.
Published: (2024)
by: Hu, Qingqiao, et al.
Published: (2024)
Vision-Language Modeling with Regularized Spatial Transformer Networks for All Weather Crosswind Landing of Aircraft
by: Pal, Debabrata, et al.
Published: (2024)
by: Pal, Debabrata, et al.
Published: (2024)
LightVLM: Acceleraing Large Multimodal Models with Pyramid Token Merging and KV Cache Compression
by: Hu, Lianyu, et al.
Published: (2025)
by: Hu, Lianyu, et al.
Published: (2025)
Similar Items
-
Unrolled Networks are Conditional Probability Flows in MRI Reconstruction
by: Qi, Kehan, et al.
Published: (2025) -
Efficient Whole Slide Pathology VQA via Token Compression
by: Lyu, Weimin, et al.
Published: (2025) -
Topo-R1: Detecting Topological Anomalies via Vision-Language Models
by: Xu, Meilong, et al.
Published: (2026) -
Semi-supervised Segmentation of Histopathology Images with Noise-Aware Topological Consistency
by: Xu, Meilong, et al.
Published: (2023) -
Bézier Meets Diffusion: Robust Generation Across Domains for Medical Image Segmentation
by: Li, Chen, et al.
Published: (2025)