Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | LASA Team, Xu, Weiwen, Chan, Hou Pong, Li, Long, Aljunied, Mahani, Yuan, Ruifeng, Wang, Jianyu, Xiao, Chenghao, Chen, Guizhen, Liu, Chaoqun, Li, Zhaodonghui, Sun, Yu, Shen, Junao, Wang, Chaojun, Tan, Jie, Zhao, Deli, Xu, Tingyang, Zhang, Hao, Rong, Yu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SeaLLMs-Audio: Large Audio-Language Models for Southeast Asia
von: Liu, Chaoqun, et al.
Veröffentlicht: (2025)
von: Liu, Chaoqun, et al.
Veröffentlicht: (2025)
Scaling Language-Centric Omnimodal Representation Learning
von: Xiao, Chenghao, et al.
Veröffentlicht: (2025)
von: Xiao, Chenghao, et al.
Veröffentlicht: (2025)
Babel: Open Multilingual Large Language Models Serving Over 90% of Global Speakers
von: Zhao, Yiran, et al.
Veröffentlicht: (2025)
von: Zhao, Yiran, et al.
Veröffentlicht: (2025)
VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning
von: Yuan, Ruifeng, et al.
Veröffentlicht: (2025)
von: Yuan, Ruifeng, et al.
Veröffentlicht: (2025)
SeaLLMs 3: Open Foundation and Chat Multilingual Large Language Models for Southeast Asian Languages
von: Zhang, Wenxuan, et al.
Veröffentlicht: (2024)
von: Zhang, Wenxuan, et al.
Veröffentlicht: (2024)
Analyzing LLMs' Knowledge Boundary Cognition Across Languages Through the Lens of Internal Representations
von: Xiao, Chenghao, et al.
Veröffentlicht: (2025)
von: Xiao, Chenghao, et al.
Veröffentlicht: (2025)
FINEREASON: Evaluating and Improving LLMs' Deliberate Reasoning through Reflective Puzzle Solving
von: Chen, Guizhen, et al.
Veröffentlicht: (2025)
von: Chen, Guizhen, et al.
Veröffentlicht: (2025)
GeoPQA: Bridging the Visual Perception Gap in MLLMs for Geometric Reasoning
von: Chen, Guizhen, et al.
Veröffentlicht: (2025)
von: Chen, Guizhen, et al.
Veröffentlicht: (2025)
M-Longdoc: A Benchmark For Multimodal Super-Long Document Understanding And A Retrieval-Aware Tuning Framework
von: Chia, Yew Ken, et al.
Veröffentlicht: (2024)
von: Chia, Yew Ken, et al.
Veröffentlicht: (2024)
ReasonMed: A 370K Multi-Agent Generated Dataset for Advancing Medical Reasoning
von: Sun, Yu, et al.
Veröffentlicht: (2025)
von: Sun, Yu, et al.
Veröffentlicht: (2025)
SeaExam and SeaBench: Benchmarking LLMs with Local Multilingual Questions in Southeast Asia
von: Liu, Chaoqun, et al.
Veröffentlicht: (2025)
von: Liu, Chaoqun, et al.
Veröffentlicht: (2025)
Domain-Expanded ASTE: Rethinking Generalization in Aspect Sentiment Triplet Extraction
von: Chia, Yew Ken, et al.
Veröffentlicht: (2023)
von: Chia, Yew Ken, et al.
Veröffentlicht: (2023)
Lingshu-Cell: A generative cellular world model for transcriptome modeling toward virtual cells
von: Zhang, Han, et al.
Veröffentlicht: (2026)
von: Zhang, Han, et al.
Veröffentlicht: (2026)
Democratizing LLMs for Low-Resource Languages by Leveraging their English Dominant Abilities with Linguistically-Diverse Prompts
von: Nguyen, Xuan-Phi, et al.
Veröffentlicht: (2023)
von: Nguyen, Xuan-Phi, et al.
Veröffentlicht: (2023)
STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs
von: Li, Zongzhao, et al.
Veröffentlicht: (2025)
von: Li, Zongzhao, et al.
Veröffentlicht: (2025)
OS-ATLAS: A Foundation Action Model for Generalist GUI Agents
von: Wu, Zhiyong, et al.
Veröffentlicht: (2024)
von: Wu, Zhiyong, et al.
Veröffentlicht: (2024)
Student-in-the-Loop Chain-of-Thought Distillation via Generation-Time Selection
von: He, Chaoqun, et al.
Veröffentlicht: (2026)
von: He, Chaoqun, et al.
Veröffentlicht: (2026)
SeaLLMs -- Large Language Models for Southeast Asia
von: Nguyen, Xuan-Phi, et al.
Veröffentlicht: (2023)
von: Nguyen, Xuan-Phi, et al.
Veröffentlicht: (2023)
LLM-R2: A Large Language Model Enhanced Rule-based Rewrite System for Boosting Query Efficiency
von: Li, Zhaodonghui, et al.
Veröffentlicht: (2024)
von: Li, Zhaodonghui, et al.
Veröffentlicht: (2024)
AMERICANO: Argument Generation with Discourse-driven Decomposition and Agent Interaction
von: Hu, Zhe, et al.
Veröffentlicht: (2023)
von: Hu, Zhe, et al.
Veröffentlicht: (2023)
Large Language Models can Contrastively Refine their Generation for Better Sentence Representation Learning
von: Wang, Huiming, et al.
Veröffentlicht: (2023)
von: Wang, Huiming, et al.
Veröffentlicht: (2023)
From Macro to Micro: Benchmarking Microscopic Spatial Intelligence on Molecules via Vision-Language Models
von: Li, Zongzhao, et al.
Veröffentlicht: (2025)
von: Li, Zongzhao, et al.
Veröffentlicht: (2025)
BRIGHT: A Collaborative Generalist-Specialist Foundation Model for Breast Pathology
von: Guo, Xiaojing, et al.
Veröffentlicht: (2026)
von: Guo, Xiaojing, et al.
Veröffentlicht: (2026)
Debate-to-Write: A Persona-Driven Multi-Agent Framework for Diverse Argument Generation
von: Hu, Zhe, et al.
Veröffentlicht: (2024)
von: Hu, Zhe, et al.
Veröffentlicht: (2024)
Praxis-VLM: Vision-Grounded Decision Making via Text-Driven Reinforcement Learning
von: Hu, Zhe, et al.
Veröffentlicht: (2025)
von: Hu, Zhe, et al.
Veröffentlicht: (2025)
Modality-Specialized Synergizers for Interleaved Vision-Language Generalists
von: Xu, Zhiyang, et al.
Veröffentlicht: (2024)
von: Xu, Zhiyang, et al.
Veröffentlicht: (2024)
Inference-Time Scaling for Generalist Reward Modeling
von: Liu, Zijun, et al.
Veröffentlicht: (2025)
von: Liu, Zijun, et al.
Veröffentlicht: (2025)
Auto-Arena: Automating LLM Evaluations with Agent Peer Battles and Committee Discussions
von: Zhao, Ruochen, et al.
Veröffentlicht: (2024)
von: Zhao, Ruochen, et al.
Veröffentlicht: (2024)
Vision Foundation Models as Generalist Tokenizers for Image Generation
von: Zheng, Anlin, et al.
Veröffentlicht: (2026)
von: Zheng, Anlin, et al.
Veröffentlicht: (2026)
Attacking and Securing Community Detection: A Game-Theoretic Framework
von: Niu, Yifan, et al.
Veröffentlicht: (2025)
von: Niu, Yifan, et al.
Veröffentlicht: (2025)
Towards Generalist Game Players: An Investigation of Foundation Models in the Game Multiverse
von: Zhang, Kuan, et al.
Veröffentlicht: (2026)
von: Zhang, Kuan, et al.
Veröffentlicht: (2026)
Chameleon: Mixed-Modal Early-Fusion Foundation Models
von: Chameleon Team
Veröffentlicht: (2024)
von: Chameleon Team
Veröffentlicht: (2024)
Reasoning Paths Optimization: Learning to Reason and Explore From Diverse Paths
von: Chia, Yew Ken, et al.
Veröffentlicht: (2024)
von: Chia, Yew Ken, et al.
Veröffentlicht: (2024)
EventRL: Enhancing Event Extraction with Outcome Supervision for Large Language Models
von: Gao, Jun, et al.
Veröffentlicht: (2024)
von: Gao, Jun, et al.
Veröffentlicht: (2024)
Towards Generalist Intelligence in Dentistry: Vision Foundation Models for Oral and Maxillofacial Radiology
von: Huang, Xinrui, et al.
Veröffentlicht: (2025)
von: Huang, Xinrui, et al.
Veröffentlicht: (2025)
Serving Chain-structured Jobs with Large Memory Footprints with Application to Large Foundation Model Serving
von: Sun, Tingyang, et al.
Veröffentlicht: (2026)
von: Sun, Tingyang, et al.
Veröffentlicht: (2026)
OS-Symphony: A Holistic Framework for Robust and Generalist Computer-Using Agent
von: Yang, Bowen, et al.
Veröffentlicht: (2026)
von: Yang, Bowen, et al.
Veröffentlicht: (2026)
MedVersa: A Generalist Foundation Model for Medical Image Interpretation
von: Zhou, Hong-Yu, et al.
Veröffentlicht: (2024)
von: Zhou, Hong-Yu, et al.
Veröffentlicht: (2024)
DiffSpectra: Molecular Structure Elucidation from Spectra using Diffusion Models
von: Wang, Liang, et al.
Veröffentlicht: (2025)
von: Wang, Liang, et al.
Veröffentlicht: (2025)
DS$^2$-ABSA: Dual-Stream Data Synthesis with Label Refinement for Few-Shot Aspect-Based Sentiment Analysis
von: Xu, Hongling, et al.
Veröffentlicht: (2024)
von: Xu, Hongling, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SeaLLMs-Audio: Large Audio-Language Models for Southeast Asia
von: Liu, Chaoqun, et al.
Veröffentlicht: (2025) -
Scaling Language-Centric Omnimodal Representation Learning
von: Xiao, Chenghao, et al.
Veröffentlicht: (2025) -
Babel: Open Multilingual Large Language Models Serving Over 90% of Global Speakers
von: Zhao, Yiran, et al.
Veröffentlicht: (2025) -
VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning
von: Yuan, Ruifeng, et al.
Veröffentlicht: (2025) -
SeaLLMs 3: Open Foundation and Chat Multilingual Large Language Models for Southeast Asian Languages
von: Zhang, Wenxuan, et al.
Veröffentlicht: (2024)