Beyond Efficiency: A Systematic Survey of Resource-Efficient Large Language Models
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Bai, Guangji, Chai, Zheng, Ling, Chen, Wang, Shiyu, Lu, Jiaying, Zhang, Nan, Shi, Tingwei, Yu, Ziyang, Zhu, Mengdan, Zhang, Yifei, Song, Xinyuan, Yang, Carl, Cheng, Yue, Zhao, Liang |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Cross-modal RAG: Sub-dimensional Text-to-Image Retrieval-Augmented Generation
par: Zhu, Mengdan, et autres
Publié: (2025)
par: Zhu, Mengdan, et autres
Publié: (2025)
StructPrune: Structured Global Pruning asymptotics with $\mathcal{O}(\sqrt{N})$ GPU Memory
par: Song, Xinyuan, et autres
Publié: (2025)
par: Song, Xinyuan, et autres
Publié: (2025)
Saliency-Bench: A Comprehensive Benchmark for Evaluating Visual Explanations
par: Zhang, Yifei, et autres
Publié: (2023)
par: Zhang, Yifei, et autres
Publié: (2023)
MEGL: Multimodal Explanation-Guided Learning
par: Zhang, Yifei, et autres
Publié: (2024)
par: Zhang, Yifei, et autres
Publié: (2024)
Measuring Spiritual Values and Bias of Large Language Models
par: Liu, Songyuan, et autres
Publié: (2024)
par: Liu, Songyuan, et autres
Publié: (2024)
SparseLLM: Towards Global Pruning for Pre-trained Language Models
par: Bai, Guangji, et autres
Publié: (2024)
par: Bai, Guangji, et autres
Publié: (2024)
Visual Attention Prompted Prediction and Learning
par: Zhang, Yifei, et autres
Publié: (2023)
par: Zhang, Yifei, et autres
Publié: (2023)
FedSpaLLM: Federated Pruning of Large Language Models
par: Bai, Guangji, et autres
Publié: (2024)
par: Bai, Guangji, et autres
Publié: (2024)
A Survey on Employing Large Language Models for Text-to-SQL Tasks
par: Shi, Liang, et autres
Publié: (2024)
par: Shi, Liang, et autres
Publié: (2024)
Continuous Temporal Domain Generalization
par: Cai, Zekun, et autres
Publié: (2024)
par: Cai, Zekun, et autres
Publié: (2024)
A Plug-and-Play Framework for Volumetric Light-Sheet Image Reconstruction
par: Gong, Yi, et autres
Publié: (2025)
par: Gong, Yi, et autres
Publié: (2025)
Efficient Inference for Large Reasoning Models: A Survey
par: Liu, Yue, et autres
Publié: (2025)
par: Liu, Yue, et autres
Publié: (2025)
FedEHR-Gen: Federated Synthetic Time-Series EHR Generation via Latent Space Alignment and Distribution-Aware Aggregation
par: Bai, Jun, et autres
Publié: (2026)
par: Bai, Jun, et autres
Publié: (2026)
Bridging Efficiency and Transparency: Explainable CoT Compression in Multimodal Large Reasoning Models
par: Wang, Yizhi, et autres
Publié: (2026)
par: Wang, Yizhi, et autres
Publié: (2026)
A Survey of Efficient Reasoning for Large Reasoning Models: Language, Multimodality, and Beyond
par: Qu, Xiaoye, et autres
Publié: (2025)
par: Qu, Xiaoye, et autres
Publié: (2025)
Structural Disentanglement of Causal and Correlated Concepts
par: Zhao, Qilong, et autres
Publié: (2024)
par: Zhao, Qilong, et autres
Publié: (2024)
Blockchain-aided wireless federated learning: Resource allocation and client scheduling
par: Li, Jun, et autres
Publié: (2024)
par: Li, Jun, et autres
Publié: (2024)
Beyond the Speculative Game: A Survey of Speculative Execution in Large Language Models
par: Zhang, Chen, et autres
Publié: (2024)
par: Zhang, Chen, et autres
Publié: (2024)
DUE: Dynamic Uncertainty-Aware Explanation Supervision via 3D Imputation
par: Zhao, Qilong, et autres
Publié: (2024)
par: Zhao, Qilong, et autres
Publié: (2024)
C3N2H5 Superalkali‐Enhanced Halide Perovskites for High‐Efficiency Photovoltaic Applications
par: Tingwei Zhou, et autres
Publié: (2025)
par: Tingwei Zhou, et autres
Publié: (2025)
Enhancing Mobile Crowdsensing Efficiency: A Coverage-aware Resource Allocation Approach
par: Fu, Yaru, et autres
Publié: (2025)
par: Fu, Yaru, et autres
Publié: (2025)
Exploring Young Adults' Mental Health Help‐Seeking Journey: Preliminary Findings on Resource Navigation Behavior
par: Jiaying Liu, et autres
Publié: (2024)
par: Jiaying Liu, et autres
Publié: (2024)
RLingua: Improving Reinforcement Learning Sample Efficiency in Robotic Manipulations With Large Language Models
par: Chen, Liangliang, et autres
Publié: (2024)
par: Chen, Liangliang, et autres
Publié: (2024)
ELAD: Explanation-Guided Large Language Models Active Distillation
par: Zhang, Yifei, et autres
Publié: (2024)
par: Zhang, Yifei, et autres
Publié: (2024)
A Review on Knowledge Graphs for Healthcare: Resources, Applications, and Promises
par: Cui, Hejie, et autres
Publié: (2023)
par: Cui, Hejie, et autres
Publié: (2023)
EfficientLLM: Efficiency in Large Language Models
par: Yuan, Zhengqing, et autres
Publié: (2025)
par: Yuan, Zhengqing, et autres
Publié: (2025)
Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMs
par: Zhu, Mengdan, et autres
Publié: (2026)
par: Zhu, Mengdan, et autres
Publié: (2026)
Continuous Domain Generalization
par: Cai, Zekun, et autres
Publié: (2025)
par: Cai, Zekun, et autres
Publié: (2025)
Uncertainty Quantification for In-Context Learning of Large Language Models
par: Ling, Chen, et autres
Publié: (2024)
par: Ling, Chen, et autres
Publié: (2024)
Learning to Recommend Multi-Agent Subgraphs from Calling Trees
par: Song, Xinyuan, et autres
Publié: (2026)
par: Song, Xinyuan, et autres
Publié: (2026)
A Survey of Multimodal Large Language Model from A Data-centric Perspective
par: Bai, Tianyi, et autres
Publié: (2024)
par: Bai, Tianyi, et autres
Publié: (2024)
AutoSurvey2: Empowering Researchers with Next Level Automated Literature Surveys
par: Wu, Siyi, et autres
Publié: (2025)
par: Wu, Siyi, et autres
Publié: (2025)
Do Large Language Models Understand Conversational Implicature -- A case study with a chinese sitcom
par: Yue, Shisen, et autres
Publié: (2024)
par: Yue, Shisen, et autres
Publié: (2024)
Large-friction and incompressible limits for pressureless Euler-Navier-Stokes flows
par: Li, Hai-Liang, et autres
Publié: (2025)
par: Li, Hai-Liang, et autres
Publié: (2025)
Hydroxyalkylation of Benzylic C─H Bonds via Heterogeneous Photocatalysis
par: Pengfei Song, et autres
Publié: (2026)
par: Pengfei Song, et autres
Publié: (2026)
COLT: Enhancing Video Large Language Models with Continual Tool Usage
par: Liu, Yuyang, et autres
Publié: (2025)
par: Liu, Yuyang, et autres
Publié: (2025)
Affective Computing in the Era of Large Language Models: A Survey from the NLP Perspective
par: Zhang, Yiqun, et autres
Publié: (2024)
par: Zhang, Yiqun, et autres
Publié: (2024)
POND: Multi-Source Time Series Domain Adaptation with Information-Aware Prompt Tuning
par: Wang, Junxiang, et autres
Publié: (2023)
par: Wang, Junxiang, et autres
Publié: (2023)
Circuit-Aware SAT Solving: Guiding CDCL via Conditional Probabilities
par: Zhu, Jiaying, et autres
Publié: (2025)
par: Zhu, Jiaying, et autres
Publié: (2025)
Gene-associated Disease Discovery Powered by Large Language Models
par: Chang, Jiayu, et autres
Publié: (2024)
par: Chang, Jiayu, et autres
Publié: (2024)
Documents similaires
-
Cross-modal RAG: Sub-dimensional Text-to-Image Retrieval-Augmented Generation
par: Zhu, Mengdan, et autres
Publié: (2025) -
StructPrune: Structured Global Pruning asymptotics with $\mathcal{O}(\sqrt{N})$ GPU Memory
par: Song, Xinyuan, et autres
Publié: (2025) -
Saliency-Bench: A Comprehensive Benchmark for Evaluating Visual Explanations
par: Zhang, Yifei, et autres
Publié: (2023) -
MEGL: Multimodal Explanation-Guided Learning
par: Zhang, Yifei, et autres
Publié: (2024) -
Measuring Spiritual Values and Bias of Large Language Models
par: Liu, Songyuan, et autres
Publié: (2024)