Merino: Entropy-driven Design for Generative Language Models on IoT Devices
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Youpeng, Lin, Ming, Tang, Huadong, Wu, Qiang, Wang, Jun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ALISA: Accelerating Large Language Model Inference via Sparsity-Aware KV Caching
von: Zhao, Youpeng, et al.
Veröffentlicht: (2024)
von: Zhao, Youpeng, et al.
Veröffentlicht: (2024)
DevPiolt: Operation Recommendation for IoT Devices at Xiaomi Home
von: Wang, Yuxiang, et al.
Veröffentlicht: (2025)
von: Wang, Yuxiang, et al.
Veröffentlicht: (2025)
Systematic Outliers in Large Language Models
von: An, Yongqi, et al.
Veröffentlicht: (2025)
von: An, Yongqi, et al.
Veröffentlicht: (2025)
Recommending Pre-Trained Models for IoT Devices
von: Patil, Parth V., et al.
Veröffentlicht: (2024)
von: Patil, Parth V., et al.
Veröffentlicht: (2024)
Split Knowledge Distillation for Large Models in IoT: Architecture, Challenges, and Solutions
von: Li, Zuguang, et al.
Veröffentlicht: (2024)
von: Li, Zuguang, et al.
Veröffentlicht: (2024)
IoT-LM: Large Multisensory Language Models for the Internet of Things
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
On the Entropy Calibration of Language Models
von: Cao, Steven, et al.
Veröffentlicht: (2025)
von: Cao, Steven, et al.
Veröffentlicht: (2025)
DTMM: Deploying TinyML Models on Extremely Weak IoT Devices with Pruning
von: Han, Lixiang, et al.
Veröffentlicht: (2024)
von: Han, Lixiang, et al.
Veröffentlicht: (2024)
Measuring Social Norms of Large Language Models
von: Yuan, Ye, et al.
Veröffentlicht: (2024)
von: Yuan, Ye, et al.
Veröffentlicht: (2024)
The Entropy Mechanism of Reinforcement Learning for Reasoning Language Models
von: Cui, Ganqu, et al.
Veröffentlicht: (2025)
von: Cui, Ganqu, et al.
Veröffentlicht: (2025)
A Systematic Survey on Large Language Models for Algorithm Design
von: Liu, Fei, et al.
Veröffentlicht: (2024)
von: Liu, Fei, et al.
Veröffentlicht: (2024)
LangTopo: Aligning Language Descriptions of Graphs with Tokenized Topological Modeling
von: Guan, Zhong, et al.
Veröffentlicht: (2024)
von: Guan, Zhong, et al.
Veröffentlicht: (2024)
Large Language Models Badly Generalize across Option Length, Problem Types, and Irrelevant Noun Replacements
von: Zhao, Guangxiang, et al.
Veröffentlicht: (2025)
von: Zhao, Guangxiang, et al.
Veröffentlicht: (2025)
Generalized Category Discovery with Large Language Models in the Loop
von: An, Wenbin, et al.
Veröffentlicht: (2023)
von: An, Wenbin, et al.
Veröffentlicht: (2023)
Large Language Models to Diffusion Finetuning
von: Cetin, Edoardo, et al.
Veröffentlicht: (2025)
von: Cetin, Edoardo, et al.
Veröffentlicht: (2025)
MixCE: Training Autoregressive Language Models by Mixing Forward and Reverse Cross-Entropies
von: Zhang, Shiyue, et al.
Veröffentlicht: (2023)
von: Zhang, Shiyue, et al.
Veröffentlicht: (2023)
Curiosity-driven Red-teaming for Large Language Models
von: Hong, Zhang-Wei, et al.
Veröffentlicht: (2024)
von: Hong, Zhang-Wei, et al.
Veröffentlicht: (2024)
Structured Agent Distillation for Large Language Model
von: Liu, Jun, et al.
Veröffentlicht: (2025)
von: Liu, Jun, et al.
Veröffentlicht: (2025)
A General Framework for Producing Interpretable Semantic Text Embeddings
von: Sun, Yiqun, et al.
Veröffentlicht: (2024)
von: Sun, Yiqun, et al.
Veröffentlicht: (2024)
E-Sparse: Boosting the Large Language Model Inference through Entropy-based N:M Sparsity
von: Li, Yun, et al.
Veröffentlicht: (2023)
von: Li, Yun, et al.
Veröffentlicht: (2023)
Entropy Aware Reward Guidance for Diffusion Language Model Alignment
von: Tejaswi, Atula, et al.
Veröffentlicht: (2026)
von: Tejaswi, Atula, et al.
Veröffentlicht: (2026)
EdgeMoE: Empowering Sparse Large Language Models on Mobile Devices
von: Yi, Rongjie, et al.
Veröffentlicht: (2023)
von: Yi, Rongjie, et al.
Veröffentlicht: (2023)
Learn while Unlearn: An Iterative Unlearning Framework for Generative Language Models
von: Tang, Haoyu, et al.
Veröffentlicht: (2024)
von: Tang, Haoyu, et al.
Veröffentlicht: (2024)
DReSS: Data-driven Regularized Structured Streamlining for Large Language Models
von: Feng, Mingkuan, et al.
Veröffentlicht: (2025)
von: Feng, Mingkuan, et al.
Veröffentlicht: (2025)
FedCCA: Client-Centric Adaptation against Data Heterogeneity in Federated Learning on IoT Devices
von: Wang, Kaile, et al.
Veröffentlicht: (2026)
von: Wang, Kaile, et al.
Veröffentlicht: (2026)
Mix Data or Merge Models? Balancing the Helpfulness, Honesty, and Harmlessness of Large Language Model via Model Merging
von: Yang, Jinluan, et al.
Veröffentlicht: (2025)
von: Yang, Jinluan, et al.
Veröffentlicht: (2025)
Pareto Multi-Objective Alignment for Language Models
von: He, Qiang, et al.
Veröffentlicht: (2025)
von: He, Qiang, et al.
Veröffentlicht: (2025)
Data-driven Discovery with Large Generative Models
von: Majumder, Bodhisattwa Prasad, et al.
Veröffentlicht: (2024)
von: Majumder, Bodhisattwa Prasad, et al.
Veröffentlicht: (2024)
MobileLLM: Optimizing Sub-billion Parameter Language Models for On-Device Use Cases
von: Liu, Zechun, et al.
Veröffentlicht: (2024)
von: Liu, Zechun, et al.
Veröffentlicht: (2024)
Generative Evaluation of Complex Reasoning in Large Language Models
von: Lin, Haowei, et al.
Veröffentlicht: (2025)
von: Lin, Haowei, et al.
Veröffentlicht: (2025)
Toward Adaptive Large Language Models Structured Pruning via Hybrid-grained Weight Importance Assessment
von: Liu, Jun, et al.
Veröffentlicht: (2024)
von: Liu, Jun, et al.
Veröffentlicht: (2024)
KwaiAgents: Generalized Information-seeking Agent System with Large Language Models
von: Pan, Haojie, et al.
Veröffentlicht: (2023)
von: Pan, Haojie, et al.
Veröffentlicht: (2023)
Clustering-driven Memory Compression for On-device Large Language Models
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2026)
von: Bohdal, Ondrej, et al.
Veröffentlicht: (2026)
ReAGent: A Model-agnostic Feature Attribution Method for Generative Language Models
von: Zhao, Zhixue, et al.
Veröffentlicht: (2024)
von: Zhao, Zhixue, et al.
Veröffentlicht: (2024)
IDEA Prune: An Integrated Enlarge-and-Prune Pipeline in Generative Language Model Pretraining
von: Li, Yixiao, et al.
Veröffentlicht: (2025)
von: Li, Yixiao, et al.
Veröffentlicht: (2025)
IoT-Based Preventive Mental Health Using Knowledge Graphs and Standards for Better Well-Being
von: Gyrard, Amelie, et al.
Veröffentlicht: (2024)
von: Gyrard, Amelie, et al.
Veröffentlicht: (2024)
Large Language Models Are Bad Dice Players: LLMs Struggle to Generate Random Numbers from Statistical Distributions
von: Zhao, Minda, et al.
Veröffentlicht: (2026)
von: Zhao, Minda, et al.
Veröffentlicht: (2026)
Recall-Extend Dynamics: Enhancing Small Language Models through Controlled Exploration and Refined Offline Integration
von: Guan, Zhong, et al.
Veröffentlicht: (2025)
von: Guan, Zhong, et al.
Veröffentlicht: (2025)
ModelGPT: Unleashing LLM's Capabilities for Tailored Model Generation
von: Tang, Zihao, et al.
Veröffentlicht: (2024)
von: Tang, Zihao, et al.
Veröffentlicht: (2024)
Multimodal Generative Engine Optimization: Rank Manipulation for Vision-Language Model Rankers
von: Du, Yixuan, et al.
Veröffentlicht: (2026)
von: Du, Yixuan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
ALISA: Accelerating Large Language Model Inference via Sparsity-Aware KV Caching
von: Zhao, Youpeng, et al.
Veröffentlicht: (2024) -
DevPiolt: Operation Recommendation for IoT Devices at Xiaomi Home
von: Wang, Yuxiang, et al.
Veröffentlicht: (2025) -
Systematic Outliers in Large Language Models
von: An, Yongqi, et al.
Veröffentlicht: (2025) -
Recommending Pre-Trained Models for IoT Devices
von: Patil, Parth V., et al.
Veröffentlicht: (2024) -
Split Knowledge Distillation for Large Models in IoT: Architecture, Challenges, and Solutions
von: Li, Zuguang, et al.
Veröffentlicht: (2024)