Curvature-Weighted Capacity Allocation: A Minimum Description Length Framework for Layer-Adaptive Large Language Model Optimization
Fuente:
arXiv
Guardado en:
| Autores principales: | Amaefuna, Theophilus, Vaidya, Hitesh, Chhabra, Anshuman, Mali, Ankur |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Rethinking Reasoning in LLMs: Neuro-Symbolic Local RetoMaton Beyond ICL and CoT
por: Mamidala, Rushitha Santhoshi, et al.
Publicado: (2025)
por: Mamidala, Rushitha Santhoshi, et al.
Publicado: (2025)
Minimum Description Length and Generalization Guarantees for Representation Learning
por: Sefidgaran, Milad, et al.
Publicado: (2024)
por: Sefidgaran, Milad, et al.
Publicado: (2024)
Golden Layers and Where to Find Them: Improved Knowledge Editing for Large Language Models Via Layer Gradient Analysis
por: Datta, Shrestha, et al.
Publicado: (2026)
por: Datta, Shrestha, et al.
Publicado: (2026)
A Theory of Machine Understanding via the Minimum Description Length Principle
por: Zhang, Canlin, et al.
Publicado: (2025)
por: Zhang, Canlin, et al.
Publicado: (2025)
Re-ranking Using Large Language Models for Mitigating Exposure to Harmful Content on Social Media Platforms
por: Oak, Rajvardhan, et al.
Publicado: (2025)
por: Oak, Rajvardhan, et al.
Publicado: (2025)
Incentivizing News Consumption on Social Media Platforms Using Large Language Models and Realistic Bot Accounts
por: Askari, Hadi, et al.
Publicado: (2024)
por: Askari, Hadi, et al.
Publicado: (2024)
First is Not Really Better Than Last: Evaluating Layer Choice and Aggregation Strategies in Language Model Data Influence Estimation
por: Vitel, Dmytro, et al.
Publicado: (2025)
por: Vitel, Dmytro, et al.
Publicado: (2025)
Large AI Models for Wireless Physical Layer
por: Guo, Jiajia, et al.
Publicado: (2025)
por: Guo, Jiajia, et al.
Publicado: (2025)
Optimized Couplings for Watermarking Large Language Models
por: Tsur, Dor, et al.
Publicado: (2025)
por: Tsur, Dor, et al.
Publicado: (2025)
Neural Minimum Weight Perfect Matching for Quantum Error Codes
por: Peled, Yotam, et al.
Publicado: (2026)
por: Peled, Yotam, et al.
Publicado: (2026)
LayerIF: Estimating Layer Quality for Large Language Models using Influence Functions
por: Askari, Hadi, et al.
Publicado: (2025)
por: Askari, Hadi, et al.
Publicado: (2025)
Towards Safer Social Media Platforms: Scalable and Performant Few-Shot Harmful Content Moderation Using Large Language Models
por: Bonagiri, Akash, et al.
Publicado: (2025)
por: Bonagiri, Akash, et al.
Publicado: (2025)
"Whose Side Are You On?" Estimating Ideology of Political and News Content Using Large Language Models and Few-shot Demonstration Selection
por: Haroon, Muhammad, et al.
Publicado: (2025)
por: Haroon, Muhammad, et al.
Publicado: (2025)
OpenRANet: Neuralized Spectrum Access by Joint Subcarrier and Power Allocation with Optimization-based Deep Learning
por: Chen, Siya, et al.
Publicado: (2024)
por: Chen, Siya, et al.
Publicado: (2024)
Deep Variable-Length Feedback Codes
por: Ding, Yu, et al.
Publicado: (2026)
por: Ding, Yu, et al.
Publicado: (2026)
The Capacity of the Weighted Read Channel
por: Yerushalmi, Omer, et al.
Publicado: (2024)
por: Yerushalmi, Omer, et al.
Publicado: (2024)
Variable-Length Stop-Feedback Coding for Minimum Age of Incorrect Information
por: Bountrogiannis, Konstantinos, et al.
Publicado: (2024)
por: Bountrogiannis, Konstantinos, et al.
Publicado: (2024)
Large Language Models for Wireless Communications: From Adaptation to Autonomy
por: Liang, Le, et al.
Publicado: (2025)
por: Liang, Le, et al.
Publicado: (2025)
Revisiting Zero-Shot Abstractive Summarization in the Era of Large Language Models from the Perspective of Position Bias
por: Chhabra, Anshuman, et al.
Publicado: (2024)
por: Chhabra, Anshuman, et al.
Publicado: (2024)
Context Channel Capacity: An Information-Theoretic Framework for Understanding Catastrophic Forgetting
por: Cheng, Ran
Publicado: (2026)
por: Cheng, Ran
Publicado: (2026)
On the Reasoning Capacity of AI Models and How to Quantify It
por: Radha, Santosh Kumar, et al.
Publicado: (2025)
por: Radha, Santosh Kumar, et al.
Publicado: (2025)
A General Deep Learning Framework for Wireless Resource Allocation under Discrete Constraints
por: Wang, Yikun, et al.
Publicado: (2026)
por: Wang, Yikun, et al.
Publicado: (2026)
The Semiotic Channel Principle: Measuring the Capacity for Meaning in LLM Communication
por: Picca, Davide
Publicado: (2025)
por: Picca, Davide
Publicado: (2025)
Beamforming and Resource Allocation for Delay Minimization in RIS-Assisted OFDM Systems
por: Ma, Yu, et al.
Publicado: (2025)
por: Ma, Yu, et al.
Publicado: (2025)
Large Language Models as Evaluators for Scientific Synthesis
por: Evans, Julia, et al.
Publicado: (2024)
por: Evans, Julia, et al.
Publicado: (2024)
Neuro-mimetic Task-free Unsupervised Online Learning with Continual Self-Organizing Maps
por: Vaidya, Hitesh, et al.
Publicado: (2024)
por: Vaidya, Hitesh, et al.
Publicado: (2024)
Context Video Semantic Transmission with Variable Length and Rate Coding over MIMO Channels
por: Xie, Bingyan, et al.
Publicado: (2025)
por: Xie, Bingyan, et al.
Publicado: (2025)
Single and Multi-Frequency Path Loss Models for Indoor Hotspot Scenario Based on Measurements Conducted at 6.75, 16.95, 28, 73 and 142 GHz
por: Poddar, Hitesh, et al.
Publicado: (2025)
por: Poddar, Hitesh, et al.
Publicado: (2025)
Achieving DNA Labeling Capacity with Minimum Labels through Extremal de Bruijn Subgraphs
por: Hofmeister, Christoph, et al.
Publicado: (2024)
por: Hofmeister, Christoph, et al.
Publicado: (2024)
Lightweight Quantum Agent for Edge Systems: Joint PQC and NOMA Resource Allocation
por: Yao, Yongtao, et al.
Publicado: (2026)
por: Yao, Yongtao, et al.
Publicado: (2026)
On the Minimum Distances of Finite-Length Lifted Product Quantum LDPC Codes
por: Raveendran, Nithin, et al.
Publicado: (2025)
por: Raveendran, Nithin, et al.
Publicado: (2025)
Large Language Models for Telecom: Forthcoming Impact on the Industry
por: Maatouk, Ali, et al.
Publicado: (2023)
por: Maatouk, Ali, et al.
Publicado: (2023)
Conversational Complexity for Assessing Risk in Large Language Models
por: Burden, John, et al.
Publicado: (2024)
por: Burden, John, et al.
Publicado: (2024)
LLMs as Noisy Channels: A Shannon Perspective on Model Capacity and Scaling Laws
por: Ouyang, Xu, et al.
Publicado: (2026)
por: Ouyang, Xu, et al.
Publicado: (2026)
Digital Twin Channel-Enabled Online Resource Allocation for 6G: Principle, Architecture and Application
por: Li, Tongjie, et al.
Publicado: (2025)
por: Li, Tongjie, et al.
Publicado: (2025)
The Generating Idempotent Is a Minimum-Weight Codeword for Some Binary BCH Codes
por: Shany, Yaron, et al.
Publicado: (2024)
por: Shany, Yaron, et al.
Publicado: (2024)
Enumeration of Minimum Weight Codewords of Pre-Transformed Polar Codes by Tree Intersection
por: Zunker, Andreas, et al.
Publicado: (2023)
por: Zunker, Andreas, et al.
Publicado: (2023)
Weight Distribution of Repeated-Root Cyclic Codes with Prime Power Lengths
por: Zhao, Wei, et al.
Publicado: (2023)
por: Zhao, Wei, et al.
Publicado: (2023)
WDMoE: Wireless Distributed Large Language Models with Mixture of Experts
por: Xue, Nan, et al.
Publicado: (2024)
por: Xue, Nan, et al.
Publicado: (2024)
Tele-LLMs: A Series of Specialized Large Language Models for Telecommunications
por: Maatouk, Ali, et al.
Publicado: (2024)
por: Maatouk, Ali, et al.
Publicado: (2024)
Ejemplares similares
-
Rethinking Reasoning in LLMs: Neuro-Symbolic Local RetoMaton Beyond ICL and CoT
por: Mamidala, Rushitha Santhoshi, et al.
Publicado: (2025) -
Minimum Description Length and Generalization Guarantees for Representation Learning
por: Sefidgaran, Milad, et al.
Publicado: (2024) -
Golden Layers and Where to Find Them: Improved Knowledge Editing for Large Language Models Via Layer Gradient Analysis
por: Datta, Shrestha, et al.
Publicado: (2026) -
A Theory of Machine Understanding via the Minimum Description Length Principle
por: Zhang, Canlin, et al.
Publicado: (2025) -
Re-ranking Using Large Language Models for Mitigating Exposure to Harmful Content on Social Media Platforms
por: Oak, Rajvardhan, et al.
Publicado: (2025)