Unveiling LLM Mechanisms Through Neural ODEs and Control Theory
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Yukun, Dong, Qi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Unraveling Text Generation in LLMs: A Stochastic Differential Equation Approach
di: Zhang, Yukun
Pubblicazione: (2024)
di: Zhang, Yukun
Pubblicazione: (2024)
Exploring LLM Reasoning Through Controlled Prompt Variations
di: Chatziveroglou, Giannis, et al.
Pubblicazione: (2025)
di: Chatziveroglou, Giannis, et al.
Pubblicazione: (2025)
Jailbreak-as-a-Service++: Unveiling Distributed AI-Driven Malicious Information Campaigns Powered by LLM Crowdsourcing
di: Yan, Yu, et al.
Pubblicazione: (2025)
di: Yan, Yu, et al.
Pubblicazione: (2025)
Boosting In-Context Learning in LLMs Through the Lens of Classical Supervised Learning
di: Gundem, Korel, et al.
Pubblicazione: (2025)
di: Gundem, Korel, et al.
Pubblicazione: (2025)
From Words to Actions: Unveiling the Theoretical Underpinnings of LLM-Driven Autonomous Systems
di: He, Jianliang, et al.
Pubblicazione: (2024)
di: He, Jianliang, et al.
Pubblicazione: (2024)
When Greedy Wins: Emergent Exploitation Bias in Meta-Bandit LLM Training
di: Chen, Sanxing, et al.
Pubblicazione: (2025)
di: Chen, Sanxing, et al.
Pubblicazione: (2025)
Explainable LLM Unlearning Through Reasoning
di: Liao, Junfeng, et al.
Pubblicazione: (2026)
di: Liao, Junfeng, et al.
Pubblicazione: (2026)
ToolFactory: Automating Tool Generation by Leveraging LLM to Understand REST API Documentations
di: Ni, Xinyi, et al.
Pubblicazione: (2025)
di: Ni, Xinyi, et al.
Pubblicazione: (2025)
Interpreting and Controlling LLM Reasoning through Integrated Policy Gradient
di: Li, Changming, et al.
Pubblicazione: (2026)
di: Li, Changming, et al.
Pubblicazione: (2026)
What's the Magic Word? A Control Theory of LLM Prompting
di: Bhargava, Aman, et al.
Pubblicazione: (2023)
di: Bhargava, Aman, et al.
Pubblicazione: (2023)
Evaluation is All You Need: Strategic Overclaiming of LLM Reasoning Capabilities Through Evaluation Design
di: Sun, Lin, et al.
Pubblicazione: (2025)
di: Sun, Lin, et al.
Pubblicazione: (2025)
NoMAD-Attention: Efficient LLM Inference on CPUs Through Multiply-add-free Attention
di: Zhang, Tianyi, et al.
Pubblicazione: (2024)
di: Zhang, Tianyi, et al.
Pubblicazione: (2024)
Unveil Multi-Picture Descriptions for Multilingual Mild Cognitive Impairment Detection via Contrastive Learning
di: Qi, Kristin, et al.
Pubblicazione: (2025)
di: Qi, Kristin, et al.
Pubblicazione: (2025)
MoECollab: Democratizing LLM Development Through Collaborative Mixture of Experts
di: Harshit
Pubblicazione: (2025)
di: Harshit
Pubblicazione: (2025)
Flow-of-Options: Diversified and Improved LLM Reasoning by Thinking Through Options
di: Nair, Lakshmi, et al.
Pubblicazione: (2025)
di: Nair, Lakshmi, et al.
Pubblicazione: (2025)
DreamPRM-Code: Function-as-Step Process Reward Model with Label Correction for LLM Coding
di: Zhang, Ruiyi, et al.
Pubblicazione: (2025)
di: Zhang, Ruiyi, et al.
Pubblicazione: (2025)
Improving Data Efficiency for LLM Reinforcement Fine-tuning Through Difficulty-targeted Online Data Selection and Rollout Replay
di: Sun, Yifan, et al.
Pubblicazione: (2025)
di: Sun, Yifan, et al.
Pubblicazione: (2025)
Comparative Analysis of Demonstration Selection Algorithms for LLM In-Context Learning
di: Shu, Dong, et al.
Pubblicazione: (2024)
di: Shu, Dong, et al.
Pubblicazione: (2024)
Crosscoding Through Time: Tracking Emergence & Consolidation Of Linguistic Representations Throughout LLM Pretraining
di: Bayazit, Deniz, et al.
Pubblicazione: (2025)
di: Bayazit, Deniz, et al.
Pubblicazione: (2025)
Continual Knowledge Updating in LLM Systems: Learning Through Multi-Timescale Memory Dynamics
di: Pattichis, Andreas, et al.
Pubblicazione: (2026)
di: Pattichis, Andreas, et al.
Pubblicazione: (2026)
Think-Augmented Function Calling: Improving LLM Parameter Accuracy Through Embedded Reasoning
di: Wei, Lei, et al.
Pubblicazione: (2026)
di: Wei, Lei, et al.
Pubblicazione: (2026)
When Refusals Fail: Unstable Safety Mechanisms in Long-Context LLM Agents
di: Hadeliya, Tsimur, et al.
Pubblicazione: (2025)
di: Hadeliya, Tsimur, et al.
Pubblicazione: (2025)
How Is LLM Reasoning Distracted by Irrelevant Context? An Analysis Using a Controlled Benchmark
di: Yang, Minglai, et al.
Pubblicazione: (2025)
di: Yang, Minglai, et al.
Pubblicazione: (2025)
DrugR: Optimizing Molecular Drugs through LLM-based Explicit Reasoning
di: Liu, Haoran, et al.
Pubblicazione: (2026)
di: Liu, Haoran, et al.
Pubblicazione: (2026)
Low-Rank Adapters Meet Neural Architecture Search for LLM Compression
di: Muñoz, J. Pablo, et al.
Pubblicazione: (2025)
di: Muñoz, J. Pablo, et al.
Pubblicazione: (2025)
Offline Reinforcement Learning for LLM Multi-Step Reasoning
di: Wang, Huaijie, et al.
Pubblicazione: (2024)
di: Wang, Huaijie, et al.
Pubblicazione: (2024)
LLM Reasoning as Trajectories: Step-Specific Representation Geometry and Correctness Signals
di: Sun, Lihao, et al.
Pubblicazione: (2026)
di: Sun, Lihao, et al.
Pubblicazione: (2026)
Prompt-prompted Adaptive Structured Pruning for Efficient LLM Generation
di: Dong, Harry, et al.
Pubblicazione: (2024)
di: Dong, Harry, et al.
Pubblicazione: (2024)
FactorLLM: Factorizing Knowledge via Mixture of Experts for Large Language Models
di: Zhao, Zhongyu, et al.
Pubblicazione: (2024)
di: Zhao, Zhongyu, et al.
Pubblicazione: (2024)
Rethinking the Role of Prompting Strategies in LLM Test-Time Scaling: A Perspective of Probability Theory
di: Liu, Yexiang, et al.
Pubblicazione: (2025)
di: Liu, Yexiang, et al.
Pubblicazione: (2025)
Selective Deficits in LLM Mental Self-Modeling in a Behavior-Based Test of Theory of Mind
di: Ackerman, Christopher
Pubblicazione: (2026)
di: Ackerman, Christopher
Pubblicazione: (2026)
BiLLM: Pushing the Limit of Post-Training Quantization for LLMs
di: Huang, Wei, et al.
Pubblicazione: (2024)
di: Huang, Wei, et al.
Pubblicazione: (2024)
Weight-of-Thought Reasoning: Exploring Neural Network Weights for Enhanced LLM Reasoning
di: Punjwani, Saif, et al.
Pubblicazione: (2025)
di: Punjwani, Saif, et al.
Pubblicazione: (2025)
The Impact of Language Mixing on Bilingual LLM Reasoning
di: Li, Yihao, et al.
Pubblicazione: (2025)
di: Li, Yihao, et al.
Pubblicazione: (2025)
Scalable LLM Reasoning Acceleration with Low-rank Distillation
di: Dong, Harry, et al.
Pubblicazione: (2025)
di: Dong, Harry, et al.
Pubblicazione: (2025)
Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation
di: Dong, Guanting, et al.
Pubblicazione: (2024)
di: Dong, Guanting, et al.
Pubblicazione: (2024)
Model-GLUE: Democratized LLM Scaling for A Large Model Zoo in the Wild
di: Zhao, Xinyu, et al.
Pubblicazione: (2024)
di: Zhao, Xinyu, et al.
Pubblicazione: (2024)
Slow-Fast Policy Optimization: Reposition-Before-Update for LLM Reasoning
di: Wang, Ziyan, et al.
Pubblicazione: (2025)
di: Wang, Ziyan, et al.
Pubblicazione: (2025)
Zero-knowledge LLM hallucination detection and mitigation through fine-grained cross-model consistency
di: Goel, Aman, et al.
Pubblicazione: (2025)
di: Goel, Aman, et al.
Pubblicazione: (2025)
RuleR: Improving LLM Controllability by Rule-based Data Recycling
di: Li, Ming, et al.
Pubblicazione: (2024)
di: Li, Ming, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Unraveling Text Generation in LLMs: A Stochastic Differential Equation Approach
di: Zhang, Yukun
Pubblicazione: (2024) -
Exploring LLM Reasoning Through Controlled Prompt Variations
di: Chatziveroglou, Giannis, et al.
Pubblicazione: (2025) -
Jailbreak-as-a-Service++: Unveiling Distributed AI-Driven Malicious Information Campaigns Powered by LLM Crowdsourcing
di: Yan, Yu, et al.
Pubblicazione: (2025) -
Boosting In-Context Learning in LLMs Through the Lens of Classical Supervised Learning
di: Gundem, Korel, et al.
Pubblicazione: (2025) -
From Words to Actions: Unveiling the Theoretical Underpinnings of LLM-Driven Autonomous Systems
di: He, Jianliang, et al.
Pubblicazione: (2024)