Gespeichert in:
| Hauptverfasser: | Zhang, Yukun, Dong, Qi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2406.16985 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Unraveling Text Generation in LLMs: A Stochastic Differential Equation Approach
von: Zhang, Yukun
Veröffentlicht: (2024)
von: Zhang, Yukun
Veröffentlicht: (2024)
Exploring LLM Reasoning Through Controlled Prompt Variations
von: Chatziveroglou, Giannis, et al.
Veröffentlicht: (2025)
von: Chatziveroglou, Giannis, et al.
Veröffentlicht: (2025)
Jailbreak-as-a-Service++: Unveiling Distributed AI-Driven Malicious Information Campaigns Powered by LLM Crowdsourcing
von: Yan, Yu, et al.
Veröffentlicht: (2025)
von: Yan, Yu, et al.
Veröffentlicht: (2025)
Boosting In-Context Learning in LLMs Through the Lens of Classical Supervised Learning
von: Gundem, Korel, et al.
Veröffentlicht: (2025)
von: Gundem, Korel, et al.
Veröffentlicht: (2025)
When Greedy Wins: Emergent Exploitation Bias in Meta-Bandit LLM Training
von: Chen, Sanxing, et al.
Veröffentlicht: (2025)
von: Chen, Sanxing, et al.
Veröffentlicht: (2025)
From Words to Actions: Unveiling the Theoretical Underpinnings of LLM-Driven Autonomous Systems
von: He, Jianliang, et al.
Veröffentlicht: (2024)
von: He, Jianliang, et al.
Veröffentlicht: (2024)
ToolFactory: Automating Tool Generation by Leveraging LLM to Understand REST API Documentations
von: Ni, Xinyi, et al.
Veröffentlicht: (2025)
von: Ni, Xinyi, et al.
Veröffentlicht: (2025)
Explainable LLM Unlearning Through Reasoning
von: Liao, Junfeng, et al.
Veröffentlicht: (2026)
von: Liao, Junfeng, et al.
Veröffentlicht: (2026)
What's the Magic Word? A Control Theory of LLM Prompting
von: Bhargava, Aman, et al.
Veröffentlicht: (2023)
von: Bhargava, Aman, et al.
Veröffentlicht: (2023)
Interpreting and Controlling LLM Reasoning through Integrated Policy Gradient
von: Li, Changming, et al.
Veröffentlicht: (2026)
von: Li, Changming, et al.
Veröffentlicht: (2026)
DrugR: Optimizing Molecular Drugs through LLM-based Explicit Reasoning
von: Liu, Haoran, et al.
Veröffentlicht: (2026)
von: Liu, Haoran, et al.
Veröffentlicht: (2026)
Evaluation is All You Need: Strategic Overclaiming of LLM Reasoning Capabilities Through Evaluation Design
von: Sun, Lin, et al.
Veröffentlicht: (2025)
von: Sun, Lin, et al.
Veröffentlicht: (2025)
Unveil Multi-Picture Descriptions for Multilingual Mild Cognitive Impairment Detection via Contrastive Learning
von: Qi, Kristin, et al.
Veröffentlicht: (2025)
von: Qi, Kristin, et al.
Veröffentlicht: (2025)
NoMAD-Attention: Efficient LLM Inference on CPUs Through Multiply-add-free Attention
von: Zhang, Tianyi, et al.
Veröffentlicht: (2024)
von: Zhang, Tianyi, et al.
Veröffentlicht: (2024)
MoECollab: Democratizing LLM Development Through Collaborative Mixture of Experts
von: Harshit
Veröffentlicht: (2025)
von: Harshit
Veröffentlicht: (2025)
Flow-of-Options: Diversified and Improved LLM Reasoning by Thinking Through Options
von: Nair, Lakshmi, et al.
Veröffentlicht: (2025)
von: Nair, Lakshmi, et al.
Veröffentlicht: (2025)
DreamPRM-Code: Function-as-Step Process Reward Model with Label Correction for LLM Coding
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2025)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2025)
Comparative Analysis of Demonstration Selection Algorithms for LLM In-Context Learning
von: Shu, Dong, et al.
Veröffentlicht: (2024)
von: Shu, Dong, et al.
Veröffentlicht: (2024)
Improving Data Efficiency for LLM Reinforcement Fine-tuning Through Difficulty-targeted Online Data Selection and Rollout Replay
von: Sun, Yifan, et al.
Veröffentlicht: (2025)
von: Sun, Yifan, et al.
Veröffentlicht: (2025)
Crosscoding Through Time: Tracking Emergence & Consolidation Of Linguistic Representations Throughout LLM Pretraining
von: Bayazit, Deniz, et al.
Veröffentlicht: (2025)
von: Bayazit, Deniz, et al.
Veröffentlicht: (2025)
Continual Knowledge Updating in LLM Systems: Learning Through Multi-Timescale Memory Dynamics
von: Pattichis, Andreas, et al.
Veröffentlicht: (2026)
von: Pattichis, Andreas, et al.
Veröffentlicht: (2026)
Think-Augmented Function Calling: Improving LLM Parameter Accuracy Through Embedded Reasoning
von: Wei, Lei, et al.
Veröffentlicht: (2026)
von: Wei, Lei, et al.
Veröffentlicht: (2026)
How Is LLM Reasoning Distracted by Irrelevant Context? An Analysis Using a Controlled Benchmark
von: Yang, Minglai, et al.
Veröffentlicht: (2025)
von: Yang, Minglai, et al.
Veröffentlicht: (2025)
Offline Reinforcement Learning for LLM Multi-Step Reasoning
von: Wang, Huaijie, et al.
Veröffentlicht: (2024)
von: Wang, Huaijie, et al.
Veröffentlicht: (2024)
When Refusals Fail: Unstable Safety Mechanisms in Long-Context LLM Agents
von: Hadeliya, Tsimur, et al.
Veröffentlicht: (2025)
von: Hadeliya, Tsimur, et al.
Veröffentlicht: (2025)
LLM Reasoning as Trajectories: Step-Specific Representation Geometry and Correctness Signals
von: Sun, Lihao, et al.
Veröffentlicht: (2026)
von: Sun, Lihao, et al.
Veröffentlicht: (2026)
Calibrating Long-form Generations from Large Language Models
von: Huang, Yukun, et al.
Veröffentlicht: (2024)
von: Huang, Yukun, et al.
Veröffentlicht: (2024)
Cite Pretrain: Retrieval-Free Knowledge Attribution for Large Language Models
von: Huang, Yukun, et al.
Veröffentlicht: (2025)
von: Huang, Yukun, et al.
Veröffentlicht: (2025)
Model-GLUE: Democratized LLM Scaling for A Large Model Zoo in the Wild
von: Zhao, Xinyu, et al.
Veröffentlicht: (2024)
von: Zhao, Xinyu, et al.
Veröffentlicht: (2024)
FactorLLM: Factorizing Knowledge via Mixture of Experts for Large Language Models
von: Zhao, Zhongyu, et al.
Veröffentlicht: (2024)
von: Zhao, Zhongyu, et al.
Veröffentlicht: (2024)
Low-Rank Adapters Meet Neural Architecture Search for LLM Compression
von: Muñoz, J. Pablo, et al.
Veröffentlicht: (2025)
von: Muñoz, J. Pablo, et al.
Veröffentlicht: (2025)
Prompt-prompted Adaptive Structured Pruning for Efficient LLM Generation
von: Dong, Harry, et al.
Veröffentlicht: (2024)
von: Dong, Harry, et al.
Veröffentlicht: (2024)
BiLLM: Pushing the Limit of Post-Training Quantization for LLMs
von: Huang, Wei, et al.
Veröffentlicht: (2024)
von: Huang, Wei, et al.
Veröffentlicht: (2024)
TrustLDM: Benchmarking Trustworthiness in Language Diffusion Models
von: Mo, Yichuan, et al.
Veröffentlicht: (2026)
von: Mo, Yichuan, et al.
Veröffentlicht: (2026)
The Impact of Language Mixing on Bilingual LLM Reasoning
von: Li, Yihao, et al.
Veröffentlicht: (2025)
von: Li, Yihao, et al.
Veröffentlicht: (2025)
Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation
von: Dong, Guanting, et al.
Veröffentlicht: (2024)
von: Dong, Guanting, et al.
Veröffentlicht: (2024)
Scalable LLM Reasoning Acceleration with Low-rank Distillation
von: Dong, Harry, et al.
Veröffentlicht: (2025)
von: Dong, Harry, et al.
Veröffentlicht: (2025)
Rethinking the Role of Prompting Strategies in LLM Test-Time Scaling: A Perspective of Probability Theory
von: Liu, Yexiang, et al.
Veröffentlicht: (2025)
von: Liu, Yexiang, et al.
Veröffentlicht: (2025)
Selective Deficits in LLM Mental Self-Modeling in a Behavior-Based Test of Theory of Mind
von: Ackerman, Christopher
Veröffentlicht: (2026)
von: Ackerman, Christopher
Veröffentlicht: (2026)
Slow-Fast Policy Optimization: Reposition-Before-Update for LLM Reasoning
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Unraveling Text Generation in LLMs: A Stochastic Differential Equation Approach
von: Zhang, Yukun
Veröffentlicht: (2024) -
Exploring LLM Reasoning Through Controlled Prompt Variations
von: Chatziveroglou, Giannis, et al.
Veröffentlicht: (2025) -
Jailbreak-as-a-Service++: Unveiling Distributed AI-Driven Malicious Information Campaigns Powered by LLM Crowdsourcing
von: Yan, Yu, et al.
Veröffentlicht: (2025) -
Boosting In-Context Learning in LLMs Through the Lens of Classical Supervised Learning
von: Gundem, Korel, et al.
Veröffentlicht: (2025) -
When Greedy Wins: Emergent Exploitation Bias in Meta-Bandit LLM Training
von: Chen, Sanxing, et al.
Veröffentlicht: (2025)