Gespeichert in:
| Hauptverfasser: | Lin, Chieh-Yen, Sun, Shao-Hua |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2605.11608 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
On Calibration of Large Language Models: From Response To Capability
von: Yang, Sin-Han, et al.
Veröffentlicht: (2026)
von: Yang, Sin-Han, et al.
Veröffentlicht: (2026)
Capturing Polysemanticity with PRISM: A Multi-Concept Feature Description Framework
von: Kopf, Laura, et al.
Veröffentlicht: (2025)
von: Kopf, Laura, et al.
Veröffentlicht: (2025)
PRISM: Parametrically Refactoring Inference for Speculative Sampling Draft Models
von: Wang, Xuliang, et al.
Veröffentlicht: (2026)
von: Wang, Xuliang, et al.
Veröffentlicht: (2026)
Decomposing and Measuring Evaluation Awareness
von: Li, Changling, et al.
Veröffentlicht: (2026)
von: Li, Changling, et al.
Veröffentlicht: (2026)
Decomposing Attention To Find Context-Sensitive Neurons
von: Gibson, Alex
Veröffentlicht: (2025)
von: Gibson, Alex
Veröffentlicht: (2025)
Decomposing Representation Space into Interpretable Subspaces with Unsupervised Learning
von: Huang, Xinting, et al.
Veröffentlicht: (2025)
von: Huang, Xinting, et al.
Veröffentlicht: (2025)
DHA: Learning Decoupled-Head Attention from Transformer Checkpoints via Adaptive Heads Fusion
von: Chen, Yilong, et al.
Veröffentlicht: (2024)
von: Chen, Yilong, et al.
Veröffentlicht: (2024)
SpreadsheetArena: Decomposing Preference in LLM Generation of Spreadsheet Workbooks
von: Kundurthy, Srivatsa, et al.
Veröffentlicht: (2026)
von: Kundurthy, Srivatsa, et al.
Veröffentlicht: (2026)
LLM Assertiveness can be Mechanistically Decomposed into Emotional and Logical Components
von: Tsujimura, Hikaru, et al.
Veröffentlicht: (2025)
von: Tsujimura, Hikaru, et al.
Veröffentlicht: (2025)
LLM Self-Correction with DeCRIM: Decompose, Critique, and Refine for Enhanced Following of Instructions with Multiple Constraints
von: Ferraz, Thomas Palmeira, et al.
Veröffentlicht: (2024)
von: Ferraz, Thomas Palmeira, et al.
Veröffentlicht: (2024)
Decomposing Elements of Problem Solving: What "Math" Does RL Teach?
von: Qin, Tian, et al.
Veröffentlicht: (2025)
von: Qin, Tian, et al.
Veröffentlicht: (2025)
Generating Pretraining Tokens from Organic Data for Data-Bound Scaling
von: Yu, Zichun, et al.
Veröffentlicht: (2026)
von: Yu, Zichun, et al.
Veröffentlicht: (2026)
Too Long, Didn't Model: Decomposing LLM Long-Context Understanding With Novels
von: Hamilton, Sil, et al.
Veröffentlicht: (2025)
von: Hamilton, Sil, et al.
Veröffentlicht: (2025)
DATA: Decomposed Attention-based Task Adaptation for Rehearsal-Free Continual Learning
von: Liao, Huanxuan, et al.
Veröffentlicht: (2025)
von: Liao, Huanxuan, et al.
Veröffentlicht: (2025)
Nemotron-Cascade: Scaling Cascaded Reinforcement Learning for General-Purpose Reasoning Models
von: Wang, Boxin, et al.
Veröffentlicht: (2025)
von: Wang, Boxin, et al.
Veröffentlicht: (2025)
Geometric-disentangelment Unlearning
von: Zhou, Duo, et al.
Veröffentlicht: (2025)
von: Zhou, Duo, et al.
Veröffentlicht: (2025)
Training Acceleration of Low-Rank Decomposed Networks using Sequential Freezing and Rank Quantization
von: Hajimolahoseini, Habib, et al.
Veröffentlicht: (2023)
von: Hajimolahoseini, Habib, et al.
Veröffentlicht: (2023)
LaMDA: Large Model Fine-Tuning via Spectrally Decomposed Low-Dimensional Adaptation
von: Azizi, Seyedarmin, et al.
Veröffentlicht: (2024)
von: Azizi, Seyedarmin, et al.
Veröffentlicht: (2024)
3DS: Medical Domain Adaptation of LLMs via Decomposed Difficulty-based Data Selection
von: Ding, Hongxin, et al.
Veröffentlicht: (2024)
von: Ding, Hongxin, et al.
Veröffentlicht: (2024)
EDoRA: Efficient Weight-Decomposed Low-Rank Adaptation via Singular Value Decomposition
von: Nasiri, Hamid, et al.
Veröffentlicht: (2025)
von: Nasiri, Hamid, et al.
Veröffentlicht: (2025)
Fourier Head: Helping Large Language Models Learn Complex Probability Distributions
von: Gillman, Nate, et al.
Veröffentlicht: (2024)
von: Gillman, Nate, et al.
Veröffentlicht: (2024)
DecomposeRL: Learning to Ask Useful, Informative, and Diverse Questions for Semi-Supervised, Traceable Claim Verification
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2026)
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2026)
Your Agent May Misevolve: Emergent Risks in Self-evolving LLM Agents
von: Shao, Shuai, et al.
Veröffentlicht: (2025)
von: Shao, Shuai, et al.
Veröffentlicht: (2025)
Multi-Head Mixture-of-Experts
von: Wu, Xun, et al.
Veröffentlicht: (2024)
von: Wu, Xun, et al.
Veröffentlicht: (2024)
Iteration Head: A Mechanistic Study of Chain-of-Thought
von: Cabannes, Vivien, et al.
Veröffentlicht: (2024)
von: Cabannes, Vivien, et al.
Veröffentlicht: (2024)
Theoretical Foundations of Scaling Law in Familial Models
von: Song, Huan, et al.
Veröffentlicht: (2025)
von: Song, Huan, et al.
Veröffentlicht: (2025)
Momentum Streams for Optimizer-Inspired Transformers
von: Gai, Jingchu, et al.
Veröffentlicht: (2026)
von: Gai, Jingchu, et al.
Veröffentlicht: (2026)
Cultural Binding Heads in Language Models
von: Floro, Avrile, et al.
Veröffentlicht: (2026)
von: Floro, Avrile, et al.
Veröffentlicht: (2026)
Mixture of Universal Experts: Scaling Virtual Width via Depth-Width Transformation
von: Chen, Yilong, et al.
Veröffentlicht: (2026)
von: Chen, Yilong, et al.
Veröffentlicht: (2026)
Beyond Size: How Gradients Shape Pruning Decisions in Large Language Models
von: Das, Rocktim Jyoti, et al.
Veröffentlicht: (2023)
von: Das, Rocktim Jyoti, et al.
Veröffentlicht: (2023)
Can an LLM Induce a Graph? Investigating Memory Drift and Context Length
von: Yousuf, Raquib Bin, et al.
Veröffentlicht: (2025)
von: Yousuf, Raquib Bin, et al.
Veröffentlicht: (2025)
Scaling Embeddings Outperforms Scaling Experts in Language Models
von: Liu, Hong, et al.
Veröffentlicht: (2026)
von: Liu, Hong, et al.
Veröffentlicht: (2026)
Which Attention Heads Matter for In-Context Learning?
von: Yin, Kayo, et al.
Veröffentlicht: (2025)
von: Yin, Kayo, et al.
Veröffentlicht: (2025)
Joint Detection of Fraud and Concept Drift inOnline Conversations with LLM-Assisted Judgment
von: Senol, Ali, et al.
Veröffentlicht: (2025)
von: Senol, Ali, et al.
Veröffentlicht: (2025)
TingIS: Real-time Risk Event Discovery from Noisy Customer Incidents at Enterprise Scale
von: Wang, Jun, et al.
Veröffentlicht: (2026)
von: Wang, Jun, et al.
Veröffentlicht: (2026)
Comprehensive Reassessment of Large-Scale Evaluation Outcomes in LLMs: A Multifaceted Statistical Approach
von: Sun, Kun, et al.
Veröffentlicht: (2024)
von: Sun, Kun, et al.
Veröffentlicht: (2024)
AutoScale: Scale-Aware Data Mixing for Pre-Training LLMs
von: Kang, Feiyang, et al.
Veröffentlicht: (2024)
von: Kang, Feiyang, et al.
Veröffentlicht: (2024)
Navigating Ideation Space: Decomposed Conceptual Representations for Positioning Scientific Ideas
von: Shen, Yuexi, et al.
Veröffentlicht: (2026)
von: Shen, Yuexi, et al.
Veröffentlicht: (2026)
DynScaling: Efficient Verifier-free Inference Scaling via Dynamic and Integrated Sampling
von: Wang, Fei, et al.
Veröffentlicht: (2025)
von: Wang, Fei, et al.
Veröffentlicht: (2025)
The Anxiety of Influence: Bloom Filters in Transformer Attention Heads
von: Balogh, Peter
Veröffentlicht: (2026)
von: Balogh, Peter
Veröffentlicht: (2026)
Ähnliche Einträge
-
On Calibration of Large Language Models: From Response To Capability
von: Yang, Sin-Han, et al.
Veröffentlicht: (2026) -
Capturing Polysemanticity with PRISM: A Multi-Concept Feature Description Framework
von: Kopf, Laura, et al.
Veröffentlicht: (2025) -
PRISM: Parametrically Refactoring Inference for Speculative Sampling Draft Models
von: Wang, Xuliang, et al.
Veröffentlicht: (2026) -
Decomposing and Measuring Evaluation Awareness
von: Li, Changling, et al.
Veröffentlicht: (2026) -
Decomposing Attention To Find Context-Sensitive Neurons
von: Gibson, Alex
Veröffentlicht: (2025)