Summarize Before You Speak with ARACH: A Training-Free Inference-Time Plug-In for Enhancing LLMs via Global Attention Reallocation
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Jingtao, Wang, Yucong, Ding, Jun, Cai, Rui, Wang, Xun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Ltri-LLM: Streaming Long Context Inference for LLMs with Training-Free Dynamic Triangular Attention Pattern
di: Tang, Hongyin, et al.
Pubblicazione: (2024)
di: Tang, Hongyin, et al.
Pubblicazione: (2024)
From Attribution to Abstention: Training-Free Attention-Based Auditing for Clinical Summarization
di: Yan, Qianqi, et al.
Pubblicazione: (2026)
di: Yan, Qianqi, et al.
Pubblicazione: (2026)
Thinking Before You Speak: A Proactive Test-time Scaling Approach
di: Liu, Cong, et al.
Pubblicazione: (2025)
di: Liu, Cong, et al.
Pubblicazione: (2025)
Know You Before You Speak: User-State Modeling for LLM Personalization in Multi-Turn Conversation
di: Luo, Jiani, et al.
Pubblicazione: (2026)
di: Luo, Jiani, et al.
Pubblicazione: (2026)
Cite Before You Speak: Enhancing Context-Response Grounding in E-commerce Conversational LLM-Agents
di: Zeng, Jingying, et al.
Pubblicazione: (2025)
di: Zeng, Jingying, et al.
Pubblicazione: (2025)
LLMs + Persona-Plug = Personalized LLMs
di: Liu, Jiongnan, et al.
Pubblicazione: (2024)
di: Liu, Jiongnan, et al.
Pubblicazione: (2024)
Think Before You Speak: Cultivating Communication Skills of Large Language Models via Inner Monologue
di: Zhou, Junkai, et al.
Pubblicazione: (2023)
di: Zhou, Junkai, et al.
Pubblicazione: (2023)
Why Not Act on What You Know? Unleashing Safety Potential of LLMs via Self-Aware Guard Enhancement
di: Ding, Peng, et al.
Pubblicazione: (2025)
di: Ding, Peng, et al.
Pubblicazione: (2025)
Look Before You Leap: Autonomous Exploration for LLM Agents
di: Ye, Ziang, et al.
Pubblicazione: (2026)
di: Ye, Ziang, et al.
Pubblicazione: (2026)
Score Before You Speak: Improving Persona Consistency in Dialogue Generation using Response Quality Scores
di: Saggar, Arpita, et al.
Pubblicazione: (2025)
di: Saggar, Arpita, et al.
Pubblicazione: (2025)
Self-Train Before You Transcribe
di: Flynn, Robert, et al.
Pubblicazione: (2024)
di: Flynn, Robert, et al.
Pubblicazione: (2024)
Talk Before You Retrieve: Agent-Led Discussions for Better RAG in Medical QA
di: Dong, Xuanzhao, et al.
Pubblicazione: (2025)
di: Dong, Xuanzhao, et al.
Pubblicazione: (2025)
Prejudge-Before-Think: Enhancing Large Language Models at Test-Time by Process Prejudge Reasoning
di: Wang, Jianing, et al.
Pubblicazione: (2025)
di: Wang, Jianing, et al.
Pubblicazione: (2025)
SyncThink: A Training-Free Strategy to Align Inference Termination with Reasoning Saturation
di: Li, Gengyang, et al.
Pubblicazione: (2026)
di: Li, Gengyang, et al.
Pubblicazione: (2026)
SPECTRUM: Speaker-Enhanced Pre-Training for Long Dialogue Summarization
di: Cho, Sangwoo, et al.
Pubblicazione: (2024)
di: Cho, Sangwoo, et al.
Pubblicazione: (2024)
Look Before You Leap: Enhancing Attention and Vigilance Regarding Harmful Content with GuidelineLLM
di: Zhang, Shaoqing, et al.
Pubblicazione: (2024)
di: Zhang, Shaoqing, et al.
Pubblicazione: (2024)
Contrastive Prompting Enhances Sentence Embeddings in LLMs through Inference-Time Steering
di: Cheng, Zifeng, et al.
Pubblicazione: (2025)
di: Cheng, Zifeng, et al.
Pubblicazione: (2025)
Understanding Before Reasoning: Enhancing Chain-of-Thought with Iterative Summarization Pre-Prompting
di: Zhu, Dong-Hai, et al.
Pubblicazione: (2025)
di: Zhu, Dong-Hai, et al.
Pubblicazione: (2025)
Infinite Retrieval: Attention Enhanced LLMs in Long-Context Processing
di: Ye, Xiaoju, et al.
Pubblicazione: (2025)
di: Ye, Xiaoju, et al.
Pubblicazione: (2025)
Plug-and-Play Training Framework for Preference Optimization
di: Ma, Jingyuan, et al.
Pubblicazione: (2024)
di: Ma, Jingyuan, et al.
Pubblicazione: (2024)
Training Matryoshka Mixture-of-Experts for Elastic Inference-Time Expert Utilization
di: Wang, Yaoxiang, et al.
Pubblicazione: (2025)
di: Wang, Yaoxiang, et al.
Pubblicazione: (2025)
With Greater Text Comes Greater Necessity: Inference-Time Training Helps Long Text Generation
di: Wang, Y., et al.
Pubblicazione: (2024)
di: Wang, Y., et al.
Pubblicazione: (2024)
ReaL: Efficient RLHF Training of Large Language Models with Parameter Reallocation
di: Mei, Zhiyu, et al.
Pubblicazione: (2024)
di: Mei, Zhiyu, et al.
Pubblicazione: (2024)
Listen and Chant Before You Read: The Ladder of Beauty in LM Pre-Training
di: Nomura, Yoshinori
Pubblicazione: (2026)
di: Nomura, Yoshinori
Pubblicazione: (2026)
Flux Attention: Context-Aware Hybrid Attention for Efficient LLMs Inference
di: Qiu, Quantong, et al.
Pubblicazione: (2026)
di: Qiu, Quantong, et al.
Pubblicazione: (2026)
AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation
di: Wang, Zijun, et al.
Pubblicazione: (2024)
di: Wang, Zijun, et al.
Pubblicazione: (2024)
Evolving LLMs' Self-Refinement Capability via Synergistic Training-Inference Optimization
di: Zeng, Yongcheng, et al.
Pubblicazione: (2025)
di: Zeng, Yongcheng, et al.
Pubblicazione: (2025)
Enhancing Faithfulness in Abstractive Summarization via Span-Level Fine-Tuning
di: Huang, Sicong, et al.
Pubblicazione: (2025)
di: Huang, Sicong, et al.
Pubblicazione: (2025)
HS-STaR: Hierarchical Sampling for Self-Taught Reasoners via Difficulty Estimation and Budget Reallocation
di: Xiong, Feng, et al.
Pubblicazione: (2025)
di: Xiong, Feng, et al.
Pubblicazione: (2025)
Hierarchical Attention Graph for Scientific Document Summarization in Global and Local Level
di: Zhao, Chenlong, et al.
Pubblicazione: (2024)
di: Zhao, Chenlong, et al.
Pubblicazione: (2024)
Verify Before You Commit: Towards Faithful Reasoning in LLM Agents via Self-Auditing
di: Yuan, Wenhao, et al.
Pubblicazione: (2026)
di: Yuan, Wenhao, et al.
Pubblicazione: (2026)
S2D2: Fast Decoding for Diffusion LLMs via Training-Free Self-Speculation
di: Han, Ligong, et al.
Pubblicazione: (2026)
di: Han, Ligong, et al.
Pubblicazione: (2026)
Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training
di: Yuan, Youliang, et al.
Pubblicazione: (2024)
di: Yuan, Youliang, et al.
Pubblicazione: (2024)
AgentEHR: Advancing Autonomous Clinical Decision-Making via Retrospective Summarization
di: Liao, Yusheng, et al.
Pubblicazione: (2026)
di: Liao, Yusheng, et al.
Pubblicazione: (2026)
Your Agent is More Brittle Than You Think: Uncovering Indirect Injection Vulnerabilities in Agentic LLMs
di: Zhu, Wenhui, et al.
Pubblicazione: (2026)
di: Zhu, Wenhui, et al.
Pubblicazione: (2026)
Steering Language Models Before They Speak: Logit-Level Interventions
di: An, Hyeseon, et al.
Pubblicazione: (2026)
di: An, Hyeseon, et al.
Pubblicazione: (2026)
Thinking Before Speaking: A Role-playing Model with Mindset
di: Zhang, Baohua, et al.
Pubblicazione: (2024)
di: Zhang, Baohua, et al.
Pubblicazione: (2024)
Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision
di: Xi, Zhiheng, et al.
Pubblicazione: (2024)
di: Xi, Zhiheng, et al.
Pubblicazione: (2024)
UniAttn: Reducing Inference Costs via Softmax Unification for Post-Training LLMs
di: Xiong, Yizhe, et al.
Pubblicazione: (2025)
di: Xiong, Yizhe, et al.
Pubblicazione: (2025)
Think Before You Prune: Self-Reflective Structured Pruning for Reasoning Language Models
di: Wang, Ziyan, et al.
Pubblicazione: (2025)
di: Wang, Ziyan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Ltri-LLM: Streaming Long Context Inference for LLMs with Training-Free Dynamic Triangular Attention Pattern
di: Tang, Hongyin, et al.
Pubblicazione: (2024) -
From Attribution to Abstention: Training-Free Attention-Based Auditing for Clinical Summarization
di: Yan, Qianqi, et al.
Pubblicazione: (2026) -
Thinking Before You Speak: A Proactive Test-time Scaling Approach
di: Liu, Cong, et al.
Pubblicazione: (2025) -
Know You Before You Speak: User-State Modeling for LLM Personalization in Multi-Turn Conversation
di: Luo, Jiani, et al.
Pubblicazione: (2026) -
Cite Before You Speak: Enhancing Context-Response Grounding in E-commerce Conversational LLM-Agents
di: Zeng, Jingying, et al.
Pubblicazione: (2025)