Bridging Latent Reasoning and Target-Language Generation via Retrieval-Transition Heads
Fuente:
arXiv
Saved in:
| Main Authors: | Patel, Shaswat, Trivedi, Vishvesh, Han, Yue, Hong, Yihuai, Choi, Eunsol |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond Single Embeddings: Capturing Diverse Targets with Multi-Query Retrieval
by: Chen, Hung-Ting, et al.
Published: (2025)
by: Chen, Hung-Ting, et al.
Published: (2025)
Open-World Evaluation for Retrieving Diverse Perspectives
by: Chen, Hung-Ting, et al.
Published: (2024)
by: Chen, Hung-Ting, et al.
Published: (2024)
CLaRa: Bridging Retrieval and Generation with Continuous Latent Reasoning
by: He, Jie, et al.
Published: (2025)
by: He, Jie, et al.
Published: (2025)
RVR: Retrieve-Verify-Retrieve for Comprehensive Question Answering
by: Qian, Deniz, et al.
Published: (2026)
by: Qian, Deniz, et al.
Published: (2026)
AmbigDocs: Reasoning across Documents on Different Entities under the Same Name
by: Lee, Yoonsang, et al.
Published: (2024)
by: Lee, Yoonsang, et al.
Published: (2024)
RARe: Retrieval Augmented Retrieval with In-Context Examples
by: Tejaswi, Atula, et al.
Published: (2024)
by: Tejaswi, Atula, et al.
Published: (2024)
Understanding Retrieval Augmentation for Long-Form Question Answering
by: Chen, Hung-Ting, et al.
Published: (2023)
by: Chen, Hung-Ting, et al.
Published: (2023)
RefreshKV: Updating Small KV Cache During Long-form Generation
by: Xu, Fangyuan, et al.
Published: (2024)
by: Xu, Fangyuan, et al.
Published: (2024)
Learning to Reason Across Parallel Samples for LLM Reasoning
by: Qi, Jianing, et al.
Published: (2025)
by: Qi, Jianing, et al.
Published: (2025)
DiffAdapt: Difficulty-Adaptive Reasoning for Token-Efficient LLM Inference
by: Liu, Xiang, et al.
Published: (2025)
by: Liu, Xiang, et al.
Published: (2025)
Complex Claim Verification with Evidence Retrieved in the Wild
by: Chen, Jifan, et al.
Published: (2023)
by: Chen, Jifan, et al.
Published: (2023)
The Reasoning-Memorization Interplay in Language Models Is Mediated by a Single Direction
by: Hong, Yihuai, et al.
Published: (2025)
by: Hong, Yihuai, et al.
Published: (2025)
From Distributional to Overton Pluralism: Investigating Large Language Model Alignment
by: Lake, Thom, et al.
Published: (2024)
by: Lake, Thom, et al.
Published: (2024)
On Language Models' Sensitivity to Suspicious Coincidences
by: Padmanabhan, Sriram, et al.
Published: (2025)
by: Padmanabhan, Sriram, et al.
Published: (2025)
StockBot 2.0: Vanilla LSTMs Outperform Transformer-based Forecasting for Stock Prices
by: Mohanty, Shaswat
Published: (2026)
by: Mohanty, Shaswat
Published: (2026)
Contrastive Learning to Improve Retrieval for Real-world Fact Checking
by: Sriram, Aniruddh, et al.
Published: (2024)
by: Sriram, Aniruddh, et al.
Published: (2024)
Exploring Design Choices for Building Language-Specific LLMs
by: Tejaswi, Atula, et al.
Published: (2024)
by: Tejaswi, Atula, et al.
Published: (2024)
Mitigating Temporal Misalignment by Discarding Outdated Facts
by: Zhang, Michael J. Q., et al.
Published: (2023)
by: Zhang, Michael J. Q., et al.
Published: (2023)
Hybrid Latent Reasoning via Reinforcement Learning
by: Yue, Zhenrui, et al.
Published: (2025)
by: Yue, Zhenrui, et al.
Published: (2025)
Bridging Relevance and Reasoning: Rationale Distillation in Retrieval-Augmented Generation
by: Jia, Pengyue, et al.
Published: (2024)
by: Jia, Pengyue, et al.
Published: (2024)
ParaSpeechCLAP: A Dual-Encoder Speech-Text Model for Rich Stylistic Language-Audio Pretraining
by: Diwan, Anuj, et al.
Published: (2026)
by: Diwan, Anuj, et al.
Published: (2026)
BAT: Learning to Reason about Spatial Sounds with Large Language Models
by: Zheng, Zhisheng, et al.
Published: (2024)
by: Zheng, Zhisheng, et al.
Published: (2024)
Not All Heads Matter: A Head-Level KV Cache Compression Method with Integrated Retrieval and Reasoning
by: Fu, Yu, et al.
Published: (2024)
by: Fu, Yu, et al.
Published: (2024)
Retrieval Heads are Dynamic
by: Lin, Yuping, et al.
Published: (2026)
by: Lin, Yuping, et al.
Published: (2026)
BRIEF: Bridging Retrieval and Inference for Multi-hop Reasoning via Compression
by: Li, Yuankai, et al.
Published: (2024)
by: Li, Yuankai, et al.
Published: (2024)
Precise In-Parameter Concept Erasure in Large Language Models
by: Gur-Arieh, Yoav, et al.
Published: (2025)
by: Gur-Arieh, Yoav, et al.
Published: (2025)
Latent Thoughts Tuning: Bridging Context and Reasoning with Fused Information in Latent Tokens
by: Liu, Weihao, et al.
Published: (2026)
by: Liu, Weihao, et al.
Published: (2026)
Rhapsody: A Dataset for Highlight Detection in Podcasts
by: Park, Younghan, et al.
Published: (2025)
by: Park, Younghan, et al.
Published: (2025)
Improving LLM-as-a-Judge Inference with the Judgment Distribution
by: Wang, Victor, et al.
Published: (2025)
by: Wang, Victor, et al.
Published: (2025)
Crafting In-context Examples according to LMs' Parametric Knowledge
by: Lee, Yoonsang, et al.
Published: (2023)
by: Lee, Yoonsang, et al.
Published: (2023)
User Feedback in Human-LLM Dialogues: A Lens to Understand Users But Noisy as a Learning Signal
by: Liu, Yuhan, et al.
Published: (2025)
by: Liu, Yuhan, et al.
Published: (2025)
Failing to Falsify: Evaluating and Mitigating Confirmation Bias in Language Models
by: Jhaveri, Ayush Rajesh, et al.
Published: (2026)
by: Jhaveri, Ayush Rajesh, et al.
Published: (2026)
Scaling Latent Reasoning via Looped Language Models
by: Zhu, Rui-Jie, et al.
Published: (2025)
by: Zhu, Rui-Jie, et al.
Published: (2025)
LatentRAG: Latent Reasoning and Retrieval for Efficient Agentic RAG
by: Zheng, Yijia, et al.
Published: (2026)
by: Zheng, Yijia, et al.
Published: (2026)
The Rise of Parameter Specialization for Knowledge Storage in Large Language Models
by: Hong, Yihuai, et al.
Published: (2025)
by: Hong, Yihuai, et al.
Published: (2025)
Modeling Future Conversation Turns to Teach LLMs to Ask Clarifying Questions
by: Zhang, Michael J. Q., et al.
Published: (2024)
by: Zhang, Michael J. Q., et al.
Published: (2024)
Retrieval-Reasoning Large Language Model-based Synthetic Clinical Trial Generation
by: Xu, Zerui, et al.
Published: (2024)
by: Xu, Zerui, et al.
Published: (2024)
PCToolkit: A Unified Plug-and-Play Prompt Compression Toolkit of Large Language Models
by: Li, Jinyi, et al.
Published: (2024)
by: Li, Jinyi, et al.
Published: (2024)
Tool Retrieval Bridge: Aligning Vague Instructions with Retriever Preferences via Bridge Model
by: Chen, Kunfeng, et al.
Published: (2026)
by: Chen, Kunfeng, et al.
Published: (2026)
Improving Contextual Faithfulness of Large Language Models via Retrieval Heads-Induced Optimization
by: Huang, Lei, et al.
Published: (2025)
by: Huang, Lei, et al.
Published: (2025)
Similar Items
-
Beyond Single Embeddings: Capturing Diverse Targets with Multi-Query Retrieval
by: Chen, Hung-Ting, et al.
Published: (2025) -
Open-World Evaluation for Retrieving Diverse Perspectives
by: Chen, Hung-Ting, et al.
Published: (2024) -
CLaRa: Bridging Retrieval and Generation with Continuous Latent Reasoning
by: He, Jie, et al.
Published: (2025) -
RVR: Retrieve-Verify-Retrieve for Comprehensive Question Answering
by: Qian, Deniz, et al.
Published: (2026) -
AmbigDocs: Reasoning across Documents on Different Entities under the Same Name
by: Lee, Yoonsang, et al.
Published: (2024)