Reinforcement Fine-Tuning for History-Aware Dense Retriever in RAG
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Yicheng, Qin, Zhen, Wu, Zhaomin, Zhang, Wenqi, Deng, Shuiguang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Personalized Federated Fine-Tuning for LLMs via Data-Driven Heterogeneous Model Architectures
by: Zhang, Yicheng, et al.
Published: (2024)
by: Zhang, Yicheng, et al.
Published: (2024)
Federated Data-Efficient Instruction Tuning for Large Language Models
by: Qin, Zhen, et al.
Published: (2024)
by: Qin, Zhen, et al.
Published: (2024)
Federated Full-Parameter Tuning of Billion-Sized Language Models with Communication Cost under 18 Kilobytes
by: Qin, Zhen, et al.
Published: (2023)
by: Qin, Zhen, et al.
Published: (2023)
Resisting Backdoor Attacks in Federated Learning via Bidirectional Elections and Individual Perspective
by: Qin, Zhen, et al.
Published: (2023)
by: Qin, Zhen, et al.
Published: (2023)
FedRAG: A Framework for Fine-Tuning Retrieval-Augmented Generation Systems
by: Fajardo, Val Andrei, et al.
Published: (2025)
by: Fajardo, Val Andrei, et al.
Published: (2025)
LARA: A Light and Anti-overfitting Retraining Approach for Unsupervised Time Series Anomaly Detection
by: Chen, Feiyi, et al.
Published: (2023)
by: Chen, Feiyi, et al.
Published: (2023)
Communication-Aware Knowledge Distillation for Federated LLM Fine-Tuning over Wireless Networks
by: Zhang, Xinlu, et al.
Published: (2025)
by: Zhang, Xinlu, et al.
Published: (2025)
Prompt Tuning for Natural Language to SQL with Embedding Fine-Tuning and RAG
by: Jang, Jisoo, et al.
Published: (2025)
by: Jang, Jisoo, et al.
Published: (2025)
VectorLiteRAG: Latency-Aware and Fine-Grained Resource Partitioning for Efficient RAG
by: Kim, Junkyum, et al.
Published: (2025)
by: Kim, Junkyum, et al.
Published: (2025)
Supervised Fine Tuning on Curated Data is Reinforcement Learning (and can be improved)
by: Qin, Chongli, et al.
Published: (2025)
by: Qin, Chongli, et al.
Published: (2025)
TADS: Task-Aware Data Selection for Multi-Task Multimodal Pre-Training
by: Cheng, Guanjie, et al.
Published: (2026)
by: Cheng, Guanjie, et al.
Published: (2026)
Retrieval & Fine-Tuning for In-Context Tabular Models
by: Thomas, Valentin, et al.
Published: (2024)
by: Thomas, Valentin, et al.
Published: (2024)
GraphRAG-R1: Graph Retrieval-Augmented Generation with Process-Constrained Reinforcement Learning
by: Yu, Chuanyue, et al.
Published: (2025)
by: Yu, Chuanyue, et al.
Published: (2025)
SAFA-SNN: Sparsity-Aware On-Device Few-Shot Class-Incremental Learning with Fast-Adaptive Structure of Spiking Neural Network
by: Zhang, Huijing, et al.
Published: (2025)
by: Zhang, Huijing, et al.
Published: (2025)
Disco-RAG: Discourse-Aware Retrieval-Augmented Generation
by: Liu, Dongqi, et al.
Published: (2026)
by: Liu, Dongqi, et al.
Published: (2026)
W-RAG: Weakly Supervised Dense Retrieval in RAG for Open-domain Question Answering
by: Nian, Jinming, et al.
Published: (2024)
by: Nian, Jinming, et al.
Published: (2024)
On the Entropy Dynamics in Reinforcement Fine-Tuning of Large Language Models
by: Wang, Shumin, et al.
Published: (2026)
by: Wang, Shumin, et al.
Published: (2026)
A Heavy-Load-Enhanced and Changeable-Periodicity-Perceived Workload Prediction Network
by: Chen, Feiyi, et al.
Published: (2023)
by: Chen, Feiyi, et al.
Published: (2023)
Entropy Polarity in Reinforcement Fine-Tuning: Direction, Asymmetry, and Control
by: Zhang, Jiazheng, et al.
Published: (2026)
by: Zhang, Jiazheng, et al.
Published: (2026)
GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification
by: Gan, Wangjie, et al.
Published: (2026)
by: Gan, Wangjie, et al.
Published: (2026)
MARFT: Multi-Agent Reinforcement Fine-Tuning
by: Liao, Junwei, et al.
Published: (2025)
by: Liao, Junwei, et al.
Published: (2025)
All Roads Lead to Likelihood: The Value of Reinforcement Learning in Fine-Tuning
by: Swamy, Gokul, et al.
Published: (2025)
by: Swamy, Gokul, et al.
Published: (2025)
FuRA: Full-Rank Parameter-Efficient Fine-Tuning with Spectral Preconditioning
by: Zhao, Yequan, et al.
Published: (2026)
by: Zhao, Yequan, et al.
Published: (2026)
RAG in the Wild: On the (In)effectiveness of LLMs with Mixture-of-Knowledge Retrieval Augmentation
by: Xu, Ran, et al.
Published: (2025)
by: Xu, Ran, et al.
Published: (2025)
Empowering Large Language Models in Wireless Communication: A Novel Dataset and Fine-Tuning Framework
by: Lin, Yushen, et al.
Published: (2025)
by: Lin, Yushen, et al.
Published: (2025)
Internalizing Curriculum Judgment for LLM Reinforcement Fine-Tuning
by: Zheng, Han, et al.
Published: (2026)
by: Zheng, Han, et al.
Published: (2026)
Your Dense Retriever is Secretly an Expeditious Reasoner
by: Zhang, Yichi, et al.
Published: (2025)
by: Zhang, Yichi, et al.
Published: (2025)
Soft Prompt Tuning for Augmenting Dense Retrieval with Large Language Models
by: Peng, Zhiyuan, et al.
Published: (2023)
by: Peng, Zhiyuan, et al.
Published: (2023)
Secure LLM Fine-Tuning via Safety-Aware Probing
by: Wu, Chengcan, et al.
Published: (2025)
by: Wu, Chengcan, et al.
Published: (2025)
Learning What Reinforcement Learning Can't: Interleaved Online Fine-Tuning for Hardest Questions
by: Ma, Lu, et al.
Published: (2025)
by: Ma, Lu, et al.
Published: (2025)
LoTA-QAF: Lossless Ternary Adaptation for Quantization-Aware Fine-Tuning
by: Chen, Junyu, et al.
Published: (2025)
by: Chen, Junyu, et al.
Published: (2025)
Curvature-Aware Safety Restoration In LLMs Fine-Tuning
by: Bach, Thong, et al.
Published: (2025)
by: Bach, Thong, et al.
Published: (2025)
WikiDBGraph: A Data Management Benchmark Suite for Collaborative Learning over Database Silos
by: Wu, Zhaomin, et al.
Published: (2025)
by: Wu, Zhaomin, et al.
Published: (2025)
Fine-Tuning vs. RAG for Multi-Hop Question Answering with Novel Knowledge
by: Yang, Zhuoyi, et al.
Published: (2026)
by: Yang, Zhuoyi, et al.
Published: (2026)
Vertical Federated Learning in Practice: The Good, the Bad, and the Ugly
by: Wu, Zhaomin, et al.
Published: (2025)
by: Wu, Zhaomin, et al.
Published: (2025)
Domain-Aware Fine-Tuning of Foundation Models
by: Kaplan, Ugur Ali, et al.
Published: (2024)
by: Kaplan, Ugur Ali, et al.
Published: (2024)
MaZO: Masked Zeroth-Order Optimization for Multi-Task Fine-Tuning of Large Language Models
by: Zhang, Zhen, et al.
Published: (2025)
by: Zhang, Zhen, et al.
Published: (2025)
Continual Fine-Tuning with Provably Accurate and Parameter-Free Task Retrieval
by: Le, Hang Thi-Thuy, et al.
Published: (2026)
by: Le, Hang Thi-Thuy, et al.
Published: (2026)
ALoFTRAG: Automatic Local Fine Tuning for Retrieval Augmented Generation
by: Devine, Peter
Published: (2025)
by: Devine, Peter
Published: (2025)
UFT: Unifying Supervised and Reinforcement Fine-Tuning
by: Liu, Mingyang, et al.
Published: (2025)
by: Liu, Mingyang, et al.
Published: (2025)
Similar Items
-
Personalized Federated Fine-Tuning for LLMs via Data-Driven Heterogeneous Model Architectures
by: Zhang, Yicheng, et al.
Published: (2024) -
Federated Data-Efficient Instruction Tuning for Large Language Models
by: Qin, Zhen, et al.
Published: (2024) -
Federated Full-Parameter Tuning of Billion-Sized Language Models with Communication Cost under 18 Kilobytes
by: Qin, Zhen, et al.
Published: (2023) -
Resisting Backdoor Attacks in Federated Learning via Bidirectional Elections and Individual Perspective
by: Qin, Zhen, et al.
Published: (2023) -
FedRAG: A Framework for Fine-Tuning Retrieval-Augmented Generation Systems
by: Fajardo, Val Andrei, et al.
Published: (2025)