Learning to Construct Knowledge through Sparse Reference Selection with Reinforcement Learning
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Yin, Shao-An |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Influencing Bandits: Arm Selection for Preference Shaping
von: Nadkarni, Viraj, et al.
Veröffentlicht: (2024)
von: Nadkarni, Viraj, et al.
Veröffentlicht: (2024)
Perceptron Collaborative Filtering
von: Chakraborty, Arya
Veröffentlicht: (2024)
von: Chakraborty, Arya
Veröffentlicht: (2024)
LLMs in the Loop: Leveraging Large Language Model Annotations for Active Learning in Low-Resource Languages
von: Kholodna, Nataliia, et al.
Veröffentlicht: (2024)
von: Kholodna, Nataliia, et al.
Veröffentlicht: (2024)
Learning Efficient and Generalizable Graph Retriever for Knowledge-Graph Question Answering
von: Yao, Tianjun, et al.
Veröffentlicht: (2025)
von: Yao, Tianjun, et al.
Veröffentlicht: (2025)
Conversion rate prediction in online advertising: modeling techniques, performance evaluation and future directions
von: Xue, Tao, et al.
Veröffentlicht: (2025)
von: Xue, Tao, et al.
Veröffentlicht: (2025)
Toward a benchmark for CTR prediction in online advertising: datasets, evaluation protocols and perspectives
von: Gao, Shan, et al.
Veröffentlicht: (2025)
von: Gao, Shan, et al.
Veröffentlicht: (2025)
Suppressing Domain-Specific Hallucination in Construction LLMs: A Knowledge Graph Foundation for GraphRAG and QLoRA on River and Sediment Control Technical Standards
von: Yasuno, Takato
Veröffentlicht: (2026)
von: Yasuno, Takato
Veröffentlicht: (2026)
Low-pass Personalized Subgraph Federated Recommendation
von: Sim, Wooseok, et al.
Veröffentlicht: (2026)
von: Sim, Wooseok, et al.
Veröffentlicht: (2026)
SuperLocalMemory V3: Information-Geometric Foundations for Zero-LLM Enterprise Agent Memory
von: Bhardwaj, Varun Pratap
Veröffentlicht: (2026)
von: Bhardwaj, Varun Pratap
Veröffentlicht: (2026)
Accelerating Complex Disease Treatment through Network Medicine and GenAI: A Case Study on Drug Repurposing for Breast Cancer
von: Hamed, Ahmed Abdeen, et al.
Veröffentlicht: (2024)
von: Hamed, Ahmed Abdeen, et al.
Veröffentlicht: (2024)
Not All Memories Age the Same: Autodiscovery of Adaptive Decay in Knowledge Graphs
von: Karhade, Mandar
Veröffentlicht: (2026)
von: Karhade, Mandar
Veröffentlicht: (2026)
ATANT v1.1: Positioning Continuity Evaluation Against Memory, Long-Context, and Agentic-Memory Benchmarks
von: Tanguturi, Samuel Sameer
Veröffentlicht: (2026)
von: Tanguturi, Samuel Sameer
Veröffentlicht: (2026)
ATANT: An Evaluation Framework for AI Continuity
von: Tanguturi, Samuel Sameer
Veröffentlicht: (2026)
von: Tanguturi, Samuel Sameer
Veröffentlicht: (2026)
Learning to Detect Relevant Contexts and Knowledge for Response Selection in Retrieval-based Dialogue Systems
von: Hua, Kai, et al.
Veröffentlicht: (2025)
von: Hua, Kai, et al.
Veröffentlicht: (2025)
Expressive Value Learning for Scalable Offline Reinforcement Learning
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025)
von: Espinosa-Dice, Nicolas, et al.
Veröffentlicht: (2025)
Bounded Ratio Reinforcement Learning
von: Ao, Yunke, et al.
Veröffentlicht: (2026)
von: Ao, Yunke, et al.
Veröffentlicht: (2026)
Taming Polysemanticity in LLMs: Provable Feature Recovery via Sparse Autoencoders
von: Chen, Siyu, et al.
Veröffentlicht: (2025)
von: Chen, Siyu, et al.
Veröffentlicht: (2025)
SURE-RAG: Sufficiency and Uncertainty-Aware Evidence Verification for Selective Retrieval-Augmented Generation
von: Qiu, Jingxi, et al.
Veröffentlicht: (2026)
von: Qiu, Jingxi, et al.
Veröffentlicht: (2026)
Why Online Reinforcement Learning is Causal
von: Schulte, Oliver, et al.
Veröffentlicht: (2024)
von: Schulte, Oliver, et al.
Veröffentlicht: (2024)
SMOSE: Sparse Mixture of Shallow Experts for Interpretable Reinforcement Learning in Continuous Control Tasks
von: Vincze, Mátyás, et al.
Veröffentlicht: (2024)
von: Vincze, Mátyás, et al.
Veröffentlicht: (2024)
Counterfactual Evaluation of Ads Ranking Models through Domain Adaptation
von: Radwan, Mohamed A., et al.
Veröffentlicht: (2024)
von: Radwan, Mohamed A., et al.
Veröffentlicht: (2024)
Understanding Goal Generalisation in Sequential Reinforcement Learning
von: Brown, Jason Ross, et al.
Veröffentlicht: (2026)
von: Brown, Jason Ross, et al.
Veröffentlicht: (2026)
Path-Coupled Bellman Flows for Distributional Reinforcement Learning
von: Xu, Boyang, et al.
Veröffentlicht: (2026)
von: Xu, Boyang, et al.
Veröffentlicht: (2026)
No More Tuning: Prioritized Multi-Task Learning with Lagrangian Differential Multiplier Methods
von: Cheng, Zhengxing, et al.
Veröffentlicht: (2024)
von: Cheng, Zhengxing, et al.
Veröffentlicht: (2024)
GAP-Net: Calibrating User Intent via Gated Adaptive Progressive Learning for CTR Prediction
von: Ke, Shenqiang, et al.
Veröffentlicht: (2026)
von: Ke, Shenqiang, et al.
Veröffentlicht: (2026)
An Idiosyncrasy of Time-discretization in Reinforcement Learning
von: De Asis, Kris, et al.
Veröffentlicht: (2024)
von: De Asis, Kris, et al.
Veröffentlicht: (2024)
Multi-Task Reinforcement Learning with Language-Encoded Gated Policy Networks
von: Arora, Rushiv
Veröffentlicht: (2025)
von: Arora, Rushiv
Veröffentlicht: (2025)
Post-Training Fairness Control: A Single-Train Framework for Dynamic Fairness in Recommendation
von: Chen, Weixin, et al.
Veröffentlicht: (2026)
von: Chen, Weixin, et al.
Veröffentlicht: (2026)
Accelerating Recommendation System Training by Leveraging Popular Choices
von: Adnan, Muhammad, et al.
Veröffentlicht: (2021)
von: Adnan, Muhammad, et al.
Veröffentlicht: (2021)
Safe Reinforcement Learning with Preference-based Constraint Inference
von: Li, Chenglin, et al.
Veröffentlicht: (2026)
von: Li, Chenglin, et al.
Veröffentlicht: (2026)
Hard Negative Mining for Domain-Specific Retrieval in Enterprise Systems
von: Meghwani, Hansa, et al.
Veröffentlicht: (2025)
von: Meghwani, Hansa, et al.
Veröffentlicht: (2025)
CORE: Towards Scalable and Efficient Causal Discovery with Reinforcement Learning
von: Sauter, Andreas W. M., et al.
Veröffentlicht: (2024)
von: Sauter, Andreas W. M., et al.
Veröffentlicht: (2024)
CogCanvas: Verbatim-Grounded Artifact Extraction for Long LLM Conversations
von: An, Tao
Veröffentlicht: (2025)
von: An, Tao
Veröffentlicht: (2025)
Graph Your Way to Inspiration: Integrating Co-Author Graphs with Retrieval-Augmented Generation for Large Language Model Based Scientific Idea Generation
von: Xie, Pengzhen, et al.
Veröffentlicht: (2025)
von: Xie, Pengzhen, et al.
Veröffentlicht: (2025)
Deep Memory Search: A Metaheuristic Approach for Optimizing Heuristic Search
von: Hedar, Abdel-Rahman, et al.
Veröffentlicht: (2024)
von: Hedar, Abdel-Rahman, et al.
Veröffentlicht: (2024)
Control Reinforcement Learning: Interpretable Token-Level Steering of LLMs via Sparse Autoencoder Features
von: Cho, Seonglae, et al.
Veröffentlicht: (2026)
von: Cho, Seonglae, et al.
Veröffentlicht: (2026)
Task Adaptation from Skills: Information Geometry, Disentanglement, and New Objectives for Unsupervised Reinforcement Learning
von: Yang, Yucheng, et al.
Veröffentlicht: (2025)
von: Yang, Yucheng, et al.
Veröffentlicht: (2025)
When Actions Disappear: Adversarial Action Removal in Self-Play Reinforcement Learning
von: Kujur, Arahan
Veröffentlicht: (2026)
von: Kujur, Arahan
Veröffentlicht: (2026)
Umbrella Reinforcement Learning -- computationally efficient tool for hard non-linear problems
von: Nuzhin, Egor E., et al.
Veröffentlicht: (2024)
von: Nuzhin, Egor E., et al.
Veröffentlicht: (2024)
GIRL: Generative Imagination Reinforcement Learning via Information-Theoretic Hallucination Control
von: Hiremath, Prakul Sunil
Veröffentlicht: (2026)
von: Hiremath, Prakul Sunil
Veröffentlicht: (2026)
Ähnliche Einträge
-
Influencing Bandits: Arm Selection for Preference Shaping
von: Nadkarni, Viraj, et al.
Veröffentlicht: (2024) -
Perceptron Collaborative Filtering
von: Chakraborty, Arya
Veröffentlicht: (2024) -
LLMs in the Loop: Leveraging Large Language Model Annotations for Active Learning in Low-Resource Languages
von: Kholodna, Nataliia, et al.
Veröffentlicht: (2024) -
Learning Efficient and Generalizable Graph Retriever for Knowledge-Graph Question Answering
von: Yao, Tianjun, et al.
Veröffentlicht: (2025) -
Conversion rate prediction in online advertising: modeling techniques, performance evaluation and future directions
von: Xue, Tao, et al.
Veröffentlicht: (2025)