Influencing Bandits: Arm Selection for Preference Shaping
Fuente:
arXiv
Saved in:
| Main Authors: | Nadkarni, Viraj, Manjunath, D., Moharir, Sharayu |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning to Construct Knowledge through Sparse Reference Selection with Reinforcement Learning
by: Yin, Shao-An
Published: (2025)
by: Yin, Shao-An
Published: (2025)
Towards Opinion Shaping: A Deep Reinforcement Learning Approach in Bot-User Interactions
by: Siahkali, Farbod, et al.
Published: (2024)
by: Siahkali, Farbod, et al.
Published: (2024)
Perceptron Collaborative Filtering
by: Chakraborty, Arya
Published: (2024)
by: Chakraborty, Arya
Published: (2024)
Conversion rate prediction in online advertising: modeling techniques, performance evaluation and future directions
by: Xue, Tao, et al.
Published: (2025)
by: Xue, Tao, et al.
Published: (2025)
Toward a benchmark for CTR prediction in online advertising: datasets, evaluation protocols and perspectives
by: Gao, Shan, et al.
Published: (2025)
by: Gao, Shan, et al.
Published: (2025)
LLMs in the Loop: Leveraging Large Language Model Annotations for Active Learning in Low-Resource Languages
by: Kholodna, Nataliia, et al.
Published: (2024)
by: Kholodna, Nataliia, et al.
Published: (2024)
Gaussian process-based online health monitoring and fault analysis of lithium-ion battery systems from field data
by: Schaeffer, Joachim, et al.
Published: (2024)
by: Schaeffer, Joachim, et al.
Published: (2024)
DSSE: a drone swarm search environment
by: Castanares, Manuel, et al.
Published: (2023)
by: Castanares, Manuel, et al.
Published: (2023)
Accuracy, Memory Efficiency and Generalization: A Comparative Study on Liquid Neural Networks and Recurrent Neural Networks
by: Zong, Shilong, et al.
Published: (2025)
by: Zong, Shilong, et al.
Published: (2025)
Low-pass Personalized Subgraph Federated Recommendation
by: Sim, Wooseok, et al.
Published: (2026)
by: Sim, Wooseok, et al.
Published: (2026)
SuperLocalMemory V3: Information-Geometric Foundations for Zero-LLM Enterprise Agent Memory
by: Bhardwaj, Varun Pratap
Published: (2026)
by: Bhardwaj, Varun Pratap
Published: (2026)
ATANT v1.1: Positioning Continuity Evaluation Against Memory, Long-Context, and Agentic-Memory Benchmarks
by: Tanguturi, Samuel Sameer
Published: (2026)
by: Tanguturi, Samuel Sameer
Published: (2026)
ATANT: An Evaluation Framework for AI Continuity
by: Tanguturi, Samuel Sameer
Published: (2026)
by: Tanguturi, Samuel Sameer
Published: (2026)
Diffusion-Based Scenario Tree Generation for Multivariate Time Series Prediction and Multistage Stochastic Optimization
by: Zarifis, Stelios, et al.
Published: (2025)
by: Zarifis, Stelios, et al.
Published: (2025)
Diffusion-Based Forecasting for Uncertainty-Aware Model Predictive Control
by: Zarifis, Stelios, et al.
Published: (2025)
by: Zarifis, Stelios, et al.
Published: (2025)
Developing the Reliable Shallow Supervised Learning for Thermal Comfort using ASHRAE RP-884 and ASHRAE Global Thermal Comfort Database II
by: Karyono, Kanisius, et al.
Published: (2023)
by: Karyono, Kanisius, et al.
Published: (2023)
Modeling Nonlinear Oscillator Networks Using Physics-Informed Hybrid Reservoir Computing
by: Shannon, Andrew, et al.
Published: (2024)
by: Shannon, Andrew, et al.
Published: (2024)
SURE-RAG: Sufficiency and Uncertainty-Aware Evidence Verification for Selective Retrieval-Augmented Generation
by: Qiu, Jingxi, et al.
Published: (2026)
by: Qiu, Jingxi, et al.
Published: (2026)
Accelerating Complex Disease Treatment through Network Medicine and GenAI: A Case Study on Drug Repurposing for Breast Cancer
by: Hamed, Ahmed Abdeen, et al.
Published: (2024)
by: Hamed, Ahmed Abdeen, et al.
Published: (2024)
Criticality and Safety Margins for Reinforcement Learning
by: Grushin, Alexander, et al.
Published: (2024)
by: Grushin, Alexander, et al.
Published: (2024)
Safety Margins for Reinforcement Learning
by: Grushin, Alexander, et al.
Published: (2023)
by: Grushin, Alexander, et al.
Published: (2023)
PCA-Driven Adaptive Sensor Triage for Edge AI Inference
by: Lade, Ankit Hemant, et al.
Published: (2026)
by: Lade, Ankit Hemant, et al.
Published: (2026)
Learning Approximate Nash Equilibria in Cooperative Multi-Agent Reinforcement Learning via Mean-Field Subsampling
by: Anand, Emile, et al.
Published: (2026)
by: Anand, Emile, et al.
Published: (2026)
Not All Memories Age the Same: Autodiscovery of Adaptive Decay in Knowledge Graphs
by: Karhade, Mandar
Published: (2026)
by: Karhade, Mandar
Published: (2026)
Post-Training Fairness Control: A Single-Train Framework for Dynamic Fairness in Recommendation
by: Chen, Weixin, et al.
Published: (2026)
by: Chen, Weixin, et al.
Published: (2026)
Accelerating Recommendation System Training by Leveraging Popular Choices
by: Adnan, Muhammad, et al.
Published: (2021)
by: Adnan, Muhammad, et al.
Published: (2021)
Differentiable Stochastic Traffic Dynamics: Physics-Informed Generative Modelling in Transportation
by: Xin, Wuping
Published: (2026)
by: Xin, Wuping
Published: (2026)
Hard Negative Mining for Domain-Specific Retrieval in Enterprise Systems
by: Meghwani, Hansa, et al.
Published: (2025)
by: Meghwani, Hansa, et al.
Published: (2025)
Learning Efficient and Generalizable Graph Retriever for Knowledge-Graph Question Answering
by: Yao, Tianjun, et al.
Published: (2025)
by: Yao, Tianjun, et al.
Published: (2025)
Generalization of Graph Neural Network Models for Distribution Grid Fault Detection
by: Karabulut, Burak, et al.
Published: (2025)
by: Karabulut, Burak, et al.
Published: (2025)
Physics-inspired Neural Networks for Parameter Learning of Adaptive Cruise Control Systems
by: Apostolakis, Theocharis, et al.
Published: (2023)
by: Apostolakis, Theocharis, et al.
Published: (2023)
Suppressing Domain-Specific Hallucination in Construction LLMs: A Knowledge Graph Foundation for GraphRAG and QLoRA on River and Sediment Control Technical Standards
by: Yasuno, Takato
Published: (2026)
by: Yasuno, Takato
Published: (2026)
CogCanvas: Verbatim-Grounded Artifact Extraction for Long LLM Conversations
by: An, Tao
Published: (2025)
by: An, Tao
Published: (2025)
Graph Your Way to Inspiration: Integrating Co-Author Graphs with Retrieval-Augmented Generation for Large Language Model Based Scientific Idea Generation
by: Xie, Pengzhen, et al.
Published: (2025)
by: Xie, Pengzhen, et al.
Published: (2025)
Deep Memory Search: A Metaheuristic Approach for Optimizing Heuristic Search
by: Hedar, Abdel-Rahman, et al.
Published: (2024)
by: Hedar, Abdel-Rahman, et al.
Published: (2024)
Democratic Preference Alignment via Sortition-Weighted RLHF
by: Sana, Suvadip, et al.
Published: (2026)
by: Sana, Suvadip, et al.
Published: (2026)
Improved Accuracy of Robot Localization Using 3-D LiDAR in a Hippocampus-Inspired Model
by: Gerstenslager, Andrew, et al.
Published: (2025)
by: Gerstenslager, Andrew, et al.
Published: (2025)
Predictive Maintenance in Photovoltaic Plants with a Big Data Approach
by: Betti, Alessandro, et al.
Published: (2019)
by: Betti, Alessandro, et al.
Published: (2019)
Counterfactual Evaluation of Ads Ranking Models through Domain Adaptation
by: Radwan, Mohamed A., et al.
Published: (2024)
by: Radwan, Mohamed A., et al.
Published: (2024)
CogRec: A Cognitive Recommender Agent Fusing Large Language Models and Soar for Explainable Recommendation
by: Hu, Jiaxin, et al.
Published: (2025)
by: Hu, Jiaxin, et al.
Published: (2025)
Similar Items
-
Learning to Construct Knowledge through Sparse Reference Selection with Reinforcement Learning
by: Yin, Shao-An
Published: (2025) -
Towards Opinion Shaping: A Deep Reinforcement Learning Approach in Bot-User Interactions
by: Siahkali, Farbod, et al.
Published: (2024) -
Perceptron Collaborative Filtering
by: Chakraborty, Arya
Published: (2024) -
Conversion rate prediction in online advertising: modeling techniques, performance evaluation and future directions
by: Xue, Tao, et al.
Published: (2025) -
Toward a benchmark for CTR prediction in online advertising: datasets, evaluation protocols and perspectives
by: Gao, Shan, et al.
Published: (2025)