Active Query Synthesis for Preference Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Nadagouda, Namrata, Ahad, Nauman, Tucker, Maegan, Davenport, Mark A. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Automata Learning from Preference and Equivalence Queries
by: Hsiung, Eric, et al.
Published: (2023)
by: Hsiung, Eric, et al.
Published: (2023)
Active Learning of General Halfspaces: Label Queries vs Membership Queries
by: Diakonikolas, Ilias, et al.
Published: (2024)
by: Diakonikolas, Ilias, et al.
Published: (2024)
Query-Policy Misalignment in Preference-Based Reinforcement Learning
by: Hu, Xiao, et al.
Published: (2023)
by: Hu, Xiao, et al.
Published: (2023)
Active Learning for Direct Preference Optimization
by: Kveton, Branislav, et al.
Published: (2025)
by: Kveton, Branislav, et al.
Published: (2025)
SGD Jittering: A Training Strategy for Robust and Accurate Model-Based Architectures
by: Guan, Peimeng, et al.
Published: (2024)
by: Guan, Peimeng, et al.
Published: (2024)
CLARIFY: Contrastive Preference Reinforcement Learning for Untangling Ambiguous Queries
by: Mu, Ni, et al.
Published: (2025)
by: Mu, Ni, et al.
Published: (2025)
Active Query Selection for Crowd-Based Reinforcement Learning
by: Erskine, Jonathan, et al.
Published: (2025)
by: Erskine, Jonathan, et al.
Published: (2025)
Preferential Multi-Objective Bayesian Optimization
by: Astudillo, Raul, et al.
Published: (2024)
by: Astudillo, Raul, et al.
Published: (2024)
Reward-Conditioned Reinforcement Learning
by: Nauman, Michal, et al.
Published: (2026)
by: Nauman, Michal, et al.
Published: (2026)
Offline Clustering of Preference Learning with Active-data Augmentation
by: Liu, Jingyuan, et al.
Published: (2025)
by: Liu, Jingyuan, et al.
Published: (2025)
Active Preference Learning for Ordering Items In- and Out-of-sample
by: Bergström, Herman, et al.
Published: (2024)
by: Bergström, Herman, et al.
Published: (2024)
Cumulative Meta-Learning from Active Learning Queries for Robustness to Spurious Correlations
by: Chew, Kin Whye, et al.
Published: (2026)
by: Chew, Kin Whye, et al.
Published: (2026)
Conformal confidence sets for biomedical image segmentation
by: Davenport, Samuel
Published: (2024)
by: Davenport, Samuel
Published: (2024)
AutoAL: Automated Active Learning with Differentiable Query Strategy Search
by: Wang, Yifeng, et al.
Published: (2024)
by: Wang, Yifeng, et al.
Published: (2024)
Cold-Start Active Preference Learning in Socio-Economic Domains
by: Fayaz-Bakhsh, Mojtaba, et al.
Published: (2025)
by: Fayaz-Bakhsh, Mojtaba, et al.
Published: (2025)
Comparison-based Active Preference Learning for Multi-dimensional Personalization
by: Oh, Minhyeon, et al.
Published: (2024)
by: Oh, Minhyeon, et al.
Published: (2024)
Learning Formal Specifications from Membership and Preference Queries
by: Shah, Ameesh, et al.
Published: (2023)
by: Shah, Ameesh, et al.
Published: (2023)
On the Theory of Risk-Aware Agents: Bridging Actor-Critic and Economics
by: Nauman, Michal, et al.
Published: (2023)
by: Nauman, Michal, et al.
Published: (2023)
Low-rank matrix completion and denoising under Poisson noise
by: McRae, Andrew D., et al.
Published: (2019)
by: McRae, Andrew D., et al.
Published: (2019)
Active Value Querying to Minimize Additive Error in Subadditive Set Function Learning
by: Černý, Martin, et al.
Published: (2026)
by: Černý, Martin, et al.
Published: (2026)
Fast and Sample Efficient Multi-Task Representation Learning in Stochastic Contextual Bandits
by: Lin, Jiabin, et al.
Published: (2024)
by: Lin, Jiabin, et al.
Published: (2024)
Nearly Optimal Active Preference Learning and Its Application to LLM Alignment
by: Zhao, Yao, et al.
Published: (2026)
by: Zhao, Yao, et al.
Published: (2026)
Querying Easily Flip-flopped Samples for Deep Active Learning
by: Cho, Seong Jin, et al.
Published: (2024)
by: Cho, Seong Jin, et al.
Published: (2024)
What Does Flow Matching Bring To TD Learning?
by: Agrawalla, Bhavya, et al.
Published: (2026)
by: Agrawalla, Bhavya, et al.
Published: (2026)
AltGDmin: Alternating GD and Minimization for Partly-Decoupled (Federated) Optimization
by: Vaswani, Namrata
Published: (2025)
by: Vaswani, Namrata
Published: (2025)
Active Preference Learning for Large Language Models
by: Muldrew, William, et al.
Published: (2024)
by: Muldrew, William, et al.
Published: (2024)
Inference-Time Personalized Alignment with a Few User Preference Queries
by: Pădurean, Victor-Alexandru, et al.
Published: (2025)
by: Pădurean, Victor-Alexandru, et al.
Published: (2025)
Sharpe Ratio-Guided Active Learning for Preference Optimization in RLHF
by: Belakaria, Syrine, et al.
Published: (2025)
by: Belakaria, Syrine, et al.
Published: (2025)
Efficient Training of Deep Networks using Guided Spectral Data Selection: A Step Toward Learning What You Need
by: Sharifi, Mohammadreza, et al.
Published: (2025)
by: Sharifi, Mohammadreza, et al.
Published: (2025)
ActiveUltraFeedback: Efficient Preference Data Generation using Active Learning
by: Melikidze, Davit, et al.
Published: (2026)
by: Melikidze, Davit, et al.
Published: (2026)
Optimizing Algorithms for Mobile Health Interventions with Active Querying Optimization
by: Rawashdeh, Aseel
Published: (2025)
by: Rawashdeh, Aseel
Published: (2025)
LORE: Jointly Learning the Intrinsic Dimensionality and Relative Similarity Structure From Ordinal Data
by: Anand, Vivek, et al.
Published: (2026)
by: Anand, Vivek, et al.
Published: (2026)
Enhancing Cost Efficiency in Active Learning with Candidate Set Query
by: Gwon, Yeho, et al.
Published: (2025)
by: Gwon, Yeho, et al.
Published: (2025)
Deep Bayesian Active Learning for Preference Modeling in Large Language Models
by: Melo, Luckeciano C., et al.
Published: (2024)
by: Melo, Luckeciano C., et al.
Published: (2024)
Batch Active Learning of Reward Functions from Human Preferences
by: Bıyık, Erdem, et al.
Published: (2024)
by: Bıyık, Erdem, et al.
Published: (2024)
Solving Inverse Problems with Model Mismatch using Untrained Neural Networks within Model-based Architectures
by: Guan, Peimeng, et al.
Published: (2024)
by: Guan, Peimeng, et al.
Published: (2024)
A Deep Positive-Negative Prototype Approach to Integrated Prototypical Discriminative Learning
by: Zarei-Sabzevar, Ramin, et al.
Published: (2025)
by: Zarei-Sabzevar, Ramin, et al.
Published: (2025)
DynaSTy: A Framework for SpatioTemporal Node Attribute Prediction in Dynamic Graphs
by: Banerji, Namrata, et al.
Published: (2026)
by: Banerji, Namrata, et al.
Published: (2026)
A Case for Validation Buffer in Pessimistic Actor-Critic
by: Nauman, Michal, et al.
Published: (2024)
by: Nauman, Michal, et al.
Published: (2024)
Reinforcement Learning from Human Feedback with Active Queries
by: Ji, Kaixuan, et al.
Published: (2024)
by: Ji, Kaixuan, et al.
Published: (2024)
Similar Items
-
Automata Learning from Preference and Equivalence Queries
by: Hsiung, Eric, et al.
Published: (2023) -
Active Learning of General Halfspaces: Label Queries vs Membership Queries
by: Diakonikolas, Ilias, et al.
Published: (2024) -
Query-Policy Misalignment in Preference-Based Reinforcement Learning
by: Hu, Xiao, et al.
Published: (2023) -
Active Learning for Direct Preference Optimization
by: Kveton, Branislav, et al.
Published: (2025) -
SGD Jittering: A Training Strategy for Robust and Accurate Model-Based Architectures
by: Guan, Peimeng, et al.
Published: (2024)