Utility-based Dueling Bandits as a Partial Monitoring Game
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gajane, Pratik, Urvoy, Tanguy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2015
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Best-of-Both-Worlds Multi-Dueling Bandits: Unified Algorithms for Stochastic and Adversarial Preferences under Condorcet and Borda Objectives
von: Akash, S, et al.
Veröffentlicht: (2026)
von: Akash, S, et al.
Veröffentlicht: (2026)
Adversarial Multi-dueling Bandits
von: Gajane, Pratik
Veröffentlicht: (2024)
von: Gajane, Pratik
Veröffentlicht: (2024)
Fairness in two-player zero-sum games with bandit feedback
von: Akash, S, et al.
Veröffentlicht: (2026)
von: Akash, S, et al.
Veröffentlicht: (2026)
Federated Linear Dueling Bandits
von: Huang, Xuhan, et al.
Veröffentlicht: (2025)
von: Huang, Xuhan, et al.
Veröffentlicht: (2025)
Online Clustering of Dueling Bandits
von: Wang, Zhiyong, et al.
Veröffentlicht: (2025)
von: Wang, Zhiyong, et al.
Veröffentlicht: (2025)
Multi-Player Approaches for Dueling Bandits
von: Raveh, Or, et al.
Veröffentlicht: (2024)
von: Raveh, Or, et al.
Veröffentlicht: (2024)
Evaluating Causal Discovery Algorithms for Path-Specific Fairness and Utility in Healthcare
von: Nagesh, Nitish, et al.
Veröffentlicht: (2026)
von: Nagesh, Nitish, et al.
Veröffentlicht: (2026)
Biased Dueling Bandits with Stochastic Delayed Feedback
von: Yi, Bongsoo, et al.
Veröffentlicht: (2024)
von: Yi, Bongsoo, et al.
Veröffentlicht: (2024)
The Sampling Complexity of Condorcet Winner Identification in Dueling Bandits
von: Saad, El Mehdi, et al.
Veröffentlicht: (2026)
von: Saad, El Mehdi, et al.
Veröffentlicht: (2026)
Conversational Dueling Bandits in Generalized Linear Models
von: Yang, Shuhua, et al.
Veröffentlicht: (2024)
von: Yang, Shuhua, et al.
Veröffentlicht: (2024)
Fusing Reward and Dueling Feedback in Stochastic Bandits
von: Wang, Xuchuang, et al.
Veröffentlicht: (2025)
von: Wang, Xuchuang, et al.
Veröffentlicht: (2025)
Linear and Neural Dueling Bandits with Delayed Feedback
von: Wang, Xiangyi, et al.
Veröffentlicht: (2026)
von: Wang, Xiangyi, et al.
Veröffentlicht: (2026)
Randomized Least Squares Value Iteration itself is Joint Differentially Private
von: Lu, Haiyang, et al.
Veröffentlicht: (2026)
von: Lu, Haiyang, et al.
Veröffentlicht: (2026)
Recycling History: Efficient Recommendations from Contextual Dueling Bandits
von: Sankagiri, Suryanarayana, et al.
Veröffentlicht: (2025)
von: Sankagiri, Suryanarayana, et al.
Veröffentlicht: (2025)
Non-Stationary Dueling Bandits Under a Weighted Borda Criterion
von: Suk, Joe, et al.
Veröffentlicht: (2024)
von: Suk, Joe, et al.
Veröffentlicht: (2024)
Feel-Good Thompson Sampling for Contextual Dueling Bandits
von: Li, Xuheng, et al.
Veröffentlicht: (2024)
von: Li, Xuheng, et al.
Veröffentlicht: (2024)
Investigating Gender Fairness in Machine Learning-driven Personalized Care for Chronic Pain
von: Gajane, Pratik, et al.
Veröffentlicht: (2024)
von: Gajane, Pratik, et al.
Veröffentlicht: (2024)
Active Human Feedback Collection via Neural Contextual Dueling Bandits
von: Verma, Arun, et al.
Veröffentlicht: (2025)
von: Verma, Arun, et al.
Veröffentlicht: (2025)
When Can We Track Significant Preference Shifts in Dueling Bandits?
von: Suk, Joe, et al.
Veröffentlicht: (2023)
von: Suk, Joe, et al.
Veröffentlicht: (2023)
Nearly Optimal Algorithms for Contextual Dueling Bandits from Adversarial Feedback
von: Di, Qiwei, et al.
Veröffentlicht: (2024)
von: Di, Qiwei, et al.
Veröffentlicht: (2024)
Neural Variance-aware Dueling Bandits with Deep Representation and Shallow Exploration
von: Oh, Youngmin, et al.
Veröffentlicht: (2025)
von: Oh, Youngmin, et al.
Veröffentlicht: (2025)
Neural Dueling Bandits: Preference-Based Optimization with Human Feedback
von: Verma, Arun, et al.
Veröffentlicht: (2024)
von: Verma, Arun, et al.
Veröffentlicht: (2024)
Variance-Aware Regret Bounds for Stochastic Contextual Dueling Bandits
von: Di, Qiwei, et al.
Veröffentlicht: (2023)
von: Di, Qiwei, et al.
Veröffentlicht: (2023)
Lipschitz Dueling Bandits over Continuous Action Spaces
von: Sharma, Mudit, et al.
Veröffentlicht: (2026)
von: Sharma, Mudit, et al.
Veröffentlicht: (2026)
Preference is More Than Comparisons: Rethinking Dueling Bandits with Augmented Human Feedback
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
von: Wang, Shengbo, et al.
Veröffentlicht: (2025)
Tabular Data Generation Models: An In-Depth Survey and Performance Benchmarks with Extensive Tuning
von: Kindji, G. Charbel N., et al.
Veröffentlicht: (2024)
von: Kindji, G. Charbel N., et al.
Veröffentlicht: (2024)
Beyond Numeric Rewards: In-Context Dueling Bandits with LLM Agents
von: Xia, Fanzeng, et al.
Veröffentlicht: (2024)
von: Xia, Fanzeng, et al.
Veröffentlicht: (2024)
Multi-User Dueling Bandits: A Fair Approach using Nash Social Welfare
von: Ahmed, Maheed H., et al.
Veröffentlicht: (2026)
von: Ahmed, Maheed H., et al.
Veröffentlicht: (2026)
Robust Linear Dueling Bandits with Post-serving Context under Unknown Delays and Adversarial Corruptions
von: Oh, Youngmin
Veröffentlicht: (2026)
von: Oh, Youngmin
Veröffentlicht: (2026)
Robust Detection of Synthetic Tabular Data under Schema Variability
von: Kindji, G. Charbel N., et al.
Veröffentlicht: (2025)
von: Kindji, G. Charbel N., et al.
Veröffentlicht: (2025)
Cross-table Synthetic Tabular Data Detection
von: Kindji, G. Charbel N., et al.
Veröffentlicht: (2024)
von: Kindji, G. Charbel N., et al.
Veröffentlicht: (2024)
Combinatorial Allocation Bandits with Nonlinear Arm Utility
von: Shibukawa, Yuki, et al.
Veröffentlicht: (2026)
von: Shibukawa, Yuki, et al.
Veröffentlicht: (2026)
Learning To Play Atari Games Using Dueling Q-Learning and Hebbian Plasticity
von: Salehin, Md Ashfaq
Veröffentlicht: (2024)
von: Salehin, Md Ashfaq
Veröffentlicht: (2024)
LLM Routing with Dueling Feedback
von: Chiang, Chao-Kai, et al.
Veröffentlicht: (2025)
von: Chiang, Chao-Kai, et al.
Veröffentlicht: (2025)
Synthetic Tabular Data Detection In the Wild
von: Kindji, G. Charbel N., et al.
Veröffentlicht: (2025)
von: Kindji, G. Charbel N., et al.
Veröffentlicht: (2025)
Datum-wise Transformer for Synthetic Tabular Data Detection in the Wild
von: Kindji, G. Charbel N., et al.
Veröffentlicht: (2025)
von: Kindji, G. Charbel N., et al.
Veröffentlicht: (2025)
Riemannian Dueling Optimization
von: Ren, Yuxuan, et al.
Veröffentlicht: (2026)
von: Ren, Yuxuan, et al.
Veröffentlicht: (2026)
Linear Bandits with Partially Observable Features
von: Kim, Wonyoung, et al.
Veröffentlicht: (2025)
von: Kim, Wonyoung, et al.
Veröffentlicht: (2025)
Tractable Multinomial Logit Contextual Bandits with Non-Linear Utilities
von: Hwang, Taehyun, et al.
Veröffentlicht: (2026)
von: Hwang, Taehyun, et al.
Veröffentlicht: (2026)
Thompson Sampling in Partially Observable Contextual Bandits
von: Park, Hongju, et al.
Veröffentlicht: (2024)
von: Park, Hongju, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Best-of-Both-Worlds Multi-Dueling Bandits: Unified Algorithms for Stochastic and Adversarial Preferences under Condorcet and Borda Objectives
von: Akash, S, et al.
Veröffentlicht: (2026) -
Adversarial Multi-dueling Bandits
von: Gajane, Pratik
Veröffentlicht: (2024) -
Fairness in two-player zero-sum games with bandit feedback
von: Akash, S, et al.
Veröffentlicht: (2026) -
Federated Linear Dueling Bandits
von: Huang, Xuhan, et al.
Veröffentlicht: (2025) -
Online Clustering of Dueling Bandits
von: Wang, Zhiyong, et al.
Veröffentlicht: (2025)