Enhancing Preference-based Linear Bandits via Human Response Time
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Shen, Zhang, Yuyang, Ren, Zhaolin, Liang, Claire, Li, Na, Shah, Julie A. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Twin-2K-500: A dataset for building digital twins of over 2,000 people based on their answers to over 500 questions
di: Toubia, Olivier, et al.
Pubblicazione: (2025)
di: Toubia, Olivier, et al.
Pubblicazione: (2025)
Designing Algorithmic Recommendations to Achieve Human-AI Complementarity
di: McLaughlin, Bryce, et al.
Pubblicazione: (2024)
di: McLaughlin, Bryce, et al.
Pubblicazione: (2024)
Wikipedia Contributions in the Wake of ChatGPT
di: Lyu, Liang, et al.
Pubblicazione: (2025)
di: Lyu, Liang, et al.
Pubblicazione: (2025)
The Economics of AI Supply Chain Regulation
di: Qian, Sihan, et al.
Pubblicazione: (2026)
di: Qian, Sihan, et al.
Pubblicazione: (2026)
Understanding the decision-making process of choice modellers
di: Nova, Gabriel, et al.
Pubblicazione: (2024)
di: Nova, Gabriel, et al.
Pubblicazione: (2024)
Inference-Time Policy Steering through Human Interactions
di: Wang, Yanwei, et al.
Pubblicazione: (2024)
di: Wang, Yanwei, et al.
Pubblicazione: (2024)
Algorithmic Feature Highlighting for Human-AI Decision-Making
di: Guo, Yifan, et al.
Pubblicazione: (2026)
di: Guo, Yifan, et al.
Pubblicazione: (2026)
Making RL with Preference-based Feedback Efficient via Randomization
di: Wu, Runzhe, et al.
Pubblicazione: (2023)
di: Wu, Runzhe, et al.
Pubblicazione: (2023)
Influencing Humans to Conform to Preference Models for RLHF
di: Hatgis-Kessell, Stephane, et al.
Pubblicazione: (2025)
di: Hatgis-Kessell, Stephane, et al.
Pubblicazione: (2025)
Inference of Human-derived Specifications of Object Placement via Demonstration
di: Cuellar, Alex, et al.
Pubblicazione: (2025)
di: Cuellar, Alex, et al.
Pubblicazione: (2025)
Hindsight PRIORs for Reward Learning from Human Preferences
di: Verma, Mudit, et al.
Pubblicazione: (2024)
di: Verma, Mudit, et al.
Pubblicazione: (2024)
Human Expertise in Algorithmic Prediction
di: Alur, Rohan, et al.
Pubblicazione: (2024)
di: Alur, Rohan, et al.
Pubblicazione: (2024)
Vibe Econometrics and the Analysis Contract
di: Ashton, Lydia
Pubblicazione: (2026)
di: Ashton, Lydia
Pubblicazione: (2026)
Learning a Canonical Basis of Human Preferences from Binary Ratings
di: Vodrahalli, Kailas, et al.
Pubblicazione: (2025)
di: Vodrahalli, Kailas, et al.
Pubblicazione: (2025)
FARPLS: A Feature-Augmented Robot Trajectory Preference Labeling System to Assist Human Labelers' Preference Elicitation
di: Lyu, Hanfang, et al.
Pubblicazione: (2024)
di: Lyu, Hanfang, et al.
Pubblicazione: (2024)
Selective Reviews of Bandit Problems in AI via a Statistical View
di: Zhou, Pengjie, et al.
Pubblicazione: (2024)
di: Zhou, Pengjie, et al.
Pubblicazione: (2024)
Dynamic Trust Calibration Using Contextual Bandits
di: Henrique, Bruno M., et al.
Pubblicazione: (2025)
di: Henrique, Bruno M., et al.
Pubblicazione: (2025)
Understanding Entrainment in Human Groups: Optimising Human-Robot Collaboration from Lessons Learned during Human-Human Collaboration
di: Schneiders, Eike, et al.
Pubblicazione: (2024)
di: Schneiders, Eike, et al.
Pubblicazione: (2024)
A Multi-Agent Conversational Bandit Approach to Online Evaluation and Selection of User-Aligned LLM Responses
di: Dai, Xiangxiang, et al.
Pubblicazione: (2025)
di: Dai, Xiangxiang, et al.
Pubblicazione: (2025)
InFiConD: Interactive No-code Fine-tuning with Concept-based Knowledge Distillation
di: Huang, Jinbin, et al.
Pubblicazione: (2024)
di: Huang, Jinbin, et al.
Pubblicazione: (2024)
The Role of AI in Peer Support for Young People: A Study of Preferences for Human- and AI-Generated Responses
di: Young, Jordyn, et al.
Pubblicazione: (2024)
di: Young, Jordyn, et al.
Pubblicazione: (2024)
Vibrotactile Preference Learning: Uncertainty-Aware Preference Learning for Personalized Vibration Feedback
di: Zhang, Rongtao, et al.
Pubblicazione: (2026)
di: Zhang, Rongtao, et al.
Pubblicazione: (2026)
How Well Do LLMs Predict Human Behavior? A Measure of their Pretrained Knowledge
di: Gao, Wayne, et al.
Pubblicazione: (2026)
di: Gao, Wayne, et al.
Pubblicazione: (2026)
MedSyn: Enhancing Diagnostics with Human-AI Collaboration
di: Sayin, Burcu, et al.
Pubblicazione: (2025)
di: Sayin, Burcu, et al.
Pubblicazione: (2025)
From Prompt Engineering to Prompt Science With Human in the Loop
di: Shah, Chirag
Pubblicazione: (2024)
di: Shah, Chirag
Pubblicazione: (2024)
Conformal Set-based Human-AI Complementarity with Multiple Experts
di: Paat, Helbert, et al.
Pubblicazione: (2025)
di: Paat, Helbert, et al.
Pubblicazione: (2025)
Towards Green Wearable Computing: A Physics-Aware Spiking Neural Network for Energy-Efficient IMU-based Human Activity Recognition
di: Zheng, Naichuan, et al.
Pubblicazione: (2026)
di: Zheng, Naichuan, et al.
Pubblicazione: (2026)
The two-way knowledge interaction interface between humans and neural networks
di: He, Zhanliang, et al.
Pubblicazione: (2024)
di: He, Zhanliang, et al.
Pubblicazione: (2024)
Comparing Exploration-Exploitation Strategies of LLMs and Humans: Insights from Standard Multi-armed Bandit Experiments
di: Zhang, Ziyuan, et al.
Pubblicazione: (2025)
di: Zhang, Ziyuan, et al.
Pubblicazione: (2025)
Understanding Impact of Human Feedback via Influence Functions
di: Min, Taywon, et al.
Pubblicazione: (2025)
di: Min, Taywon, et al.
Pubblicazione: (2025)
Semiparametric Preference Optimization: Your Language Model is Secretly a Single-Index Model
di: Kallus, Nathan
Pubblicazione: (2025)
di: Kallus, Nathan
Pubblicazione: (2025)
Enhancing Human Experience in Human-Agent Collaboration: A Human-Centered Modeling Approach Based on Positive Human Gain
di: Gao, Yiming, et al.
Pubblicazione: (2024)
di: Gao, Yiming, et al.
Pubblicazione: (2024)
LookALike: Human Mimicry based collaborative decision making
di: Karanjai, Rabimba, et al.
Pubblicazione: (2024)
di: Karanjai, Rabimba, et al.
Pubblicazione: (2024)
When Trust Collides: Decoding Human-LLM Cooperation Dynamics through the Prisoner's Dilemma
di: Jiang, Guanxuan, et al.
Pubblicazione: (2025)
di: Jiang, Guanxuan, et al.
Pubblicazione: (2025)
Integrating Human Expertise in Continuous Spaces: A Novel Interactive Bayesian Optimization Framework with Preference Expected Improvement
di: Feith, Nikolaus, et al.
Pubblicazione: (2024)
di: Feith, Nikolaus, et al.
Pubblicazione: (2024)
Socratic: Enhancing Human Teamwork via AI-enabled Coaching
di: Seo, Sangwon, et al.
Pubblicazione: (2025)
di: Seo, Sangwon, et al.
Pubblicazione: (2025)
In Pursuit of Predictive Models of Human Preferences Toward AI Teammates
di: Siu, Ho Chit, et al.
Pubblicazione: (2025)
di: Siu, Ho Chit, et al.
Pubblicazione: (2025)
Problem Solving Through Human-AI Preference-Based Cooperation
di: Dutta, Subhabrata, et al.
Pubblicazione: (2024)
di: Dutta, Subhabrata, et al.
Pubblicazione: (2024)
Predictive AI Can Support Human Learning while Preserving Error Diversity
di: He, Vivianna Fang, et al.
Pubblicazione: (2025)
di: He, Vivianna Fang, et al.
Pubblicazione: (2025)
Transformers Handle Endogeneity in In-Context Linear Regression
di: Liang, Haodong, et al.
Pubblicazione: (2024)
di: Liang, Haodong, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Twin-2K-500: A dataset for building digital twins of over 2,000 people based on their answers to over 500 questions
di: Toubia, Olivier, et al.
Pubblicazione: (2025) -
Designing Algorithmic Recommendations to Achieve Human-AI Complementarity
di: McLaughlin, Bryce, et al.
Pubblicazione: (2024) -
Wikipedia Contributions in the Wake of ChatGPT
di: Lyu, Liang, et al.
Pubblicazione: (2025) -
The Economics of AI Supply Chain Regulation
di: Qian, Sihan, et al.
Pubblicazione: (2026) -
Understanding the decision-making process of choice modellers
di: Nova, Gabriel, et al.
Pubblicazione: (2024)