Initializing Services in Interactive ML Systems for Diverse Users
Fuente:
arXiv
Saved in:
| Main Authors: | Bose, Avinandan, Curmei, Mihaela, Jiang, Daniel L., Morgenstern, Jamie, Dean, Sarah, Ratliff, Lillian J., Fazel, Maryam |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Emergent specialization from participation dynamics and multi-learner retraining
by: Dean, Sarah, et al.
Published: (2022)
by: Dean, Sarah, et al.
Published: (2022)
Offline Multi-task Transfer RL with Representational Penalization
by: Bose, Avinandan, et al.
Published: (2024)
by: Bose, Avinandan, et al.
Published: (2024)
Dynamics of Learning under User Choice: Overspecialization and Peer-Model Probing
by: Narang, Adhyyan, et al.
Published: (2026)
by: Narang, Adhyyan, et al.
Published: (2026)
LoRe: Personalizing LLMs via Low-Rank Reward Modeling
by: Bose, Avinandan, et al.
Published: (2025)
by: Bose, Avinandan, et al.
Published: (2025)
PrefDisco: Benchmarking Proactive Personalized Reasoning
by: Li, Shuyue Stella, et al.
Published: (2025)
by: Li, Shuyue Stella, et al.
Published: (2025)
Safe Probabilistic Planning for Human-Robot Interaction using Conformal Risk Control
by: Gonzales, Jake, et al.
Published: (2026)
by: Gonzales, Jake, et al.
Published: (2026)
Cold-Start Personalization via Training-Free Priors from Structured World Models
by: Bose, Avinandan, et al.
Published: (2026)
by: Bose, Avinandan, et al.
Published: (2026)
Sample Complexity Reduction via Policy Difference Estimation in Tabular Reinforcement Learning
by: Narang, Adhyyan, et al.
Published: (2024)
by: Narang, Adhyyan, et al.
Published: (2024)
dUltra: Ultra-Fast Diffusion Language Models via Reinforcement Learning
by: Chen, Shirui, et al.
Published: (2025)
by: Chen, Shirui, et al.
Published: (2025)
Keeping up with dynamic attackers: Certifying robustness to adaptive online data poisoning
by: Bose, Avinandan, et al.
Published: (2025)
by: Bose, Avinandan, et al.
Published: (2025)
Learning from Streaming Data when Users Choose
by: Su, Jinyan, et al.
Published: (2024)
by: Su, Jinyan, et al.
Published: (2024)
LLM-DaaS: LLM-driven Drone-as-a-Service Operations from Text User Requests
by: Wassim, Lillian, et al.
Published: (2024)
by: Wassim, Lillian, et al.
Published: (2024)
Hybrid Preference Optimization for Alignment: Provably Faster Convergence Rates by Combining Offline Preferences with Online Exploration
by: Bose, Avinandan, et al.
Published: (2024)
by: Bose, Avinandan, et al.
Published: (2024)
Learning Optimal Tax Design in Nonatomic Congestion Games
by: Cui, Qiwen, et al.
Published: (2024)
by: Cui, Qiwen, et al.
Published: (2024)
Effect of Adaptation Rate and Cost Display in a Human-AI Interaction Game
by: Isa, Jason T., et al.
Published: (2024)
by: Isa, Jason T., et al.
Published: (2024)
Do LLMs Favor LLMs? Quantifying Interaction Effects in Peer Review
by: Sharma, Vibhhu, et al.
Published: (2026)
by: Sharma, Vibhhu, et al.
Published: (2026)
Online SuBmodular + SuPermodular (BP) Maximization with Bandit Feedback
by: Narang, Adhyyan, et al.
Published: (2022)
by: Narang, Adhyyan, et al.
Published: (2022)
Follower Agnostic Methods for Stackelberg Games
by: Maheshwari, Chinmay, et al.
Published: (2023)
by: Maheshwari, Chinmay, et al.
Published: (2023)
A Black-box Approach for Non-stationary Multi-agent Reinforcement Learning
by: Jiang, Haozhe, et al.
Published: (2023)
by: Jiang, Haozhe, et al.
Published: (2023)
A Flexible Method for Behaviorally Measuring Alignment Between Human and Artificial Intelligence Using Representational Similarity Analysis
by: Ogg, Mattson, et al.
Published: (2024)
by: Ogg, Mattson, et al.
Published: (2024)
Welfare-Centric Clustering
by: Zhang, Claire Jie, et al.
Published: (2025)
by: Zhang, Claire Jie, et al.
Published: (2025)
AgentInit: Initializing LLM-based Multi-Agent Systems via Diversity and Expertise Orchestration for Effective and Efficient Collaboration
by: Tian, Chunhao, et al.
Published: (2025)
by: Tian, Chunhao, et al.
Published: (2025)
AutoML Systems For Medical Imaging
by: Jidney, Tasmia Tahmida, et al.
Published: (2023)
by: Jidney, Tasmia Tahmida, et al.
Published: (2023)
How Do Data Owners Say No? A Case Study of Data Consent Mechanisms in Web-Scraped Vision-Language AI Training Datasets
by: Lee, Chung Peng, et al.
Published: (2025)
by: Lee, Chung Peng, et al.
Published: (2025)
RecUserSim: A Realistic and Diverse User Simulator for Evaluating Conversational Recommender Systems
by: Chen, Luyu, et al.
Published: (2025)
by: Chen, Luyu, et al.
Published: (2025)
Datasets for Navigating Sensitive Topics in Recommendation Systems
by: Kovacs, Amelia, et al.
Published: (2025)
by: Kovacs, Amelia, et al.
Published: (2025)
Understanding User Preferences in Explainable Artificial Intelligence: A Survey and a Mapping Function Proposal
by: Hashemi, Maryam, et al.
Published: (2023)
by: Hashemi, Maryam, et al.
Published: (2023)
Efficient Agent Evaluation via Diversity-Guided User Simulation
by: Nakash, Itay, et al.
Published: (2026)
by: Nakash, Itay, et al.
Published: (2026)
An AI Agent Execution Environment to Safeguard User Data
by: Stanley, Robert, et al.
Published: (2026)
by: Stanley, Robert, et al.
Published: (2026)
Fair Clustering: Critique, Caveats, and Future Directions
by: Dickerson, John, et al.
Published: (2024)
by: Dickerson, John, et al.
Published: (2024)
AI/ML in 3GPP 5G Advanced -- Services and Architecture
by: Taksande, Pradnya, et al.
Published: (2025)
by: Taksande, Pradnya, et al.
Published: (2025)
AdaptoML-UX: An Adaptive User-centered GUI-based AutoML Toolkit for Non-AI Experts and HCI Researchers
by: Gomaa, Amr, et al.
Published: (2024)
by: Gomaa, Amr, et al.
Published: (2024)
SafePickle: Robust and Generic ML Detection of Malicious Pickle-based ML Models
by: Ohayon, Hillel, et al.
Published: (2026)
by: Ohayon, Hillel, et al.
Published: (2026)
TOPReward: Token Probabilities as Hidden Zero-Shot Rewards for Robotics
by: Chen, Shirui, et al.
Published: (2026)
by: Chen, Shirui, et al.
Published: (2026)
Vortex: Hosting ML Inference and Knowledge Retrieval Services With Tight Latency and Throughput Requirements
by: Yang, Yuting, et al.
Published: (2025)
by: Yang, Yuting, et al.
Published: (2025)
Finite Sample Identification of Partially Observed Bilinear Dynamical Systems
by: Sattar, Yahya, et al.
Published: (2025)
by: Sattar, Yahya, et al.
Published: (2025)
The Interaction Layer: An Exploration for Co-Designing User-LLM Interactions in Parental Wellbeing Support Systems
by: Viswanathan, Sruthi, et al.
Published: (2024)
by: Viswanathan, Sruthi, et al.
Published: (2024)
Deep Reinforcement Learning for Personalized Diagnostic Decision Pathways Using Electronic Health Records: A Comparative Study on Anemia and Systemic Lupus Erythematosus
by: Muyama, Lillian, et al.
Published: (2024)
by: Muyama, Lillian, et al.
Published: (2024)
Initial Steps in Integrating Large Reasoning and Action Models for Service Composition
by: Georgievski, Ilche, et al.
Published: (2025)
by: Georgievski, Ilche, et al.
Published: (2025)
Gating is Weighting: Understanding Gated Linear Attention through In-context Learning
by: Li, Yingcong, et al.
Published: (2025)
by: Li, Yingcong, et al.
Published: (2025)
Similar Items
-
Emergent specialization from participation dynamics and multi-learner retraining
by: Dean, Sarah, et al.
Published: (2022) -
Offline Multi-task Transfer RL with Representational Penalization
by: Bose, Avinandan, et al.
Published: (2024) -
Dynamics of Learning under User Choice: Overspecialization and Peer-Model Probing
by: Narang, Adhyyan, et al.
Published: (2026) -
LoRe: Personalizing LLMs via Low-Rank Reward Modeling
by: Bose, Avinandan, et al.
Published: (2025) -
PrefDisco: Benchmarking Proactive Personalized Reasoning
by: Li, Shuyue Stella, et al.
Published: (2025)