Understanding Impact of Human Feedback via Influence Functions
Fuente:
arXiv
Saved in:
| Main Authors: | Min, Taywon, Lee, Haeone, Kwon, Yongchan, Lee, Kimin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Benchmarking Mobile Device Control Agents across Diverse Configurations
by: Lee, Juyong, et al.
Published: (2024)
by: Lee, Juyong, et al.
Published: (2024)
State Your Intention to Steer Your Attention: An AI Assistant for Intentional Digital Living
by: Choi, Juheon, et al.
Published: (2025)
by: Choi, Juheon, et al.
Published: (2025)
RuleEdit: Failure-Guided Human-AI Model Editing with Prospective Impact Preview
by: Lee, Min Hun, et al.
Published: (2026)
by: Lee, Min Hun, et al.
Published: (2026)
From Accuracy to Readiness: Metrics and Benchmarks for Human-AI Decision-Making
by: Lee, Min Hun
Published: (2026)
by: Lee, Min Hun
Published: (2026)
Voice "Cloning" is Style Transfer
by: Zhou, Kaitlyn, et al.
Published: (2026)
by: Zhou, Kaitlyn, et al.
Published: (2026)
Towards Interactive Reinforcement Learning with Intrinsic Feedback
by: Poole, Benjamin, et al.
Published: (2021)
by: Poole, Benjamin, et al.
Published: (2021)
Towards Uncertainty Aware Task Delegation and Human-AI Collaborative Decision-Making
by: Lee, Min Hun, et al.
Published: (2025)
by: Lee, Min Hun, et al.
Published: (2025)
Improving Health Professionals' Onboarding with AI and XAI for Trustworthy Human-AI Collaborative Decision Making
by: Lee, Min Hun, et al.
Published: (2024)
by: Lee, Min Hun, et al.
Published: (2024)
Sentiment Analysis in Learning Management Systems Understanding Student Feedback at Scale
by: Almutairi, Mohammed
Published: (2025)
by: Almutairi, Mohammed
Published: (2025)
Interactive Example-based Explanations to Improve Health Professionals' Onboarding with AI for Human-AI Collaborative Decision Making
by: Lee, Min Hun, et al.
Published: (2024)
by: Lee, Min Hun, et al.
Published: (2024)
Influencing Humans to Conform to Preference Models for RLHF
by: Hatgis-Kessell, Stephane, et al.
Published: (2025)
by: Hatgis-Kessell, Stephane, et al.
Published: (2025)
Cognitive Exoskeleton: Augmenting Human Cognition with an AI-Mediated Intelligent Visual Feedback
by: Xu, Songlin, et al.
Published: (2025)
by: Xu, Songlin, et al.
Published: (2025)
MENTOR: Guiding Hierarchical Reinforcement Learning with Human Feedback and Dynamic Distance Constraint
by: Zhou, Xinglin, et al.
Published: (2024)
by: Zhou, Xinglin, et al.
Published: (2024)
Unintended Misalignment from Agentic Fine-Tuning: Risks and Mitigation
by: Hahm, Dongyoon, et al.
Published: (2025)
by: Hahm, Dongyoon, et al.
Published: (2025)
Can Interpretability Layouts Influence Human Perception of Offensive Sentences?
by: Santos, Thiago Freitas dos, et al.
Published: (2024)
by: Santos, Thiago Freitas dos, et al.
Published: (2024)
Making RL with Preference-based Feedback Efficient via Randomization
by: Wu, Runzhe, et al.
Published: (2023)
by: Wu, Runzhe, et al.
Published: (2023)
Reversing the Lens: Using Explainable AI to Understand Human Expertise
by: Rahman, Roussel, et al.
Published: (2025)
by: Rahman, Roussel, et al.
Published: (2025)
Investigating an Intelligent System to Monitor \& Explain Abnormal Activity Patterns of Older Adults
by: Lee, Min Hun, et al.
Published: (2025)
by: Lee, Min Hun, et al.
Published: (2025)
Evaluation of Human-Understandability of Global Model Explanations using Decision Tree
by: Sivaprasad, Adarsa, et al.
Published: (2023)
by: Sivaprasad, Adarsa, et al.
Published: (2023)
Does Explanation Correctness Matter? Linking Computational XAI Evaluation to Human Understanding
by: Baer, Gregor, et al.
Published: (2026)
by: Baer, Gregor, et al.
Published: (2026)
Trustworthy Human-AI Collaboration: Reinforcement Learning with Human Feedback and Physics Knowledge for Safe Autonomous Driving
by: Huang, Zilin, et al.
Published: (2024)
by: Huang, Zilin, et al.
Published: (2024)
Uni-RLHF: Universal Platform and Benchmark Suite for Reinforcement Learning with Diverse Human Feedback
by: Yuan, Yifu, et al.
Published: (2024)
by: Yuan, Yifu, et al.
Published: (2024)
Catch Me if You Search: When Contextual Web Search Results Affect the Detection of Hallucinations
by: Nahar, Mahjabin, et al.
Published: (2025)
by: Nahar, Mahjabin, et al.
Published: (2025)
Off-Policy Selection for Initiating Human-Centric Experimental Design
by: Gao, Ge, et al.
Published: (2024)
by: Gao, Ge, et al.
Published: (2024)
Improving Dialogue Agents by Decomposing One Global Explicit Annotation with Local Implicit Multimodal Feedback
by: Lee, Dong Won, et al.
Published: (2024)
by: Lee, Dong Won, et al.
Published: (2024)
Reassessing Evaluation Functions in Algorithmic Recourse: An Empirical Study from a Human-Centered Perspective
by: Tominaga, Tomu, et al.
Published: (2024)
by: Tominaga, Tomu, et al.
Published: (2024)
Advancing Human-Machine Teaming: Concepts, Challenges, and Applications
by: Chen, Dian, et al.
Published: (2025)
by: Chen, Dian, et al.
Published: (2025)
Introducing User Feedback-based Counterfactual Explanations (UFCE)
by: Suffian, Muhammad, et al.
Published: (2024)
by: Suffian, Muhammad, et al.
Published: (2024)
Mobile Fitting Room: On-device Virtual Try-on via Diffusion Models
by: Blalock, Justin, et al.
Published: (2024)
by: Blalock, Justin, et al.
Published: (2024)
Co-Creative Learning via Metropolis-Hastings Interaction between Humans and AI
by: Okumura, Ryota, et al.
Published: (2025)
by: Okumura, Ryota, et al.
Published: (2025)
Vibrotactile Preference Learning: Uncertainty-Aware Preference Learning for Personalized Vibration Feedback
by: Zhang, Rongtao, et al.
Published: (2026)
by: Zhang, Rongtao, et al.
Published: (2026)
Learning Social Cost Functions for Human-Aware Path Planning
by: Eirale, Andrea, et al.
Published: (2024)
by: Eirale, Andrea, et al.
Published: (2024)
Principled Reinforcement Learning with Human Feedback from Pairwise or $K$-wise Comparisons
by: Zhu, Banghua, et al.
Published: (2023)
by: Zhu, Banghua, et al.
Published: (2023)
On the Utility of Accounting for Human Beliefs about AI Intention in Human-AI Collaboration
by: Yu, Guanghui, et al.
Published: (2024)
by: Yu, Guanghui, et al.
Published: (2024)
Human Expertise in Algorithmic Prediction
by: Alur, Rohan, et al.
Published: (2024)
by: Alur, Rohan, et al.
Published: (2024)
Human-Computer Interaction and Human-AI Collaboration in Advanced Air Mobility: A Comprehensive Review
by: Sagirli, Fatma Yamac, et al.
Published: (2024)
by: Sagirli, Fatma Yamac, et al.
Published: (2024)
Quality over Quantity: Demonstration Curation via Influence Functions for Data-Centric Robot Learning
by: Lee, Haeone, et al.
Published: (2026)
by: Lee, Haeone, et al.
Published: (2026)
Align When They Want, Complement When They Need! Human-Centered Ensembles for Adaptive Human-AI Collaboration
by: Amin, Hasan, et al.
Published: (2026)
by: Amin, Hasan, et al.
Published: (2026)
HybridQuestion: Human-AI Collaboration for Identifying High-Impact Research Questions
by: Zhao, Keyu, et al.
Published: (2025)
by: Zhao, Keyu, et al.
Published: (2025)
The Hardness of Achieving Impact in AI for Social Impact Research: A Ground-Level View of Challenges & Opportunities
by: Majumdar, Aditya, et al.
Published: (2025)
by: Majumdar, Aditya, et al.
Published: (2025)
Similar Items
-
Benchmarking Mobile Device Control Agents across Diverse Configurations
by: Lee, Juyong, et al.
Published: (2024) -
State Your Intention to Steer Your Attention: An AI Assistant for Intentional Digital Living
by: Choi, Juheon, et al.
Published: (2025) -
RuleEdit: Failure-Guided Human-AI Model Editing with Prospective Impact Preview
by: Lee, Min Hun, et al.
Published: (2026) -
From Accuracy to Readiness: Metrics and Benchmarks for Human-AI Decision-Making
by: Lee, Min Hun
Published: (2026) -
Voice "Cloning" is Style Transfer
by: Zhou, Kaitlyn, et al.
Published: (2026)