Reinforcement Learning from Human Feedback: Whose Culture, Whose Values, Whose Perspectives?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Barman, Kristian González, Lohse, Simon, de Regt, Henk |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
When AI Writes, Whose Voice Remains? Quantifying Cultural Marker Erasure Across World English Varieties in Large Language Models
von: Navneet, Satyam Kumar, et al.
Veröffentlicht: (2026)
von: Navneet, Satyam Kumar, et al.
Veröffentlicht: (2026)
Towards a Benchmark for Scientific Understanding in Humans and Machines
von: Barman, Kristian Gonzalez, et al.
Veröffentlicht: (2023)
von: Barman, Kristian Gonzalez, et al.
Veröffentlicht: (2023)
Whose Knowledge is Valued?: Epistemic Injustice in CSCW Applications
von: Ajmani, Leah Hope, et al.
Veröffentlicht: (2024)
von: Ajmani, Leah Hope, et al.
Veröffentlicht: (2024)
Whose Knowledge Counts? Co-Designing Community-Centered AI Auditing Tools with Educators in Hawai`i
von: Zhao, Dora, et al.
Veröffentlicht: (2026)
von: Zhao, Dora, et al.
Veröffentlicht: (2026)
Whose Preferences? Differences in Fairness Preferences and Their Impact on the Fairness of AI Utilizing Human Feedback
von: Lerner, Emilia Agis, et al.
Veröffentlicht: (2024)
von: Lerner, Emilia Agis, et al.
Veröffentlicht: (2024)
Whose Good, Whose Place? The Moral Geography of Agentic AI for Social Good
von: Nemkova, Poli, et al.
Veröffentlicht: (2026)
von: Nemkova, Poli, et al.
Veröffentlicht: (2026)
When AI Speaks, Whose Values Does It Express? A Cross-Cultural Audit of Individualism-Collectivism Bias in Large Language Models
von: Venkata, Pruthvinath Jeripity
Veröffentlicht: (2026)
von: Venkata, Pruthvinath Jeripity
Veröffentlicht: (2026)
Culturally Grounded Personas in Large Language Models: Characterization and Alignment with Socio-Psychological Value Frameworks
von: Greco, Candida M., et al.
Veröffentlicht: (2026)
von: Greco, Candida M., et al.
Veröffentlicht: (2026)
LearnLens: LLM-Enabled Personalised, Curriculum-Grounded Feedback with Educators in the Loop
von: Zhao, Runcong, et al.
Veröffentlicht: (2025)
von: Zhao, Runcong, et al.
Veröffentlicht: (2025)
Culturally-Attuned Moral Machines: Implicit Learning of Human Value Systems by AI through Inverse Reinforcement Learning
von: Oliveira, Nigini, et al.
Veröffentlicht: (2023)
von: Oliveira, Nigini, et al.
Veröffentlicht: (2023)
Evaluating the Application of Large Language Models to Generate Feedback in Programming Education
von: Jacobs, Sven, et al.
Veröffentlicht: (2024)
von: Jacobs, Sven, et al.
Veröffentlicht: (2024)
Bottom-Up Perspectives on AI Governance: Insights from User Reviews of AI Products
von: Pasch, Stefan
Veröffentlicht: (2025)
von: Pasch, Stefan
Veröffentlicht: (2025)
TALES: A Taxonomy and Analysis of Cultural Representations in LLM-generated Stories
von: Bhagat, Kirti, et al.
Veröffentlicht: (2025)
von: Bhagat, Kirti, et al.
Veröffentlicht: (2025)
Beyond Models: A Framework for Contextual and Cultural Intelligence in African AI Deployment
von: Ndlovu, Qness
Veröffentlicht: (2025)
von: Ndlovu, Qness
Veröffentlicht: (2025)
Civil Society in the Loop: Feedback-Driven Adaptation of (L)LM-Assisted Classification in an Open-Source Telegram Monitoring Tool
von: Pustet, Milena, et al.
Veröffentlicht: (2025)
von: Pustet, Milena, et al.
Veröffentlicht: (2025)
The Cognitive Divergence: AI Context Windows, Human Attention Decline, and the Delegation Feedback Loop
von: Eliav, Netanel
Veröffentlicht: (2026)
von: Eliav, Netanel
Veröffentlicht: (2026)
Imperfectly Cooperative Human-AI Interactions: Comparing the Impacts of Human and AI Attributes in Simulated and User Studies
von: Cohen, Myke C., et al.
Veröffentlicht: (2026)
von: Cohen, Myke C., et al.
Veröffentlicht: (2026)
Minion: A Technology Probe to Explore How Users Negotiate Harmful Value Conflicts with AI Companions
von: Fan, Xianzhe, et al.
Veröffentlicht: (2024)
von: Fan, Xianzhe, et al.
Veröffentlicht: (2024)
TUX: Measuring Human--AI Tacit Understanding
von: Li, Yueshen, et al.
Veröffentlicht: (2026)
von: Li, Yueshen, et al.
Veröffentlicht: (2026)
Are You the A-hole? A Fair, Multi-Perspective Ethical Reasoning Framework
von: Munir, Sheza, et al.
Veröffentlicht: (2026)
von: Munir, Sheza, et al.
Veröffentlicht: (2026)
Whose Journey Matters? Investigating Identity Biases in Large Language Models (LLMs) for Travel Planning Assistance
von: Ren, Ruiping, et al.
Veröffentlicht: (2024)
von: Ren, Ruiping, et al.
Veröffentlicht: (2024)
Human-Centred LLM Privacy Audits: Findings and Frictions
von: Staufer, Dimitri, et al.
Veröffentlicht: (2026)
von: Staufer, Dimitri, et al.
Veröffentlicht: (2026)
Human Decision-making is Susceptible to AI-driven Manipulation
von: Sabour, Sahand, et al.
Veröffentlicht: (2025)
von: Sabour, Sahand, et al.
Veröffentlicht: (2025)
Human Preferences for Constructive Interactions in Language Model Alignment
von: Kyrychenko, Yara, et al.
Veröffentlicht: (2025)
von: Kyrychenko, Yara, et al.
Veröffentlicht: (2025)
Prompt Engineering Techniques for Mitigating Cultural Bias Against Arabs and Muslims in Large Language Models: A Systematic Review
von: Asseri, Bushra, et al.
Veröffentlicht: (2025)
von: Asseri, Bushra, et al.
Veröffentlicht: (2025)
ChatBench: From Static Benchmarks to Human-AI Evaluation
von: Chang, Serina, et al.
Veröffentlicht: (2025)
von: Chang, Serina, et al.
Veröffentlicht: (2025)
Mind the Style: Impact of Communication Style on Human-Chatbot Interaction
von: Derner, Erik, et al.
Veröffentlicht: (2026)
von: Derner, Erik, et al.
Veröffentlicht: (2026)
GPT-4's One-Dimensional Mapping of Morality: How the Accuracy of Country-Estimates Depends on Moral Domain
von: Strimling, Pontus, et al.
Veröffentlicht: (2024)
von: Strimling, Pontus, et al.
Veröffentlicht: (2024)
Dialogue Systems for Emotional Support via Value Reinforcement
von: Kim, Juhee, et al.
Veröffentlicht: (2025)
von: Kim, Juhee, et al.
Veröffentlicht: (2025)
Human-Centric NLP or AI-Centric Illusion?: A Critical Investigation
von: Spencer, Piyapath T
Veröffentlicht: (2024)
von: Spencer, Piyapath T
Veröffentlicht: (2024)
STAR: SocioTechnical Approach to Red Teaming Language Models
von: Weidinger, Laura, et al.
Veröffentlicht: (2024)
von: Weidinger, Laura, et al.
Veröffentlicht: (2024)
Large Language Models Show Human-like Social Desirability Biases in Survey Responses
von: Salecha, Aadesh, et al.
Veröffentlicht: (2024)
von: Salecha, Aadesh, et al.
Veröffentlicht: (2024)
Conversational DNA: A New Visual Language for Understanding Dialogue Structure in Human and AI
von: Lin, Baihan
Veröffentlicht: (2025)
von: Lin, Baihan
Veröffentlicht: (2025)
Alignment Drift in Long-Term Human-LLM Interaction: A Mechanism-Oriented Framework
von: Yao, Xintong
Veröffentlicht: (2026)
von: Yao, Xintong
Veröffentlicht: (2026)
Beyond One-Way Influence: Bidirectional Opinion Dynamics in Multi-Turn Human-LLM Interactions
von: Jiang, Yuyang, et al.
Veröffentlicht: (2025)
von: Jiang, Yuyang, et al.
Veröffentlicht: (2025)
GenAI Against Humanity: Nefarious Applications of Generative Artificial Intelligence and Large Language Models
von: Ferrara, Emilio
Veröffentlicht: (2023)
von: Ferrara, Emilio
Veröffentlicht: (2023)
Leading Across the Spectrum of Human-AI Relationships: A Conceptual Framework for Increasingly Heterogeneous Teams
von: Jadad, Alejandro R.
Veröffentlicht: (2026)
von: Jadad, Alejandro R.
Veröffentlicht: (2026)
What Do LLMs Associate with Your Name? A Human-Centered Black-Box Audit of Personal Data
von: Staufer, Dimitri, et al.
Veröffentlicht: (2026)
von: Staufer, Dimitri, et al.
Veröffentlicht: (2026)
How AI Ideas Affect the Creativity, Diversity, and Evolution of Human Ideas: Evidence From a Large, Dynamic Experiment
von: Ashkinaze, Joshua, et al.
Veröffentlicht: (2024)
von: Ashkinaze, Joshua, et al.
Veröffentlicht: (2024)
Learning in Blocks: A Multi Agent Debate Assisted Personalized Adaptive Learning Framework for Language Learning
von: Scaria, Nicy, et al.
Veröffentlicht: (2026)
von: Scaria, Nicy, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
When AI Writes, Whose Voice Remains? Quantifying Cultural Marker Erasure Across World English Varieties in Large Language Models
von: Navneet, Satyam Kumar, et al.
Veröffentlicht: (2026) -
Towards a Benchmark for Scientific Understanding in Humans and Machines
von: Barman, Kristian Gonzalez, et al.
Veröffentlicht: (2023) -
Whose Knowledge is Valued?: Epistemic Injustice in CSCW Applications
von: Ajmani, Leah Hope, et al.
Veröffentlicht: (2024) -
Whose Knowledge Counts? Co-Designing Community-Centered AI Auditing Tools with Educators in Hawai`i
von: Zhao, Dora, et al.
Veröffentlicht: (2026) -
Whose Preferences? Differences in Fairness Preferences and Their Impact on the Fairness of AI Utilizing Human Feedback
von: Lerner, Emilia Agis, et al.
Veröffentlicht: (2024)