Position: Towards Bidirectional Human-AI Alignment
Fuente:
arXiv
Saved in:
| Main Authors: | Shen, Hua, Knearem, Tiffany, Ghosh, Reshmi, Alkiek, Kenan, Krishna, Kundan, Liu, Yachuan, Ma, Ziqiao, Petridis, Savvas, Peng, Yi-Hao, Qiwei, Li, Rakshit, Sushrita, Si, Chenglei, Xie, Yutong, Bigham, Jeffrey P., Bentley, Frank, Chai, Joyce, Lipton, Zachary, Mei, Qiaozhu, Mihalcea, Rada, Terry, Michael, Yang, Diyi, Morris, Meredith Ringel, Resnick, Paul, Jurgens, David |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Big Reasoning with Small Models: Instruction Retrieval at Inference Time
by: Alkiek, Kenan, et al.
Published: (2025)
by: Alkiek, Kenan, et al.
Published: (2025)
Neurobiber: Fast and Interpretable Stylistic Feature Extraction
by: Alkiek, Kenan, et al.
Published: (2025)
by: Alkiek, Kenan, et al.
Published: (2025)
Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions
by: Rakshit, Sushrita, et al.
Published: (2026)
by: Rakshit, Sushrita, et al.
Published: (2026)
VET: A Framework for Analyzing AI Discourse
by: Morris, Meredith Ringel
Published: (2026)
by: Morris, Meredith Ringel
Published: (2026)
ValueCompass: A Framework for Measuring Contextual Value Alignment Between Human and LLMs
by: Shen, Hua, et al.
Published: (2024)
by: Shen, Hua, et al.
Published: (2024)
Towards Implicit Bias Detection and Mitigation in Multi-Agent LLM Interactions
by: Borah, Angana, et al.
Published: (2024)
by: Borah, Angana, et al.
Published: (2024)
Rethinking Table Instruction Tuning
by: Deng, Naihao, et al.
Published: (2025)
by: Deng, Naihao, et al.
Published: (2025)
Are Human Interactions Replicable by Generative Agents? A Case Study on Pronoun Usage in Hierarchical Interactions
by: Deng, Naihao, et al.
Published: (2025)
by: Deng, Naihao, et al.
Published: (2025)
Whose wife is it anyway? Assessing bias against same-gender relationships in machine translation
by: Stewart, Ian, et al.
Published: (2024)
by: Stewart, Ian, et al.
Published: (2024)
The GenUI Study: Exploring the Design of Generative UI Tools to Support UX Practitioners and Beyond
by: Chen, Xiang 'Anthony', et al.
Published: (2025)
by: Chen, Xiang 'Anthony', et al.
Published: (2025)
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
by: Si, Chenglei, et al.
Published: (2024)
by: Si, Chenglei, et al.
Published: (2024)
The Ideation-Execution Gap: Execution Outcomes of LLM-Generated versus Human Research Ideas
by: Si, Chenglei, et al.
Published: (2025)
by: Si, Chenglei, et al.
Published: (2025)
GenAudit: Fixing Factual Errors in Language Model Outputs with Evidence
by: Krishna, Kundan, et al.
Published: (2024)
by: Krishna, Kundan, et al.
Published: (2024)
Emotionally-Aware Agents for Dispute Resolution
by: Rakshit, Sushrita, et al.
Published: (2025)
by: Rakshit, Sushrita, et al.
Published: (2025)
KODIS: A Multicultural Dispute Resolution Dialogue Corpus
by: Hale, James, et al.
Published: (2025)
by: Hale, James, et al.
Published: (2025)
The Age of Curiosity Meets the Age of AI: Benchmarking Child Safety in Large Language Models
by: Arif, Samee, et al.
Published: (2026)
by: Arif, Samee, et al.
Published: (2026)
Mind the (Belief) Gap: Group Identity in the World of LLMs
by: Borah, Angana, et al.
Published: (2025)
by: Borah, Angana, et al.
Published: (2025)
Towards Region-aware Bias Evaluation Metrics
by: Borah, Angana, et al.
Published: (2024)
by: Borah, Angana, et al.
Published: (2024)
MAiDE-up: Multilingual Deception Detection of GPT-generated Hotel Reviews
by: Ignat, Oana, et al.
Published: (2024)
by: Ignat, Oana, et al.
Published: (2024)
The Curious Case of Curiosity across Human Cultures and LLMs
by: Borah, Angana, et al.
Published: (2025)
by: Borah, Angana, et al.
Published: (2025)
Uplifting Lower-Income Data: Strategies for Socioeconomic Perspective Shifts in Large Multi-modal Models
by: Nwatu, Joan, et al.
Published: (2024)
by: Nwatu, Joan, et al.
Published: (2024)
Generative Ghosts: Anticipating Benefits and Risks of AI Afterlives
by: Morris, Meredith Ringel, et al.
Published: (2024)
by: Morris, Meredith Ringel, et al.
Published: (2024)
Temple under the national flag: History, legitimacy, and everyday politics in a religious landscape of rural northwest China
by: Yachuan Zhao
Published: (2025)
by: Yachuan Zhao
Published: (2025)
Towards Dog Bark Decoding: Leveraging Human Speech Processing for Automated Bark Classification
by: Abzaliev, Artem, et al.
Published: (2024)
by: Abzaliev, Artem, et al.
Published: (2024)
Cross-cultural Inspiration Detection and Analysis in Real and LLM-generated Social Media Data
by: Ignat, Oana, et al.
Published: (2024)
by: Ignat, Oana, et al.
Published: (2024)
Patient-Centered RAG for Oncology Visit Aid Following the Ottawa Decision Guide
by: Liu, Siyang, et al.
Published: (2025)
by: Liu, Siyang, et al.
Published: (2025)
Persuasion at Play: Understanding Misinformation Dynamics in Demographic-Aware Human-LLM Interactions
by: Borah, Angana, et al.
Published: (2025)
by: Borah, Angana, et al.
Published: (2025)
The Call for Socially Aware Language Technologies
by: Yang, Diyi, et al.
Published: (2024)
by: Yang, Diyi, et al.
Published: (2024)
ExAnte: A Benchmark for Ex-Ante Inference in Large Language Models
by: Liu, Yachuan, et al.
Published: (2025)
by: Liu, Yachuan, et al.
Published: (2025)
Human Action Co-occurrence in Lifestyle Vlogs using Graph Link Prediction
by: Ignat, Oana, et al.
Published: (2023)
by: Ignat, Oana, et al.
Published: (2023)
The Power of Many: Multi-Agent Multimodal Models for Cultural Image Captioning
by: Bai, Longju, et al.
Published: (2024)
by: Bai, Longju, et al.
Published: (2024)
Annotations on a Budget: Leveraging Geo-Data Similarity to Balance Model Performance and Annotation Cost
by: Ignat, Oana, et al.
Published: (2024)
by: Ignat, Oana, et al.
Published: (2024)
DOTRAG: Retrieval-Time Reasoning Along Paths
by: Moore, Larnell, et al.
Published: (2026)
by: Moore, Larnell, et al.
Published: (2026)
Free Lunch for User Experience: Crowdsourcing Agents for Scalable User Studies
by: Liu, Siyang, et al.
Published: (2025)
by: Liu, Siyang, et al.
Published: (2025)
Culture Affordance Atlas: Reconciling Object Diversity Through Functional Mapping
by: Nwatu, Joan, et al.
Published: (2025)
by: Nwatu, Joan, et al.
Published: (2025)
One Word at a Time: Incremental Completion Decomposition Breaks LLM Safety
by: Arif, Samee, et al.
Published: (2026)
by: Arif, Samee, et al.
Published: (2026)
Towards Algorithmic Fidelity: Mental Health Representation across Demographics in Synthetic vs. Human-generated Data
by: Mori, Shinka, et al.
Published: (2024)
by: Mori, Shinka, et al.
Published: (2024)
LUCid: Redefining Relevance For Lifelong Personalization
by: Okite, Chimaobi, et al.
Published: (2026)
by: Okite, Chimaobi, et al.
Published: (2026)
Belief-Sim: Towards Belief-Driven Simulation of Demographic Misinformation Susceptibility
by: Borah, Angana, et al.
Published: (2026)
by: Borah, Angana, et al.
Published: (2026)
Evaluating the Factuality of Zero-shot Summarizers Across Varied Domains
by: Ramprasad, Sanjana, et al.
Published: (2024)
by: Ramprasad, Sanjana, et al.
Published: (2024)
Similar Items
-
Big Reasoning with Small Models: Instruction Retrieval at Inference Time
by: Alkiek, Kenan, et al.
Published: (2025) -
Neurobiber: Fast and Interpretable Stylistic Feature Extraction
by: Alkiek, Kenan, et al.
Published: (2025) -
Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions
by: Rakshit, Sushrita, et al.
Published: (2026) -
VET: A Framework for Analyzing AI Discourse
by: Morris, Meredith Ringel
Published: (2026) -
ValueCompass: A Framework for Measuring Contextual Value Alignment Between Human and LLMs
by: Shen, Hua, et al.
Published: (2024)