CURATe: Benchmarking Personalised Alignment of Conversational AI Assistants
Fuente:
arXiv
Salvato in:
| Autori principali: | Alberts, Lize, Ellis, Benjamin, Lupu, Andrei, Foerster, Jakob |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
"I followed what felt right, not what I was told": Autonomy, Coaching, and Recognizing Bias Through AI-Mediated Dialogue
di: Taheri, Atieh, et al.
Pubblicazione: (2026)
di: Taheri, Atieh, et al.
Pubblicazione: (2026)
AI-Spectra: A Visual Dashboard for Model Multiplicity to Enhance Informed and Transparent Decision-Making
di: Eerlings, Gilles, et al.
Pubblicazione: (2024)
di: Eerlings, Gilles, et al.
Pubblicazione: (2024)
What Would GPT Click: Practical Effects of Human-AI Behavioral Misalignment and the Cost of Synthetic Participants in User Experience
di: Kuric, Eduard, et al.
Pubblicazione: (2026)
di: Kuric, Eduard, et al.
Pubblicazione: (2026)
Think, Act, Learn: A Framework for Autonomous Robotic Agents using Closed-Loop Large Language Models
di: Menon, Anjali R., et al.
Pubblicazione: (2025)
di: Menon, Anjali R., et al.
Pubblicazione: (2025)
Should agentic conversational AI change how we think about ethics? Characterising an interactional ethics centred on respect
di: Alberts, Lize, et al.
Pubblicazione: (2024)
di: Alberts, Lize, et al.
Pubblicazione: (2024)
NeuroState-Bench: A Human-Calibrated Benchmark for Commitment Integrity in LLM Agent Profiles
di: Jia, Xiao
Pubblicazione: (2026)
di: Jia, Xiao
Pubblicazione: (2026)
Approximating Discrimination Within Models When Faced With Several Non-Binary Sensitive Attributes
di: Bian, Yijun, et al.
Pubblicazione: (2024)
di: Bian, Yijun, et al.
Pubblicazione: (2024)
Does Machine Bring in Extra Bias in Learning? Approximating Fairness in Models Promptly
di: Bian, Yijun, et al.
Pubblicazione: (2024)
di: Bian, Yijun, et al.
Pubblicazione: (2024)
Inducing Causal World Models in LLMs for Zero-Shot Physical Reasoning
di: Sharma, Aditya, et al.
Pubblicazione: (2025)
di: Sharma, Aditya, et al.
Pubblicazione: (2025)
Playing telephone with generative models: "verification disability," "compelled reliance," and accessibility in data visualization
di: Elavsky, Frank, et al.
Pubblicazione: (2025)
di: Elavsky, Frank, et al.
Pubblicazione: (2025)
Communicative Agents for Slideshow Storytelling Video Generation based on LLMs
di: Fan, Jingxing, et al.
Pubblicazione: (2025)
di: Fan, Jingxing, et al.
Pubblicazione: (2025)
Reflective Verbal Reward Design for Pluralistic Alignment
di: Blair, Carter, et al.
Pubblicazione: (2025)
di: Blair, Carter, et al.
Pubblicazione: (2025)
AI and My Values: User Perceptions of LLMs' Ability to Extract, Embody, and Explain Human Values from Casual Conversations
di: Yun, Bhada, et al.
Pubblicazione: (2026)
di: Yun, Bhada, et al.
Pubblicazione: (2026)
Sample-Efficient Language Model for Hinglish Conversational AI
di: Singh, Sakshi, et al.
Pubblicazione: (2025)
di: Singh, Sakshi, et al.
Pubblicazione: (2025)
Co-Writing with AI, on Human Terms: Aligning Research with User Demands Across the Writing Process
di: Reza, Mohi, et al.
Pubblicazione: (2025)
di: Reza, Mohi, et al.
Pubblicazione: (2025)
Emergence of Goal-Directed Behaviors via Active Inference with Self-Prior
di: Kim, Dongmin, et al.
Pubblicazione: (2025)
di: Kim, Dongmin, et al.
Pubblicazione: (2025)
Grounded Gesture Generation: Language, Motion, and Space
di: Deichler, Anna, et al.
Pubblicazione: (2025)
di: Deichler, Anna, et al.
Pubblicazione: (2025)
Approaches to Semantic Textual Similarity in Slovak Language: From Algorithms to Transformers
di: Radosky, Lukas, et al.
Pubblicazione: (2026)
di: Radosky, Lukas, et al.
Pubblicazione: (2026)
Vis-CoT: A Human-in-the-Loop Framework for Interactive Visualization and Intervention in LLM Chain-of-Thought Reasoning
di: Pather, Kaviraj, et al.
Pubblicazione: (2025)
di: Pather, Kaviraj, et al.
Pubblicazione: (2025)
Memory-Augmented State Machine Prompting: A Novel LLM Agent Framework for Real-Time Strategy Games
di: Qi, Runnan, et al.
Pubblicazione: (2025)
di: Qi, Runnan, et al.
Pubblicazione: (2025)
Evaluation of Hate Speech Detection Using Large Language Models and Geographical Contextualization
di: Zahid, Anwar Hossain, et al.
Pubblicazione: (2025)
di: Zahid, Anwar Hossain, et al.
Pubblicazione: (2025)
Navigational Thinking as an Emerging Paradigm of Computer Science in the Age of Generative AI
di: Levin, Ilya
Pubblicazione: (2026)
di: Levin, Ilya
Pubblicazione: (2026)
NeuroChat: A Neuroadaptive AI Chatbot for Customizing Learning Experiences
di: Baradari, Dünya, et al.
Pubblicazione: (2025)
di: Baradari, Dünya, et al.
Pubblicazione: (2025)
Conversation Tree Architecture: A Structured Framework for Context-Aware Multi-Branch LLM Conversations
di: Hemanth, Pranav, et al.
Pubblicazione: (2026)
di: Hemanth, Pranav, et al.
Pubblicazione: (2026)
Emotion-Attended Stateful Memory (EASM):The Architecture for Hyper-Personalization at Scale
di: Kotecha, Vineet, et al.
Pubblicazione: (2026)
di: Kotecha, Vineet, et al.
Pubblicazione: (2026)
Benchmarking Educational LLMs with Analytics: A Case Study on Gender Bias in Feedback
di: Du, Yishan, et al.
Pubblicazione: (2025)
di: Du, Yishan, et al.
Pubblicazione: (2025)
TwinVoice: A Multi-dimensional Benchmark Towards Digital Twins via LLM Persona Simulation
di: Du, Bangde, et al.
Pubblicazione: (2025)
di: Du, Bangde, et al.
Pubblicazione: (2025)
Who's Asking? Investigating Bias Through the Lens of Disability Framed Queries in LLMs
di: Hari, Vishnu, et al.
Pubblicazione: (2025)
di: Hari, Vishnu, et al.
Pubblicazione: (2025)
AI for Accessible Education: Personalized Audio-Based Learning for Blind Students
di: Yang, Crystal, et al.
Pubblicazione: (2025)
di: Yang, Crystal, et al.
Pubblicazione: (2025)
Intersymbolic AI: Interlinking Symbolic AI and Subsymbolic AI
di: Platzer, André
Pubblicazione: (2024)
di: Platzer, André
Pubblicazione: (2024)
Semantic Modeling for World-Centered Architectures
di: Mantsivoda, Andrei, et al.
Pubblicazione: (2026)
di: Mantsivoda, Andrei, et al.
Pubblicazione: (2026)
Study on the Helpfulness of Explainable Artificial Intelligence
di: Labarta, Tobias, et al.
Pubblicazione: (2024)
di: Labarta, Tobias, et al.
Pubblicazione: (2024)
Causal Dimensionality of Transformer Representations: Measurement, Scaling, and Layer Structure
di: Sarkar, Nilesh, et al.
Pubblicazione: (2026)
di: Sarkar, Nilesh, et al.
Pubblicazione: (2026)
An Automatic Text Classification Method Based on Hierarchical Taxonomies, Neural Networks and Document Embedding: The NETHIC Tool
di: Lomasto, Luigi, et al.
Pubblicazione: (2026)
di: Lomasto, Luigi, et al.
Pubblicazione: (2026)
Active Inference with a Self-Prior in the Mirror-Mark Task
di: Kim, Dongmin, et al.
Pubblicazione: (2026)
di: Kim, Dongmin, et al.
Pubblicazione: (2026)
h4rm3l: A language for Composable Jailbreak Attack Synthesis
di: Doumbouya, Moussa Koulako Bala, et al.
Pubblicazione: (2024)
di: Doumbouya, Moussa Koulako Bala, et al.
Pubblicazione: (2024)
Improving Fairness with Ensemble Combination: Margin-Dependent Bounds
di: Bian, Yijun
Pubblicazione: (2023)
di: Bian, Yijun
Pubblicazione: (2023)
SafeAnchor: Preventing Cumulative Safety Erosion in Continual Domain Adaptation of Large Language Models
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
di: Guo, Dongxin, et al.
Pubblicazione: (2026)
Informed AI Regulation: Comparing the Ethical Frameworks of Leading LLM Chatbots Using an Ethics-Based Audit to Assess Moral Reasoning and Normative Values
di: Chun, Jon, et al.
Pubblicazione: (2024)
di: Chun, Jon, et al.
Pubblicazione: (2024)
Reframing linguistic bootstrapping as joint inference using visually-grounded grammar induction models
di: Portelance, Eva, et al.
Pubblicazione: (2024)
di: Portelance, Eva, et al.
Pubblicazione: (2024)
Documenti analoghi
-
"I followed what felt right, not what I was told": Autonomy, Coaching, and Recognizing Bias Through AI-Mediated Dialogue
di: Taheri, Atieh, et al.
Pubblicazione: (2026) -
AI-Spectra: A Visual Dashboard for Model Multiplicity to Enhance Informed and Transparent Decision-Making
di: Eerlings, Gilles, et al.
Pubblicazione: (2024) -
What Would GPT Click: Practical Effects of Human-AI Behavioral Misalignment and the Cost of Synthetic Participants in User Experience
di: Kuric, Eduard, et al.
Pubblicazione: (2026) -
Think, Act, Learn: A Framework for Autonomous Robotic Agents using Closed-Loop Large Language Models
di: Menon, Anjali R., et al.
Pubblicazione: (2025) -
Should agentic conversational AI change how we think about ethics? Characterising an interactional ethics centred on respect
di: Alberts, Lize, et al.
Pubblicazione: (2024)