Advancing Social Intelligence in AI Agents: Technical Challenges and Open Questions
Fuente:
arXiv
Salvato in:
| Autori principali: | Mathur, Leena, Liang, Paul Pu, Morency, Louis-Philippe |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Social Genome: Grounded Social Reasoning Abilities of Multimodal Models
di: Mathur, Leena, et al.
Pubblicazione: (2025)
di: Mathur, Leena, et al.
Pubblicazione: (2025)
Improving Dialogue Agents by Decomposing One Global Explicit Annotation with Local Implicit Multimodal Feedback
di: Lee, Dong Won, et al.
Pubblicazione: (2024)
di: Lee, Dong Won, et al.
Pubblicazione: (2024)
Social Caption: Evaluating Social Understanding in Multimodal Models
di: Thumu, Bhaavanaa, et al.
Pubblicazione: (2026)
di: Thumu, Bhaavanaa, et al.
Pubblicazione: (2026)
HybridQuestion: Human-AI Collaboration for Identifying High-Impact Research Questions
di: Zhao, Keyu, et al.
Pubblicazione: (2025)
di: Zhao, Keyu, et al.
Pubblicazione: (2025)
Advanced Machine Learning Techniques for Social Support Detection on Social Media
di: Kolesnikova, Olga, et al.
Pubblicazione: (2025)
di: Kolesnikova, Olga, et al.
Pubblicazione: (2025)
Almost AI, Almost Human: The Challenge of Detecting AI-Polished Writing
di: Saha, Shoumik, et al.
Pubblicazione: (2025)
di: Saha, Shoumik, et al.
Pubblicazione: (2025)
ETS: Open Vocabulary Electroencephalography-To-Text Decoding and Sentiment Classification
di: Masry, Mohamed, et al.
Pubblicazione: (2025)
di: Masry, Mohamed, et al.
Pubblicazione: (2025)
Processes Matter: How ML/GAI Approaches Could Support Open Qualitative Coding of Online Discourse Datasets
di: Chen, John, et al.
Pubblicazione: (2025)
di: Chen, John, et al.
Pubblicazione: (2025)
Training AI Co-Scientists Using Rubric Rewards
di: Goel, Shashwat, et al.
Pubblicazione: (2025)
di: Goel, Shashwat, et al.
Pubblicazione: (2025)
Survey of User Interface Design and Interaction Techniques in Generative AI Applications
di: Luera, Reuben, et al.
Pubblicazione: (2024)
di: Luera, Reuben, et al.
Pubblicazione: (2024)
LLAMADRS: Evaluating Open-Source LLMs on Real Clinical Interviews--To Reason or Not to Reason?
di: Kebe, Gaoussou Youssouf, et al.
Pubblicazione: (2025)
di: Kebe, Gaoussou Youssouf, et al.
Pubblicazione: (2025)
Depression detection from Social Media Bangla Text Using Recurrent Neural Networks
di: Ahmed, Sultan, et al.
Pubblicazione: (2024)
di: Ahmed, Sultan, et al.
Pubblicazione: (2024)
A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI Autonomy
di: Zou, Henry Peng, et al.
Pubblicazione: (2025)
di: Zou, Henry Peng, et al.
Pubblicazione: (2025)
(Ir)rationality in AI: State of the Art, Research Challenges and Open Questions
di: Macmillan-Scott, Olivia, et al.
Pubblicazione: (2023)
di: Macmillan-Scott, Olivia, et al.
Pubblicazione: (2023)
SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents
di: Zhou, Xuhui, et al.
Pubblicazione: (2023)
di: Zhou, Xuhui, et al.
Pubblicazione: (2023)
Probing the Multi-turn Planning Capabilities of LLMs via 20 Question Games
di: Zhang, Yizhe, et al.
Pubblicazione: (2023)
di: Zhang, Yizhe, et al.
Pubblicazione: (2023)
Are Open-Weight LLMs Ready for Social Media Moderation? A Comparative Study on Bluesky
di: Chou, Hsuan-Yu, et al.
Pubblicazione: (2026)
di: Chou, Hsuan-Yu, et al.
Pubblicazione: (2026)
HEMM: Holistic Evaluation of Multimodal Foundation Models
di: Liang, Paul Pu, et al.
Pubblicazione: (2024)
di: Liang, Paul Pu, et al.
Pubblicazione: (2024)
Agent Laboratory: Using LLM Agents as Research Assistants
di: Schmidgall, Samuel, et al.
Pubblicazione: (2025)
di: Schmidgall, Samuel, et al.
Pubblicazione: (2025)
Evaluating the Usability of LLMs in Threat Intelligence Enrichment
di: Srikanth, Sanchana, et al.
Pubblicazione: (2024)
di: Srikanth, Sanchana, et al.
Pubblicazione: (2024)
Evaluating Explanations: An Explanatory Virtues Framework for Mechanistic Interpretability -- The Strange Science Part I.ii
di: Ayonrinde, Kola, et al.
Pubblicazione: (2025)
di: Ayonrinde, Kola, et al.
Pubblicazione: (2025)
Detecting and Preventing Harmful Behaviors in AI Companions: Development and Evaluation of the SHIELD Supervisory System
di: Ben-Zion, Ziv, et al.
Pubblicazione: (2025)
di: Ben-Zion, Ziv, et al.
Pubblicazione: (2025)
AutoGLM: Autonomous Foundation Agents for GUIs
di: Liu, Xiao, et al.
Pubblicazione: (2024)
di: Liu, Xiao, et al.
Pubblicazione: (2024)
Open-Source Large Language Models as Multilingual Crowdworkers: Synthesizing Open-Domain Dialogues in Several Languages With No Examples in Targets and No Machine Translation
di: Njifenjou, Ahmed, et al.
Pubblicazione: (2025)
di: Njifenjou, Ahmed, et al.
Pubblicazione: (2025)
EmoAgent: Assessing and Safeguarding Human-AI Interaction for Mental Health Safety
di: Qiu, Jiahao, et al.
Pubblicazione: (2025)
di: Qiu, Jiahao, et al.
Pubblicazione: (2025)
PsychoGAT: A Novel Psychological Measurement Paradigm through Interactive Fiction Games with LLM Agents
di: Yang, Qisen, et al.
Pubblicazione: (2024)
di: Yang, Qisen, et al.
Pubblicazione: (2024)
Humanizing LLMs: A Survey of Psychological Measurements with Tools, Datasets, and Human-Agent Applications
di: Dong, Wenhan, et al.
Pubblicazione: (2025)
di: Dong, Wenhan, et al.
Pubblicazione: (2025)
Beyond Semantic Similarity: A Component-Wise Evaluation Framework for Medical Question Answering Systems with Health Equity Implications
di: Sakib, Abu Noman Md, et al.
Pubblicazione: (2026)
di: Sakib, Abu Noman Md, et al.
Pubblicazione: (2026)
A Computational Method for Measuring "Open Codes" in Qualitative Analysis
di: Chen, John, et al.
Pubblicazione: (2024)
di: Chen, John, et al.
Pubblicazione: (2024)
Deep Representation Learning for Open Vocabulary Electroencephalography-to-Text Decoding
di: Amrani, Hamza, et al.
Pubblicazione: (2023)
di: Amrani, Hamza, et al.
Pubblicazione: (2023)
Future of Work with AI Agents: Auditing Automation and Augmentation Potential across the U.S. Workforce
di: Shao, Yijia, et al.
Pubblicazione: (2025)
di: Shao, Yijia, et al.
Pubblicazione: (2025)
From Measurement to Expertise: Empathetic Expert Adapters for Context-Based Empathy in Conversational AI Agents
di: Shayegani, Erfan, et al.
Pubblicazione: (2025)
di: Shayegani, Erfan, et al.
Pubblicazione: (2025)
Can Generative AI Support Patients' & Caregivers' Informational Needs? Towards Task-Centric Evaluation Of AI Systems
di: Rajagopal, Shreya, et al.
Pubblicazione: (2024)
di: Rajagopal, Shreya, et al.
Pubblicazione: (2024)
A Framework for Evaluating LLMs Under Task Indeterminacy
di: Guerdan, Luke, et al.
Pubblicazione: (2024)
di: Guerdan, Luke, et al.
Pubblicazione: (2024)
Leveraging Variation Theory in Counterfactual Data Augmentation for Optimized Active Learning
di: Gebreegziabher, Simret Araya, et al.
Pubblicazione: (2024)
di: Gebreegziabher, Simret Araya, et al.
Pubblicazione: (2024)
Multi-Level Feedback Generation with Large Language Models for Empowering Novice Peer Counselors
di: Chaszczewicz, Alicja, et al.
Pubblicazione: (2024)
di: Chaszczewicz, Alicja, et al.
Pubblicazione: (2024)
'Simulacrum of Stories': Examining Large Language Models as Qualitative Research Participants
di: Kapania, Shivani, et al.
Pubblicazione: (2024)
di: Kapania, Shivani, et al.
Pubblicazione: (2024)
JailbreakHunter: A Visual Analytics Approach for Jailbreak Prompts Discovery from Large-Scale Human-LLM Conversational Datasets
di: Jin, Zhihua, et al.
Pubblicazione: (2024)
di: Jin, Zhihua, et al.
Pubblicazione: (2024)
Sample-Efficient Human Evaluation of Large Language Models via Maximum Discrepancy Competition
di: Feng, Kehua, et al.
Pubblicazione: (2024)
di: Feng, Kehua, et al.
Pubblicazione: (2024)
Exploring Empty Spaces: Human-in-the-Loop Data Augmentation
di: Yeh, Catherine, et al.
Pubblicazione: (2024)
di: Yeh, Catherine, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Social Genome: Grounded Social Reasoning Abilities of Multimodal Models
di: Mathur, Leena, et al.
Pubblicazione: (2025) -
Improving Dialogue Agents by Decomposing One Global Explicit Annotation with Local Implicit Multimodal Feedback
di: Lee, Dong Won, et al.
Pubblicazione: (2024) -
Social Caption: Evaluating Social Understanding in Multimodal Models
di: Thumu, Bhaavanaa, et al.
Pubblicazione: (2026) -
HybridQuestion: Human-AI Collaboration for Identifying High-Impact Research Questions
di: Zhao, Keyu, et al.
Pubblicazione: (2025) -
Advanced Machine Learning Techniques for Social Support Detection on Social Media
di: Kolesnikova, Olga, et al.
Pubblicazione: (2025)