Social Genome: Grounded Social Reasoning Abilities of Multimodal Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mathur, Leena, Qian, Marian, Liang, Paul Pu, Morency, Louis-Philippe |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Advancing Social Intelligence in AI Agents: Technical Challenges and Open Questions
von: Mathur, Leena, et al.
Veröffentlicht: (2024)
von: Mathur, Leena, et al.
Veröffentlicht: (2024)
Social Caption: Evaluating Social Understanding in Multimodal Models
von: Thumu, Bhaavanaa, et al.
Veröffentlicht: (2026)
von: Thumu, Bhaavanaa, et al.
Veröffentlicht: (2026)
HEMM: Holistic Evaluation of Multimodal Foundation Models
von: Liang, Paul Pu, et al.
Veröffentlicht: (2024)
von: Liang, Paul Pu, et al.
Veröffentlicht: (2024)
SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents
von: Zhou, Xuhui, et al.
Veröffentlicht: (2023)
von: Zhou, Xuhui, et al.
Veröffentlicht: (2023)
IoT-LM: Large Multisensory Language Models for the Internet of Things
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
MultiIoT: Benchmarking Machine Learning for the Internet of Things
von: Mo, Shentong, et al.
Veröffentlicht: (2023)
von: Mo, Shentong, et al.
Veröffentlicht: (2023)
Omitted Variable Bias in Language Models Under Distribution Shift
von: Lin, Victoria, et al.
Veröffentlicht: (2026)
von: Lin, Victoria, et al.
Veröffentlicht: (2026)
Optimizing Language Models for Human Preferences is a Causal Inference Problem
von: Lin, Victoria, et al.
Veröffentlicht: (2024)
von: Lin, Victoria, et al.
Veröffentlicht: (2024)
Multimodal Learning Without Labeled Multimodal Data: Guarantees and Applications
von: Liang, Paul Pu, et al.
Veröffentlicht: (2023)
von: Liang, Paul Pu, et al.
Veröffentlicht: (2023)
OpenFace 3.0: A Lightweight Multitask System for Comprehensive Facial Behavior Analysis
von: Hu, Jiewen, et al.
Veröffentlicht: (2025)
von: Hu, Jiewen, et al.
Veröffentlicht: (2025)
MMoE: Enhancing Multimodal Models with Mixtures of Multimodal Interaction Experts
von: Yu, Haofei, et al.
Veröffentlicht: (2023)
von: Yu, Haofei, et al.
Veröffentlicht: (2023)
Improving Dialogue Agents by Decomposing One Global Explicit Annotation with Local Implicit Multimodal Feedback
von: Lee, Dong Won, et al.
Veröffentlicht: (2024)
von: Lee, Dong Won, et al.
Veröffentlicht: (2024)
Social Determinants of Health Prediction for ICD-9 Code with Reasoning Models
von: Khan, Sharim, et al.
Veröffentlicht: (2025)
von: Khan, Sharim, et al.
Veröffentlicht: (2025)
ORION: Teaching Language Models to Reason Efficiently in the Language of Thought
von: Tanmay, Kumar, et al.
Veröffentlicht: (2025)
von: Tanmay, Kumar, et al.
Veröffentlicht: (2025)
On the Reasoning Abilities of Masked Diffusion Language Models
von: Svete, Anej, et al.
Veröffentlicht: (2025)
von: Svete, Anej, et al.
Veröffentlicht: (2025)
Towards Reasoning Ability of Small Language Models
von: Srivastava, Gaurav, et al.
Veröffentlicht: (2025)
von: Srivastava, Gaurav, et al.
Veröffentlicht: (2025)
MultiMed: Massively Multimodal and Multitask Medical Understanding
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
von: Mo, Shentong, et al.
Veröffentlicht: (2024)
Understanding Reasoning Ability of Language Models From the Perspective of Reasoning Paths Aggregation
von: Wang, Xinyi, et al.
Veröffentlicht: (2024)
von: Wang, Xinyi, et al.
Veröffentlicht: (2024)
Reasoning-Grounded Natural Language Explanations for Language Models
von: Cahlik, Vojtech, et al.
Veröffentlicht: (2025)
von: Cahlik, Vojtech, et al.
Veröffentlicht: (2025)
Evaluating and Advancing Multimodal Large Language Models in Perception Ability Lens
von: Chen, Feng, et al.
Veröffentlicht: (2024)
von: Chen, Feng, et al.
Veröffentlicht: (2024)
Grounding Multilingual Multimodal LLMs With Cultural Knowledge
von: Nyandwi, Jean de Dieu, et al.
Veröffentlicht: (2025)
von: Nyandwi, Jean de Dieu, et al.
Veröffentlicht: (2025)
GeoReasoner: Reasoning On Geospatially Grounded Context For Natural Language Understanding
von: Yan, Yibo, et al.
Veröffentlicht: (2024)
von: Yan, Yibo, et al.
Veröffentlicht: (2024)
Relative Kinetic Utility for Reasoning-Aware Structural Pruning in Large Language Models
von: Qian, Tianhao
Veröffentlicht: (2026)
von: Qian, Tianhao
Veröffentlicht: (2026)
A Vision for Multisensory Intelligence: Sensing, Science, and Synergy
von: Liang, Paul Pu
Veröffentlicht: (2026)
von: Liang, Paul Pu
Veröffentlicht: (2026)
$π^2$: Structure-Originated Reasoning Data Improves Long-Context Reasoning Ability of Large Language Models
von: Do, Quyet V., et al.
Veröffentlicht: (2026)
von: Do, Quyet V., et al.
Veröffentlicht: (2026)
Why Reasoning Matters? A Survey of Advancements in Multimodal Reasoning (v1)
von: Bi, Jing, et al.
Veröffentlicht: (2025)
von: Bi, Jing, et al.
Veröffentlicht: (2025)
Emergent Abilities in Reduced-Scale Generative Language Models
von: Muckatira, Sherin, et al.
Veröffentlicht: (2024)
von: Muckatira, Sherin, et al.
Veröffentlicht: (2024)
DeLTa: A Decoding Strategy based on Logit Trajectory Prediction Improves Factuality and Reasoning Ability
von: He, Yunzhen, et al.
Veröffentlicht: (2025)
von: He, Yunzhen, et al.
Veröffentlicht: (2025)
Large Multimodal Models for Low-Resource Languages: A Survey
von: Lupascu, Marian, et al.
Veröffentlicht: (2025)
von: Lupascu, Marian, et al.
Veröffentlicht: (2025)
An Empirical Study of Data Ability Boundary in LLMs' Math Reasoning
von: Chen, Zui, et al.
Veröffentlicht: (2024)
von: Chen, Zui, et al.
Veröffentlicht: (2024)
Do Large Language Models Have Compositional Ability? An Investigation into Limitations and Scalability
von: Xu, Zhuoyan, et al.
Veröffentlicht: (2024)
von: Xu, Zhuoyan, et al.
Veröffentlicht: (2024)
Foundations of Multisensory Artificial Intelligence
von: Liang, Paul Pu
Veröffentlicht: (2024)
von: Liang, Paul Pu
Veröffentlicht: (2024)
GSR-BENCH: A Benchmark for Grounded Spatial Reasoning Evaluation via Multimodal LLMs
von: Rajabi, Navid, et al.
Veröffentlicht: (2024)
von: Rajabi, Navid, et al.
Veröffentlicht: (2024)
Assessing the Emergent Symbolic Reasoning Abilities of Llama Large Language Models
von: Petruzzellis, Flavio, et al.
Veröffentlicht: (2024)
von: Petruzzellis, Flavio, et al.
Veröffentlicht: (2024)
Enhancing Multi-Step Reasoning Abilities of Language Models through Direct Q-Function Optimization
von: Ji, Kaixuan, et al.
Veröffentlicht: (2024)
von: Ji, Kaixuan, et al.
Veröffentlicht: (2024)
Aligning Dialogue Agents with Global Feedback via Large Language Model Multimodal Reward Decomposition
von: Lee, Dong Won, et al.
Veröffentlicht: (2025)
von: Lee, Dong Won, et al.
Veröffentlicht: (2025)
What If the TV Was Off? Examining Counterfactual Reasoning Abilities of Multi-modal Language Models
von: Zhang, Letian, et al.
Veröffentlicht: (2023)
von: Zhang, Letian, et al.
Veröffentlicht: (2023)
Modeling Multimodal Social Interactions: New Challenges and Baselines with Densely Aligned Representations
von: Lee, Sangmin, et al.
Veröffentlicht: (2024)
von: Lee, Sangmin, et al.
Veröffentlicht: (2024)
Graphically Speaking: Unmasking Abuse in Social Media with Conversation Insights
von: Nouri, Célia, et al.
Veröffentlicht: (2025)
von: Nouri, Célia, et al.
Veröffentlicht: (2025)
AgentGen: Enhancing Planning Abilities for Large Language Model based Agent via Environment and Task Generation
von: Hu, Mengkang, et al.
Veröffentlicht: (2024)
von: Hu, Mengkang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Advancing Social Intelligence in AI Agents: Technical Challenges and Open Questions
von: Mathur, Leena, et al.
Veröffentlicht: (2024) -
Social Caption: Evaluating Social Understanding in Multimodal Models
von: Thumu, Bhaavanaa, et al.
Veröffentlicht: (2026) -
HEMM: Holistic Evaluation of Multimodal Foundation Models
von: Liang, Paul Pu, et al.
Veröffentlicht: (2024) -
SOTOPIA: Interactive Evaluation for Social Intelligence in Language Agents
von: Zhou, Xuhui, et al.
Veröffentlicht: (2023) -
IoT-LM: Large Multisensory Language Models for the Internet of Things
von: Mo, Shentong, et al.
Veröffentlicht: (2024)