S$^3$IT: A Benchmark for Spatially Situated Social Intelligence Test
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Zhe, Yang, Xueyuan, Lu, Yujie, Zhang, Zhenliang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Situated Ground Truths: Enhancing Bias-Aware AI by Situating Data Labels with SituAnnotate
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2024)
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2024)
Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situations
von: Park, Eunkyu, et al.
Veröffentlicht: (2025)
von: Park, Eunkyu, et al.
Veröffentlicht: (2025)
Situational Awareness as the Imperative Capability for Disaster Resilience in the Era of Complex Hazards and Artificial Intelligence
von: Pak, Hongrak, et al.
Veröffentlicht: (2025)
von: Pak, Hongrak, et al.
Veröffentlicht: (2025)
Social Intelligence Data Infrastructure: Structuring the Present and Navigating the Future
von: Li, Minzhi, et al.
Veröffentlicht: (2024)
von: Li, Minzhi, et al.
Veröffentlicht: (2024)
TestAgent: An Adaptive and Intelligent Expert for Human Assessment
von: Yu, Junhao, et al.
Veröffentlicht: (2025)
von: Yu, Junhao, et al.
Veröffentlicht: (2025)
An Optimized Evacuation Plan for an Active-Shooter Situation Constrained by Network Capacity
von: Lavalle-Rivera, Joseph, et al.
Veröffentlicht: (2025)
von: Lavalle-Rivera, Joseph, et al.
Veröffentlicht: (2025)
A Comprehensive Survey of Artificial Intelligence Techniques for Talent Analytics
von: Qin, Chuan, et al.
Veröffentlicht: (2023)
von: Qin, Chuan, et al.
Veröffentlicht: (2023)
Fairness Testing of Large Language Models in Role-Playing
von: Li, Xinyue, et al.
Veröffentlicht: (2024)
von: Li, Xinyue, et al.
Veröffentlicht: (2024)
Dialogue with the Machine and Dialogue with the Art World: Evaluating Generative AI for Culturally-Situated Creativity
von: Qadri, Rida, et al.
Veröffentlicht: (2024)
von: Qadri, Rida, et al.
Veröffentlicht: (2024)
Social Bias in Popular Question-Answering Benchmarks
von: Kraft, Angelie, et al.
Veröffentlicht: (2025)
von: Kraft, Angelie, et al.
Veröffentlicht: (2025)
InterveneBench: Benchmarking LLMs for Intervention Reasoning and Causal Study Design in Real Social Systems
von: Shi, Shaojie, et al.
Veröffentlicht: (2026)
von: Shi, Shaojie, et al.
Veröffentlicht: (2026)
Arti-"fickle" Intelligence: Using LLMs as a Tool for Inference in the Political and Social Sciences
von: Argyle, Lisa P., et al.
Veröffentlicht: (2025)
von: Argyle, Lisa P., et al.
Veröffentlicht: (2025)
PluRule: A Benchmark for Moderating Pluralistic Communities on Social Media
von: Kachwala, Zoher, et al.
Veröffentlicht: (2026)
von: Kachwala, Zoher, et al.
Veröffentlicht: (2026)
STOP! Benchmarking Large Language Models with Sensitivity Testing on Offensive Progressions
von: Morabito, Robert, et al.
Veröffentlicht: (2024)
von: Morabito, Robert, et al.
Veröffentlicht: (2024)
Intelligent Computing Social Modeling and Methodological Innovations in Political Science in the Era of Large Language Models
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
von: Wang, Zhenyu, et al.
Veröffentlicht: (2024)
Exploring Multimodal Challenges in Toxic Chinese Detection: Taxonomy, Benchmark, and Findings
von: Yang, Shujian, et al.
Veröffentlicht: (2025)
von: Yang, Shujian, et al.
Veröffentlicht: (2025)
Artificial Intelligence for Collective Intelligence: A National-Scale Research Strategy
von: Bullock, Seth, et al.
Veröffentlicht: (2024)
von: Bullock, Seth, et al.
Veröffentlicht: (2024)
The Interplay of Learning, Analytics, and Artificial Intelligence in Education: A Vision for Hybrid Intelligence
von: Cukurova, Mutlu
Veröffentlicht: (2024)
von: Cukurova, Mutlu
Veröffentlicht: (2024)
ML-EAT: A Multilevel Embedding Association Test for Interpretable and Transparent Social Science
von: Wolfe, Robert, et al.
Veröffentlicht: (2024)
von: Wolfe, Robert, et al.
Veröffentlicht: (2024)
From Mimicry to True Intelligence (TI) -- A New Paradigm for Artificial General Intelligence
von: Subasioglu, Meltem, et al.
Veröffentlicht: (2025)
von: Subasioglu, Meltem, et al.
Veröffentlicht: (2025)
Artificial Intelligence / Human Intelligence: Who Controls Whom?
von: Jacquemot, Charlotte
Veröffentlicht: (2025)
von: Jacquemot, Charlotte
Veröffentlicht: (2025)
Preparing for the Intelligence Explosion
von: MacAskill, William, et al.
Veröffentlicht: (2025)
von: MacAskill, William, et al.
Veröffentlicht: (2025)
Corporations Constitute Intelligence
von: Abiri, Gilad
Veröffentlicht: (2026)
von: Abiri, Gilad
Veröffentlicht: (2026)
ClinBench-HPB: A Clinical Benchmark for Evaluating LLMs in Hepato-Pancreato-Biliary Diseases
von: Li, Yuchong, et al.
Veröffentlicht: (2025)
von: Li, Yuchong, et al.
Veröffentlicht: (2025)
A Framework for the Private Governance of Frontier Artificial Intelligence
von: Ball, Dean W.
Veröffentlicht: (2025)
von: Ball, Dean W.
Veröffentlicht: (2025)
Agency in Artificial Intelligence Systems
von: Das, Parashar
Veröffentlicht: (2025)
von: Das, Parashar
Veröffentlicht: (2025)
Hybrid Intelligence for Digital Humanities
von: de Boer, Victor, et al.
Veröffentlicht: (2024)
von: de Boer, Victor, et al.
Veröffentlicht: (2024)
AI-enhanced Collective Intelligence
von: Cui, Hao, et al.
Veröffentlicht: (2024)
von: Cui, Hao, et al.
Veröffentlicht: (2024)
Benchmarking Large Language Models on Homework Assessment in Circuit Analysis
von: Chen, Liangliang, et al.
Veröffentlicht: (2025)
von: Chen, Liangliang, et al.
Veröffentlicht: (2025)
Explainable Ethical Assessment on Human Behaviors by Generating Conflicting Social Norms
von: Sun, Yuxi, et al.
Veröffentlicht: (2025)
von: Sun, Yuxi, et al.
Veröffentlicht: (2025)
Preconditioned Test-Time Adaptation for Out-of-Distribution Debiasing in Narrative Generation
von: Shen, Hanwen, et al.
Veröffentlicht: (2026)
von: Shen, Hanwen, et al.
Veröffentlicht: (2026)
Digital Domination: A Case for Republican Liberty in Artificial Intelligence
von: Hamilton, Matthew David
Veröffentlicht: (2025)
von: Hamilton, Matthew David
Veröffentlicht: (2025)
Ethics Readiness of Artificial Intelligence: A Practical Evaluation Method
von: Adomaitis, Laurynas, et al.
Veröffentlicht: (2025)
von: Adomaitis, Laurynas, et al.
Veröffentlicht: (2025)
A debate game about societal impacts of Artificial Intelligence
von: Adam, Carole, et al.
Veröffentlicht: (2026)
von: Adam, Carole, et al.
Veröffentlicht: (2026)
MIRA: A Bilingual Benchmark for Medical Information Response Audit
von: Xu, Mengyu, et al.
Veröffentlicht: (2026)
von: Xu, Mengyu, et al.
Veröffentlicht: (2026)
White Men Lead, Black Women Help? Benchmarking and Mitigating Language Agency Social Biases in LLMs
von: Wan, Yixin, et al.
Veröffentlicht: (2024)
von: Wan, Yixin, et al.
Veröffentlicht: (2024)
Social World Model-Augmented Mechanism Design Policy Learning
von: Zhang, Xiaoyuan, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaoyuan, et al.
Veröffentlicht: (2025)
AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies
von: Zeng, Yi, et al.
Veröffentlicht: (2024)
von: Zeng, Yi, et al.
Veröffentlicht: (2024)
Trustworthy Intelligent Education: A Systematic Perspective on Progress, Challenges, and Future Directions
von: Yu, Xiaoshan, et al.
Veröffentlicht: (2026)
von: Yu, Xiaoshan, et al.
Veröffentlicht: (2026)
Citizenship Challenges in Artificial Intelligence Education
von: Romero, Margarida
Veröffentlicht: (2025)
von: Romero, Margarida
Veröffentlicht: (2025)
Ähnliche Einträge
-
Situated Ground Truths: Enhancing Bias-Aware AI by Situating Data Labels with SituAnnotate
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2024) -
Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situations
von: Park, Eunkyu, et al.
Veröffentlicht: (2025) -
Situational Awareness as the Imperative Capability for Disaster Resilience in the Era of Complex Hazards and Artificial Intelligence
von: Pak, Hongrak, et al.
Veröffentlicht: (2025) -
Social Intelligence Data Infrastructure: Structuring the Present and Navigating the Future
von: Li, Minzhi, et al.
Veröffentlicht: (2024) -
TestAgent: An Adaptive and Intelligent Expert for Human Assessment
von: Yu, Junhao, et al.
Veröffentlicht: (2025)