Do Machines Fail Like Humans? A Human-Centred Out-of-Distribution Spectrum for Mapping Error Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Binxia, Luo, Xiaoliang, Dickens, Luke, Mok, Robert M. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Measuring Error Alignment for Decision-Making Systems
von: Xu, Binxia, et al.
Veröffentlicht: (2024)
von: Xu, Binxia, et al.
Veröffentlicht: (2024)
On Benchmarking Human-Like Intelligence in Machines
von: Ying, Lance, et al.
Veröffentlicht: (2025)
von: Ying, Lance, et al.
Veröffentlicht: (2025)
Assessing Social Alignment: Do Personality-Prompted Large Language Models Behave Like Humans?
von: Zakazov, Ivan, et al.
Veröffentlicht: (2024)
von: Zakazov, Ivan, et al.
Veröffentlicht: (2024)
A Machine With Human-Like Memory Systems
von: Kim, Taewoon, et al.
Veröffentlicht: (2022)
von: Kim, Taewoon, et al.
Veröffentlicht: (2022)
LLMs Do Not Grade Essays Like Humans
von: Mathew, Jerin George, et al.
Veröffentlicht: (2026)
von: Mathew, Jerin George, et al.
Veröffentlicht: (2026)
Representation-based Broad Hallucination Detectors Fail to Generalize Out of Distribution
von: Dubanowska, Zuzanna, et al.
Veröffentlicht: (2025)
von: Dubanowska, Zuzanna, et al.
Veröffentlicht: (2025)
Map Imagination Like Blind Humans: Group Diffusion Model for Robotic Map Generation
von: Song, Qijin, et al.
Veröffentlicht: (2024)
von: Song, Qijin, et al.
Veröffentlicht: (2024)
Is Human-Like Text Liked by Humans? Multilingual Human Detection and Preference Against AI
von: Wang, Yuxia, et al.
Veröffentlicht: (2025)
von: Wang, Yuxia, et al.
Veröffentlicht: (2025)
Taming False Positives in Out-of-Distribution Detection with Human Feedback
von: Vishwakarma, Harit, et al.
Veröffentlicht: (2024)
von: Vishwakarma, Harit, et al.
Veröffentlicht: (2024)
Teaching Values to Machines: Simulating Human-Like Behavior in LLMs
von: Yehudai, Asaf, et al.
Veröffentlicht: (2026)
von: Yehudai, Asaf, et al.
Veröffentlicht: (2026)
How Likely Do LLMs with CoT Mimic Human Reasoning?
von: Bao, Guangsheng, et al.
Veröffentlicht: (2024)
von: Bao, Guangsheng, et al.
Veröffentlicht: (2024)
Language Models Exhibit Inconsistent Biases Towards Algorithmic Agents and Human Experts
von: Bo, Jessica Y., et al.
Veröffentlicht: (2026)
von: Bo, Jessica Y., et al.
Veröffentlicht: (2026)
Measuring AI Alignment with Human Flourishing
von: Hilliard, Elizabeth, et al.
Veröffentlicht: (2025)
von: Hilliard, Elizabeth, et al.
Veröffentlicht: (2025)
Adaptive Scoring and Thresholding with Human Feedback for Robust Out-of-Distribution Detection
von: Yamada, Daisuke, et al.
Veröffentlicht: (2025)
von: Yamada, Daisuke, et al.
Veröffentlicht: (2025)
Do LLMs Share Human-Like Biases? Causal Reasoning Under Prior Knowledge, Irrelevant Context, and Varying Compute Budgets
von: Dettki, Hanna M., et al.
Veröffentlicht: (2026)
von: Dettki, Hanna M., et al.
Veröffentlicht: (2026)
Do Large Language Models Learn Human-Like Strategic Preferences?
von: Roberts, Jesse, et al.
Veröffentlicht: (2024)
von: Roberts, Jesse, et al.
Veröffentlicht: (2024)
A Translation of Probabilistic Event Calculus into Markov Decision Processes
von: Xu, Lyris, et al.
Veröffentlicht: (2025)
von: Xu, Lyris, et al.
Veröffentlicht: (2025)
Getting More Juice Out of the SFT Data: Reward Learning from Human Demonstration Improves SFT for LLM Alignment
von: Li, Jiaxiang, et al.
Veröffentlicht: (2024)
von: Li, Jiaxiang, et al.
Veröffentlicht: (2024)
Connecting Concept Convexity and Human-Machine Alignment in Deep Neural Networks
von: Dorszewski, Teresa, et al.
Veröffentlicht: (2024)
von: Dorszewski, Teresa, et al.
Veröffentlicht: (2024)
Learning Human-Like Badminton Skills for Humanoid Robots
von: Chen, Yeke, et al.
Veröffentlicht: (2026)
von: Chen, Yeke, et al.
Veröffentlicht: (2026)
Why Do Multi-Agent LLM Systems Fail?
von: Cemri, Mert, et al.
Veröffentlicht: (2025)
von: Cemri, Mert, et al.
Veröffentlicht: (2025)
Advancing Frontiers in SLAM: A Survey of Symbolic Representation and Human-Machine Teaming in Environmental Mapping
von: Colelough, Brandon Curtis
Veröffentlicht: (2024)
von: Colelough, Brandon Curtis
Veröffentlicht: (2024)
Psychometric Alignment: Capturing Human Knowledge Distributions via Language Models
von: He-Yueya, Joy, et al.
Veröffentlicht: (2024)
von: He-Yueya, Joy, et al.
Veröffentlicht: (2024)
Benchmarking and Improving Monitors for Out-Of-Distribution Alignment Failure in LLMs
von: Feng, Dylan, et al.
Veröffentlicht: (2026)
von: Feng, Dylan, et al.
Veröffentlicht: (2026)
Constitutive Components for Human-Like Autonomous Artificial Intelligence
von: Yamada, Kazunori D
Veröffentlicht: (2025)
von: Yamada, Kazunori D
Veröffentlicht: (2025)
Uncovering the Computational Ingredients of Human-Like Representations in LLMs
von: Studdiford, Zach, et al.
Veröffentlicht: (2025)
von: Studdiford, Zach, et al.
Veröffentlicht: (2025)
Do LLMs Exhibit Human-Like Reasoning? Evaluating Theory of Mind in LLMs for Open-Ended Responses
von: Amirizaniani, Maryam, et al.
Veröffentlicht: (2024)
von: Amirizaniani, Maryam, et al.
Veröffentlicht: (2024)
LLMs for Qualitative Data Analysis Fail on Security-specificComments in Human Experiments
von: Camporese, Maria, et al.
Veröffentlicht: (2026)
von: Camporese, Maria, et al.
Veröffentlicht: (2026)
Negating Negatives: Alignment with Human Negative Samples via Distributional Dispreference Optimization
von: Duan, Shitong, et al.
Veröffentlicht: (2024)
von: Duan, Shitong, et al.
Veröffentlicht: (2024)
DMA: Online RAG Alignment with Human Feedback
von: Bai, Yu, et al.
Veröffentlicht: (2025)
von: Bai, Yu, et al.
Veröffentlicht: (2025)
Towards Human-Like Trajectory Prediction for Autonomous Driving: A Behavior-Centric Approach
von: Liao, Haicheng, et al.
Veröffentlicht: (2025)
von: Liao, Haicheng, et al.
Veröffentlicht: (2025)
Human/AI Collective Intelligence for Deliberative Democracy: A Human-Centred Design Approach
von: De Liddo, Anna, et al.
Veröffentlicht: (2026)
von: De Liddo, Anna, et al.
Veröffentlicht: (2026)
Adapting Like Humans: A Metacognitive Agent with Test-time Reasoning
von: Li, Yang, et al.
Veröffentlicht: (2025)
von: Li, Yang, et al.
Veröffentlicht: (2025)
CLHA: A Simple yet Effective Contrastive Learning Framework for Human Alignment
von: Fang, Feiteng, et al.
Veröffentlicht: (2024)
von: Fang, Feiteng, et al.
Veröffentlicht: (2024)
Simulating Human-Like Learning Dynamics with LLM-Empowered Agents
von: Yuan, Yu, et al.
Veröffentlicht: (2025)
von: Yuan, Yu, et al.
Veröffentlicht: (2025)
VisDoT : Enhancing Visual Reasoning through Human-Like Interpretation Grounding and Decomposition of Thought
von: Lee, Eunsoo, et al.
Veröffentlicht: (2026)
von: Lee, Eunsoo, et al.
Veröffentlicht: (2026)
A Practical Analysis of Human Alignment with *PO
von: Ahrabian, Kian, et al.
Veröffentlicht: (2024)
von: Ahrabian, Kian, et al.
Veröffentlicht: (2024)
Quantum Adiabatic Generation of Human-Like Passwords
von: Mücke, Sascha, et al.
Veröffentlicht: (2025)
von: Mücke, Sascha, et al.
Veröffentlicht: (2025)
Aligning Generalisation Between Humans and Machines
von: Ilievski, Filip, et al.
Veröffentlicht: (2024)
von: Ilievski, Filip, et al.
Veröffentlicht: (2024)
Can Large Language Models Express Uncertainty Like Human?
von: Tao, Linwei, et al.
Veröffentlicht: (2025)
von: Tao, Linwei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Measuring Error Alignment for Decision-Making Systems
von: Xu, Binxia, et al.
Veröffentlicht: (2024) -
On Benchmarking Human-Like Intelligence in Machines
von: Ying, Lance, et al.
Veröffentlicht: (2025) -
Assessing Social Alignment: Do Personality-Prompted Large Language Models Behave Like Humans?
von: Zakazov, Ivan, et al.
Veröffentlicht: (2024) -
A Machine With Human-Like Memory Systems
von: Kim, Taewoon, et al.
Veröffentlicht: (2022) -
LLMs Do Not Grade Essays Like Humans
von: Mathew, Jerin George, et al.
Veröffentlicht: (2026)