Exploring Failure Cases in Multimodal Reasoning About Physical Dynamics
Fuente:
arXiv
Salvato in:
| Autori principali: | Ghaffari, Sadaf, Krishnaswamy, Nikhil |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Cross-Lingual Transfer Robustness to Lower-Resource Languages on Adversarial Datasets
di: Manafi, Shadi, et al.
Pubblicazione: (2024)
di: Manafi, Shadi, et al.
Pubblicazione: (2024)
Frictive Policy Optimization for LLMs: Epistemic Intervention, Risk-Sensitive Control, and Reflective Alignment
di: Pustejovsky, James, et al.
Pubblicazione: (2026)
di: Pustejovsky, James, et al.
Pubblicazione: (2026)
Dynamic Epistemic Friction in Dialogue
di: Obiso, Timothy, et al.
Pubblicazione: (2025)
di: Obiso, Timothy, et al.
Pubblicazione: (2025)
The Impact of Background Speech on Interruption Detection in Collaborative Groups
di: Bradford, Mariah, et al.
Pubblicazione: (2025)
di: Bradford, Mariah, et al.
Pubblicazione: (2025)
Okay, Let's Do This! Modeling Event Coreference with Generated Rationales and Knowledge Distillation
di: Nath, Abhijnan, et al.
Pubblicazione: (2024)
di: Nath, Abhijnan, et al.
Pubblicazione: (2024)
Frictional Agent Alignment Framework: Slow Down and Don't Break Things
di: Nath, Abhijnan, et al.
Pubblicazione: (2025)
di: Nath, Abhijnan, et al.
Pubblicazione: (2025)
CRAFT: Grounded Multi-Agent Coordination Under Partial Information
di: Nath, Abhijnan, et al.
Pubblicazione: (2026)
di: Nath, Abhijnan, et al.
Pubblicazione: (2026)
Collaborate, Deliberate, Evaluate: How LLM Alignment Affects Coordinated Multi-Agent Outcomes
di: Nath, Abhijnan, et al.
Pubblicazione: (2025)
di: Nath, Abhijnan, et al.
Pubblicazione: (2025)
Simultaneous Reward Distillation and Preference Learning: Get You a Language Model Who Can Do Both
di: Nath, Abhijnan, et al.
Pubblicazione: (2024)
di: Nath, Abhijnan, et al.
Pubblicazione: (2024)
Common Ground Tracking in Multimodal Dialogue
di: Khebour, Ibrahim, et al.
Pubblicazione: (2024)
di: Khebour, Ibrahim, et al.
Pubblicazione: (2024)
Multimodal Cross-Document Event Coreference Resolution Using Linear Semantic Transfer and Mixed-Modality Ensembles
di: Nath, Abhijnan, et al.
Pubblicazione: (2024)
di: Nath, Abhijnan, et al.
Pubblicazione: (2024)
Computational Thought Experiments for a More Rigorous Philosophy and Science of the Mind
di: Oved, Iris, et al.
Pubblicazione: (2024)
di: Oved, Iris, et al.
Pubblicazione: (2024)
Dissecting Failure Dynamics in Large Language Model Reasoning
di: Zhu, Wei, et al.
Pubblicazione: (2026)
di: Zhu, Wei, et al.
Pubblicazione: (2026)
Reasoning Models Will Sometimes Lie About Their Reasoning
di: Walden, William, et al.
Pubblicazione: (2026)
di: Walden, William, et al.
Pubblicazione: (2026)
Exploring and Evaluating Multimodal Knowledge Reasoning Consistency of Multimodal Large Language Models
di: Jia, Boyu, et al.
Pubblicazione: (2025)
di: Jia, Boyu, et al.
Pubblicazione: (2025)
TRACE: Real-Time Multimodal Common Ground Tracking in Situated Collaborative Dialogues
di: VanderHoeven, Hannah, et al.
Pubblicazione: (2025)
di: VanderHoeven, Hannah, et al.
Pubblicazione: (2025)
Any Other Thoughts, Hedgehog? Linking Deliberation Chains in Collaborative Dialogues
di: Nath, Abhijnan, et al.
Pubblicazione: (2024)
di: Nath, Abhijnan, et al.
Pubblicazione: (2024)
StressEval: Failure-Driven Dynamic Benchmarking for Knowledge-Intensive Reasoning in Large Language Models
di: Chen, Yongrui, et al.
Pubblicazione: (2026)
di: Chen, Yongrui, et al.
Pubblicazione: (2026)
Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning
di: Wang, Yiqi, et al.
Pubblicazione: (2024)
di: Wang, Yiqi, et al.
Pubblicazione: (2024)
Distributed Partial Information Puzzles: Examining Common Ground Construction Under Epistemic Asymmetry
di: Zhu, Yifan, et al.
Pubblicazione: (2026)
di: Zhu, Yifan, et al.
Pubblicazione: (2026)
Reasoning About the Unsaid: Misinformation Detection with Omission-Aware Graph Inference
di: Wang, Zhengjia, et al.
Pubblicazione: (2025)
di: Wang, Zhengjia, et al.
Pubblicazione: (2025)
Guided Verifier: Collaborative Multimodal Reasoning via Dynamic Process Supervision
di: Sun, Lingzhuang, et al.
Pubblicazione: (2026)
di: Sun, Lingzhuang, et al.
Pubblicazione: (2026)
PhysicsArena: The First Multimodal Physics Reasoning Benchmark Exploring Variable, Process, and Solution Dimensions
di: Dai, Song, et al.
Pubblicazione: (2025)
di: Dai, Song, et al.
Pubblicazione: (2025)
The Reasoning Error About Reasoning: Why Different Types of Reasoning Require Different Representational Structures
di: Wu, Yiling
Pubblicazione: (2026)
di: Wu, Yiling
Pubblicazione: (2026)
Unsupervised Neural Network for Automated Classification of Surgical Urgency Levels in Medical Transcriptions
di: Tabatabaee, Sadaf, et al.
Pubblicazione: (2026)
di: Tabatabaee, Sadaf, et al.
Pubblicazione: (2026)
Garbage In, Reasoning Out? Why Benchmark Scores are Unreliable and What to Do About It
di: Mousavi, Seyed Mahed, et al.
Pubblicazione: (2025)
di: Mousavi, Seyed Mahed, et al.
Pubblicazione: (2025)
Two Failures of Self-Consistency in the Multi-Step Reasoning of LLMs
di: Chen, Angelica, et al.
Pubblicazione: (2023)
di: Chen, Angelica, et al.
Pubblicazione: (2023)
Reasoning Within the Mind: Dynamic Multimodal Interleaving in Latent Space
di: Liu, Chengzhi, et al.
Pubblicazione: (2025)
di: Liu, Chengzhi, et al.
Pubblicazione: (2025)
Should We be Pedantic About Reasoning Errors in Machine Translation?
di: Bao, Calvin, et al.
Pubblicazione: (2026)
di: Bao, Calvin, et al.
Pubblicazione: (2026)
How do Humans and Language Models Reason About Creativity? A Comparative Analysis
di: Laverghetta Jr., Antonio, et al.
Pubblicazione: (2025)
di: Laverghetta Jr., Antonio, et al.
Pubblicazione: (2025)
Multi-Physics: A Comprehensive Benchmark for Multimodal LLMs Reasoning on Chinese Multi-Subject Physics Problems
di: Luo, Zhongze, et al.
Pubblicazione: (2025)
di: Luo, Zhongze, et al.
Pubblicazione: (2025)
Reasoning About Exceptional Behavior At the Level of Java Bytecode
di: Paganoni, Marco, et al.
Pubblicazione: (2024)
di: Paganoni, Marco, et al.
Pubblicazione: (2024)
Failure Modes of LLMs for Causal Reasoning on Narratives
di: Yamin, Khurram, et al.
Pubblicazione: (2024)
di: Yamin, Khurram, et al.
Pubblicazione: (2024)
Mind with Eyes: from Language Reasoning to Multimodal Reasoning
di: Lin, Zhiyu, et al.
Pubblicazione: (2025)
di: Lin, Zhiyu, et al.
Pubblicazione: (2025)
Reasoning-Table: Exploring Reinforcement Learning for Table Reasoning
di: Lei, Fangyu, et al.
Pubblicazione: (2025)
di: Lei, Fangyu, et al.
Pubblicazione: (2025)
Exploring Defeasibility in Causal Reasoning
di: Cui, Shaobo, et al.
Pubblicazione: (2024)
di: Cui, Shaobo, et al.
Pubblicazione: (2024)
Exploring the Potential of Multimodal LLM with Knowledge-Intensive Multimodal ASR
di: Wang, Minghan, et al.
Pubblicazione: (2024)
di: Wang, Minghan, et al.
Pubblicazione: (2024)
Bidirectional Human-AI Learning in Real-Time Disoriented Balancing
di: Mannan, Sheikh, et al.
Pubblicazione: (2024)
di: Mannan, Sheikh, et al.
Pubblicazione: (2024)
Can LLMs Reason About Trust?: A Pilot Study
di: Debnath, Anushka, et al.
Pubblicazione: (2025)
di: Debnath, Anushka, et al.
Pubblicazione: (2025)
Teaching and Evaluating LLMs to Reason About Polymer Design Related Tasks
di: Mohanty, Dikshya, et al.
Pubblicazione: (2026)
di: Mohanty, Dikshya, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Cross-Lingual Transfer Robustness to Lower-Resource Languages on Adversarial Datasets
di: Manafi, Shadi, et al.
Pubblicazione: (2024) -
Frictive Policy Optimization for LLMs: Epistemic Intervention, Risk-Sensitive Control, and Reflective Alignment
di: Pustejovsky, James, et al.
Pubblicazione: (2026) -
Dynamic Epistemic Friction in Dialogue
di: Obiso, Timothy, et al.
Pubblicazione: (2025) -
The Impact of Background Speech on Interruption Detection in Collaborative Groups
di: Bradford, Mariah, et al.
Pubblicazione: (2025) -
Okay, Let's Do This! Modeling Event Coreference with Generated Rationales and Knowledge Distillation
di: Nath, Abhijnan, et al.
Pubblicazione: (2024)