Towards A Litmus Test for Common Sense
Fuente:
arXiv
Saved in:
| Main Author: | Latapie, Hugo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Common Sense Is All You Need
by: Latapie, Hugo
Published: (2025)
by: Latapie, Hugo
Published: (2025)
Mathematical Proof as a Litmus Test: Revealing Failure Modes of Advanced Large Reasoning Models
by: Guo, Dadi, et al.
Published: (2025)
by: Guo, Dadi, et al.
Published: (2025)
A Retrieve-and-Read Framework for Knowledge Graph Link Prediction
by: Pahuja, Vardaan, et al.
Published: (2022)
by: Pahuja, Vardaan, et al.
Published: (2022)
EconEvals: Benchmarks and Litmus Tests for Economic Decision-Making by LLM Agents
by: Fish, Sara, et al.
Published: (2025)
by: Fish, Sara, et al.
Published: (2025)
LIBERO-X: Robustness Litmus for Vision-Language-Action Models
by: Wang, Guodong, et al.
Published: (2026)
by: Wang, Guodong, et al.
Published: (2026)
Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas
by: Chiu, Yu Ying, et al.
Published: (2025)
by: Chiu, Yu Ying, et al.
Published: (2025)
Litmus (Re)Agent: A Benchmark and Agentic System for Predictive Evaluation of Multilingual Models
by: Mittal, Avni, et al.
Published: (2026)
by: Mittal, Avni, et al.
Published: (2026)
Towards a Common Framework for Autoformalization
by: Mensfelt, Agnieszka, et al.
Published: (2025)
by: Mensfelt, Agnieszka, et al.
Published: (2025)
How to Understand Named Entities: Using Common Sense for News Captioning
by: Xu, Ning, et al.
Published: (2024)
by: Xu, Ning, et al.
Published: (2024)
Middleware for LLMs: Tools Are Instrumental for Language Agents in Complex Environments
by: Gu, Yu, et al.
Published: (2024)
by: Gu, Yu, et al.
Published: (2024)
Navigating Semantic Relations: Challenges for Language Models in Abstract Common-Sense Reasoning
by: Gawin, Cole, et al.
Published: (2025)
by: Gawin, Cole, et al.
Published: (2025)
Common Sense vs. Morality: The Curious Case of Narrative Focus Bias in LLMs
by: Purkayastha, Saugata, et al.
Published: (2026)
by: Purkayastha, Saugata, et al.
Published: (2026)
WinoWhat: A Parallel Corpus of Paraphrased WinoGrande Sentences with Common Sense Categorization
by: Gevers, Ine, et al.
Published: (2025)
by: Gevers, Ine, et al.
Published: (2025)
HDReason: Algorithm-Hardware Codesign for Hyperdimensional Knowledge Graph Reasoning
by: Chen, Hanning, et al.
Published: (2024)
by: Chen, Hanning, et al.
Published: (2024)
WorkArena++: Towards Compositional Planning and Reasoning-based Common Knowledge Work Tasks
by: Boisvert, Léo, et al.
Published: (2024)
by: Boisvert, Léo, et al.
Published: (2024)
IPPON: Common Sense Guided Informative Path Planning for Object Goal Navigation
by: Qu, Kaixian, et al.
Published: (2024)
by: Qu, Kaixian, et al.
Published: (2024)
Large Language Models as Common-Sense Heuristics
by: Borro, Andrey, et al.
Published: (2025)
by: Borro, Andrey, et al.
Published: (2025)
Common Belief Revisited
by: Ågotnes, Thomas
Published: (2026)
by: Ågotnes, Thomas
Published: (2026)
FRIDA to the Rescue! Analyzing Synthetic Data Effectiveness in Object-Based Common Sense Reasoning for Disaster Response
by: Shichman, Mollie, et al.
Published: (2025)
by: Shichman, Mollie, et al.
Published: (2025)
The Einstein Test: Towards a Practical Test of a Machine's Ability to Exhibit Superintelligence
by: Benrimoh, David, et al.
Published: (2025)
by: Benrimoh, David, et al.
Published: (2025)
SemEval-2024 Task 9: BRAINTEASER: A Novel Task Defying Common Sense
by: Jiang, Yifan, et al.
Published: (2024)
by: Jiang, Yifan, et al.
Published: (2024)
VQPy: An Object-Oriented Approach to Modern Video Analytics
by: Yu, Shan, et al.
Published: (2023)
by: Yu, Shan, et al.
Published: (2023)
A Proposal to Extend the Common Model of Cognition with Metacognition
by: Laird, John, et al.
Published: (2025)
by: Laird, John, et al.
Published: (2025)
A Reliable Common-Sense Reasoning Socialbot Built Using LLMs and Goal-Directed ASP
by: Zeng, Yankai, et al.
Published: (2024)
by: Zeng, Yankai, et al.
Published: (2024)
Towards Reliable Testing of Machine Unlearning
by: Mazhar, Anna, et al.
Published: (2026)
by: Mazhar, Anna, et al.
Published: (2026)
Cognitive Architecture Toward Common Ground Sharing Among Humans and Generative AIs: Trial on Model-Model Interactions in Tangram Naming Task
by: Morita, Junya, et al.
Published: (2023)
by: Morita, Junya, et al.
Published: (2023)
Co-Producing AI: Toward an Augmented, Participatory Lifecycle
by: Mushkani, Rashid, et al.
Published: (2025)
by: Mushkani, Rashid, et al.
Published: (2025)
Common Practices and Taxonomy in Deep Multi-view Fusion for Remote Sensing Applications
by: Mena, Francisco, et al.
Published: (2022)
by: Mena, Francisco, et al.
Published: (2022)
Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning
by: NVIDIA, et al.
Published: (2025)
by: NVIDIA, et al.
Published: (2025)
Where Common Knowledge Cannot Be Formed, Common Belief Can -- Planning with Multi-Agent Belief Using Group Justified Perspectives
by: Hu, Guang, et al.
Published: (2024)
by: Hu, Guang, et al.
Published: (2024)
A Learning Search Algorithm for the Restricted Longest Common Subsequence Problem
by: Djukanović, Marko, et al.
Published: (2024)
by: Djukanović, Marko, et al.
Published: (2024)
EvolMathEval: Towards Evolvable Benchmarks for Mathematical Reasoning via Evolutionary Testing
by: Wang, Shengbo, et al.
Published: (2025)
by: Wang, Shengbo, et al.
Published: (2025)
A Pluggable Common Sense-Enhanced Framework for Knowledge Graph Completion
by: Niu, Guanglin, et al.
Published: (2024)
by: Niu, Guanglin, et al.
Published: (2024)
CommonIT: Commonality-Aware Instruction Tuning for Large Language Models via Data Partitions
by: Rao, Jun, et al.
Published: (2024)
by: Rao, Jun, et al.
Published: (2024)
Voice-Enabled AI Agents can Perform Common Scams
by: Fang, Richard, et al.
Published: (2024)
by: Fang, Richard, et al.
Published: (2024)
On Solving the Multiple Variable Gapped Longest Common Subsequence Problem
by: Djukanović, Marko, et al.
Published: (2026)
by: Djukanović, Marko, et al.
Published: (2026)
Towards Practical Multi-label Causal Discovery in High-Dimensional Event Sequences via One-Shot Graph Aggregation
by: Math, Hugo, et al.
Published: (2025)
by: Math, Hugo, et al.
Published: (2025)
Towards A Cultural Intelligence and Values Inferences Quality Benchmark for Community Values and Common Knowledge
by: Johnson, Brittany, et al.
Published: (2025)
by: Johnson, Brittany, et al.
Published: (2025)
A Proposal for Extending the Common Model of Cognition to Emotion
by: Rosenbloom, Paul S., et al.
Published: (2024)
by: Rosenbloom, Paul S., et al.
Published: (2024)
SocRATES: Towards Automated Scenario-based Testing of Social Navigation Algorithms
by: Marpally, Shashank Rao, et al.
Published: (2024)
by: Marpally, Shashank Rao, et al.
Published: (2024)
Similar Items
-
Common Sense Is All You Need
by: Latapie, Hugo
Published: (2025) -
Mathematical Proof as a Litmus Test: Revealing Failure Modes of Advanced Large Reasoning Models
by: Guo, Dadi, et al.
Published: (2025) -
A Retrieve-and-Read Framework for Knowledge Graph Link Prediction
by: Pahuja, Vardaan, et al.
Published: (2022) -
EconEvals: Benchmarks and Litmus Tests for Economic Decision-Making by LLM Agents
by: Fish, Sara, et al.
Published: (2025) -
LIBERO-X: Robustness Litmus for Vision-Language-Action Models
by: Wang, Guodong, et al.
Published: (2026)