Towards Quantifying Commonsense Reasoning with Mechanistic Insights
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Joshi, Abhinav, Ahmad, Areeb, Shukla, Divyaksh, Modi, Ashutosh |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Beyond Components: Singular Vector-Based Interpretability of Transformer Circuits
von: Ahmad, Areeb, et al.
Veröffentlicht: (2025)
von: Ahmad, Areeb, et al.
Veröffentlicht: (2025)
Calibration Across Layers: Understanding Calibration Evolution in LLMs
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
COLD: Causal reasOning in cLosed Daily activities
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
Calibration vs Decision Making: Revisiting the Reliability Paradox in Unlearned Language Models
von: Shukla, Divyaksh, et al.
Veröffentlicht: (2026)
von: Shukla, Divyaksh, et al.
Veröffentlicht: (2026)
IITK at SemEval-2024 Task 10: Who is the speaker? Improving Emotion Recognition and Flip Reasoning in Conversations via Speaker Embeddings
von: Patel, Shubham, et al.
Veröffentlicht: (2024)
von: Patel, Shubham, et al.
Veröffentlicht: (2024)
Towards Robust Evaluation of Unlearning in LLMs via Data Transformations
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
CoMuMDR: Code-mixed Multi-modal Multi-domain corpus for Discourse paRsing in conversations
von: Shukla, Divyaksh, et al.
Veröffentlicht: (2025)
von: Shukla, Divyaksh, et al.
Veröffentlicht: (2025)
Geometry of Decision Making in Language Models
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
IL-TUR: Benchmark for Indian Legal Text Understanding and Reasoning
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
POSESTITCH-SLT: Linguistically Inspired Pose-Stitching for End-to-End Sign Language Translation
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
IITK at SemEval-2024 Task 2: Exploring the Capabilities of LLMs for Safe Biomedical Natural Language Inference for Clinical Trials
von: Mandal, Shreyasi, et al.
Veröffentlicht: (2024)
von: Mandal, Shreyasi, et al.
Veröffentlicht: (2024)
LoRMA: Low-Rank Multiplicative Adaptation for LLMs
von: Bihany, Harsh, et al.
Veröffentlicht: (2025)
von: Bihany, Harsh, et al.
Veröffentlicht: (2025)
IITK at SemEval-2024 Task 4: Hierarchical Embeddings for Detection of Persuasion Techniques in Memes
von: Chikoti, Shreenaga, et al.
Veröffentlicht: (2024)
von: Chikoti, Shreenaga, et al.
Veröffentlicht: (2024)
iSign: A Benchmark for Indian Sign Language Processing
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024)
IITK at SemEval-2024 Task 1: Contrastive Learning and Autoencoders for Semantic Textual Relatedness in Multilingual Texts
von: Basak, Udvas, et al.
Veröffentlicht: (2024)
von: Basak, Udvas, et al.
Veröffentlicht: (2024)
BookSQL: A Large Scale Text-to-SQL Dataset for Accounting Domain
von: Kumar, Rahul, et al.
Veröffentlicht: (2024)
von: Kumar, Rahul, et al.
Veröffentlicht: (2024)
mCSQA: Multilingual Commonsense Reasoning Dataset with Unified Creation Strategy by Language Models and Humans
von: Sakai, Yusuke, et al.
Veröffentlicht: (2024)
von: Sakai, Yusuke, et al.
Veröffentlicht: (2024)
Generation and De-Identification of Indian Clinical Discharge Summaries using LLMs
von: Singh, Sanjeet, et al.
Veröffentlicht: (2024)
von: Singh, Sanjeet, et al.
Veröffentlicht: (2024)
IL-PCSR: Legal Corpus for Prior Case and Statute Retrieval
von: Paul, Shounak, et al.
Veröffentlicht: (2025)
von: Paul, Shounak, et al.
Veröffentlicht: (2025)
Towards Operationalizing Right to Data Protection
von: Java, Abhinav, et al.
Veröffentlicht: (2024)
von: Java, Abhinav, et al.
Veröffentlicht: (2024)
HyperDAS: Towards Automating Mechanistic Interpretability with Hypernetworks
von: Sun, Jiuding, et al.
Veröffentlicht: (2025)
von: Sun, Jiuding, et al.
Veröffentlicht: (2025)
Reasoning Circuits in Language Models: A Mechanistic Interpretation of Syllogistic Inference
von: Kim, Geonhee, et al.
Veröffentlicht: (2024)
von: Kim, Geonhee, et al.
Veröffentlicht: (2024)
The Reasoning Trap -- Logical Reasoning as a Mechanistic Pathway to Situational Awareness
von: Sahoo, Subramanyam, et al.
Veröffentlicht: (2026)
von: Sahoo, Subramanyam, et al.
Veröffentlicht: (2026)
EtiCor++: Towards Understanding Etiquettical Bias in LLMs
von: Dwivedi, Ashutosh, et al.
Veröffentlicht: (2025)
von: Dwivedi, Ashutosh, et al.
Veröffentlicht: (2025)
EchoChain: A Full-Duplex Benchmark for State-Update Reasoning Under Interruptions
von: Modi, Smit Nautambhai, et al.
Veröffentlicht: (2026)
von: Modi, Smit Nautambhai, et al.
Veröffentlicht: (2026)
Mechanistic?
von: Saphra, Naomi, et al.
Veröffentlicht: (2024)
von: Saphra, Naomi, et al.
Veröffentlicht: (2024)
Intrinsic Self-Correction in LLMs: Towards Explainable Prompting via Mechanistic Interpretability
von: Lee, Yu-Ting, et al.
Veröffentlicht: (2025)
von: Lee, Yu-Ting, et al.
Veröffentlicht: (2025)
Towards Understanding Safety Alignment: A Mechanistic Perspective from Safety Neurons
von: Chen, Jianhui, et al.
Veröffentlicht: (2024)
von: Chen, Jianhui, et al.
Veröffentlicht: (2024)
seqBench: A Tunable Benchmark to Quantify Sequential Reasoning Limits of LLMs
von: Ramezanali, Mohammad, et al.
Veröffentlicht: (2025)
von: Ramezanali, Mohammad, et al.
Veröffentlicht: (2025)
HLDC: Hindi Legal Documents Corpus
von: Kapoor, Arnav, et al.
Veröffentlicht: (2022)
von: Kapoor, Arnav, et al.
Veröffentlicht: (2022)
Effective Reasoning Chains Reduce Intrinsic Dimensionality
von: Prasad, Archiki, et al.
Veröffentlicht: (2026)
von: Prasad, Archiki, et al.
Veröffentlicht: (2026)
Does Cross-Cultural Alignment Change the Commonsense Morality of Language Models?
von: Jinnai, Yuu
Veröffentlicht: (2024)
von: Jinnai, Yuu
Veröffentlicht: (2024)
Learning to Reason with Insight for Informal Theorem Proving
von: Li, Yunhe, et al.
Veröffentlicht: (2026)
von: Li, Yunhe, et al.
Veröffentlicht: (2026)
X-Reasoner: Towards Generalizable Reasoning Across Modalities and Domains
von: Liu, Qianchu, et al.
Veröffentlicht: (2025)
von: Liu, Qianchu, et al.
Veröffentlicht: (2025)
MolReasoner: Toward Effective and Interpretable Reasoning for Molecular LLMs
von: Zhao, Guojiang, et al.
Veröffentlicht: (2025)
von: Zhao, Guojiang, et al.
Veröffentlicht: (2025)
Teaching Transformers Causal Reasoning through Axiomatic Training
von: Vashishtha, Aniket, et al.
Veröffentlicht: (2024)
von: Vashishtha, Aniket, et al.
Veröffentlicht: (2024)
Towards Effective Code-Integrated Reasoning
von: Bai, Fei, et al.
Veröffentlicht: (2025)
von: Bai, Fei, et al.
Veröffentlicht: (2025)
Circuit Insights: Towards Interpretability Beyond Activations
von: Golimblevskaia, Elena, et al.
Veröffentlicht: (2025)
von: Golimblevskaia, Elena, et al.
Veröffentlicht: (2025)
Mechanistic Fine-tuning for In-context Learning
von: Cho, Hakaze, et al.
Veröffentlicht: (2025)
von: Cho, Hakaze, et al.
Veröffentlicht: (2025)
MIB: A Mechanistic Interpretability Benchmark
von: Mueller, Aaron, et al.
Veröffentlicht: (2025)
von: Mueller, Aaron, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Beyond Components: Singular Vector-Based Interpretability of Transformer Circuits
von: Ahmad, Areeb, et al.
Veröffentlicht: (2025) -
Calibration Across Layers: Understanding Calibration Evolution in LLMs
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025) -
COLD: Causal reasOning in cLosed Daily activities
von: Joshi, Abhinav, et al.
Veröffentlicht: (2024) -
Calibration vs Decision Making: Revisiting the Reliability Paradox in Unlearned Language Models
von: Shukla, Divyaksh, et al.
Veröffentlicht: (2026) -
IITK at SemEval-2024 Task 10: Who is the speaker? Improving Emotion Recognition and Flip Reasoning in Conversations via Speaker Embeddings
von: Patel, Shubham, et al.
Veröffentlicht: (2024)