Flying Pigs, FaR and Beyond: Evaluating LLM Reasoning in Counterfactual Worlds
Fuente:
arXiv
Saved in:
| Main Authors: | Joishy, Anish R, Balappanawar, Ishwar B, Bonagiri, Vamshi Krishna, Gaur, Manas, Thirunarayan, Krishnaprasad, Kumaraguru, Ponnurangam |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Measuring Moral Inconsistencies in Large Language Models
by: Bonagiri, Vamshi Krishna, et al.
Published: (2024)
by: Bonagiri, Vamshi Krishna, et al.
Published: (2024)
SaGE: Evaluating Moral Consistency in Large Language Models
by: Bonagiri, Vamshi Krishna, et al.
Published: (2024)
by: Bonagiri, Vamshi Krishna, et al.
Published: (2024)
LLM Vocabulary Compression for Low-Compute Environments
by: Vennam, Sreeram, et al.
Published: (2024)
by: Vennam, Sreeram, et al.
Published: (2024)
Intrinsic Guardrails: How Semantic Geometry of Personality Interacts with Emergent Misalignment in LLMs
by: Aneja, Krishak, et al.
Published: (2026)
by: Aneja, Krishak, et al.
Published: (2026)
COBIAS: Assessing the Contextual Reliability of Bias Benchmarks for Language Models
by: Govil, Priyanshul, et al.
Published: (2024)
by: Govil, Priyanshul, et al.
Published: (2024)
K-PERM: Personalized Response Generation Using Dynamic Knowledge Retrieval and Persona-Adaptive Queries
by: Raj, Kanak, et al.
Published: (2023)
by: Raj, Kanak, et al.
Published: (2023)
Check Yourself Before You Wreck Yourself: Selectively Quitting Improves LLM Agent Safety
by: Bonagiri, Vamshi Krishna, et al.
Published: (2025)
by: Bonagiri, Vamshi Krishna, et al.
Published: (2025)
Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability
by: Aggarwal, Yash, et al.
Published: (2026)
by: Aggarwal, Yash, et al.
Published: (2026)
Towards Infusing Auxiliary Knowledge for Distracted Driver Detection
by: Balappanawar, Ishwar B, et al.
Published: (2024)
by: Balappanawar, Ishwar B, et al.
Published: (2024)
From Human Judgements to Predictive Models: Unravelling Acceptability in Code-Mixed Sentences
by: Kodali, Prashant, et al.
Published: (2024)
by: Kodali, Prashant, et al.
Published: (2024)
Fact-and-Reflection (FaR) Improves Confidence Calibration of Large Language Models
by: Zhao, Xinran, et al.
Published: (2024)
by: Zhao, Xinran, et al.
Published: (2024)
A Practical Exercise in Adapting SIFT Using FHE Primitives
by: Balappanawar, Ishwar B, et al.
Published: (2024)
by: Balappanawar, Ishwar B, et al.
Published: (2024)
LABELING COPILOT: A Deep Research Agent for Automated Data Curation in Computer Vision
by: Ganguly, Debargha, et al.
Published: (2025)
by: Ganguly, Debargha, et al.
Published: (2025)
I Can't Believe It's Corrupt: Evaluating Corruption in Multi-Agent Governance Systems
by: P, Vedanta S, et al.
Published: (2026)
by: P, Vedanta S, et al.
Published: (2026)
FaR: Enhancing Multi-Concept Text-to-Image Diffusion via Concept Fusion and Localized Refinement
by: Tran, Gia-Nghia, et al.
Published: (2025)
by: Tran, Gia-Nghia, et al.
Published: (2025)
Do LLMs Adhere to Label Definitions? Examining Their Receptivity to External Label Definitions
by: Mohammadi, Seyedali, et al.
Published: (2025)
by: Mohammadi, Seyedali, et al.
Published: (2025)
Beyond Memorization: Testing LLM Reasoning on Unseen Theory of Computation Tasks
by: Shelat, Shlok, et al.
Published: (2026)
by: Shelat, Shlok, et al.
Published: (2026)
Can Language Models Falsify? Evaluating Algorithmic Reasoning with Counterexample Creation
by: Sinha, Shiven, et al.
Published: (2025)
by: Sinha, Shiven, et al.
Published: (2025)
Personal Narratives Empower Politically Disinclined Individuals to Engage in Political Discussions
by: Chebrolu, Tejasvi, et al.
Published: (2025)
by: Chebrolu, Tejasvi, et al.
Published: (2025)
Structured Definitions and Segmentations for Legal Reasoning in LLMs: A Study on Indian Legal Data
by: Khatri, Mann, et al.
Published: (2025)
by: Khatri, Mann, et al.
Published: (2025)
PrivacyBench: A Conversational Benchmark for Evaluating Privacy in Personalized AI
by: Mukhopadhyay, Srija, et al.
Published: (2025)
by: Mukhopadhyay, Srija, et al.
Published: (2025)
TAMAS: Benchmarking Adversarial Risks in Multi-Agent LLM Systems
by: Kavathekar, Ishan, et al.
Published: (2025)
by: Kavathekar, Ishan, et al.
Published: (2025)
X-posing Free Speech: Examining the Impact of Moderation Relaxation on Online Social Networks
by: Arun, Arvindh, et al.
Published: (2024)
by: Arun, Arvindh, et al.
Published: (2024)
Analyzing Patterns and Influence of Advertising in Print Newspapers
by: Vardhan, N Harsha, et al.
Published: (2025)
by: Vardhan, N Harsha, et al.
Published: (2025)
Long-context Non-factoid Question Answering in Indic Languages
by: Mishra, Ritwik, et al.
Published: (2025)
by: Mishra, Ritwik, et al.
Published: (2025)
Who's the Evil Twin? Differential Auditing for Undesired Behavior
by: Balappanawar, Ishwar, et al.
Published: (2025)
by: Balappanawar, Ishwar, et al.
Published: (2025)
Sample Complexity of Causal Identification with Temporal Heterogeneity
by: Rathod, Ameya, et al.
Published: (2026)
by: Rathod, Ameya, et al.
Published: (2026)
Causal Reasoning Favors Encoders: On The Limits of Decoder-Only Models
by: Roy, Amartya, et al.
Published: (2025)
by: Roy, Amartya, et al.
Published: (2025)
CausalFlow: Causal Attribution and Counterfactual Repair for LLM Agent Failures
by: Bonagiri, Akash, et al.
Published: (2026)
by: Bonagiri, Akash, et al.
Published: (2026)
Small Models, Big Tasks: An Exploratory Empirical Study on Small Language Models for Function Calling
by: Kavathekar, Ishan, et al.
Published: (2025)
by: Kavathekar, Ishan, et al.
Published: (2025)
Ketto and the Science of Giving: A Data-Driven Investigation of Crowdfunding for India
by: Chandra, Karuna, et al.
Published: (2025)
by: Chandra, Karuna, et al.
Published: (2025)
MetaGMT: Improving Actionable Interpretability of Graph Multilinear Networks via Meta-Learning Filtration
by: Bhattacharya, Rishabh, et al.
Published: (2025)
by: Bhattacharya, Rishabh, et al.
Published: (2025)
Higher Order Structures For Graph Explanations
by: Sinha, Akshit, et al.
Published: (2024)
by: Sinha, Akshit, et al.
Published: (2024)
Rethinking Thinking Tokens: Understanding Why They Underperform in Practice
by: Vennam, Sreeram, et al.
Published: (2024)
by: Vennam, Sreeram, et al.
Published: (2024)
A thermodynamic approach to nonlinear ultrasonics for material state awareness and prognosis
by: Chillara, Vamshi Krishna
Published: (2016)
by: Chillara, Vamshi Krishna
Published: (2016)
Framing the Fray: Evaluating Conflict Frames in Indian Election News Coverage
by: Chebrolu, Tejasvi, et al.
Published: (2023)
by: Chebrolu, Tejasvi, et al.
Published: (2023)
Mind the Gap: Pitfalls of LLM Alignment with Asian Public Opinion
by: Shankar, Hari, et al.
Published: (2026)
by: Shankar, Hari, et al.
Published: (2026)
Generation-Time vs. Post-hoc Citation: A Holistic Evaluation of LLM Attribution
by: Saxena, Yash, et al.
Published: (2025)
by: Saxena, Yash, et al.
Published: (2025)
SPIRIT: Short-term Prediction of solar IRradIance for zero-shot Transfer learning using Foundation Models
by: Mishra, Aditya, et al.
Published: (2025)
by: Mishra, Aditya, et al.
Published: (2025)
Multilingual Coreference Resolution in Low-resource South Asian Languages
by: Mishra, Ritwik, et al.
Published: (2024)
by: Mishra, Ritwik, et al.
Published: (2024)
Similar Items
-
Measuring Moral Inconsistencies in Large Language Models
by: Bonagiri, Vamshi Krishna, et al.
Published: (2024) -
SaGE: Evaluating Moral Consistency in Large Language Models
by: Bonagiri, Vamshi Krishna, et al.
Published: (2024) -
LLM Vocabulary Compression for Low-Compute Environments
by: Vennam, Sreeram, et al.
Published: (2024) -
Intrinsic Guardrails: How Semantic Geometry of Personality Interacts with Emergent Misalignment in LLMs
by: Aneja, Krishak, et al.
Published: (2026) -
COBIAS: Assessing the Contextual Reliability of Bias Benchmarks for Language Models
by: Govil, Priyanshul, et al.
Published: (2024)