Do LLMs Have the Generalization Ability in Conducting Causal Inference?
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Chen, Zhao, Dongming, Wang, Bo, He, Ruifang, Hou, Yuexian |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Explicit vs. Implicit: Investigating Social Bias in Large Language Models through Self-Reflection
por: Zhao, Yachao, et al.
Publicado: (2025)
por: Zhao, Yachao, et al.
Publicado: (2025)
MORPHEUS: Modeling Role from Personalized Dialogue History by Exploring and Utilizing Latent Space
por: Tang, Yihong, et al.
Publicado: (2024)
por: Tang, Yihong, et al.
Publicado: (2024)
RoleBreak: Character Hallucination as a Jailbreak Attack in Role-Playing Systems
por: Tang, Yihong, et al.
Publicado: (2024)
por: Tang, Yihong, et al.
Publicado: (2024)
Think Twice: A Human-like Two-stage Conversational Agent for Emotional Response Generation
por: Qian, Yushan, et al.
Publicado: (2023)
por: Qian, Yushan, et al.
Publicado: (2023)
LLMs Are Prone to Fallacies in Causal Inference
por: Joshi, Nitish, et al.
Publicado: (2024)
por: Joshi, Nitish, et al.
Publicado: (2024)
TactfulToM: Do LLMs Have the Theory of Mind Ability to Understand White Lies?
por: Liu, Yiwei, et al.
Publicado: (2025)
por: Liu, Yiwei, et al.
Publicado: (2025)
Modeling Implicit Conflict Monitoring Mechanisms against Stereotypes in LLMs
por: Zhang, Jingshen, et al.
Publicado: (2026)
por: Zhang, Jingshen, et al.
Publicado: (2026)
Understanding the Self-Reflection Mechanisms of LLMs through Biased Attitude Associations
por: Zhang, Jingshen, et al.
Publicado: (2026)
por: Zhang, Jingshen, et al.
Publicado: (2026)
Memorization $\neq$ Understanding: Do Large Language Models Have the Ability of Scenario Cognition?
por: Ma, Boxiang, et al.
Publicado: (2025)
por: Ma, Boxiang, et al.
Publicado: (2025)
Syntax-Aware Complex-Valued Neural Machine Translation
por: Liu, Yang, et al.
Publicado: (2023)
por: Liu, Yang, et al.
Publicado: (2023)
TUMS: Enhancing Tool-use Abilities of LLMs with Multi-structure Handlers
por: He, Aiyao, et al.
Publicado: (2025)
por: He, Aiyao, et al.
Publicado: (2025)
Eliciting Causal Abilities in Large Language Models for Reasoning Tasks
por: Wang, Yajing, et al.
Publicado: (2024)
por: Wang, Yajing, et al.
Publicado: (2024)
Beyond Surface Structure: A Causal Assessment of LLMs' Comprehension Ability
por: Han, Yujin, et al.
Publicado: (2024)
por: Han, Yujin, et al.
Publicado: (2024)
Do Large Language Models Have Compositional Ability? An Investigation into Limitations and Scalability
por: Xu, Zhuoyan, et al.
Publicado: (2024)
por: Xu, Zhuoyan, et al.
Publicado: (2024)
Enhancing Retrieval-Augmented LMs with a Two-stage Consistency Learning Compressor
por: Xu, Chuankai, et al.
Publicado: (2024)
por: Xu, Chuankai, et al.
Publicado: (2024)
Have LLMs Reopened the Pandora's Box of AI-Generated Fake News?
por: Wang, Xinyu, et al.
Publicado: (2024)
por: Wang, Xinyu, et al.
Publicado: (2024)
Quantifying the Reasoning Abilities of LLMs on Real-world Clinical Cases
por: Qiu, Pengcheng, et al.
Publicado: (2025)
por: Qiu, Pengcheng, et al.
Publicado: (2025)
Effective Distillation of Table-based Reasoning Ability from LLMs
por: Yang, Bohao, et al.
Publicado: (2023)
por: Yang, Bohao, et al.
Publicado: (2023)
Do LLMs Know When to NOT Answer? Investigating Abstention Abilities of Large Language Models
por: Madhusudhan, Nishanth, et al.
Publicado: (2024)
por: Madhusudhan, Nishanth, et al.
Publicado: (2024)
Do LLMs and VLMs Share Neurons for Inference? Evidence and Mechanisms of Cross-Modal Transfer
por: Cui, Chenhang, et al.
Publicado: (2026)
por: Cui, Chenhang, et al.
Publicado: (2026)
Distill Visual Chart Reasoning Ability from LLMs to MLLMs
por: He, Wei, et al.
Publicado: (2024)
por: He, Wei, et al.
Publicado: (2024)
Do LLMs Have Distinct and Consistent Personality? TRAIT: Personality Testset designed for LLMs with Psychometrics
por: Lee, Seungbeen, et al.
Publicado: (2024)
por: Lee, Seungbeen, et al.
Publicado: (2024)
Structured Thinking Matters: Improving LLMs Generalization in Causal Inference Tasks
por: Sun, Wentao, et al.
Publicado: (2025)
por: Sun, Wentao, et al.
Publicado: (2025)
A Survey on Enhancing Causal Reasoning Ability of Large Language Models
por: Li, Xin, et al.
Publicado: (2025)
por: Li, Xin, et al.
Publicado: (2025)
Route Before Retrieve: Activating Latent Routing Abilities of LLMs for RAG vs. Long-Context Selection
por: Chen, Yiwen, et al.
Publicado: (2026)
por: Chen, Yiwen, et al.
Publicado: (2026)
ZPD-SCA: Unveiling the Blind Spots of LLMs in Assessing Students' Cognitive Abilities
por: Dong, Wenhan, et al.
Publicado: (2025)
por: Dong, Wenhan, et al.
Publicado: (2025)
Understanding the Ability of LLMs to Handle Character-Level Perturbation
por: Zhuo, Anyuan, et al.
Publicado: (2025)
por: Zhuo, Anyuan, et al.
Publicado: (2025)
Evaluating the Generalization Ability of Quantized LLMs: Benchmark, Analysis, and Toolbox
por: Liu, Yijun, et al.
Publicado: (2024)
por: Liu, Yijun, et al.
Publicado: (2024)
Infusing Hierarchical Guidance into Prompt Tuning: A Parameter-Efficient Framework for Multi-level Implicit Discourse Relation Recognition
por: Zhao, Haodong, et al.
Publicado: (2024)
por: Zhao, Haodong, et al.
Publicado: (2024)
Predictable Emergent Abilities of LLMs: Proxy Tasks Are All You Need
por: Zhang, Bo-Wen, et al.
Publicado: (2024)
por: Zhang, Bo-Wen, et al.
Publicado: (2024)
Doing Good or Doing Right? Exploring the Weakness of Commonsense Causal Reasoning Models
por: Han, Mingyue, et al.
Publicado: (2021)
por: Han, Mingyue, et al.
Publicado: (2021)
Extracting and Combining Abilities For Building Multi-lingual Ability-enhanced Large Language Models
por: Chen, Zhipeng, et al.
Publicado: (2024)
por: Chen, Zhipeng, et al.
Publicado: (2024)
Measuring Bargaining Abilities of LLMs: A Benchmark and A Buyer-Enhancement Method
por: Xia, Tian, et al.
Publicado: (2024)
por: Xia, Tian, et al.
Publicado: (2024)
LongDPO: Unlock Better Long-form Generation Abilities for LLMs via Critique-augmented Stepwise Information
por: Ping, Bowen, et al.
Publicado: (2025)
por: Ping, Bowen, et al.
Publicado: (2025)
Do Large Language Models Have an English Accent? Evaluating and Improving the Naturalness of Multilingual LLMs
por: Guo, Yanzhu, et al.
Publicado: (2024)
por: Guo, Yanzhu, et al.
Publicado: (2024)
Do LLMs Think Fast and Slow? A Causal Study on Sentiment Analysis
por: Lyu, Zhiheng, et al.
Publicado: (2024)
por: Lyu, Zhiheng, et al.
Publicado: (2024)
Enhancing Reasoning Abilities of Small LLMs with Cognitive Alignment
por: Cai, Wenrui, et al.
Publicado: (2025)
por: Cai, Wenrui, et al.
Publicado: (2025)
Rethinking the Understanding Ability across LLMs through Mutual Information
por: Wang, Shaojie, et al.
Publicado: (2025)
por: Wang, Shaojie, et al.
Publicado: (2025)
ROUGE-K: Do Your Summaries Have Keywords?
por: Takeshita, Sotaro, et al.
Publicado: (2024)
por: Takeshita, Sotaro, et al.
Publicado: (2024)
NACL: A General and Effective KV Cache Eviction Framework for LLMs at Inference Time
por: Chen, Yilong, et al.
Publicado: (2024)
por: Chen, Yilong, et al.
Publicado: (2024)
Ejemplares similares
-
Explicit vs. Implicit: Investigating Social Bias in Large Language Models through Self-Reflection
por: Zhao, Yachao, et al.
Publicado: (2025) -
MORPHEUS: Modeling Role from Personalized Dialogue History by Exploring and Utilizing Latent Space
por: Tang, Yihong, et al.
Publicado: (2024) -
RoleBreak: Character Hallucination as a Jailbreak Attack in Role-Playing Systems
por: Tang, Yihong, et al.
Publicado: (2024) -
Think Twice: A Human-like Two-stage Conversational Agent for Emotional Response Generation
por: Qian, Yushan, et al.
Publicado: (2023) -
LLMs Are Prone to Fallacies in Causal Inference
por: Joshi, Nitish, et al.
Publicado: (2024)