Latent Causal Probing: A Formal Perspective on Probing with Causal Models of Data
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Jin, Charles, Rinard, Martin |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Emergent Representations of Program Semantics in Language Models Trained on Programs
par: Jin, Charles, et autres
Publié: (2023)
par: Jin, Charles, et autres
Publié: (2023)
Probing Causality Manipulation of Large Language Models
par: Zhang, Chenyang, et autres
Publié: (2024)
par: Zhang, Chenyang, et autres
Publié: (2024)
Nuance Matters: Probing Epistemic Consistency in Causal Reasoning
par: Cui, Shaobo, et autres
Publié: (2024)
par: Cui, Shaobo, et autres
Publié: (2024)
How Reliable are Causal Probing Interventions?
par: Canby, Marc, et autres
Publié: (2024)
par: Canby, Marc, et autres
Publié: (2024)
Improving Causal Interventions in Amnesic Probing with Mean Projection or LEACE
par: Dobrzeniecka, Alicja, et autres
Publié: (2025)
par: Dobrzeniecka, Alicja, et autres
Publié: (2025)
Polarity-Aware Probing for Quantifying Latent Alignment in Language Models
par: Sadiekh, Sabrina, et autres
Publié: (2025)
par: Sadiekh, Sabrina, et autres
Publié: (2025)
Discovering Hierarchical Latent Capabilities of Language Models via Causal Representation Learning
par: Jin, Jikai, et autres
Publié: (2025)
par: Jin, Jikai, et autres
Publié: (2025)
Do BERT Embeddings Encode Narrative Dimensions? A Token-Level Probing Analysis of Time, Space, Causality, and Character in Fiction
par: Bei, Beicheng, et autres
Publié: (2026)
par: Bei, Beicheng, et autres
Publié: (2026)
Probing Language Models for Pre-training Data Detection
par: Liu, Zhenhua, et autres
Publié: (2024)
par: Liu, Zhenhua, et autres
Publié: (2024)
CausalEval: Towards Better Causal Reasoning in Language Models
par: Yu, Longxuan, et autres
Publié: (2024)
par: Yu, Longxuan, et autres
Publié: (2024)
Look Within, Why LLMs Hallucinate: A Causal Perspective
par: Li, He, et autres
Publié: (2024)
par: Li, He, et autres
Publié: (2024)
Dynamics Within Latent Chain-of-Thought: An Empirical Study of Causal Structure
par: Li, Zirui, et autres
Publié: (2026)
par: Li, Zirui, et autres
Publié: (2026)
A Meta-Learning Perspective on Transformers for Causal Language Modeling
par: Wu, Xinbo, et autres
Publié: (2023)
par: Wu, Xinbo, et autres
Publié: (2023)
CausalDetox: Causal Head Selection and Intervention for Language Model Detoxification
par: Wang, Yian, et autres
Publié: (2026)
par: Wang, Yian, et autres
Publié: (2026)
H-Probes: Extracting Hierarchical Structures From Latent Representations of Language Models
par: Dawes, Cutter, et autres
Publié: (2026)
par: Dawes, Cutter, et autres
Publié: (2026)
Probing for Arithmetic Errors in Language Models
par: Sun, Yucheng, et autres
Publié: (2025)
par: Sun, Yucheng, et autres
Publié: (2025)
InfoCausalQA:Can Models Perform Non-explicit Causal Reasoning Based on Infographic?
par: Ka, Keummin, et autres
Publié: (2025)
par: Ka, Keummin, et autres
Publié: (2025)
CAT: Causal Attention Tuning For Injecting Fine-grained Causal Knowledge into Large Language Models
par: Han, Kairong, et autres
Publié: (2025)
par: Han, Kairong, et autres
Publié: (2025)
Probing the Robustness of Large Language Models Safety to Latent Perturbations
par: Gu, Tianle, et autres
Publié: (2025)
par: Gu, Tianle, et autres
Publié: (2025)
CausalARC: Abstract Reasoning with Causal World Models
par: Maasch, Jacqueline, et autres
Publié: (2025)
par: Maasch, Jacqueline, et autres
Publié: (2025)
NoisyCausal: A Benchmark for Evaluating Causal Reasoning Under Structured Noise
par: Xu, Zhi, et autres
Publié: (2026)
par: Xu, Zhi, et autres
Publié: (2026)
Causal Order: The Key to Leveraging Imperfect Experts in Causal Inference
par: Vashishtha, Aniket, et autres
Publié: (2023)
par: Vashishtha, Aniket, et autres
Publié: (2023)
Causality for Natural Language Processing
par: Jin, Zhijing
Publié: (2025)
par: Jin, Zhijing
Publié: (2025)
Causal Inference with Large Language Model: A Survey
par: Ma, Jing
Publié: (2024)
par: Ma, Jing
Publié: (2024)
Probing and Steering Evaluation Awareness of Language Models
par: Nguyen, Jord, et autres
Publié: (2025)
par: Nguyen, Jord, et autres
Publié: (2025)
Probing Neural Topology of Large Language Models
par: Zheng, Yu, et autres
Publié: (2025)
par: Zheng, Yu, et autres
Publié: (2025)
Probing Persona-Dependent Preferences in Language Models
par: Gilg, Oscar, et autres
Publié: (2026)
par: Gilg, Oscar, et autres
Publié: (2026)
Generative Framework for Personalized Persuasion: Inferring Causal, Counterfactual, and Latent Knowledge
par: Zeng, Donghuo, et autres
Publié: (2025)
par: Zeng, Donghuo, et autres
Publié: (2025)
WikiCausal: Corpus and Evaluation Framework for Causal Knowledge Graph Construction
par: Hassanzadeh, Oktie
Publié: (2024)
par: Hassanzadeh, Oktie
Publié: (2024)
CausalAbstain: Enhancing Multilingual LLMs with Causal Reasoning for Trustworthy Abstention
par: Sun, Yuxi, et autres
Publié: (2025)
par: Sun, Yuxi, et autres
Publié: (2025)
Are the Values of LLMs Structurally Aligned with Humans? A Causal Perspective
par: Kang, Yipeng, et autres
Publié: (2024)
par: Kang, Yipeng, et autres
Publié: (2024)
Explaining Black-box Language Models with Knowledge Probing Systems: A Post-hoc Explanation Perspective
par: Zhao, Yunxiao, et autres
Publié: (2025)
par: Zhao, Yunxiao, et autres
Publié: (2025)
Causal Agent based on Large Language Model
par: Han, Kairong, et autres
Publié: (2024)
par: Han, Kairong, et autres
Publié: (2024)
A Causal Graph Approach to Oppositional Narrative Analysis
par: Revilla, Diego, et autres
Publié: (2026)
par: Revilla, Diego, et autres
Publié: (2026)
ReProbe: Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models
par: Ni, Jingwei, et autres
Publié: (2025)
par: Ni, Jingwei, et autres
Publié: (2025)
Probing Experts' Perspectives on AI-Assisted Public Speaking Training
par: Fourati, Nesrine, et autres
Publié: (2025)
par: Fourati, Nesrine, et autres
Publié: (2025)
Large Language Models and Causal Inference in Collaboration: A Survey
par: Liu, Xiaoyu, et autres
Publié: (2024)
par: Liu, Xiaoyu, et autres
Publié: (2024)
Probing the Robustness of Theory of Mind in Large Language Models
par: Nickel, Christian, et autres
Publié: (2024)
par: Nickel, Christian, et autres
Publié: (2024)
SAFER: Probing Safety in Reward Models with Sparse Autoencoder
par: Shi, Wei, et autres
Publié: (2025)
par: Shi, Wei, et autres
Publié: (2025)
Probing the Limits of Stylistic Alignment in Vision-Language Models
par: Farajidizaji, Asma, et autres
Publié: (2025)
par: Farajidizaji, Asma, et autres
Publié: (2025)
Documents similaires
-
Emergent Representations of Program Semantics in Language Models Trained on Programs
par: Jin, Charles, et autres
Publié: (2023) -
Probing Causality Manipulation of Large Language Models
par: Zhang, Chenyang, et autres
Publié: (2024) -
Nuance Matters: Probing Epistemic Consistency in Causal Reasoning
par: Cui, Shaobo, et autres
Publié: (2024) -
How Reliable are Causal Probing Interventions?
par: Canby, Marc, et autres
Publié: (2024) -
Improving Causal Interventions in Amnesic Probing with Mean Projection or LEACE
par: Dobrzeniecka, Alicja, et autres
Publié: (2025)