Causal Interventions Reveal Shared Structure Across English Filler-Gap Constructions
Fuente:
arXiv
Saved in:
| Main Authors: | Boguraev, Sasha, Potts, Christopher, Mahowald, Kyle |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Causal Drawbridges: Characterizing Gradient Blocking of Syntactic Islands in Transformer LMs
by: Boguraev, Sasha, et al.
Published: (2026)
by: Boguraev, Sasha, et al.
Published: (2026)
Models Can and Should Embrace the Communicative Nature of Human-Generated Math
by: Boguraev, Sasha, et al.
Published: (2024)
by: Boguraev, Sasha, et al.
Published: (2024)
France or Spain or Germany or France: A Neural Account of Non-Redundant Redundant Disjunctions
by: Boguraev, Sasha, et al.
Published: (2026)
by: Boguraev, Sasha, et al.
Published: (2026)
Decrypting Cryptic Crosswords: Semantically Complex Wordplay Puzzles as a Target for NLP
by: Rozner, Josh, et al.
Published: (2021)
by: Rozner, Josh, et al.
Published: (2021)
Mission: Impossible Language Models
by: Kallini, Julie, et al.
Published: (2024)
by: Kallini, Julie, et al.
Published: (2024)
Emergent Introspection in AI is Content-Agnostic
by: Lederman, Harvey, et al.
Published: (2026)
by: Lederman, Harvey, et al.
Published: (2026)
The Counterexample Game: Iterated Conceptual Analysis and Repair in Language Models
by: Drucker, Daniel, et al.
Published: (2026)
by: Drucker, Daniel, et al.
Published: (2026)
Language Models Fail to Introspect About Their Knowledge of Language
by: Song, Siyuan, et al.
Published: (2025)
by: Song, Siyuan, et al.
Published: (2025)
Experimental Contexts Can Facilitate Robust Semantic Property Inference in Language Models, but Inconsistently
by: Misra, Kanishka, et al.
Published: (2024)
by: Misra, Kanishka, et al.
Published: (2024)
Privileged Self-Access Matters for Introspection in AI
by: Song, Siyuan, et al.
Published: (2025)
by: Song, Siyuan, et al.
Published: (2025)
semantic-features: A User-Friendly Tool for Studying Contextual Word Embeddings in Interpretable Semantic Spaces
by: Ranganathan, Jwalanthi, et al.
Published: (2025)
by: Ranganathan, Jwalanthi, et al.
Published: (2025)
Language models align with human judgments on key grammatical constructions
by: Hu, Jennifer, et al.
Published: (2024)
by: Hu, Jennifer, et al.
Published: (2024)
Constructions are Revealed in Word Distributions
by: Rozner, Joshua, et al.
Published: (2025)
by: Rozner, Joshua, et al.
Published: (2025)
What Can String Probability Tell Us About Grammaticality?
by: Hu, Jennifer, et al.
Published: (2025)
by: Hu, Jennifer, et al.
Published: (2025)
Is It JUST Semantics? A Case Study of Discourse Particle Understanding in LLMs
by: Sheffield, William, et al.
Published: (2025)
by: Sheffield, William, et al.
Published: (2025)
CausalGym: Benchmarking causal interpretability methods on linguistic tasks
by: Arora, Aryaman, et al.
Published: (2024)
by: Arora, Aryaman, et al.
Published: (2024)
Counterfactual Simulation Training for Chain-of-Thought Faithfulness
by: Hase, Peter, et al.
Published: (2026)
by: Hase, Peter, et al.
Published: (2026)
Dissociating language and thought in large language models
by: Mahowald, Kyle, et al.
Published: (2023)
by: Mahowald, Kyle, et al.
Published: (2023)
CausalDetox: Causal Head Selection and Intervention for Language Model Detoxification
by: Wang, Yian, et al.
Published: (2026)
by: Wang, Yian, et al.
Published: (2026)
Internal Causal Mechanisms Robustly Predict Language Model Out-of-Distribution Behaviors
by: Huang, Jing, et al.
Published: (2025)
by: Huang, Jing, et al.
Published: (2025)
Causal Interventions on Causal Paths: Mapping GPT-2's Reasoning From Syntax to Semantics
by: Lee, Isabelle, et al.
Published: (2024)
by: Lee, Isabelle, et al.
Published: (2024)
Estimating Causal Effects of Text Interventions Leveraging LLMs
by: Guo, Siyi, et al.
Published: (2024)
by: Guo, Siyi, et al.
Published: (2024)
WikiCausal: Corpus and Evaluation Framework for Causal Knowledge Graph Construction
by: Hassanzadeh, Oktie
Published: (2024)
by: Hassanzadeh, Oktie
Published: (2024)
Phonetic Perturbations Reveal Tokenizer-Rooted Safety Gaps in LLMs
by: Aswal, Darpan, et al.
Published: (2025)
by: Aswal, Darpan, et al.
Published: (2025)
Revealing the Numeracy Gap: An Empirical Investigation of Text Embedding Models
by: Deng, Ningyuan, et al.
Published: (2025)
by: Deng, Ningyuan, et al.
Published: (2025)
Improving Causal Interventions in Amnesic Probing with Mean Projection or LEACE
by: Dobrzeniecka, Alicja, et al.
Published: (2025)
by: Dobrzeniecka, Alicja, et al.
Published: (2025)
Event Causality Identification with Synthetic Control
by: Wang, Haoyu, et al.
Published: (2025)
by: Wang, Haoyu, et al.
Published: (2025)
Causal Intervention-Based Memory Selection for Long-Horizon LLM Agents
by: Srivastava, Saksham Sahai
Published: (2026)
by: Srivastava, Saksham Sahai
Published: (2026)
Why LLMs Fail at Causal Discovery and How Interventional Agents Escape
by: Roy, Amartya, et al.
Published: (2026)
by: Roy, Amartya, et al.
Published: (2026)
Debiasing Reward Models via Causally Motivated Inference-Time Intervention
by: Shinoda, Kazutoshi, et al.
Published: (2026)
by: Shinoda, Kazutoshi, et al.
Published: (2026)
Sentiment Analysis Across Languages: Evaluation Before and After Machine Translation to English
by: Kathunia, Aekansh, et al.
Published: (2024)
by: Kathunia, Aekansh, et al.
Published: (2024)
How Reliable are Causal Probing Interventions?
by: Canby, Marc, et al.
Published: (2024)
by: Canby, Marc, et al.
Published: (2024)
Information Gain-Guided Causal Intervention for Autonomous Debiasing Large Language Models
by: Sun, Zhouhao, et al.
Published: (2025)
by: Sun, Zhouhao, et al.
Published: (2025)
Sparse Feature Coactivation Reveals Causal Semantic Modules in Large Language Models
by: Deng, Ruixuan, et al.
Published: (2025)
by: Deng, Ruixuan, et al.
Published: (2025)
Retrieval Augmented Spelling Correction for E-Commerce Applications
by: Guo, Xuan, et al.
Published: (2024)
by: Guo, Xuan, et al.
Published: (2024)
Bridging the Data Gap: Creating a Hindi Text Summarization Dataset from the English XSUM
by: Katwe, Praveenkumar, et al.
Published: (2026)
by: Katwe, Praveenkumar, et al.
Published: (2026)
Findings of the BlackboxNLP 2025 Shared Task: Localizing Circuits and Causal Variables in Language Models
by: Arad, Dana, et al.
Published: (2025)
by: Arad, Dana, et al.
Published: (2025)
NoisyCausal: A Benchmark for Evaluating Causal Reasoning Under Structured Noise
by: Xu, Zhi, et al.
Published: (2026)
by: Xu, Zhi, et al.
Published: (2026)
Layers at Similar Depths Generate Similar Activations Across LLM Architectures
by: Wolfram, Christopher, et al.
Published: (2025)
by: Wolfram, Christopher, et al.
Published: (2025)
Exploring the Role of Reasoning Structures for Constructing Proofs in Multi-Step Natural Language Reasoning with Large Language Models
by: Zheng, Zi'ou, et al.
Published: (2024)
by: Zheng, Zi'ou, et al.
Published: (2024)
Similar Items
-
Causal Drawbridges: Characterizing Gradient Blocking of Syntactic Islands in Transformer LMs
by: Boguraev, Sasha, et al.
Published: (2026) -
Models Can and Should Embrace the Communicative Nature of Human-Generated Math
by: Boguraev, Sasha, et al.
Published: (2024) -
France or Spain or Germany or France: A Neural Account of Non-Redundant Redundant Disjunctions
by: Boguraev, Sasha, et al.
Published: (2026) -
Decrypting Cryptic Crosswords: Semantically Complex Wordplay Puzzles as a Target for NLP
by: Rozner, Josh, et al.
Published: (2021) -
Mission: Impossible Language Models
by: Kallini, Julie, et al.
Published: (2024)