Transformer-based Causal Language Models Perform Clustering
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Xinbo, Varshney, Lav R. |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Meta-Learning Perspective on Transformers for Causal Language Modeling
by: Wu, Xinbo, et al.
Published: (2023)
by: Wu, Xinbo, et al.
Published: (2023)
SwitchCIT: Switching for Continual Instruction Tuning
by: Wu, Xinbo, et al.
Published: (2024)
by: Wu, Xinbo, et al.
Published: (2024)
Hallucination Basins: A Dynamic Framework for Understanding and Controlling LLM Hallucinations
by: Cherukuri, Kalyan, et al.
Published: (2026)
by: Cherukuri, Kalyan, et al.
Published: (2026)
Efficient Model-Agnostic Multi-Group Equivariant Networks
by: Baltaji, Razan, et al.
Published: (2023)
by: Baltaji, Razan, et al.
Published: (2023)
Persona Inconstancy in Multi-Agent LLM Collaboration: Conformity, Confabulation, and Impersonation
by: Baltaji, Razan, et al.
Published: (2024)
by: Baltaji, Razan, et al.
Published: (2024)
Concealment of Intent: A Game-Theoretic Analysis
by: Wu, Xinbo, et al.
Published: (2025)
by: Wu, Xinbo, et al.
Published: (2025)
Skip-It? Theoretical Conditions for Layer Skipping in Vision-Language Models
by: Hartman, Max, et al.
Published: (2025)
by: Hartman, Max, et al.
Published: (2025)
Fed-SB: A Silver Bullet for Extreme Communication Efficiency and Performance in (Private) Federated LoRA Fine-Tuning
by: Singhal, Raghav, et al.
Published: (2025)
by: Singhal, Raghav, et al.
Published: (2025)
A Theoretical Game of Attacks via Compositional Skills
by: Wu, Xinbo, et al.
Published: (2026)
by: Wu, Xinbo, et al.
Published: (2026)
Causal Agent based on Large Language Model
by: Han, Kairong, et al.
Published: (2024)
by: Han, Kairong, et al.
Published: (2024)
DeepInsert: Early Layer Bypass for Efficient and Performant Multimodal Understanding
by: Choraria, Moulik, et al.
Published: (2025)
by: Choraria, Moulik, et al.
Published: (2025)
Indic-TunedLens: Interpreting Multilingual Models in Indian Languages
by: Panchal, Mihir, et al.
Published: (2026)
by: Panchal, Mihir, et al.
Published: (2026)
CausalEval: Towards Better Causal Reasoning in Language Models
by: Yu, Longxuan, et al.
Published: (2024)
by: Yu, Longxuan, et al.
Published: (2024)
Evaluating Large Language Models for Financial Reasoning: A CFA-Based Benchmark Study
by: Yao, Xuan, et al.
Published: (2025)
by: Yao, Xuan, et al.
Published: (2025)
Many LLMs Are More Utilitarian Than One
by: Keshmirian, Anita, et al.
Published: (2025)
by: Keshmirian, Anita, et al.
Published: (2025)
Transformational Creativity in Science: A Graphical Theory
by: Schapiro, Samuel, et al.
Published: (2025)
by: Schapiro, Samuel, et al.
Published: (2025)
Watermarking Discrete Diffusion Language Models
by: Bagchi, Avi, et al.
Published: (2025)
by: Bagchi, Avi, et al.
Published: (2025)
Causality for Large Language Models
by: Wu, Anpeng, et al.
Published: (2024)
by: Wu, Anpeng, et al.
Published: (2024)
Causal Language Control in Multilingual Transformers via Sparse Feature Steering
by: Chou, Cheng-Ting, et al.
Published: (2025)
by: Chou, Cheng-Ting, et al.
Published: (2025)
Large Language Models for Causal Discovery: Current Landscape and Future Directions
by: Wan, Guangya, et al.
Published: (2024)
by: Wan, Guangya, et al.
Published: (2024)
CausalDetox: Causal Head Selection and Intervention for Language Model Detoxification
by: Wang, Yian, et al.
Published: (2026)
by: Wang, Yian, et al.
Published: (2026)
InfoCausalQA:Can Models Perform Non-explicit Causal Reasoning Based on Infographic?
by: Ka, Keummin, et al.
Published: (2025)
by: Ka, Keummin, et al.
Published: (2025)
CausalVLBench: Benchmarking Visual Causal Reasoning in Large Vision-Language Models
by: Komanduri, Aneesh, et al.
Published: (2025)
by: Komanduri, Aneesh, et al.
Published: (2025)
Large Language Models and Causal Inference in Collaboration: A Survey
by: Liu, Xiaoyu, et al.
Published: (2024)
by: Liu, Xiaoyu, et al.
Published: (2024)
Transformer-based Language Models for Reasoning in the Description Logic ALCQ
by: Poulis, Angelos, et al.
Published: (2024)
by: Poulis, Angelos, et al.
Published: (2024)
Probing Causality Manipulation of Large Language Models
by: Zhang, Chenyang, et al.
Published: (2024)
by: Zhang, Chenyang, et al.
Published: (2024)
Deriving Strategic Market Insights with Large Language Models: A Benchmark for Forward Counterfactual Generation
by: Ong, Keane, et al.
Published: (2025)
by: Ong, Keane, et al.
Published: (2025)
Efficient Length-Generalizable Attention via Causal Retrieval for Long-Context Language Modeling
by: Hu, Xiang, et al.
Published: (2024)
by: Hu, Xiang, et al.
Published: (2024)
Natural Language Satisfiability: Exploring the Problem Distribution and Evaluating Transformer-based Language Models
by: Madusanka, Tharindu, et al.
Published: (2025)
by: Madusanka, Tharindu, et al.
Published: (2025)
Progtuning: Progressive Fine-tuning Framework for Transformer-based Language Models
by: Ji, Xiaoshuang, et al.
Published: (2025)
by: Ji, Xiaoshuang, et al.
Published: (2025)
AI Content Self-Detection for Transformer-based Large Language Models
by: Caiado, Antônio Junior Alves, et al.
Published: (2023)
by: Caiado, Antônio Junior Alves, et al.
Published: (2023)
Large Language Models for Constrained-Based Causal Discovery
by: Cohrs, Kai-Hendrik, et al.
Published: (2024)
by: Cohrs, Kai-Hendrik, et al.
Published: (2024)
Causal Inference with Large Language Model: A Survey
by: Ma, Jing
Published: (2024)
by: Ma, Jing
Published: (2024)
Exploration of Masked and Causal Language Modelling for Text Generation
by: Micheletti, Nicolo, et al.
Published: (2024)
by: Micheletti, Nicolo, et al.
Published: (2024)
CAT: Causal Attention Tuning For Injecting Fine-grained Causal Knowledge into Large Language Models
by: Han, Kairong, et al.
Published: (2025)
by: Han, Kairong, et al.
Published: (2025)
Knowledge Graph Structure as Prompt: Improving Small Language Models Capabilities for Knowledge-based Causal Discovery
by: Susanti, Yuni, et al.
Published: (2024)
by: Susanti, Yuni, et al.
Published: (2024)
Multi-LogiEval: Towards Evaluating Multi-Step Logical Reasoning Ability of Large Language Models
by: Patel, Nisarg, et al.
Published: (2024)
by: Patel, Nisarg, et al.
Published: (2024)
PRISM: A Transformer-based Language Model of Structured Clinical Event Data
by: Levine, Lionel, et al.
Published: (2025)
by: Levine, Lionel, et al.
Published: (2025)
LogicBench: Towards Systematic Evaluation of Logical Reasoning Ability of Large Language Models
by: Parmar, Mihir, et al.
Published: (2024)
by: Parmar, Mihir, et al.
Published: (2024)
Causal-Guided Active Learning for Debiasing Large Language Models
by: Du, Li, et al.
Published: (2024)
by: Du, Li, et al.
Published: (2024)
Similar Items
-
A Meta-Learning Perspective on Transformers for Causal Language Modeling
by: Wu, Xinbo, et al.
Published: (2023) -
SwitchCIT: Switching for Continual Instruction Tuning
by: Wu, Xinbo, et al.
Published: (2024) -
Hallucination Basins: A Dynamic Framework for Understanding and Controlling LLM Hallucinations
by: Cherukuri, Kalyan, et al.
Published: (2026) -
Efficient Model-Agnostic Multi-Group Equivariant Networks
by: Baltaji, Razan, et al.
Published: (2023) -
Persona Inconstancy in Multi-Agent LLM Collaboration: Conformity, Confabulation, and Impersonation
by: Baltaji, Razan, et al.
Published: (2024)