Sparse Auto-Encoders and Holism about Large Language Models
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Grindrod, Jumbly |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Modelling Language using Large Language Models
von: Grindrod, Jumbly
Veröffentlicht: (2024)
von: Grindrod, Jumbly
Veröffentlicht: (2024)
Word Meanings in Transformer Language Models
von: Grindrod, Jumbly, et al.
Veröffentlicht: (2025)
von: Grindrod, Jumbly, et al.
Veröffentlicht: (2025)
Large language models and linguistic intentionality
von: Grindrod, Jumbly
Veröffentlicht: (2024)
von: Grindrod, Jumbly
Veröffentlicht: (2024)
Transformers, Contextualism, and Polysemy
von: Grindrod, Jumbly
Veröffentlicht: (2024)
von: Grindrod, Jumbly
Veröffentlicht: (2024)
Distributional Semantics, Holism, and the Instability of Meaning
von: Grindrod, Jumbly, et al.
Veröffentlicht: (2024)
von: Grindrod, Jumbly, et al.
Veröffentlicht: (2024)
HalluSAE: Detecting Hallucinations in Large Language Models via Sparse Auto-Encoders
von: Chen, Boshui, et al.
Veröffentlicht: (2026)
von: Chen, Boshui, et al.
Veröffentlicht: (2026)
Layer-Wise Evolution of Representations in Fine-Tuned Transformers: Insights from Sparse AutoEncoders
von: Nadipalli, Suneel
Veröffentlicht: (2025)
von: Nadipalli, Suneel
Veröffentlicht: (2025)
On the Representations of Entities in Auto-regressive Large Language Models
von: Morand, Victor, et al.
Veröffentlicht: (2025)
von: Morand, Victor, et al.
Veröffentlicht: (2025)
Large Language Models Are Overparameterized Text Encoders
von: K, Thennal D, et al.
Veröffentlicht: (2024)
von: K, Thennal D, et al.
Veröffentlicht: (2024)
LLM2Vec: Large Language Models Are Secretly Powerful Text Encoders
von: BehnamGhader, Parishad, et al.
Veröffentlicht: (2024)
von: BehnamGhader, Parishad, et al.
Veröffentlicht: (2024)
AutoRE: Document-Level Relation Extraction with Large Language Models
von: Xue, Lilong, et al.
Veröffentlicht: (2024)
von: Xue, Lilong, et al.
Veröffentlicht: (2024)
Quasi-symbolic Semantic Geometry over Transformer-based Variational AutoEncoder
von: Zhang, Yingji, et al.
Veröffentlicht: (2022)
von: Zhang, Yingji, et al.
Veröffentlicht: (2022)
AutoDAN: Generating Stealthy Jailbreak Prompts on Aligned Large Language Models
von: Liu, Xiaogeng, et al.
Veröffentlicht: (2023)
von: Liu, Xiaogeng, et al.
Veröffentlicht: (2023)
RePrompT: Recurrent Prompt Tuning for Integrating Structured EHR Encoders with Large Language Models
von: Moghaddam, Arya Hadizadeh, et al.
Veröffentlicht: (2026)
von: Moghaddam, Arya Hadizadeh, et al.
Veröffentlicht: (2026)
LabelFusion: Fusing Large Language Models with Transformer Encoders for Robust Financial News Classification
von: Schlee, Michael, et al.
Veröffentlicht: (2025)
von: Schlee, Michael, et al.
Veröffentlicht: (2025)
Large Language Models are Powerful Electronic Health Record Encoders
von: Hegselmann, Stefan, et al.
Veröffentlicht: (2025)
von: Hegselmann, Stefan, et al.
Veröffentlicht: (2025)
Dial-MAE: ConTextual Masked Auto-Encoder for Retrieval-based Dialogue Systems
von: Su, Zhenpeng, et al.
Veröffentlicht: (2023)
von: Su, Zhenpeng, et al.
Veröffentlicht: (2023)
Correlation Dimension of Auto-Regressive Large Language Models
von: Du, Xin, et al.
Veröffentlicht: (2025)
von: Du, Xin, et al.
Veröffentlicht: (2025)
Auto-Search and Refinement: An Automated Framework for Gender Bias Mitigation in Large Language Models
von: Xu, Yue, et al.
Veröffentlicht: (2025)
von: Xu, Yue, et al.
Veröffentlicht: (2025)
Pruning Large Language Models with Semi-Structural Adaptive Sparse Training
von: Huang, Weiyu, et al.
Veröffentlicht: (2024)
von: Huang, Weiyu, et al.
Veröffentlicht: (2024)
LLM+AL: Bridging Large Language Models and Action Languages for Complex Reasoning about Actions
von: Ishay, Adam, et al.
Veröffentlicht: (2025)
von: Ishay, Adam, et al.
Veröffentlicht: (2025)
What is an "Abstract Reasoner"? Revisiting Experiments and Arguments about Large Language Models
von: Yun, Tian, et al.
Veröffentlicht: (2025)
von: Yun, Tian, et al.
Veröffentlicht: (2025)
Language Models as Hierarchy Encoders
von: He, Yuan, et al.
Veröffentlicht: (2024)
von: He, Yuan, et al.
Veröffentlicht: (2024)
AutoMix: Automatically Mixing Language Models
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2023)
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2023)
Language Lives in Sparse Dimensions: Toward Interpretable and Efficient Multilingual Control for Large Language Models
von: Zhong, Chengzhi, et al.
Veröffentlicht: (2025)
von: Zhong, Chengzhi, et al.
Veröffentlicht: (2025)
Enabling Precise Topic Alignment in Large Language Models Via Sparse Autoencoders
von: Joshi, Ananya, et al.
Veröffentlicht: (2025)
von: Joshi, Ananya, et al.
Veröffentlicht: (2025)
Sparse Feature Coactivation Reveals Causal Semantic Modules in Large Language Models
von: Deng, Ruixuan, et al.
Veröffentlicht: (2025)
von: Deng, Ruixuan, et al.
Veröffentlicht: (2025)
SMILE-Next: Teaching Large Language Models to Detect, Classify, and Reason about Laughter
von: Jung-Mok, Lee, et al.
Veröffentlicht: (2026)
von: Jung-Mok, Lee, et al.
Veröffentlicht: (2026)
A Causal Language Modeling Detour Improves Encoder Continued Pretraining
von: Touchent, Rian, et al.
Veröffentlicht: (2026)
von: Touchent, Rian, et al.
Veröffentlicht: (2026)
AutoSurvey: Large Language Models Can Automatically Write Surveys
von: Wang, Yidong, et al.
Veröffentlicht: (2024)
von: Wang, Yidong, et al.
Veröffentlicht: (2024)
AutoFlow: Automated Workflow Generation for Large Language Model Agents
von: Li, Zelong, et al.
Veröffentlicht: (2024)
von: Li, Zelong, et al.
Veröffentlicht: (2024)
Clustering Discourses: Racial Biases in Short Stories about Women Generated by Large Language Models
von: Bonil, Gustavo, et al.
Veröffentlicht: (2025)
von: Bonil, Gustavo, et al.
Veröffentlicht: (2025)
AutoSCORE: Enhancing Automated Scoring with Multi-Agent Large Language Models via Structured Component Recognition
von: Wang, Yun, et al.
Veröffentlicht: (2025)
von: Wang, Yun, et al.
Veröffentlicht: (2025)
AutoPsyC: Automatic Recognition of Psychodynamic Conflicts from Semi-structured Interviews with Large Language Models
von: Hossain, Sayed Muddashir, et al.
Veröffentlicht: (2025)
von: Hossain, Sayed Muddashir, et al.
Veröffentlicht: (2025)
Scaling Sparse Fine-Tuning to Large Language Models
von: Ansell, Alan, et al.
Veröffentlicht: (2024)
von: Ansell, Alan, et al.
Veröffentlicht: (2024)
What are Models Thinking about? Understanding Large Language Model Hallucinations "Psychology" through Model Inner State Analysis
von: Wang, Peiran, et al.
Veröffentlicht: (2025)
von: Wang, Peiran, et al.
Veröffentlicht: (2025)
SparseDoctor: Towards Efficient Chat Doctor with Mixture of Experts Enhanced Large Language Models
von: Zhang, Jianbin, et al.
Veröffentlicht: (2025)
von: Zhang, Jianbin, et al.
Veröffentlicht: (2025)
Group-SAE: Efficient Training of Sparse Autoencoders for Large Language Models via Layer Groups
von: Ghilardi, Davide, et al.
Veröffentlicht: (2024)
von: Ghilardi, Davide, et al.
Veröffentlicht: (2024)
LBM: Hierarchical Large Auto-Bidding Model via Reasoning and Acting
von: Li, Yewen, et al.
Veröffentlicht: (2026)
von: Li, Yewen, et al.
Veröffentlicht: (2026)
NAG: A Unified Native Architecture for Encoder-free Text-Graph Modeling in Language Models
von: Gong, Haisong, et al.
Veröffentlicht: (2026)
von: Gong, Haisong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Modelling Language using Large Language Models
von: Grindrod, Jumbly
Veröffentlicht: (2024) -
Word Meanings in Transformer Language Models
von: Grindrod, Jumbly, et al.
Veröffentlicht: (2025) -
Large language models and linguistic intentionality
von: Grindrod, Jumbly
Veröffentlicht: (2024) -
Transformers, Contextualism, and Polysemy
von: Grindrod, Jumbly
Veröffentlicht: (2024) -
Distributional Semantics, Holism, and the Instability of Meaning
von: Grindrod, Jumbly, et al.
Veröffentlicht: (2024)