Language models struggle with compartmentalization
Fuente:
arXiv
Saved in:
| Main Authors: | Howe, Thomas Vincent, Wingate, David |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Features that Make a Difference: Leveraging Gradients for Improved Dictionary Learning
by: Olmo, Jeffrey, et al.
Published: (2024)
by: Olmo, Jeffrey, et al.
Published: (2024)
Out of One, Many: Using Language Models to Simulate Human Samples
by: Argyle, Lisa P., et al.
Published: (2022)
by: Argyle, Lisa P., et al.
Published: (2022)
Learning Dynamics in Continual Pre-Training for Large Language Models
by: Wang, Xingjin, et al.
Published: (2025)
by: Wang, Xingjin, et al.
Published: (2025)
Scaling Law with Learning Rate Annealing
by: Tissue, Howe, et al.
Published: (2024)
by: Tissue, Howe, et al.
Published: (2024)
Domain2Vec: Vectorizing Datasets to Find the Optimal Data Mixture without Training
by: Zhang, Mozhi, et al.
Published: (2025)
by: Zhang, Mozhi, et al.
Published: (2025)
Language models scale reliably with over-training and on downstream tasks
by: Gadre, Samir Yitzhak, et al.
Published: (2024)
by: Gadre, Samir Yitzhak, et al.
Published: (2024)
Dataset Scale and Societal Consistency Mediate Facial Impression Bias in Vision-Language AI
by: Wolfe, Robert, et al.
Published: (2024)
by: Wolfe, Robert, et al.
Published: (2024)
Leveraging Zero-Shot Prompting for Efficient Language Model Distillation
by: Vöge, Lukas, et al.
Published: (2024)
by: Vöge, Lukas, et al.
Published: (2024)
ELIXIR: Efficient and LIghtweight model for eXplaIning Recommendations
by: Kabongo, Ben, et al.
Published: (2025)
by: Kabongo, Ben, et al.
Published: (2025)
Exploring Gender Bias in Large Language Models: An In-depth Dive into the German Language
by: Gnadt, Kristin, et al.
Published: (2025)
by: Gnadt, Kristin, et al.
Published: (2025)
Cross-lingual transfer of multilingual models on low resource African Languages
by: Thangaraj, Harish, et al.
Published: (2024)
by: Thangaraj, Harish, et al.
Published: (2024)
Game of LLMs: Discovering Structural Constructs in Activities using Large Language Models
by: Hiremath, Shruthi K., et al.
Published: (2024)
by: Hiremath, Shruthi K., et al.
Published: (2024)
Scaling Laws for Discriminative Classification in Large Language Models
by: Wyatte, Dean, et al.
Published: (2024)
by: Wyatte, Dean, et al.
Published: (2024)
Neural Proto-Language Reconstruction
by: Cui, Chenxuan, et al.
Published: (2024)
by: Cui, Chenxuan, et al.
Published: (2024)
Language-conditioned world model improves policy generalization by reading environmental descriptions
by: Nguyen, Anh, et al.
Published: (2025)
by: Nguyen, Anh, et al.
Published: (2025)
Continuous Language Model Interpolation for Dynamic and Controllable Text Generation
by: Kangaslahti, Sara, et al.
Published: (2024)
by: Kangaslahti, Sara, et al.
Published: (2024)
Mechanistic Anomaly Detection for "Quirky" Language Models
by: Johnston, David O., et al.
Published: (2025)
by: Johnston, David O., et al.
Published: (2025)
Additive Large Language Models for Semi-Structured Text
by: K, Karthikeyan, et al.
Published: (2025)
by: K, Karthikeyan, et al.
Published: (2025)
Using Source-Side Confidence Estimation for Reliable Translation into Unfamiliar Languages
by: Sible, Kenneth J., et al.
Published: (2025)
by: Sible, Kenneth J., et al.
Published: (2025)
No Need to Talk: Asynchronous Mixture of Language Models
by: Filippova, Anastasiia, et al.
Published: (2024)
by: Filippova, Anastasiia, et al.
Published: (2024)
Erasing Conceptual Knowledge from Language Models
by: Gandikota, Rohit, et al.
Published: (2024)
by: Gandikota, Rohit, et al.
Published: (2024)
GNN: Graph Neural Network and Large Language Model for Data Discovery
by: Hoang, Thomas
Published: (2024)
by: Hoang, Thomas
Published: (2024)
Automated Multi-Language to English Machine Translation Using Generative Pre-Trained Transformers
by: Pelofske, Elijah, et al.
Published: (2024)
by: Pelofske, Elijah, et al.
Published: (2024)
BitFit: Simple Parameter-efficient Fine-tuning for Transformer-based Masked Language-models
by: Ben-Zaken, Elad, et al.
Published: (2021)
by: Ben-Zaken, Elad, et al.
Published: (2021)
The Zero Body Problem: Probing LLM Use of Sensory Language
by: Hicke, Rebecca M. M., et al.
Published: (2025)
by: Hicke, Rebecca M. M., et al.
Published: (2025)
Need a Small Specialized Language Model? Plan Early!
by: Grangier, David, et al.
Published: (2024)
by: Grangier, David, et al.
Published: (2024)
Uncertainty Distillation: Teaching Language Models to Express Semantic Confidence
by: Hager, Sophia, et al.
Published: (2025)
by: Hager, Sophia, et al.
Published: (2025)
Learning to Decode Collaboratively with Multiple Language Models
by: Shen, Shannon Zejiang, et al.
Published: (2024)
by: Shen, Shannon Zejiang, et al.
Published: (2024)
Data Augmentations for Improved (Large) Language Model Generalization
by: Feder, Amir, et al.
Published: (2023)
by: Feder, Amir, et al.
Published: (2023)
Learning and Forgetting Unsafe Examples in Large Language Models
by: Zhao, Jiachen, et al.
Published: (2023)
by: Zhao, Jiachen, et al.
Published: (2023)
Command R7B Arabic: A Small, Enterprise Focused, Multilingual, and Culturally Aware Arabic LLM
by: Alnumay, Yazeed, et al.
Published: (2025)
by: Alnumay, Yazeed, et al.
Published: (2025)
Task-Adaptive Pretrained Language Models via Clustered-Importance Sampling
by: Grangier, David, et al.
Published: (2024)
by: Grangier, David, et al.
Published: (2024)
Composable Interventions for Language Models
by: Kolbeinsson, Arinbjorn, et al.
Published: (2024)
by: Kolbeinsson, Arinbjorn, et al.
Published: (2024)
Function Vectors in Large Language Models
by: Todd, Eric, et al.
Published: (2023)
by: Todd, Eric, et al.
Published: (2023)
Taxonomy, Opportunities, and Challenges of Representation Engineering for Large Language Models
by: Wehner, Jan, et al.
Published: (2025)
by: Wehner, Jan, et al.
Published: (2025)
Benchmarking Distilled Language Models: Performance and Efficiency in Resource-Constrained Settings
by: Wani, Sachin Gopal, et al.
Published: (2026)
by: Wani, Sachin Gopal, et al.
Published: (2026)
Training Bilingual LMs with Data Constraints in the Targeted Language
by: Seto, Skyler, et al.
Published: (2024)
by: Seto, Skyler, et al.
Published: (2024)
Open-ended VQA benchmarking of Vision-Language models by exploiting Classification datasets and their semantic hierarchy
by: Ging, Simon, et al.
Published: (2024)
by: Ging, Simon, et al.
Published: (2024)
Steering Language Models With Activation Engineering
by: Turner, Alexander Matt, et al.
Published: (2023)
by: Turner, Alexander Matt, et al.
Published: (2023)
Training Language Models to Explain Their Own Computations
by: Li, Belinda Z., et al.
Published: (2025)
by: Li, Belinda Z., et al.
Published: (2025)
Similar Items
-
Features that Make a Difference: Leveraging Gradients for Improved Dictionary Learning
by: Olmo, Jeffrey, et al.
Published: (2024) -
Out of One, Many: Using Language Models to Simulate Human Samples
by: Argyle, Lisa P., et al.
Published: (2022) -
Learning Dynamics in Continual Pre-Training for Large Language Models
by: Wang, Xingjin, et al.
Published: (2025) -
Scaling Law with Learning Rate Annealing
by: Tissue, Howe, et al.
Published: (2024) -
Domain2Vec: Vectorizing Datasets to Find the Optimal Data Mixture without Training
by: Zhang, Mozhi, et al.
Published: (2025)