Isolating Culture Neurons in Multilingual Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Namazifard, Danial, Poech, Lukas Galke |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Confidence and Calibration of Activation Oracles for Reliable Interpretation of Language Model Internals
von: Torrielli, Federico, et al.
Veröffentlicht: (2026)
von: Torrielli, Federico, et al.
Veröffentlicht: (2026)
Comparative Study of Multilingual Idioms and Similes in Large Language Models
von: Khoshtab, Paria, et al.
Veröffentlicht: (2024)
von: Khoshtab, Paria, et al.
Veröffentlicht: (2024)
Chain of Summaries: Summarization Through Iterative Questioning
von: Brach, William, et al.
Veröffentlicht: (2025)
von: Brach, William, et al.
Veröffentlicht: (2025)
SDUs DAISY: A Benchmark for Danish Culture
von: Nielsen, Jacob, et al.
Veröffentlicht: (2026)
von: Nielsen, Jacob, et al.
Veröffentlicht: (2026)
Training Language Models to Use Prolog as a Tool
von: Mellgren, Niklas, et al.
Veröffentlicht: (2025)
von: Mellgren, Niklas, et al.
Veröffentlicht: (2025)
A Computational Approach to Language Contact -- A Case Study of Persian
von: Basirat, Ali, et al.
Veröffentlicht: (2026)
von: Basirat, Ali, et al.
Veröffentlicht: (2026)
DaLA: Danish Linguistic Acceptability Evaluation Guided by Real World Errors
von: Barmina, Gianluca, et al.
Veröffentlicht: (2025)
von: Barmina, Gianluca, et al.
Veröffentlicht: (2025)
Tokenization and Morphology in Multilingual Language Models: A Comparative Analysis of mT5 and ByT5
von: Dang, Thao Anh, et al.
Veröffentlicht: (2024)
von: Dang, Thao Anh, et al.
Veröffentlicht: (2024)
Emergent Languages in Populations of Language Model Agents: From Token Efficiency to Oversight Evasion
von: Beltoft, Stine Lyngsø, et al.
Veröffentlicht: (2026)
von: Beltoft, Stine Lyngsø, et al.
Veröffentlicht: (2026)
Language-Specific Neurons: The Key to Multilingual Capabilities in Large Language Models
von: Tang, Tianyi, et al.
Veröffentlicht: (2024)
von: Tang, Tianyi, et al.
Veröffentlicht: (2024)
ChronoMedKG: A Temporally-Grounded Biomedical Knowledge Graph and Benchmark for Clinical Reasoning
von: Ahmed, Md Shamim, et al.
Veröffentlicht: (2026)
von: Ahmed, Md Shamim, et al.
Veröffentlicht: (2026)
The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment
von: Brach, William, et al.
Veröffentlicht: (2026)
von: Brach, William, et al.
Veröffentlicht: (2026)
Efficient Continual Learning for Small Language Models with a Discrete Key-Value Bottleneck
von: Diera, Andor, et al.
Veröffentlicht: (2024)
von: Diera, Andor, et al.
Veröffentlicht: (2024)
Disentangling Language and Culture for Evaluating Multilingual Large Language Models
von: Ying, Jiahao, et al.
Veröffentlicht: (2025)
von: Ying, Jiahao, et al.
Veröffentlicht: (2025)
Neuron-Level Analysis of Cultural Understanding in Large Language Models
von: Yamamoto, Taisei, et al.
Veröffentlicht: (2025)
von: Yamamoto, Taisei, et al.
Veröffentlicht: (2025)
Learning and communication pressures in neural networks: Lessons from emergent communication
von: Galke, Lukas, et al.
Veröffentlicht: (2024)
von: Galke, Lukas, et al.
Veröffentlicht: (2024)
Isotropy Matters: Soft-ZCA Whitening of Embeddings for Semantic Code Search
von: Diera, Andor, et al.
Veröffentlicht: (2024)
von: Diera, Andor, et al.
Veröffentlicht: (2024)
CRANE: Causal Relevance Analysis of Language-Specific Neurons in Multilingual Large Language Models
von: Le, Yifan, et al.
Veröffentlicht: (2026)
von: Le, Yifan, et al.
Veröffentlicht: (2026)
The Provenance Gap in Clinical AI: Evidence-Traceable Temporal Knowledge Graphs for Rare Disease Reasoning
von: Ahmed, Md Shamim, et al.
Veröffentlicht: (2026)
von: Ahmed, Md Shamim, et al.
Veröffentlicht: (2026)
Benchmarking Large Language Models for Persian: A Preliminary Study Focusing on ChatGPT
von: Abaskohi, Amirhossein, et al.
Veröffentlicht: (2024)
von: Abaskohi, Amirhossein, et al.
Veröffentlicht: (2024)
SommBench: Assessing Sommelier Expertise of Language Models
von: Brach, William, et al.
Veröffentlicht: (2026)
von: Brach, William, et al.
Veröffentlicht: (2026)
Four Shades of Life Sciences: A Dataset for Disinformation Detection in the Life Sciences
von: Seidlmayer, Eva, et al.
Veröffentlicht: (2025)
von: Seidlmayer, Eva, et al.
Veröffentlicht: (2025)
Multilingual Large Language Models and Curse of Multilinguality
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2024)
Multilingual Knowledge Editing with Language-Agnostic Factual Neurons
von: Zhang, Xue, et al.
Veröffentlicht: (2024)
von: Zhang, Xue, et al.
Veröffentlicht: (2024)
What makes a language easy to deep-learn? Deep neural networks and humans similarly benefit from compositional structure
von: Galke, Lukas, et al.
Veröffentlicht: (2023)
von: Galke, Lukas, et al.
Veröffentlicht: (2023)
Pruning Multilingual Large Language Models for Multilingual Inference
von: Kim, Hwichan, et al.
Veröffentlicht: (2024)
von: Kim, Hwichan, et al.
Veröffentlicht: (2024)
When are 1.58 bits enough? A Bottom-up Exploration of BitNet Quantization
von: Nielsen, Jacob, et al.
Veröffentlicht: (2024)
von: Nielsen, Jacob, et al.
Veröffentlicht: (2024)
Language Surgery in Multilingual Large Language Models
von: Lopo, Joanito Agili, et al.
Veröffentlicht: (2025)
von: Lopo, Joanito Agili, et al.
Veröffentlicht: (2025)
Not Everything That Counts Can Be Counted: A Case for Safe Qualitative AI
von: Beltoft, Stine, et al.
Veröffentlicht: (2025)
von: Beltoft, Stine, et al.
Veröffentlicht: (2025)
On the Multilingual Ability of Decoder-based Pre-trained Language Models: Finding and Controlling Language-Specific Neurons
von: Kojima, Takeshi, et al.
Veröffentlicht: (2024)
von: Kojima, Takeshi, et al.
Veröffentlicht: (2024)
Language over Content: Tracing Cultural Understanding in Multilingual Large Language Models
von: Cho, Seungho, et al.
Veröffentlicht: (2025)
von: Cho, Seungho, et al.
Veröffentlicht: (2025)
On Relation-Specific Neurons in Large Language Models
von: Liu, Yihong, et al.
Veröffentlicht: (2025)
von: Liu, Yihong, et al.
Veröffentlicht: (2025)
Dictionary Insertion Prompting for Multilingual Reasoning on Multilingual Large Language Models
von: Lu, Hongyuan, et al.
Veröffentlicht: (2024)
von: Lu, Hongyuan, et al.
Veröffentlicht: (2024)
Quantifying Language Disparities in Multilingual Large Language Models
von: Hu, Songbo, et al.
Veröffentlicht: (2025)
von: Hu, Songbo, et al.
Veröffentlicht: (2025)
Let's Focus on Neuron: Neuron-Level Supervised Fine-tuning for Large Language Model
von: Xu, Haoyun, et al.
Veröffentlicht: (2024)
von: Xu, Haoyun, et al.
Veröffentlicht: (2024)
Neuron-Level Sequential Editing for Large Language Models
von: Jiang, Houcheng, et al.
Veröffentlicht: (2024)
von: Jiang, Houcheng, et al.
Veröffentlicht: (2024)
Multilingual Hallucination Gaps in Large Language Models
von: Chataigner, Cléa, et al.
Veröffentlicht: (2024)
von: Chataigner, Cléa, et al.
Veröffentlicht: (2024)
Multilingual Jailbreak Challenges in Large Language Models
von: Deng, Yue, et al.
Veröffentlicht: (2023)
von: Deng, Yue, et al.
Veröffentlicht: (2023)
XTRUST: On the Multilingual Trustworthiness of Large Language Models
von: Li, Yahan, et al.
Veröffentlicht: (2024)
von: Li, Yahan, et al.
Veröffentlicht: (2024)
AlignX: Advancing Multilingual Large Language Models with Multilingual Representation Alignment
von: Bu, Mengyu, et al.
Veröffentlicht: (2025)
von: Bu, Mengyu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Confidence and Calibration of Activation Oracles for Reliable Interpretation of Language Model Internals
von: Torrielli, Federico, et al.
Veröffentlicht: (2026) -
Comparative Study of Multilingual Idioms and Similes in Large Language Models
von: Khoshtab, Paria, et al.
Veröffentlicht: (2024) -
Chain of Summaries: Summarization Through Iterative Questioning
von: Brach, William, et al.
Veröffentlicht: (2025) -
SDUs DAISY: A Benchmark for Danish Culture
von: Nielsen, Jacob, et al.
Veröffentlicht: (2026) -
Training Language Models to Use Prolog as a Tool
von: Mellgren, Niklas, et al.
Veröffentlicht: (2025)