Interpreting Arithmetic Mechanism in Large Language Models through Comparative Neuron Analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yu, Zeping, Ananiadou, Sophia |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Neuron-Level Knowledge Attribution in Large Language Models
von: Yu, Zeping, et al.
Veröffentlicht: (2023)
von: Yu, Zeping, et al.
Veröffentlicht: (2023)
Understanding and Mitigating Gender Bias in LLMs via Interpretable Neuron Editing
von: Yu, Zeping, et al.
Veröffentlicht: (2025)
von: Yu, Zeping, et al.
Veröffentlicht: (2025)
Locate-then-Merge: Neuron-Level Parameter Fusion for Mitigating Catastrophic Forgetting in Multimodal LLMs
von: Yu, Zeping, et al.
Veröffentlicht: (2025)
von: Yu, Zeping, et al.
Veröffentlicht: (2025)
Understanding Multimodal LLMs: the Mechanistic Interpretability of Llava in Visual Question Answering
von: Yu, Zeping, et al.
Veröffentlicht: (2024)
von: Yu, Zeping, et al.
Veröffentlicht: (2024)
How do Large Language Models Learn In-Context? Query and Key Matrices of In-Context Heads are Two Towers for Metric Learning
von: Yu, Zeping, et al.
Veröffentlicht: (2024)
von: Yu, Zeping, et al.
Veröffentlicht: (2024)
Back Attention: Understanding and Enhancing Multi-Hop Reasoning in Large Language Models
von: Yu, Zeping, et al.
Veröffentlicht: (2025)
von: Yu, Zeping, et al.
Veröffentlicht: (2025)
Towards Interpretable Mental Health Analysis with Large Language Models
von: Yang, Kailai, et al.
Veröffentlicht: (2023)
von: Yang, Kailai, et al.
Veröffentlicht: (2023)
MentaLLaMA: Interpretable Mental Health Analysis on Social Media with Large Language Models
von: Yang, Kailai, et al.
Veröffentlicht: (2023)
von: Yang, Kailai, et al.
Veröffentlicht: (2023)
The Lay Person's Guide to Biomedicine: Orchestrating Large Language Models
von: Luo, Zheheng, et al.
Veröffentlicht: (2024)
von: Luo, Zheheng, et al.
Veröffentlicht: (2024)
Implicit Graph, Explicit Retrieval: Towards Efficient and Interpretable Long-horizon Memory for Large Language Models
von: Zhang, Xin, et al.
Veröffentlicht: (2026)
von: Zhang, Xin, et al.
Veröffentlicht: (2026)
EmoLLMs: A Series of Emotional Large Language Models and Annotation Tools for Comprehensive Affective Analysis
von: Liu, Zhiwei, et al.
Veröffentlicht: (2024)
von: Liu, Zhiwei, et al.
Veröffentlicht: (2024)
Emotion Detection for Misinformation: A Review
von: Liu, Zhiwei, et al.
Veröffentlicht: (2023)
von: Liu, Zhiwei, et al.
Veröffentlicht: (2023)
Are Large Language Models True Healthcare Jacks-of-All-Trades? Benchmarking Across Health Professions Beyond Physician Exams
von: Luo, Zheheng, et al.
Veröffentlicht: (2024)
von: Luo, Zheheng, et al.
Veröffentlicht: (2024)
Religious Bias Landscape in Language and Text-to-Image Models: Analysis, Detection, and Debiasing Strategies
von: Abrar, Ajwad, et al.
Veröffentlicht: (2025)
von: Abrar, Ajwad, et al.
Veröffentlicht: (2025)
From n-gram to Attention: How Model Architectures Learn and Propagate Bias in Language Modeling
von: Kabir, Mohsinul, et al.
Veröffentlicht: (2025)
von: Kabir, Mohsinul, et al.
Veröffentlicht: (2025)
Interpreting and Improving Large Language Models in Arithmetic Calculation
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
von: Zhang, Wei, et al.
Veröffentlicht: (2024)
ConspEmoLLM: Conspiracy Theory Detection Using an Emotion-Based Large Language Model
von: Liu, Zhiwei, et al.
Veröffentlicht: (2024)
von: Liu, Zhiwei, et al.
Veröffentlicht: (2024)
FMDLlama: Financial Misinformation Detection based on Large Language Models
von: Liu, Zhiwei, et al.
Veröffentlicht: (2024)
von: Liu, Zhiwei, et al.
Veröffentlicht: (2024)
Language Arithmetics: Towards Systematic Language Neuron Identification and Manipulation
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025)
von: Gurgurov, Daniil, et al.
Veröffentlicht: (2025)
Factual consistency evaluation of summarization in the Era of large language models
von: Luo, Zheheng, et al.
Veröffentlicht: (2024)
von: Luo, Zheheng, et al.
Veröffentlicht: (2024)
Disentangled VAD Representations via a Variational Framework for Political Stance Detection
von: Xu, Beiyu, et al.
Veröffentlicht: (2025)
von: Xu, Beiyu, et al.
Veröffentlicht: (2025)
DOCS: Quantifying Weight Similarity for Deeper Insights into Large Language Models
von: Min, Zeping, et al.
Veröffentlicht: (2025)
von: Min, Zeping, et al.
Veröffentlicht: (2025)
Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large Language Models
von: Zhang, Hengyuan, et al.
Veröffentlicht: (2026)
von: Zhang, Hengyuan, et al.
Veröffentlicht: (2026)
Break the Checkbox: Challenging Closed-Style Evaluations of Cultural Alignment in LLMs
von: Kabir, Mohsinul, et al.
Veröffentlicht: (2025)
von: Kabir, Mohsinul, et al.
Veröffentlicht: (2025)
MMNeuron: Discovering Neuron-Level Domain-Specific Interpretation in Multimodal Large Language Model
von: Huo, Jiahao, et al.
Veröffentlicht: (2024)
von: Huo, Jiahao, et al.
Veröffentlicht: (2024)
Flash Interpretability: Decoding Specialised Feature Neurons in Large Language Models with the LM-Head
von: Davies, Harry J
Veröffentlicht: (2025)
von: Davies, Harry J
Veröffentlicht: (2025)
Plutus: Benchmarking Large Language Models in Low-Resource Greek Finance
von: Peng, Xueqing, et al.
Veröffentlicht: (2025)
von: Peng, Xueqing, et al.
Veröffentlicht: (2025)
The Devil is in the Neurons: Interpreting and Mitigating Social Biases in Pre-trained Language Models
von: Liu, Yan, et al.
Veröffentlicht: (2024)
von: Liu, Yan, et al.
Veröffentlicht: (2024)
Neuron-Level Analysis of Cultural Understanding in Large Language Models
von: Yamamoto, Taisei, et al.
Veröffentlicht: (2025)
von: Yamamoto, Taisei, et al.
Veröffentlicht: (2025)
On Representational Dissociation of Language and Arithmetic in Large Language Models
von: Kisako, Riku, et al.
Veröffentlicht: (2025)
von: Kisako, Riku, et al.
Veröffentlicht: (2025)
Exploring the Integration of Large Language Models into Automatic Speech Recognition Systems: An Empirical Study
von: Min, Zeping, et al.
Veröffentlicht: (2023)
von: Min, Zeping, et al.
Veröffentlicht: (2023)
MetaAligner: Towards Generalizable Multi-Objective Alignment of Language Models
von: Yang, Kailai, et al.
Veröffentlicht: (2024)
von: Yang, Kailai, et al.
Veröffentlicht: (2024)
Unveiling Factual Recall Behaviors of Large Language Models through Knowledge Neurons
von: Wang, Yifei, et al.
Veröffentlicht: (2024)
von: Wang, Yifei, et al.
Veröffentlicht: (2024)
ConspEmoLLM-v2: A robust and stable model to detect sentiment-transformed conspiracy theories
von: Liu, Zhiwei, et al.
Veröffentlicht: (2025)
von: Liu, Zhiwei, et al.
Veröffentlicht: (2025)
Rumor Detection by Multi-task Suffix Learning based on Time-series Dual Sentiments
von: Liu, Zhiwei, et al.
Veröffentlicht: (2025)
von: Liu, Zhiwei, et al.
Veröffentlicht: (2025)
No Language is an Island: Unifying Chinese and English in Financial Large Language Models, Instruction Data, and Benchmarks
von: Hu, Gang, et al.
Veröffentlicht: (2024)
von: Hu, Gang, et al.
Veröffentlicht: (2024)
Ebisu: Benchmarking Large Language Models in Japanese Finance
von: Peng, Xueqing, et al.
Veröffentlicht: (2026)
von: Peng, Xueqing, et al.
Veröffentlicht: (2026)
Neuron-Level Differentiation of Memorization and Generalization in Large Language Models
von: Huang, Ko-Wei, et al.
Veröffentlicht: (2024)
von: Huang, Ko-Wei, et al.
Veröffentlicht: (2024)
MMAFFBen: A Multilingual and Multimodal Affective Analysis Benchmark for Evaluating LLMs and VLMs
von: Liu, Zhiwei, et al.
Veröffentlicht: (2025)
von: Liu, Zhiwei, et al.
Veröffentlicht: (2025)
Know-MRI: A Knowledge Mechanisms Revealer&Interpreter for Large Language Models
von: Liu, Jiaxiang, et al.
Veröffentlicht: (2025)
von: Liu, Jiaxiang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Neuron-Level Knowledge Attribution in Large Language Models
von: Yu, Zeping, et al.
Veröffentlicht: (2023) -
Understanding and Mitigating Gender Bias in LLMs via Interpretable Neuron Editing
von: Yu, Zeping, et al.
Veröffentlicht: (2025) -
Locate-then-Merge: Neuron-Level Parameter Fusion for Mitigating Catastrophic Forgetting in Multimodal LLMs
von: Yu, Zeping, et al.
Veröffentlicht: (2025) -
Understanding Multimodal LLMs: the Mechanistic Interpretability of Llava in Visual Question Answering
von: Yu, Zeping, et al.
Veröffentlicht: (2024) -
How do Large Language Models Learn In-Context? Query and Key Matrices of In-Context Heads are Two Towers for Metric Learning
von: Yu, Zeping, et al.
Veröffentlicht: (2024)