Know-MRI: A Knowledge Mechanisms Revealer&Interpreter for Large Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Jiaxiang, Xing, Boxuan, Yuan, Chenhao, Zhang, Chenxiang, Wu, Di, Huang, Xiusheng, Yu, Haida, Lang, Chuhan, Cao, Pengfei, Zhao, Jun, Liu, Kang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Capability Localization: Capabilities Can be Localized rather than Individual Knowledge
di: Huang, Xiusheng, et al.
Pubblicazione: (2025)
di: Huang, Xiusheng, et al.
Pubblicazione: (2025)
Reasons and Solutions for the Decline in Model Performance after Editing
di: Huang, Xiusheng, et al.
Pubblicazione: (2024)
di: Huang, Xiusheng, et al.
Pubblicazione: (2024)
Commonsense Knowledge Editing Based on Free-Text in LLMs
di: Huang, Xiusheng, et al.
Pubblicazione: (2024)
di: Huang, Xiusheng, et al.
Pubblicazione: (2024)
RWKU: Benchmarking Real-World Knowledge Unlearning for Large Language Models
di: Jin, Zhuoran, et al.
Pubblicazione: (2024)
di: Jin, Zhuoran, et al.
Pubblicazione: (2024)
Cutting Off the Head Ends the Conflict: A Mechanism for Interpreting and Mitigating Knowledge Conflicts in Language Models
di: Jin, Zhuoran, et al.
Pubblicazione: (2024)
di: Jin, Zhuoran, et al.
Pubblicazione: (2024)
Knowledge in Superposition: Unveiling the Failures of Lifelong Knowledge Editing for Large Language Models
di: Hu, Chenhui, et al.
Pubblicazione: (2024)
di: Hu, Chenhui, et al.
Pubblicazione: (2024)
GraphWalker: Agentic Knowledge Graph Question Answering via Synthetic Trajectory Curriculum
di: Xu, Shuwen, et al.
Pubblicazione: (2026)
di: Xu, Shuwen, et al.
Pubblicazione: (2026)
Enhancing Large Language Models with Pseudo- and Multisource- Knowledge Graphs for Open-ended Question Answering
di: Liu, Jiaxiang, et al.
Pubblicazione: (2024)
di: Liu, Jiaxiang, et al.
Pubblicazione: (2024)
LINKED: Eliciting, Filtering and Integrating Knowledge in Large Language Model for Commonsense Reasoning
di: Li, Jiachun, et al.
Pubblicazione: (2024)
di: Li, Jiachun, et al.
Pubblicazione: (2024)
Revealing the Deceptiveness of Knowledge Editing: A Mechanistic Analysis of Superficial Editing
di: Xie, Jiakuan, et al.
Pubblicazione: (2025)
di: Xie, Jiakuan, et al.
Pubblicazione: (2025)
Towards Robust Knowledge Unlearning: An Adversarial Framework for Assessing and Improving Unlearning Robustness in Large Language Models
di: Yuan, Hongbang, et al.
Pubblicazione: (2024)
di: Yuan, Hongbang, et al.
Pubblicazione: (2024)
Cracking Factual Knowledge: A Comprehensive Analysis of Degenerate Knowledge Neurons in Large Language Models
di: Chen, Yuheng, et al.
Pubblicazione: (2024)
di: Chen, Yuheng, et al.
Pubblicazione: (2024)
Unlocking the Future: Exploring Look-Ahead Planning Mechanistic Interpretability in Large Language Models
di: Men, Tianyi, et al.
Pubblicazione: (2024)
di: Men, Tianyi, et al.
Pubblicazione: (2024)
Learn to Refuse: Making Large Language Models More Controllable and Reliable through Knowledge Scope Limitation and Refusal Mechanism
di: Cao, Lang
Pubblicazione: (2023)
di: Cao, Lang
Pubblicazione: (2023)
One Mind, Many Tongues: A Deep Dive into Language-Agnostic Knowledge Neurons in Large Language Models
di: Cao, Pengfei, et al.
Pubblicazione: (2024)
di: Cao, Pengfei, et al.
Pubblicazione: (2024)
Mutual Enhancement Between Global Tokens and Patch Tokens: From Theory to Practice
di: Huang, Xiusheng, et al.
Pubblicazione: (2026)
di: Huang, Xiusheng, et al.
Pubblicazione: (2026)
Bridging Quantum and Semiclassical Volume: A Numerical Study of Coherent State Matrix Elements in Loop Quantum Gravity
di: Li, Haida, et al.
Pubblicazione: (2026)
di: Li, Haida, et al.
Pubblicazione: (2026)
Beyond Expectation Values: Generalized Semiclassical Expansions for Matrix Elements of Gauge Coherent States
di: Li, Haida, et al.
Pubblicazione: (2026)
di: Li, Haida, et al.
Pubblicazione: (2026)
Focus on Your Question! Interpreting and Mitigating Toxic CoT Problems in Commonsense Reasoning
di: Li, Jiachun, et al.
Pubblicazione: (2024)
di: Li, Jiachun, et al.
Pubblicazione: (2024)
Towards Atoms of Large Language Models
di: Hu, Chenhui, et al.
Pubblicazione: (2025)
di: Hu, Chenhui, et al.
Pubblicazione: (2025)
Task-Stratified Knowledge Scaling Laws for Post-Training Quantized Large Language Models
di: Zhou, Chenxi, et al.
Pubblicazione: (2025)
di: Zhou, Chenxi, et al.
Pubblicazione: (2025)
EAQEC codes from s-Galois hulls decomposition of linear codes
di: Li, Hui, et al.
Pubblicazione: (2024)
di: Li, Hui, et al.
Pubblicazione: (2024)
Linear complementary pairs of codes over a finite non-commutative Frobenius ring
di: Bhowmick, Sanjit, et al.
Pubblicazione: (2024)
di: Bhowmick, Sanjit, et al.
Pubblicazione: (2024)
Beyond Under-Alignment: Atomic Preference Enhanced Factuality Tuning for Large Language Models
di: Yuan, Hongbang, et al.
Pubblicazione: (2024)
di: Yuan, Hongbang, et al.
Pubblicazione: (2024)
The Knowledge Microscope: Features as Better Analytical Lenses than Neurons
di: Chen, Yuheng, et al.
Pubblicazione: (2025)
di: Chen, Yuheng, et al.
Pubblicazione: (2025)
If an LLM Were a Character, Would It Know Its Own Story? Evaluating Lifelong Learning in LLMs
di: Fan, Siqi, et al.
Pubblicazione: (2025)
di: Fan, Siqi, et al.
Pubblicazione: (2025)
WilKE: Wise-Layer Knowledge Editor for Lifelong Knowledge Editing
di: Hu, Chenhui, et al.
Pubblicazione: (2024)
di: Hu, Chenhui, et al.
Pubblicazione: (2024)
Delta Knowledge Distillation for Large Language Models
di: Cao, Yihan, et al.
Pubblicazione: (2025)
di: Cao, Yihan, et al.
Pubblicazione: (2025)
EvoEdit: Lifelong Free-Text Knowledge Editing through Latent Perturbation Augmentation and Knowledge-driven Parameter Fusion
di: Cao, Pengfei, et al.
Pubblicazione: (2025)
di: Cao, Pengfei, et al.
Pubblicazione: (2025)
Whispers that Shake Foundations: Analyzing and Mitigating False Premise Hallucinations in Large Language Models
di: Yuan, Hongbang, et al.
Pubblicazione: (2024)
di: Yuan, Hongbang, et al.
Pubblicazione: (2024)
The First to Know: How Token Distributions Reveal Hidden Knowledge in Large Vision-Language Models?
di: Zhao, Qinyu, et al.
Pubblicazione: (2024)
di: Zhao, Qinyu, et al.
Pubblicazione: (2024)
Knowledge Localization: Mission Not Accomplished? Enter Query Localization!
di: Chen, Yuheng, et al.
Pubblicazione: (2024)
di: Chen, Yuheng, et al.
Pubblicazione: (2024)
Towards Adaptive Mechanism Activation in Language Agent
di: Huang, Ziyang, et al.
Pubblicazione: (2024)
di: Huang, Ziyang, et al.
Pubblicazione: (2024)
Tug-of-War Between Knowledge: Exploring and Resolving Knowledge Conflicts in Retrieval-Augmented Language Models
di: Jin, Zhuoran, et al.
Pubblicazione: (2024)
di: Jin, Zhuoran, et al.
Pubblicazione: (2024)
Know More, Know Clearer: A Meta-Cognitive Framework for Knowledge Augmentation in Large Language Models
di: Chen, Hao, et al.
Pubblicazione: (2026)
di: Chen, Hao, et al.
Pubblicazione: (2026)
Deflection of Massive Spin-$\frac{1}{2}$ Particles around Kerr Black Hole
di: Li, Haida, et al.
Pubblicazione: (2025)
di: Li, Haida, et al.
Pubblicazione: (2025)
Particle Deflections around Microscopic Loop Quantum Black Holes with Rigorous Quantum Parameters
di: Li, Haida, et al.
Pubblicazione: (2025)
di: Li, Haida, et al.
Pubblicazione: (2025)
Gravitational Lensing Effect from The Revised Deser-Woodard Nonlocal Gravity
di: Li, Haida, et al.
Pubblicazione: (2026)
di: Li, Haida, et al.
Pubblicazione: (2026)
Performance Law of Large Language Models
di: Wu, Chuhan, et al.
Pubblicazione: (2024)
di: Wu, Chuhan, et al.
Pubblicazione: (2024)
KnowGPT: Knowledge Graph based Prompting for Large Language Models
di: Zhang, Qinggang, et al.
Pubblicazione: (2023)
di: Zhang, Qinggang, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Capability Localization: Capabilities Can be Localized rather than Individual Knowledge
di: Huang, Xiusheng, et al.
Pubblicazione: (2025) -
Reasons and Solutions for the Decline in Model Performance after Editing
di: Huang, Xiusheng, et al.
Pubblicazione: (2024) -
Commonsense Knowledge Editing Based on Free-Text in LLMs
di: Huang, Xiusheng, et al.
Pubblicazione: (2024) -
RWKU: Benchmarking Real-World Knowledge Unlearning for Large Language Models
di: Jin, Zhuoran, et al.
Pubblicazione: (2024) -
Cutting Off the Head Ends the Conflict: A Mechanism for Interpreting and Mitigating Knowledge Conflicts in Language Models
di: Jin, Zhuoran, et al.
Pubblicazione: (2024)