Mamba Knockout for Unraveling Factual Information Flow
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Endy, Nir, Grosbard, Idan Daniel, Ran-Milo, Yuval, Slutzky, Yonatan, Tshuva, Itay, Giryes, Raja |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Do Neural Networks Need Gradient Descent to Generalize? A Theoretical Study
von: Alexander, Yotam, et al.
Veröffentlicht: (2025)
von: Alexander, Yotam, et al.
Veröffentlicht: (2025)
Bridging the Visual Gap: Fine-Tuning Multimodal Models with Knowledge-Adapted Captions
von: Yanuka, Moran, et al.
Veröffentlicht: (2024)
von: Yanuka, Moran, et al.
Veröffentlicht: (2024)
ZOQO: Zero-Order Quantized Optimization
von: Bar, Noga, et al.
Veröffentlicht: (2025)
von: Bar, Noga, et al.
Veröffentlicht: (2025)
Mechanistically Interpretable Neural Encoding Reveals Fine-Grained Functional Selectivity in Human Visual Cortex
von: Grosbard, Idan Daniel, et al.
Veröffentlicht: (2026)
von: Grosbard, Idan Daniel, et al.
Veröffentlicht: (2026)
Provable Benefits of Complex Parameterizations for Structured State Space Models
von: Ran-Milo, Yuval, et al.
Veröffentlicht: (2024)
von: Ran-Milo, Yuval, et al.
Veröffentlicht: (2024)
Planted in Pretraining, Swayed by Finetuning: A Case Study on the Origins of Cognitive Biases in LLMs
von: Itzhak, Itay, et al.
Veröffentlicht: (2025)
von: Itzhak, Itay, et al.
Veröffentlicht: (2025)
Jamba: A Hybrid Transformer-Mamba Language Model
von: Lieber, Opher, et al.
Veröffentlicht: (2024)
von: Lieber, Opher, et al.
Veröffentlicht: (2024)
From Feelings to Metrics: Understanding and Formalizing How Users Vibe-Test LLMs
von: Itzhak, Itay, et al.
Veröffentlicht: (2026)
von: Itzhak, Itay, et al.
Veröffentlicht: (2026)
Attention Sinks Are Provably Necessary in Softmax Transformers: Evidence from Trigger-Conditional Tasks
von: Ran-Milo, Yuval
Veröffentlicht: (2026)
von: Ran-Milo, Yuval
Veröffentlicht: (2026)
Jamba-1.5: Hybrid Transformer-Mamba Models at Scale
von: Jamba Team, et al.
Veröffentlicht: (2024)
von: Jamba Team, et al.
Veröffentlicht: (2024)
Instructed to Bias: Instruction-Tuned Language Models Exhibit Emergent Cognitive Bias
von: Itzhak, Itay, et al.
Veröffentlicht: (2023)
von: Itzhak, Itay, et al.
Veröffentlicht: (2023)
Block Sparse Flash Attention
von: Ohayon, Daniel, et al.
Veröffentlicht: (2025)
von: Ohayon, Daniel, et al.
Veröffentlicht: (2025)
Pruning at Initialization -- A Sketching Perspective
von: Bar, Noga, et al.
Veröffentlicht: (2023)
von: Bar, Noga, et al.
Veröffentlicht: (2023)
Understanding Finetuning for Factual Knowledge Extraction
von: Ghosal, Gaurav, et al.
Veröffentlicht: (2024)
von: Ghosal, Gaurav, et al.
Veröffentlicht: (2024)
Leveraging Prototypical Representations for Mitigating Social Bias without Demographic Information
von: Iskander, Shadi, et al.
Veröffentlicht: (2024)
von: Iskander, Shadi, et al.
Veröffentlicht: (2024)
Assessing Factual Music Comprehension in Large Audio Language Models
von: Lin, Daniel Chenyu, et al.
Veröffentlicht: (2025)
von: Lin, Daniel Chenyu, et al.
Veröffentlicht: (2025)
Conformal Language Model Reasoning with Coherent Factuality
von: Rubin-Toles, Maxon, et al.
Veröffentlicht: (2025)
von: Rubin-Toles, Maxon, et al.
Veröffentlicht: (2025)
Persuasion Tokens for Editing Factual Knowledge in LLMs
von: Youssef, Paul, et al.
Veröffentlicht: (2026)
von: Youssef, Paul, et al.
Veröffentlicht: (2026)
Factual Consistency of Multilingual Pretrained Language Models
von: Fierro, Constanza, et al.
Veröffentlicht: (2022)
von: Fierro, Constanza, et al.
Veröffentlicht: (2022)
Language Models Need Inductive Biases to Count Inductively
von: Chang, Yingshan, et al.
Veröffentlicht: (2024)
von: Chang, Yingshan, et al.
Veröffentlicht: (2024)
Diverse Subset Selection via Norm-Based Sampling and Orthogonality
von: Bar, Noga, et al.
Veröffentlicht: (2024)
von: Bar, Noga, et al.
Veröffentlicht: (2024)
The Implicit Bias of Structured State Space Models Can Be Poisoned With Clean Labels
von: Slutzky, Yonatan, et al.
Veröffentlicht: (2024)
von: Slutzky, Yonatan, et al.
Veröffentlicht: (2024)
Verify with Caution: The Pitfalls of Relying on Imperfect Factuality Metrics
von: Godbole, Ameya, et al.
Veröffentlicht: (2025)
von: Godbole, Ameya, et al.
Veröffentlicht: (2025)
LoFTI: Localization and Factuality Transfer to Indian Locales
von: Simon, Sona Elza, et al.
Veröffentlicht: (2024)
von: Simon, Sona Elza, et al.
Veröffentlicht: (2024)
Temporally Consistent Factuality Probing for Large Language Models
von: Bajpai, Ashutosh, et al.
Veröffentlicht: (2024)
von: Bajpai, Ashutosh, et al.
Veröffentlicht: (2024)
Factual Confidence of LLMs: on Reliability and Robustness of Current Estimators
von: Mahaut, Matéo, et al.
Veröffentlicht: (2024)
von: Mahaut, Matéo, et al.
Veröffentlicht: (2024)
Zero-shot Factual Consistency Evaluation Across Domains
von: Agarwal, Raunak
Veröffentlicht: (2024)
von: Agarwal, Raunak
Veröffentlicht: (2024)
MenakBERT -- Hebrew Diacriticizer
von: Cohen, Ido, et al.
Veröffentlicht: (2024)
von: Cohen, Ido, et al.
Veröffentlicht: (2024)
Knowledge-Level Consistency Reinforcement Learning: Dual-Fact Alignment for Long-Form Factuality
von: Li, Junliang, et al.
Veröffentlicht: (2025)
von: Li, Junliang, et al.
Veröffentlicht: (2025)
Conformal Linguistic Calibration: Trading-off between Factuality and Specificity
von: Jiang, Zhengping, et al.
Veröffentlicht: (2025)
von: Jiang, Zhengping, et al.
Veröffentlicht: (2025)
SIFiD: Reassess Summary Factual Inconsistency Detection with LLM
von: Yang, Jiuding, et al.
Veröffentlicht: (2024)
von: Yang, Jiuding, et al.
Veröffentlicht: (2024)
Alexpaca: Learning Factual Clarification Question Generation Without Examples
von: Toles, Matthew, et al.
Veröffentlicht: (2023)
von: Toles, Matthew, et al.
Veröffentlicht: (2023)
GenAudit: Fixing Factual Errors in Language Model Outputs with Evidence
von: Krishna, Kundan, et al.
Veröffentlicht: (2024)
von: Krishna, Kundan, et al.
Veröffentlicht: (2024)
Alternate Preference Optimization for Unlearning Factual Knowledge in Large Language Models
von: Mekala, Anmol, et al.
Veröffentlicht: (2024)
von: Mekala, Anmol, et al.
Veröffentlicht: (2024)
Summing Up the Facts: Additive Mechanisms Behind Factual Recall in LLMs
von: Chughtai, Bilal, et al.
Veröffentlicht: (2024)
von: Chughtai, Bilal, et al.
Veröffentlicht: (2024)
Through a Compressed Lens: Investigating The Impact of Quantization on Factual Knowledge Recall
von: Wang, Qianli, et al.
Veröffentlicht: (2025)
von: Wang, Qianli, et al.
Veröffentlicht: (2025)
The Illusionist's Prompt: Exposing the Factual Vulnerabilities of Large Language Models with Linguistic Nuances
von: Wang, Yining, et al.
Veröffentlicht: (2025)
von: Wang, Yining, et al.
Veröffentlicht: (2025)
Who's Asking? Evaluating LLM Robustness to Inquiry Personas in Factual Question Answering
von: Akpinar, Nil-Jana, et al.
Veröffentlicht: (2025)
von: Akpinar, Nil-Jana, et al.
Veröffentlicht: (2025)
Identifying Factual Inconsistencies in Summaries: Grounding LLM Inference via Task Taxonomy
von: Xu, Liyan, et al.
Veröffentlicht: (2024)
von: Xu, Liyan, et al.
Veröffentlicht: (2024)
Language Models with Conformal Factuality Guarantees
von: Mohri, Christopher, et al.
Veröffentlicht: (2024)
von: Mohri, Christopher, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Do Neural Networks Need Gradient Descent to Generalize? A Theoretical Study
von: Alexander, Yotam, et al.
Veröffentlicht: (2025) -
Bridging the Visual Gap: Fine-Tuning Multimodal Models with Knowledge-Adapted Captions
von: Yanuka, Moran, et al.
Veröffentlicht: (2024) -
ZOQO: Zero-Order Quantized Optimization
von: Bar, Noga, et al.
Veröffentlicht: (2025) -
Mechanistically Interpretable Neural Encoding Reveals Fine-Grained Functional Selectivity in Human Visual Cortex
von: Grosbard, Idan Daniel, et al.
Veröffentlicht: (2026) -
Provable Benefits of Complex Parameterizations for Structured State Space Models
von: Ran-Milo, Yuval, et al.
Veröffentlicht: (2024)