Faithfulness Measurable Masked Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Madsen, Andreas, Reddy, Siva, Chandar, Sarath |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Are self-explanations from Large Language Models faithful?
di: Madsen, Andreas, et al.
Pubblicazione: (2024)
di: Madsen, Andreas, et al.
Pubblicazione: (2024)
Interpretability Needs a New Paradigm
di: Madsen, Andreas, et al.
Pubblicazione: (2024)
di: Madsen, Andreas, et al.
Pubblicazione: (2024)
New Faithfulness-Centric Interpretability Paradigms for Natural Language Processing
di: Madsen, Andreas
Pubblicazione: (2024)
di: Madsen, Andreas
Pubblicazione: (2024)
The Markovian Thinker: Architecture-Agnostic Linear Scaling of Reasoning
di: Aghajohari, Milad, et al.
Pubblicazione: (2025)
di: Aghajohari, Milad, et al.
Pubblicazione: (2025)
Robust Infidelity: When Faithfulness Measures on Masked Language Models Are Misleading
di: Crothers, Evan, et al.
Pubblicazione: (2023)
di: Crothers, Evan, et al.
Pubblicazione: (2023)
Effect of Document Packing on the Latent Multi-Hop Reasoning Capabilities of Large Language Models
di: Prato, Gabriele, et al.
Pubblicazione: (2025)
di: Prato, Gabriele, et al.
Pubblicazione: (2025)
Do Large Language Models Know How Much They Know?
di: Prato, Gabriele, et al.
Pubblicazione: (2025)
di: Prato, Gabriele, et al.
Pubblicazione: (2025)
Investigating the Multilingual Calibration Effects of Language Model Instruction-Tuning
di: Huang, Jerry, et al.
Pubblicazione: (2026)
di: Huang, Jerry, et al.
Pubblicazione: (2026)
Towards Practical Tool Usage for Continually Learning LLMs
di: Huang, Jerry, et al.
Pubblicazione: (2024)
di: Huang, Jerry, et al.
Pubblicazione: (2024)
Too Big to Fool: Resisting Deception in Language Models
di: Samsami, Mohammad Reza, et al.
Pubblicazione: (2024)
di: Samsami, Mohammad Reza, et al.
Pubblicazione: (2024)
Should We Attend More or Less? Modulating Attention for Fairness
di: Zayed, Abdelrahman, et al.
Pubblicazione: (2023)
di: Zayed, Abdelrahman, et al.
Pubblicazione: (2023)
Why Don't Prompt-Based Fairness Metrics Correlate?
di: Zayed, Abdelrahman, et al.
Pubblicazione: (2024)
di: Zayed, Abdelrahman, et al.
Pubblicazione: (2024)
Mapping Faithful Reasoning in Language Models
di: Li, Jiazheng, et al.
Pubblicazione: (2025)
di: Li, Jiazheng, et al.
Pubblicazione: (2025)
Do Robot Snakes Dream like Electric Sheep? Investigating the Effects of Architectural Inductive Biases on Hallucination
di: Huang, Jerry, et al.
Pubblicazione: (2024)
di: Huang, Jerry, et al.
Pubblicazione: (2024)
Revisiting Replay and Gradient Alignment for Continual Pre-Training of Large Language Models
di: Abbes, Istabrak, et al.
Pubblicazione: (2025)
di: Abbes, Istabrak, et al.
Pubblicazione: (2025)
Walk the Talk? Measuring the Faithfulness of Large Language Model Explanations
di: Matton, Katie, et al.
Pubblicazione: (2025)
di: Matton, Katie, et al.
Pubblicazione: (2025)
MetaFaith: Faithful Natural Language Uncertainty Expression in LLMs
di: Liu, Gabrielle Kaili-May, et al.
Pubblicazione: (2025)
di: Liu, Gabrielle Kaili-May, et al.
Pubblicazione: (2025)
FaithLM: Towards Faithful Explanations for Large Language Models
di: Chuang, Yu-Neng, et al.
Pubblicazione: (2024)
di: Chuang, Yu-Neng, et al.
Pubblicazione: (2024)
Forecasting Downstream Performance of LLMs With Proxy Metrics
di: Patel, Arkil, et al.
Pubblicazione: (2026)
di: Patel, Arkil, et al.
Pubblicazione: (2026)
NeuroFaith: Evaluating LLM Self-Explanation Faithfulness via Internal Representation Alignment
di: Bhan, Milan, et al.
Pubblicazione: (2025)
di: Bhan, Milan, et al.
Pubblicazione: (2025)
Build the web for agents, not agents for the web
di: Lù, Xing Han, et al.
Pubblicazione: (2025)
di: Lù, Xing Han, et al.
Pubblicazione: (2025)
RePro: Training Language Models to Faithfully Recycle the Web for Pretraining
di: Yu, Zichun, et al.
Pubblicazione: (2025)
di: Yu, Zichun, et al.
Pubblicazione: (2025)
Retrieval-Augmented and Knowledge-Grounded Language Models for Faithful Clinical Medicine
di: Liu, Fenglin, et al.
Pubblicazione: (2022)
di: Liu, Fenglin, et al.
Pubblicazione: (2022)
Characterizing Large Language Model Geometry Helps Solve Toxicity Detection and Generation
di: Balestriero, Randall, et al.
Pubblicazione: (2023)
di: Balestriero, Randall, et al.
Pubblicazione: (2023)
Representation Deficiency in Masked Language Modeling
di: Meng, Yu, et al.
Pubblicazione: (2023)
di: Meng, Yu, et al.
Pubblicazione: (2023)
Contextual Text Denoising with Masked Language Models
di: Sun, Yifu, et al.
Pubblicazione: (2019)
di: Sun, Yifu, et al.
Pubblicazione: (2019)
Scaling Beyond Masked Diffusion Language Models
di: Sahoo, Subham Sekhar, et al.
Pubblicazione: (2026)
di: Sahoo, Subham Sekhar, et al.
Pubblicazione: (2026)
FaithEval: Can Your Language Model Stay Faithful to Context, Even If "The Moon is Made of Marshmallows"
di: Ming, Yifei, et al.
Pubblicazione: (2024)
di: Ming, Yifei, et al.
Pubblicazione: (2024)
WebLINX: Real-World Website Navigation with Multi-Turn Dialogue
di: Lù, Xing Han, et al.
Pubblicazione: (2024)
di: Lù, Xing Han, et al.
Pubblicazione: (2024)
Measuring What LLMs Think They Do: SHAP Faithfulness and Deployability on Financial Tabular Classification
di: AlMarri, Saeed, et al.
Pubblicazione: (2025)
di: AlMarri, Saeed, et al.
Pubblicazione: (2025)
Small Encoders Can Rival Large Decoders in Detecting Groundedness
di: Abbes, Istabrak, et al.
Pubblicazione: (2025)
di: Abbes, Istabrak, et al.
Pubblicazione: (2025)
Dynamic Attention-Guided Context Decoding for Mitigating Context Faithfulness Hallucinations in Large Language Models
di: Huang, Yanwen, et al.
Pubblicazione: (2025)
di: Huang, Yanwen, et al.
Pubblicazione: (2025)
Measuring Chain-of-Thought Monitorability Through Faithfulness and Verbosity
di: Meek, Austin, et al.
Pubblicazione: (2025)
di: Meek, Austin, et al.
Pubblicazione: (2025)
BRIDGE: Predicting Human Task Completion Time From Model Performance
di: Liu, Fengyuan, et al.
Pubblicazione: (2026)
di: Liu, Fengyuan, et al.
Pubblicazione: (2026)
Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning
di: Jia, Jinghan, et al.
Pubblicazione: (2026)
di: Jia, Jinghan, et al.
Pubblicazione: (2026)
Towards Probabilistically-Sound Beam Search with Masked Language Models
di: Brooks, Creston, et al.
Pubblicazione: (2024)
di: Brooks, Creston, et al.
Pubblicazione: (2024)
Diffusion-State Policy Optimization for Masked Diffusion Language Models
di: Oba, Daisuke, et al.
Pubblicazione: (2026)
di: Oba, Daisuke, et al.
Pubblicazione: (2026)
DOS: Dependency-Oriented Sampler for Masked Diffusion Language Models
di: Zhou, Xueyu, et al.
Pubblicazione: (2026)
di: Zhou, Xueyu, et al.
Pubblicazione: (2026)
Reconsidering Positional Supervision in Masked Diffusion Language Model Training
di: Ye, Mengyu, et al.
Pubblicazione: (2026)
di: Ye, Mengyu, et al.
Pubblicazione: (2026)
Contextual Graph Transformer: A Small Language Model for Enhanced Engineering Document Information Extraction
di: Reddy, Karan, et al.
Pubblicazione: (2025)
di: Reddy, Karan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Are self-explanations from Large Language Models faithful?
di: Madsen, Andreas, et al.
Pubblicazione: (2024) -
Interpretability Needs a New Paradigm
di: Madsen, Andreas, et al.
Pubblicazione: (2024) -
New Faithfulness-Centric Interpretability Paradigms for Natural Language Processing
di: Madsen, Andreas
Pubblicazione: (2024) -
The Markovian Thinker: Architecture-Agnostic Linear Scaling of Reasoning
di: Aghajohari, Milad, et al.
Pubblicazione: (2025) -
Robust Infidelity: When Faithfulness Measures on Masked Language Models Are Misleading
di: Crothers, Evan, et al.
Pubblicazione: (2023)