On Behalf of the Stakeholders: Trends in NLP Model Interpretability in the Era of LLMs
Fuente:
arXiv
Salvato in:
| Autori principali: | Calderon, Nitay, Reichart, Roi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
di: Calderon, Nitay, et al.
Pubblicazione: (2025)
di: Calderon, Nitay, et al.
Pubblicazione: (2025)
LIBERTy: A Causal Framework for Benchmarking Concept-Based Explanations of LLMs with Structural Counterfactuals
di: Toker, Gilat, et al.
Pubblicazione: (2026)
di: Toker, Gilat, et al.
Pubblicazione: (2026)
The Colorful Future of LLMs: Evaluating and Improving LLMs as Emotional Supporters for Queer Youth
di: Lissak, Shir, et al.
Pubblicazione: (2024)
di: Lissak, Shir, et al.
Pubblicazione: (2024)
NL-Eye: Abductive NLI for Images
di: Ventura, Mor, et al.
Pubblicazione: (2024)
di: Ventura, Mor, et al.
Pubblicazione: (2024)
Multi-Domain Explainability of Preferences
di: Calderon, Nitay, et al.
Pubblicazione: (2025)
di: Calderon, Nitay, et al.
Pubblicazione: (2025)
Are LLMs Better than Reported? Detecting Label Errors and Mitigating Their Effect on Model Performance
di: Nahum, Omer, et al.
Pubblicazione: (2024)
di: Nahum, Omer, et al.
Pubblicazione: (2024)
MEETING DELEGATE: Benchmarking LLMs on Attending Meetings on Our Behalf
di: Hu, Lingxiang, et al.
Pubblicazione: (2025)
di: Hu, Lingxiang, et al.
Pubblicazione: (2025)
Measuring the Robustness of NLP Models to Domain Shifts
di: Calderon, Nitay, et al.
Pubblicazione: (2023)
di: Calderon, Nitay, et al.
Pubblicazione: (2023)
Systematic Biases in LLM Simulations of Debates
di: Taubenfeld, Amir, et al.
Pubblicazione: (2024)
di: Taubenfeld, Amir, et al.
Pubblicazione: (2024)
Predicting Decisions of AI Agents from Limited Interaction through Text-Tabular Modeling
di: Shapira, Eilam, et al.
Pubblicazione: (2026)
di: Shapira, Eilam, et al.
Pubblicazione: (2026)
Navigating Cultural Chasms: Exploring and Unlocking the Cultural POV of Text-To-Image Models
di: Ventura, Mor, et al.
Pubblicazione: (2023)
di: Ventura, Mor, et al.
Pubblicazione: (2023)
Dementia Through Different Eyes: Explainable Modeling of Human and LLM Perceptions for Early Awareness
di: Peled-Cohen, Lotem, et al.
Pubblicazione: (2025)
di: Peled-Cohen, Lotem, et al.
Pubblicazione: (2025)
Can LLMs Replace Economic Choice Prediction Labs? The Case of Language-based Persuasion Games
di: Shapira, Eilam, et al.
Pubblicazione: (2024)
di: Shapira, Eilam, et al.
Pubblicazione: (2024)
DeLeaker: Dynamic Inference-Time Reweighting For Semantic Leakage Mitigation in Text-to-Image Models
di: Ventura, Mor, et al.
Pubblicazione: (2025)
di: Ventura, Mor, et al.
Pubblicazione: (2025)
The Poisoned Apple Effect: Strategic Manipulation of Mediated Markets via Technology Expansion of AI Agents
di: Shapira, Eilam, et al.
Pubblicazione: (2026)
di: Shapira, Eilam, et al.
Pubblicazione: (2026)
AdaptiVocab: Enhancing LLM Efficiency in Focused Domains through Lightweight Vocabulary Adaptation
di: Nakash, Itay, et al.
Pubblicazione: (2025)
di: Nakash, Itay, et al.
Pubblicazione: (2025)
MASE: Interpretable NLP Models via Model-Agnostic Saliency Estimation
di: Yang, Zhou, et al.
Pubblicazione: (2025)
di: Yang, Zhou, et al.
Pubblicazione: (2025)
Empty Shelves or Lost Keys? Recall Is the Bottleneck for Parametric Factuality
di: Calderon, Nitay, et al.
Pubblicazione: (2026)
di: Calderon, Nitay, et al.
Pubblicazione: (2026)
A Systematic Review of NLP for Dementia -- Tasks, Datasets and Opportunities
di: Peled-Cohen, Lotem, et al.
Pubblicazione: (2024)
di: Peled-Cohen, Lotem, et al.
Pubblicazione: (2024)
Donors and Recipients: On Asymmetric Transfer Across Tasks and Languages with Parameter-Efficient Fine-Tuning
di: Dymkiewicz, Kajetan, et al.
Pubblicazione: (2025)
di: Dymkiewicz, Kajetan, et al.
Pubblicazione: (2025)
Pretrained LLMs Learn Multiple Types of Uncertainty
di: Cohen, Roi, et al.
Pubblicazione: (2025)
di: Cohen, Roi, et al.
Pubblicazione: (2025)
Advancing NLP Security by Leveraging LLMs as Adversarial Engines
di: Srinivasan, Sudarshan, et al.
Pubblicazione: (2024)
di: Srinivasan, Sudarshan, et al.
Pubblicazione: (2024)
Ground Truth Generation for Multilingual Historical NLP using LLMs
di: Gladstone, Clovis, et al.
Pubblicazione: (2025)
di: Gladstone, Clovis, et al.
Pubblicazione: (2025)
LLMs Know More Than They Show: On the Intrinsic Representation of LLM Hallucinations
di: Orgad, Hadas, et al.
Pubblicazione: (2024)
di: Orgad, Hadas, et al.
Pubblicazione: (2024)
Rethinking Interpretability in the Era of Large Language Models
di: Singh, Chandan, et al.
Pubblicazione: (2024)
di: Singh, Chandan, et al.
Pubblicazione: (2024)
From Transformers to LLMs: A Systematic Survey of Efficiency Considerations in NLP
di: Ansar, Wazib, et al.
Pubblicazione: (2024)
di: Ansar, Wazib, et al.
Pubblicazione: (2024)
From Noise to Signal to Selbstzweck: Reframing Human Label Variation in the Era of Post-training in NLP
di: Xu, Shanshan, et al.
Pubblicazione: (2025)
di: Xu, Shanshan, et al.
Pubblicazione: (2025)
Moral Mazes in the Era of LLMs
di: Nguyen, Dang, et al.
Pubblicazione: (2026)
di: Nguyen, Dang, et al.
Pubblicazione: (2026)
Mind Your Theory: Theory of Mind Goes Deeper Than Reasoning
di: Wagner, Eitan, et al.
Pubblicazione: (2024)
di: Wagner, Eitan, et al.
Pubblicazione: (2024)
Extreme Speech Classification in the Era of LLMs: Exploring Open-Source and Proprietary Models
di: Mahajan, Sarthak, et al.
Pubblicazione: (2025)
di: Mahajan, Sarthak, et al.
Pubblicazione: (2025)
Cognitive BASIC: An In-Model Interpreted Reasoning Language for LLMs
di: Kramer, Oliver
Pubblicazione: (2025)
di: Kramer, Oliver
Pubblicazione: (2025)
Machine-Assisted Grading of Nationwide School-Leaving Essay Exams with LLMs and Statistical NLP
di: Karjus, Andres, et al.
Pubblicazione: (2026)
di: Karjus, Andres, et al.
Pubblicazione: (2026)
Replace, Don't Expand: Mitigating Context Dilution in Multi-Hop RAG via Fixed-Budget Evidence Assembly
di: Lahmy, Moshe, et al.
Pubblicazione: (2025)
di: Lahmy, Moshe, et al.
Pubblicazione: (2025)
LLMs in Interpreting Legal Documents
di: Corbo, Simone
Pubblicazione: (2025)
di: Corbo, Simone
Pubblicazione: (2025)
Recommender Systems in the Era of Large Language Models (LLMs)
di: Zhao, Zihuai, et al.
Pubblicazione: (2023)
di: Zhao, Zihuai, et al.
Pubblicazione: (2023)
The Pitfalls of Publishing in the Age of LLMs: Strange and Surprising Adventures with a High-Impact NLP Journal
di: Verma, Rakesh M., et al.
Pubblicazione: (2024)
di: Verma, Rakesh M., et al.
Pubblicazione: (2024)
Do BERT-Like Bidirectional Models Still Perform Better on Text Classification in the Era of LLMs?
di: Zhang, Junyan, et al.
Pubblicazione: (2025)
di: Zhang, Junyan, et al.
Pubblicazione: (2025)
Large Language Models Meet NLP: A Survey
di: Qin, Libo, et al.
Pubblicazione: (2024)
di: Qin, Libo, et al.
Pubblicazione: (2024)
Deanthropomorphising NLP: Can a Language Model Be Conscious?
di: Shardlow, Matthew, et al.
Pubblicazione: (2022)
di: Shardlow, Matthew, et al.
Pubblicazione: (2022)
Interpretable AI for Time-Series: Multi-Model Heatmap Fusion with Global Attention and NLP-Generated Explanations
di: Francis, Jiztom Kavalakkatt, et al.
Pubblicazione: (2025)
di: Francis, Jiztom Kavalakkatt, et al.
Pubblicazione: (2025)
Documenti analoghi
-
The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
di: Calderon, Nitay, et al.
Pubblicazione: (2025) -
LIBERTy: A Causal Framework for Benchmarking Concept-Based Explanations of LLMs with Structural Counterfactuals
di: Toker, Gilat, et al.
Pubblicazione: (2026) -
The Colorful Future of LLMs: Evaluating and Improving LLMs as Emotional Supporters for Queer Youth
di: Lissak, Shir, et al.
Pubblicazione: (2024) -
NL-Eye: Abductive NLI for Images
di: Ventura, Mor, et al.
Pubblicazione: (2024) -
Multi-Domain Explainability of Preferences
di: Calderon, Nitay, et al.
Pubblicazione: (2025)