Latent Concept-based Explanation of NLP Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Yu, Xuemin, Dalvi, Fahim, Durrani, Nadir, Nouri, Marzia, Sajjad, Hassan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Discovering Salient Neurons in Deep NLP Models
von: Durrani, Nadir, et al.
Veröffentlicht: (2022)
von: Durrani, Nadir, et al.
Veröffentlicht: (2022)
Scaling up Discovery of Latent Concepts in Deep NLP Models
von: Hawasly, Majd, et al.
Veröffentlicht: (2023)
von: Hawasly, Majd, et al.
Veröffentlicht: (2023)
Editing Across Languages: A Survey of Multilingual Knowledge Editing
von: Durrani, Nadir, et al.
Veröffentlicht: (2025)
von: Durrani, Nadir, et al.
Veröffentlicht: (2025)
An Exploration of Knowledge Editing for Arabic
von: Mousi, Basel, et al.
Veröffentlicht: (2025)
von: Mousi, Basel, et al.
Veröffentlicht: (2025)
Beyond the Leaderboard: Understanding Performance Disparities in Large Language Models via Model Diffing
von: Boughorbel, Sabri, et al.
Veröffentlicht: (2025)
von: Boughorbel, Sabri, et al.
Veröffentlicht: (2025)
Once Correct, Still Wrong: Counterfactual Hallucination in Multilingual Vision-Language Models
von: Mousi, Basel, et al.
Veröffentlicht: (2026)
von: Mousi, Basel, et al.
Veröffentlicht: (2026)
From Words to Waves: Analyzing Concept Formation in Speech and Text-Based Foundation Models
von: Ersoy, Asım, et al.
Veröffentlicht: (2025)
von: Ersoy, Asım, et al.
Veröffentlicht: (2025)
Exploring Alignment in Shared Cross-lingual Spaces
von: Mousi, Basel, et al.
Veröffentlicht: (2024)
von: Mousi, Basel, et al.
Veröffentlicht: (2024)
Vector Quantized Latent Concepts: A Scalable Alternative to Clustering-Based Concept Discovery
von: Yu, Xuemin, et al.
Veröffentlicht: (2026)
von: Yu, Xuemin, et al.
Veröffentlicht: (2026)
Cross-Layer Discrete Concept Discovery for Interpreting Language Models
von: Garg, Ankur, et al.
Veröffentlicht: (2025)
von: Garg, Ankur, et al.
Veröffentlicht: (2025)
The Landscape of Arabic Large Language Models (ALLMs): A New Era for Arabic Language Technology
von: Al-Khalifa, Shahad, et al.
Veröffentlicht: (2025)
von: Al-Khalifa, Shahad, et al.
Veröffentlicht: (2025)
AraDiCE: Benchmarks for Dialectal and Cultural Capabilities in LLMs
von: Mousi, Basel, et al.
Veröffentlicht: (2024)
von: Mousi, Basel, et al.
Veröffentlicht: (2024)
There Is More to Refusal in Large Language Models than a Single Direction
von: Joad, Faaiz, et al.
Veröffentlicht: (2026)
von: Joad, Faaiz, et al.
Veröffentlicht: (2026)
Towards Faithful Model Explanation in NLP: A Survey
von: Lyu, Qing, et al.
Veröffentlicht: (2022)
von: Lyu, Qing, et al.
Veröffentlicht: (2022)
OASIS: A Multilingual and Multimodal Dataset for Culturally Grounded Spoken Visual QA
von: Alam, Firoj, et al.
Veröffentlicht: (2025)
von: Alam, Firoj, et al.
Veröffentlicht: (2025)
Faithfulness and the Notion of Adversarial Sensitivity in NLP Explanations
von: Manna, Supriya, et al.
Veröffentlicht: (2024)
von: Manna, Supriya, et al.
Veröffentlicht: (2024)
Understanding Syntactic Generalization in Structure-inducing Language Models
von: Arps, David, et al.
Veröffentlicht: (2025)
von: Arps, David, et al.
Veröffentlicht: (2025)
Latent Factor Models Meets Instructions: Goal-conditioned Latent Factor Discovery without Task Supervision
von: Xie, Zhouhang, et al.
Veröffentlicht: (2025)
von: Xie, Zhouhang, et al.
Veröffentlicht: (2025)
Large Language Models on Wikipedia-Style Survey Generation: an Evaluation in NLP Concepts
von: Gao, Fan, et al.
Veröffentlicht: (2023)
von: Gao, Fan, et al.
Veröffentlicht: (2023)
LLMeBench: A Flexible Framework for Accelerating LLMs Benchmarking
von: Dalvi, Fahim, et al.
Veröffentlicht: (2023)
von: Dalvi, Fahim, et al.
Veröffentlicht: (2023)
Leveraging Large Language Models for Concept Graph Recovery and Question Answering in NLP Education
von: Yang, Rui, et al.
Veröffentlicht: (2024)
von: Yang, Rui, et al.
Veröffentlicht: (2024)
CHiRPE: A Step Towards Real-World Clinical NLP with Clinician-Oriented Model Explanations
von: Fong, Stephanie, et al.
Veröffentlicht: (2026)
von: Fong, Stephanie, et al.
Veröffentlicht: (2026)
Resolving Lexical Bias in Model Editing
von: Rizwan, Hammad, et al.
Veröffentlicht: (2024)
von: Rizwan, Hammad, et al.
Veröffentlicht: (2024)
On Evaluating Explanation Utility for Human-AI Decision Making in NLP
von: Chaleshtori, Fateme Hashemi, et al.
Veröffentlicht: (2024)
von: Chaleshtori, Fateme Hashemi, et al.
Veröffentlicht: (2024)
Long-form evaluation of model editing
von: Rosati, Domenic, et al.
Veröffentlicht: (2024)
von: Rosati, Domenic, et al.
Veröffentlicht: (2024)
Latent Concept Disentanglement in Transformer-based Language Models
von: Hong, Guan Zhe, et al.
Veröffentlicht: (2025)
von: Hong, Guan Zhe, et al.
Veröffentlicht: (2025)
UrBLiMP: A Benchmark for Evaluating the Linguistic Competence of Large Language Models in Urdu
von: Adeeba, Farah, et al.
Veröffentlicht: (2025)
von: Adeeba, Farah, et al.
Veröffentlicht: (2025)
Multilingual Nonce Dependency Treebanks: Understanding how Language Models represent and process syntactic structure
von: Arps, David, et al.
Veröffentlicht: (2023)
von: Arps, David, et al.
Veröffentlicht: (2023)
DIALECTBENCH: A NLP Benchmark for Dialects, Varieties, and Closely-Related Languages
von: Faisal, Fahim, et al.
Veröffentlicht: (2024)
von: Faisal, Fahim, et al.
Veröffentlicht: (2024)
CUICurate: A GraphRAG-based Framework for Automated Clinical Concept Curation for NLP applications
von: Blake, Victoria, et al.
Veröffentlicht: (2026)
von: Blake, Victoria, et al.
Veröffentlicht: (2026)
LAraBench: Benchmarking Arabic AI with Large Language Models
von: Abdelali, Ahmed, et al.
Veröffentlicht: (2023)
von: Abdelali, Ahmed, et al.
Veröffentlicht: (2023)
Quantifying the Capabilities of LLMs across Scale and Precision
von: Badshah, Sher, et al.
Veröffentlicht: (2024)
von: Badshah, Sher, et al.
Veröffentlicht: (2024)
Interpreting the Effects of Quantization on LLMs
von: Singh, Manpreet, et al.
Veröffentlicht: (2025)
von: Singh, Manpreet, et al.
Veröffentlicht: (2025)
From Code-Centric to Concept-Centric: Teaching NLP with LLM-Assisted "Vibe Coding"
von: Al-Khalifa, Hend
Veröffentlicht: (2026)
von: Al-Khalifa, Hend
Veröffentlicht: (2026)
Enhancing the Comprehensibility of Text Explanations via Unsupervised Concept Discovery
von: Sun, Yifan, et al.
Veröffentlicht: (2025)
von: Sun, Yifan, et al.
Veröffentlicht: (2025)
Robust Explanations for User Trust in Enterprise NLP Systems
von: Zhang, Guilin, et al.
Veröffentlicht: (2026)
von: Zhang, Guilin, et al.
Veröffentlicht: (2026)
TALE: A Tool-Augmented Framework for Reference-Free Evaluation of Large Language Models
von: Badshah, Sher, et al.
Veröffentlicht: (2025)
von: Badshah, Sher, et al.
Veröffentlicht: (2025)
An Efficient Approach for Studying Cross-Lingual Transfer in Multilingual Language Models
von: Faisal, Fahim, et al.
Veröffentlicht: (2024)
von: Faisal, Fahim, et al.
Veröffentlicht: (2024)
ConSim: Measuring Concept-Based Explanations' Effectiveness with Automated Simulatability
von: Poché, Antonin, et al.
Veröffentlicht: (2025)
von: Poché, Antonin, et al.
Veröffentlicht: (2025)
The Nature of NLP: Analyzing Contributions in NLP Papers
von: Pramanick, Aniket, et al.
Veröffentlicht: (2024)
von: Pramanick, Aniket, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Discovering Salient Neurons in Deep NLP Models
von: Durrani, Nadir, et al.
Veröffentlicht: (2022) -
Scaling up Discovery of Latent Concepts in Deep NLP Models
von: Hawasly, Majd, et al.
Veröffentlicht: (2023) -
Editing Across Languages: A Survey of Multilingual Knowledge Editing
von: Durrani, Nadir, et al.
Veröffentlicht: (2025) -
An Exploration of Knowledge Editing for Arabic
von: Mousi, Basel, et al.
Veröffentlicht: (2025) -
Beyond the Leaderboard: Understanding Performance Disparities in Large Language Models via Model Diffing
von: Boughorbel, Sabri, et al.
Veröffentlicht: (2025)