Does It Make Sense to Explain a Black Box With Another Black Box?
Fuente:
arXiv
Salvato in:
| Autori principali: | Delaunay, Julien, Galárraga, Luis, Largouët, Christine |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Surfing the modeling of PoS taggers in low-resource scenarios
di: Ferro, Manuel Vilares, et al.
Pubblicazione: (2024)
di: Ferro, Manuel Vilares, et al.
Pubblicazione: (2024)
Composition of Relational Features with an Application to Explaining Black-Box Predictors
di: Srinivasan, Ashwin, et al.
Pubblicazione: (2022)
di: Srinivasan, Ashwin, et al.
Pubblicazione: (2022)
Pivot Language for Low-Resource Machine Translation
di: Talwar, Abhimanyu, et al.
Pubblicazione: (2025)
di: Talwar, Abhimanyu, et al.
Pubblicazione: (2025)
Uncertainty Estimation and Quantification for LLMs: A Simple Supervised Approach
di: Liu, Linyu, et al.
Pubblicazione: (2024)
di: Liu, Linyu, et al.
Pubblicazione: (2024)
Aligning Black-box Language Models with Human Judgments
di: Burg, Gerrit J. J. van den, et al.
Pubblicazione: (2025)
di: Burg, Gerrit J. J. van den, et al.
Pubblicazione: (2025)
Evaluating Large Language Models for Public Health Classification and Extraction Tasks
di: Harris, Joshua, et al.
Pubblicazione: (2024)
di: Harris, Joshua, et al.
Pubblicazione: (2024)
QuAILoRA: Quantization-Aware Initialization for LoRA
di: Lawton, Neal, et al.
Pubblicazione: (2024)
di: Lawton, Neal, et al.
Pubblicazione: (2024)
Review GIDE -- Restaurant Review Gastrointestinal Illness Detection and Extraction with Large Language Models
di: Laurence, Timothy, et al.
Pubblicazione: (2025)
di: Laurence, Timothy, et al.
Pubblicazione: (2025)
CausalSent: Interpretable Sentiment Classification with RieszNet
di: Frees, Daniel, et al.
Pubblicazione: (2025)
di: Frees, Daniel, et al.
Pubblicazione: (2025)
Healthy LLMs? Benchmarking LLM Knowledge of UK Government Public Health Information
di: Harris, Joshua, et al.
Pubblicazione: (2025)
di: Harris, Joshua, et al.
Pubblicazione: (2025)
Optimization Strategies for Enhancing Resource Efficiency in Transformers & Large Language Models
di: Wallace, Tom, et al.
Pubblicazione: (2025)
di: Wallace, Tom, et al.
Pubblicazione: (2025)
Recent Advances in Named Entity Recognition: A Comprehensive Survey and Comparative Study
di: Keraghel, Imed, et al.
Pubblicazione: (2024)
di: Keraghel, Imed, et al.
Pubblicazione: (2024)
LLM-Assisted Crisis Management: Building Advanced LLM Platforms for Effective Emergency Response and Public Collaboration
di: Otal, Hakan T., et al.
Pubblicazione: (2024)
di: Otal, Hakan T., et al.
Pubblicazione: (2024)
Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring
di: Aksoy, Sinan G., et al.
Pubblicazione: (2026)
di: Aksoy, Sinan G., et al.
Pubblicazione: (2026)
Template-Based Probes Are Imperfect Lenses for Counterfactual Bias Evaluation in LLMs
di: Kohankhaki, Farnaz, et al.
Pubblicazione: (2024)
di: Kohankhaki, Farnaz, et al.
Pubblicazione: (2024)
Clustering in pure-attention hardmax transformers and its role in sentiment analysis
di: Alcalde, Albert, et al.
Pubblicazione: (2024)
di: Alcalde, Albert, et al.
Pubblicazione: (2024)
A Language Model-Driven Semi-Supervised Ensemble Framework for Illicit Market Detection Across Deep/Dark Web and Social Platforms
di: Yazdanjue, Navid, et al.
Pubblicazione: (2025)
di: Yazdanjue, Navid, et al.
Pubblicazione: (2025)
Evaluating Named Entity Recognition: A comparative analysis of mono- and multilingual transformer models on a novel Brazilian corporate earnings call transcripts dataset
di: Abilio, Ramon, et al.
Pubblicazione: (2024)
di: Abilio, Ramon, et al.
Pubblicazione: (2024)
Self-Attention as Transport: Limits of Symmetric Spectral Diagnostics
di: Dahlem, Dominik, et al.
Pubblicazione: (2026)
di: Dahlem, Dominik, et al.
Pubblicazione: (2026)
OptPO: Optimal Rollout Allocation for Test-time Policy Optimization
di: Wang, Youkang, et al.
Pubblicazione: (2025)
di: Wang, Youkang, et al.
Pubblicazione: (2025)
LEGAL-UQA: A Low-Resource Urdu-English Dataset for Legal Question Answering
di: Faisal, Faizan, et al.
Pubblicazione: (2024)
di: Faisal, Faizan, et al.
Pubblicazione: (2024)
Measuring Faithfulness and Abstention: An Automated Pipeline for Evaluating LLM-Generated 3-ply Case-Based Legal Arguments
di: Zhang, Li, et al.
Pubblicazione: (2025)
di: Zhang, Li, et al.
Pubblicazione: (2025)
Data filtering methods for training language models
di: Shevchenko, Egor, et al.
Pubblicazione: (2026)
di: Shevchenko, Egor, et al.
Pubblicazione: (2026)
SVDq: 1.25-bit and 410x Key Cache Compression for LLM Attention
di: Yankun, Hong, et al.
Pubblicazione: (2025)
di: Yankun, Hong, et al.
Pubblicazione: (2025)
Mediator: Memory-efficient LLM Merging with Less Parameter Conflicts and Uncertainty Based Routing
di: Lai, Kunfeng, et al.
Pubblicazione: (2025)
di: Lai, Kunfeng, et al.
Pubblicazione: (2025)
Which Pieces Does Unigram Tokenization Really Need?
di: Land, Sander, et al.
Pubblicazione: (2025)
di: Land, Sander, et al.
Pubblicazione: (2025)
In-Context Explainers: Harnessing LLMs for Explaining Black Box Models
di: Kroeger, Nicholas, et al.
Pubblicazione: (2023)
di: Kroeger, Nicholas, et al.
Pubblicazione: (2023)
Cache-to-Cache: Direct Semantic Communication Between Large Language Models
di: Fu, Tianyu, et al.
Pubblicazione: (2025)
di: Fu, Tianyu, et al.
Pubblicazione: (2025)
LLM Questionnaire Completion for Automatic Psychiatric Assessment
di: Rosenman, Gony, et al.
Pubblicazione: (2024)
di: Rosenman, Gony, et al.
Pubblicazione: (2024)
Automatic identification of diagnosis from hospital discharge letters via weakly-supervised Natural Language Processing
di: Torri, Vittorio, et al.
Pubblicazione: (2024)
di: Torri, Vittorio, et al.
Pubblicazione: (2024)
A Flexible Large Language Models Guardrail Development Methodology Applied to Off-Topic Prompt Detection
di: Chua, Gabriel, et al.
Pubblicazione: (2024)
di: Chua, Gabriel, et al.
Pubblicazione: (2024)
Preserving Empirical Probabilities in BERT for Small-sample Clinical Entity Recognition
di: Rehman, Abdul, et al.
Pubblicazione: (2024)
di: Rehman, Abdul, et al.
Pubblicazione: (2024)
Zero- and Few-Shot Prompting with LLMs: A Comparative Study with Fine-tuned Models for Bangla Sentiment Analysis
di: Hasan, Md. Arid, et al.
Pubblicazione: (2023)
di: Hasan, Md. Arid, et al.
Pubblicazione: (2023)
Position-aware Automatic Circuit Discovery
di: Haklay, Tal, et al.
Pubblicazione: (2025)
di: Haklay, Tal, et al.
Pubblicazione: (2025)
$FastDoc$: Domain-Specific Fast Continual Pre-training Technique using Document-Level Metadata and Taxonomy
di: Nandy, Abhilash, et al.
Pubblicazione: (2023)
di: Nandy, Abhilash, et al.
Pubblicazione: (2023)
FairLangProc: A Python package for fairness in NLP
di: Pérez-Peralta, Arturo, et al.
Pubblicazione: (2025)
di: Pérez-Peralta, Arturo, et al.
Pubblicazione: (2025)
A-VERT: Agnostic Verification with Embedding Ranking Targets
di: Aguirre, Nicolás, et al.
Pubblicazione: (2025)
di: Aguirre, Nicolás, et al.
Pubblicazione: (2025)
Unilogit: Robust Machine Unlearning for LLMs Using Uniform-Target Self-Distillation
di: Vasilev, Stefan, et al.
Pubblicazione: (2025)
di: Vasilev, Stefan, et al.
Pubblicazione: (2025)
Comparison of Modern Multilingual Text Embedding Techniques for Hate Speech Detection Task
di: Vaiciukynas, Evaldas, et al.
Pubblicazione: (2026)
di: Vaiciukynas, Evaldas, et al.
Pubblicazione: (2026)
Parameter-Efficient Transformer Embeddings
di: Ndubuaku, Henry, et al.
Pubblicazione: (2025)
di: Ndubuaku, Henry, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Surfing the modeling of PoS taggers in low-resource scenarios
di: Ferro, Manuel Vilares, et al.
Pubblicazione: (2024) -
Composition of Relational Features with an Application to Explaining Black-Box Predictors
di: Srinivasan, Ashwin, et al.
Pubblicazione: (2022) -
Pivot Language for Low-Resource Machine Translation
di: Talwar, Abhimanyu, et al.
Pubblicazione: (2025) -
Uncertainty Estimation and Quantification for LLMs: A Simple Supervised Approach
di: Liu, Linyu, et al.
Pubblicazione: (2024) -
Aligning Black-box Language Models with Human Judgments
di: Burg, Gerrit J. J. van den, et al.
Pubblicazione: (2025)