Progressive Inference: Explaining Decoder-Only Sequence Classification Models Using Intermediate Predictions
Fuente:
arXiv
Guardado en:
| Autores principales: | Kariyappa, Sanjay, Lécué, Freddy, Mishra, Saumitra, Pond, Christopher, Magazzeni, Daniele, Veloso, Manuela |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Interpreting Language Reward Models via Contrastive Explanations
por: Jiang, Junqi, et al.
Publicado: (2024)
por: Jiang, Junqi, et al.
Publicado: (2024)
Counterfactual Metarules for Local and Global Recourse
por: Bewley, Tom, et al.
Publicado: (2024)
por: Bewley, Tom, et al.
Publicado: (2024)
Sequential Harmful Shift Detection Without Labels
por: Amoukou, Salim I., et al.
Publicado: (2024)
por: Amoukou, Salim I., et al.
Publicado: (2024)
ShapShift: Explaining Model Prediction Shifts with Subgroup Conditional Shapley Values
por: Bewley, Tom, et al.
Publicado: (2026)
por: Bewley, Tom, et al.
Publicado: (2026)
Quantifying Prediction Consistency Under Fine-Tuning Multiplicity in Tabular LLMs
por: Hamman, Faisal, et al.
Publicado: (2024)
por: Hamman, Faisal, et al.
Publicado: (2024)
Entropic Projection Alignment: Estimating, Explaining, and Improving Model Performance Under Distribution Shift
por: Amoukou, Salim I., et al.
Publicado: (2026)
por: Amoukou, Salim I., et al.
Publicado: (2026)
Stronger Enforcement of Instruction Hierarchy via Augmented Intermediate Representations
por: Kariyappa, Sanjay, et al.
Publicado: (2025)
por: Kariyappa, Sanjay, et al.
Publicado: (2025)
Robust Counterfactual Explanations for Neural Networks With Probabilistic Guarantees
por: Hamman, Faisal, et al.
Publicado: (2023)
por: Hamman, Faisal, et al.
Publicado: (2023)
Temporal Fairness in Decision Making Problems
por: Torres, Manuel R., et al.
Publicado: (2024)
por: Torres, Manuel R., et al.
Publicado: (2024)
To Steer or Not to Steer? Mechanistic Error Reduction with Abstention for Language Models
por: Hedström, Anna, et al.
Publicado: (2025)
por: Hedström, Anna, et al.
Publicado: (2025)
Deep Reinforcement Learning and Mean-Variance Strategies for Responsible Portfolio Optimization
por: Acero, Fernando, et al.
Publicado: (2024)
por: Acero, Fernando, et al.
Publicado: (2024)
REFRESH: Responsible and Efficient Feature Reselection Guided by SHAP Values
por: Sharma, Shubham, et al.
Publicado: (2024)
por: Sharma, Shubham, et al.
Publicado: (2024)
SideQuest: Model-Driven KV Cache Management for Long-Horizon Agentic Reasoning
por: Kariyappa, Sanjay, et al.
Publicado: (2026)
por: Kariyappa, Sanjay, et al.
Publicado: (2026)
Are Logistic Models Really Interpretable?
por: Dervovic, Danial, et al.
Publicado: (2024)
por: Dervovic, Danial, et al.
Publicado: (2024)
Beyond Manual Planning: Seating Allocation for Large Organizations
por: Ipsen, Anton, et al.
Publicado: (2026)
por: Ipsen, Anton, et al.
Publicado: (2026)
Capacity Planning and Scheduling for Jobs with Uncertainty in Resource Usage and Duration
por: Patra, Sunandita, et al.
Publicado: (2025)
por: Patra, Sunandita, et al.
Publicado: (2025)
Accelerating Cutting-Plane Algorithms via Reinforcement Learning Surrogates
por: Mana, Kyle, et al.
Publicado: (2023)
por: Mana, Kyle, et al.
Publicado: (2023)
Ensemble Methods for Sequence Classification with Hidden Markov Models
por: Kawawa-Beaudan, Maxime, et al.
Publicado: (2024)
por: Kawawa-Beaudan, Maxime, et al.
Publicado: (2024)
Knowledge-Aware Neuron Interpretation for Scene Classification
por: Guan, Yong, et al.
Publicado: (2024)
por: Guan, Yong, et al.
Publicado: (2024)
The Effect of Data Poisoning on Counterfactual Explanations
por: Artelt, André, et al.
Publicado: (2024)
por: Artelt, André, et al.
Publicado: (2024)
Fair Wasserstein Coresets
por: Xiong, Zikai, et al.
Publicado: (2023)
por: Xiong, Zikai, et al.
Publicado: (2023)
Transforming Decoder-Only Transformers for Accurate WiFi-Telemetry Based Indoor Localization
por: Bhatia, Nayan Sanjay, et al.
Publicado: (2025)
por: Bhatia, Nayan Sanjay, et al.
Publicado: (2025)
Behavioral Sequence Modeling with Ensemble Learning
por: Kawawa-Beaudan, Maxime, et al.
Publicado: (2024)
por: Kawawa-Beaudan, Maxime, et al.
Publicado: (2024)
Intelligent Execution through Plan Analysis
por: Borrajo, Daniel, et al.
Publicado: (2024)
por: Borrajo, Daniel, et al.
Publicado: (2024)
Indoor Localization using Compact, Telemetry-Agnostic, Transfer-Learning Enabled Decoder-Only Transformer
por: Bhatia, Nayan Sanjay, et al.
Publicado: (2025)
por: Bhatia, Nayan Sanjay, et al.
Publicado: (2025)
Explaining Explaining
por: Nirenburg, Sergei, et al.
Publicado: (2024)
por: Nirenburg, Sergei, et al.
Publicado: (2024)
Cache What Lasts: Token Retention for Memory-Bounded KV Cache in LLMs
por: Bui, Ngoc, et al.
Publicado: (2025)
por: Bui, Ngoc, et al.
Publicado: (2025)
Transferring Visual Explainability of Self-Explaining Models to Prediction-Only Models without Additional Training
por: Yoshikawa, Yuya, et al.
Publicado: (2025)
por: Yoshikawa, Yuya, et al.
Publicado: (2025)
Meta-RAG on Large Codebases Using Code Summarization
por: Tawosi, Vali, et al.
Publicado: (2025)
por: Tawosi, Vali, et al.
Publicado: (2025)
Counterfactual Reasoning in Automated Planning
por: Pozanco, Alberto, et al.
Publicado: (2026)
por: Pozanco, Alberto, et al.
Publicado: (2026)
On Computing Plans with Uniform Action Costs
por: Pozanco, Alberto, et al.
Publicado: (2024)
por: Pozanco, Alberto, et al.
Publicado: (2024)
Interpretable LLM-based Table Question Answering
por: Nguyen, Giang, et al.
Publicado: (2024)
por: Nguyen, Giang, et al.
Publicado: (2024)
The Unseen Threat: Residual Knowledge in Machine Unlearning under Perturbed Samples
por: Hsu, Hsiang, et al.
Publicado: (2026)
por: Hsu, Hsiang, et al.
Publicado: (2026)
Teaching LLMs Music Theory with In-Context Learning and Chain-of-Thought Prompting: Pedagogical Strategies for Machines
por: Pond, Liam, et al.
Publicado: (2025)
por: Pond, Liam, et al.
Publicado: (2025)
ReasoningBomb: A Stealthy Denial-of-Service Attack by Inducing Pathologically Long Reasoning in Large Reasoning Models
por: Liu, Xiaogeng, et al.
Publicado: (2026)
por: Liu, Xiaogeng, et al.
Publicado: (2026)
Transformer Model for Alzheimer's Disease Progression Prediction Using Longitudinal Visit Sequences
por: Moghaddami, Mahdi, et al.
Publicado: (2025)
por: Moghaddami, Mahdi, et al.
Publicado: (2025)
KnowHalu: Hallucination Detection via Multi-Form Knowledge Based Factual Checking
por: Zhang, Jiawei, et al.
Publicado: (2024)
por: Zhang, Jiawei, et al.
Publicado: (2024)
Explaining Time Series Classification Predictions via Causal Attributions
por: Alcaraz, Juan Miguel Lopez, et al.
Publicado: (2024)
por: Alcaraz, Juan Miguel Lopez, et al.
Publicado: (2024)
TacoERE: Cluster-aware Compression for Event Relation Extraction
por: Guan, Yong, et al.
Publicado: (2024)
por: Guan, Yong, et al.
Publicado: (2024)
ABLE: Using Adversarial Pairs to Construct Local Models for Explaining Model Predictions
por: Khadka, Krishna, et al.
Publicado: (2025)
por: Khadka, Krishna, et al.
Publicado: (2025)
Ejemplares similares
-
Interpreting Language Reward Models via Contrastive Explanations
por: Jiang, Junqi, et al.
Publicado: (2024) -
Counterfactual Metarules for Local and Global Recourse
por: Bewley, Tom, et al.
Publicado: (2024) -
Sequential Harmful Shift Detection Without Labels
por: Amoukou, Salim I., et al.
Publicado: (2024) -
ShapShift: Explaining Model Prediction Shifts with Subgroup Conditional Shapley Values
por: Bewley, Tom, et al.
Publicado: (2026) -
Quantifying Prediction Consistency Under Fine-Tuning Multiplicity in Tabular LLMs
por: Hamman, Faisal, et al.
Publicado: (2024)