Select or Project? Evaluating Lower-dimensional Vectors for LLM Training Data Explanations
Fuente:
arXiv
Saved in:
| Main Authors: | Hinterleitner, Lukas, Schoenegger, Loris, Roth, Benjamin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Compact Example-Based Explanations for Language Models
by: Schoenegger, Loris, et al.
Published: (2026)
by: Schoenegger, Loris, et al.
Published: (2026)
Influence-driven Curriculum Learning for Pre-training on Limited Data
by: Schoenegger, Loris, et al.
Published: (2025)
by: Schoenegger, Loris, et al.
Published: (2025)
Rubrik's Cube: Testing a New Rubric for Evaluating Explanations on the CUBE dataset
by: Galvan-Sosa, Diana, et al.
Published: (2025)
by: Galvan-Sosa, Diana, et al.
Published: (2025)
LLM-GLOBE: A Benchmark Evaluating the Cultural Values Embedded in LLM Output
by: Karinshak, Elise, et al.
Published: (2024)
by: Karinshak, Elise, et al.
Published: (2024)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
by: Ashuach, Tomer, et al.
Published: (2025)
by: Ashuach, Tomer, et al.
Published: (2025)
SAGE: Hierarchical LLM-Based Literary Evaluation through Ontology-Grounded Interpretive Dimensions
by: Wang, Tianyu, et al.
Published: (2026)
by: Wang, Tianyu, et al.
Published: (2026)
Improving the OOD Performance of Closed-Source LLMs on NLI Through Strategic Data Selection
by: Stacey, Joe, et al.
Published: (2025)
by: Stacey, Joe, et al.
Published: (2025)
Verbosity Tradeoffs and the Impact of Scale on the Faithfulness of LLM Self-Explanations
by: Siegel, Noah Y., et al.
Published: (2025)
by: Siegel, Noah Y., et al.
Published: (2025)
Hidden Failures in Robustness: Why Supervised Uncertainty Quantification Needs Better Evaluation
by: Stacey, Joe, et al.
Published: (2026)
by: Stacey, Joe, et al.
Published: (2026)
Evaluating Input Feature Explanations through a Unified Diagnostic Evaluation Framework
by: Sun, Jingyi, et al.
Published: (2024)
by: Sun, Jingyi, et al.
Published: (2024)
Learning and communication pressures in neural networks: Lessons from emergent communication
by: Galke, Lukas, et al.
Published: (2024)
by: Galke, Lukas, et al.
Published: (2024)
Enhancing Paraphrase Type Generation: The Impact of DPO and RLHF Evaluated with Human-Ranked Data
by: Lübbers, Christopher Lee
Published: (2025)
by: Lübbers, Christopher Lee
Published: (2025)
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
by: Hashemi, Helia, et al.
Published: (2024)
by: Hashemi, Helia, et al.
Published: (2024)
Calibrated Confidence Estimation for Tabular Question Answering
by: Voss, Lukas
Published: (2026)
by: Voss, Lukas
Published: (2026)
The Sufficiency-Conciseness Trade-off in LLM Self-Explanation from an Information Bottleneck Perspective
by: Zahedzadeh, Ali, et al.
Published: (2026)
by: Zahedzadeh, Ali, et al.
Published: (2026)
HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns
by: Wang, Xintao, et al.
Published: (2026)
by: Wang, Xintao, et al.
Published: (2026)
What makes a language easy to deep-learn? Deep neural networks and humans similarly benefit from compositional structure
by: Galke, Lukas, et al.
Published: (2023)
by: Galke, Lukas, et al.
Published: (2023)
Tokenization and Morphology in Multilingual Language Models: A Comparative Analysis of mT5 and ByT5
by: Dang, Thao Anh, et al.
Published: (2024)
by: Dang, Thao Anh, et al.
Published: (2024)
Evaluating the efficacy of LLM Safety Solutions : The Palit Benchmark Dataset
by: Palit, Sayon, et al.
Published: (2025)
by: Palit, Sayon, et al.
Published: (2025)
Revisiting Word Embeddings in the LLM Era
by: Mahajan, Yash, et al.
Published: (2024)
by: Mahajan, Yash, et al.
Published: (2024)
Interactive Text-to-SQL Generation via Editable Step-by-Step Explanations
by: Tian, Yuan, et al.
Published: (2023)
by: Tian, Yuan, et al.
Published: (2023)
Distinguishing Ignorance from Error in LLM Hallucinations
by: Simhi, Adi, et al.
Published: (2024)
by: Simhi, Adi, et al.
Published: (2024)
Decoding-Free Sampling Strategies for LLM Marginalization
by: Pohl, David, et al.
Published: (2025)
by: Pohl, David, et al.
Published: (2025)
Efficient Reasoning via Thought-Training and Thought-Free Inference
by: Wu, Canhui, et al.
Published: (2025)
by: Wu, Canhui, et al.
Published: (2025)
GroUSE: A Benchmark to Evaluate Evaluators in Grounded Question Answering
by: Muller, Sacha, et al.
Published: (2024)
by: Muller, Sacha, et al.
Published: (2024)
Masked by Consensus: Disentangling Privileged Knowledge in LLM Correctness
by: Ashuach, Tomer, et al.
Published: (2026)
by: Ashuach, Tomer, et al.
Published: (2026)
LUCID: LLM-Generated Utterances for Complex and Interesting Dialogues
by: Stacey, Joe, et al.
Published: (2024)
by: Stacey, Joe, et al.
Published: (2024)
Automatic Task Detection and Heterogeneous LLM Speculative Decoding
by: Ge, Danying, et al.
Published: (2025)
by: Ge, Danying, et al.
Published: (2025)
Overcoming Black-box Attack Inefficiency with Hybrid and Dynamic Select Algorithms
by: Belde, Abhinay Shankar, et al.
Published: (2025)
by: Belde, Abhinay Shankar, et al.
Published: (2025)
Policy-driven Knowledge Selection and Response Generation for Document-grounded Dialogue
by: Ma, Longxuan, et al.
Published: (2024)
by: Ma, Longxuan, et al.
Published: (2024)
Intent Classification for Bank Chatbots through LLM Fine-Tuning
by: Lajčinová, Bibiána, et al.
Published: (2024)
by: Lajčinová, Bibiána, et al.
Published: (2024)
Can LLM Graph Reasoning Generalize beyond Pattern Memorization?
by: Zhang, Yizhuo, et al.
Published: (2024)
by: Zhang, Yizhuo, et al.
Published: (2024)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
by: Saji, Alan, et al.
Published: (2025)
by: Saji, Alan, et al.
Published: (2025)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
by: Peters, Sydney, et al.
Published: (2025)
by: Peters, Sydney, et al.
Published: (2025)
LLMBridge: An LLM Pipeline for End-to-end Referential Bridging Resolution in English
by: Levine, Lauren, et al.
Published: (2026)
by: Levine, Lauren, et al.
Published: (2026)
Whose Facts Win? LLM Source Preferences under Knowledge Conflicts
by: Schuster, Jakob, et al.
Published: (2026)
by: Schuster, Jakob, et al.
Published: (2026)
SEPTQ: A Simple and Effective Post-Training Quantization Paradigm for Large Language Models
by: Liu, Han, et al.
Published: (2026)
by: Liu, Han, et al.
Published: (2026)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
by: Oketunji, Abiodun Finbarrs
Published: (2023)
by: Oketunji, Abiodun Finbarrs
Published: (2023)
Augmenting Dialog with Think-Aloud Utterances for Modeling Individual Personality Traits by LLM
by: Ishikura, Seiya, et al.
Published: (2025)
by: Ishikura, Seiya, et al.
Published: (2025)
Extracting Structured Insights from Financial News: An Augmented LLM Driven Approach
by: Dolphin, Rian, et al.
Published: (2024)
by: Dolphin, Rian, et al.
Published: (2024)
Similar Items
-
Compact Example-Based Explanations for Language Models
by: Schoenegger, Loris, et al.
Published: (2026) -
Influence-driven Curriculum Learning for Pre-training on Limited Data
by: Schoenegger, Loris, et al.
Published: (2025) -
Rubrik's Cube: Testing a New Rubric for Evaluating Explanations on the CUBE dataset
by: Galvan-Sosa, Diana, et al.
Published: (2025) -
LLM-GLOBE: A Benchmark Evaluating the Cultural Values Embedded in LLM Output
by: Karinshak, Elise, et al.
Published: (2024) -
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
by: Ashuach, Tomer, et al.
Published: (2025)