Explaining Text Similarity in Transformer Models
Fuente:
arXiv
Saved in:
| Main Authors: | Vasileiou, Alexandros, Eberle, Oliver |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cat, Rat, Meow: On the Alignment of Language Model and Human Term-Similarity Judgments
by: Linhardt, Lorenz, et al.
Published: (2025)
by: Linhardt, Lorenz, et al.
Published: (2025)
LLMs Explain't: A Post-Mortem on Semantic Interpretability in Transformer Models
by: Abdelhalim, Alhassan, et al.
Published: (2026)
by: Abdelhalim, Alhassan, et al.
Published: (2026)
Explaining Text Classifiers with Counterfactual Representations
by: Lemberger, Pirmin, et al.
Published: (2024)
by: Lemberger, Pirmin, et al.
Published: (2024)
Description-Based Text Similarity
by: Ravfogel, Shauli, et al.
Published: (2023)
by: Ravfogel, Shauli, et al.
Published: (2023)
Learning to Explain: Supervised Token Attribution from Transformer Attention Patterns
by: Mihaila, George
Published: (2026)
by: Mihaila, George
Published: (2026)
Position Information Emerges in Causal Transformers Without Positional Encodings via Similarity of Nearby Embeddings
by: Zuo, Chunsheng, et al.
Published: (2024)
by: Zuo, Chunsheng, et al.
Published: (2024)
Fairness Definitions in Language Models Explained
by: Yin, Zhipeng, et al.
Published: (2024)
by: Yin, Zhipeng, et al.
Published: (2024)
On the Effectiveness of Large Language Models in Automating Categorization of Scientific Texts
by: Shahi, Gautam Kishore, et al.
Published: (2025)
by: Shahi, Gautam Kishore, et al.
Published: (2025)
The Impact of Quantization on the Robustness of Transformer-based Text Classifiers
by: Neshaei, Seyed Parsa, et al.
Published: (2024)
by: Neshaei, Seyed Parsa, et al.
Published: (2024)
Transforming Chatbot Text: A Sequence-to-Sequence Approach
by: Reddy, Natesh, et al.
Published: (2025)
by: Reddy, Natesh, et al.
Published: (2025)
Generative or Discriminative? Revisiting Text Classification in the Era of Transformers
by: Kasa, Siva Rajesh, et al.
Published: (2025)
by: Kasa, Siva Rajesh, et al.
Published: (2025)
Analyzing Similarity Metrics for Data Selection for Language Model Pretraining
by: Sam, Dylan, et al.
Published: (2025)
by: Sam, Dylan, et al.
Published: (2025)
JoPA:Explaining Large Language Model's Generation via Joint Prompt Attribution
by: Chang, Yurui, et al.
Published: (2024)
by: Chang, Yurui, et al.
Published: (2024)
NeuronScope: A Multi-Agent Framework for Explaining Polysemantic Neurons in Language Models
by: Liu, Weiqi, et al.
Published: (2026)
by: Liu, Weiqi, et al.
Published: (2026)
Merging Text Transformer Models from Different Initializations
by: Verma, Neha, et al.
Published: (2024)
by: Verma, Neha, et al.
Published: (2024)
Explaining the role of Intrinsic Dimensionality in Adversarial Training
by: Altinisik, Enes, et al.
Published: (2024)
by: Altinisik, Enes, et al.
Published: (2024)
SENTRA: Selected-Next-Token Transformer for LLM Text Detection
by: Plyler, Mitchell, et al.
Published: (2025)
by: Plyler, Mitchell, et al.
Published: (2025)
Rethinking Layer Relevance in Large Language Models Beyond Cosine Similarity
by: Hinostroza, Cristian, et al.
Published: (2026)
by: Hinostroza, Cristian, et al.
Published: (2026)
Training Language Models to Explain Their Own Computations
by: Li, Belinda Z., et al.
Published: (2025)
by: Li, Belinda Z., et al.
Published: (2025)
Explaining Large Language Models with gSMILE
by: Dehghani, Zeinab, et al.
Published: (2025)
by: Dehghani, Zeinab, et al.
Published: (2025)
Similarity-Distance-Magnitude Activations
by: Schmaltz, Allen
Published: (2025)
by: Schmaltz, Allen
Published: (2025)
Explaining Length Bias in LLM-Based Preference Evaluations
by: Hu, Zhengyu, et al.
Published: (2024)
by: Hu, Zhengyu, et al.
Published: (2024)
TransformLLM: Adapting Large Language Models via LLM-Transformed Reading Comprehension Text
by: Arbel, Iftach, et al.
Published: (2024)
by: Arbel, Iftach, et al.
Published: (2024)
Capturing Polysemanticity with PRISM: A Multi-Concept Feature Description Framework
by: Kopf, Laura, et al.
Published: (2025)
by: Kopf, Laura, et al.
Published: (2025)
Similarity-Distance-Magnitude Universal Verification
by: Schmaltz, Allen
Published: (2025)
by: Schmaltz, Allen
Published: (2025)
Do Sentence Transformers Learn Quasi-Geospatial Concepts from General Text?
by: Ilyankou, Ilya, et al.
Published: (2024)
by: Ilyankou, Ilya, et al.
Published: (2024)
L3Cube-MahaSTS: A Marathi Sentence Similarity Dataset and Models
by: Mirashi, Aishwarya, et al.
Published: (2025)
by: Mirashi, Aishwarya, et al.
Published: (2025)
TextLap: Customizing Language Models for Text-to-Layout Planning
by: Chen, Jian, et al.
Published: (2024)
by: Chen, Jian, et al.
Published: (2024)
Explainable AI: Context-Aware Layer-Wise Integrated Gradients for Explaining Transformer Models
by: Mersha, Melkamu Abay, et al.
Published: (2026)
by: Mersha, Melkamu Abay, et al.
Published: (2026)
An Enhanced Text Compression Approach Using Transformer-based Language Models
by: Rahman, Chowdhury Mofizur, et al.
Published: (2024)
by: Rahman, Chowdhury Mofizur, et al.
Published: (2024)
Barriers to Universal Reasoning With Transformers (And How to Overcome Them)
by: Kraus, Oliver, et al.
Published: (2026)
by: Kraus, Oliver, et al.
Published: (2026)
Efficient Prompt Caching via Embedding Similarity
by: Zhu, Hanlin, et al.
Published: (2024)
by: Zhu, Hanlin, et al.
Published: (2024)
Similar Phrases for Cause of Actions of Civil Cases
by: Huang, Ho-Chien, et al.
Published: (2024)
by: Huang, Ho-Chien, et al.
Published: (2024)
SWSC: Shared Weight for Similar Channel in LLM
by: Zeng, Binrui, et al.
Published: (2025)
by: Zeng, Binrui, et al.
Published: (2025)
Retrieve to Explain: Evidence-driven Predictions for Explainable Drug Target Identification
by: Patel, Ravi, et al.
Published: (2024)
by: Patel, Ravi, et al.
Published: (2024)
Explaining Datasets in Words: Statistical Models with Natural Language Parameters
by: Zhong, Ruiqi, et al.
Published: (2024)
by: Zhong, Ruiqi, et al.
Published: (2024)
Explingo: Explaining AI Predictions using Large Language Models
by: Zytek, Alexandra, et al.
Published: (2024)
by: Zytek, Alexandra, et al.
Published: (2024)
Explaining Large Language Models Decisions Using Shapley Values
by: Mohammadi, Behnam
Published: (2024)
by: Mohammadi, Behnam
Published: (2024)
Frequency Explains the Inverse Correlation of Large Language Models' Size, Training Data Amount, and Surprisal's Fit to Reading Times
by: Oh, Byung-Doh, et al.
Published: (2024)
by: Oh, Byung-Doh, et al.
Published: (2024)
Aligning Attention with Human Rationales for Self-Explaining Hate Speech Detection
by: Eilertsen, Brage, et al.
Published: (2025)
by: Eilertsen, Brage, et al.
Published: (2025)
Similar Items
-
Cat, Rat, Meow: On the Alignment of Language Model and Human Term-Similarity Judgments
by: Linhardt, Lorenz, et al.
Published: (2025) -
LLMs Explain't: A Post-Mortem on Semantic Interpretability in Transformer Models
by: Abdelhalim, Alhassan, et al.
Published: (2026) -
Explaining Text Classifiers with Counterfactual Representations
by: Lemberger, Pirmin, et al.
Published: (2024) -
Description-Based Text Similarity
by: Ravfogel, Shauli, et al.
Published: (2023) -
Learning to Explain: Supervised Token Attribution from Transformer Attention Patterns
by: Mihaila, George
Published: (2026)