Summaries as Centroids for Interpretable and Scalable Text Clustering
Fuente:
arXiv
Saved in:
| Main Author: | Diaz-Rodriguez, Jairo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Consistent Kernel Change-Point Detection under m-Dependence for Text Segmentation
by: Diaz-Rodriguez, Jairo, et al.
Published: (2025)
by: Diaz-Rodriguez, Jairo, et al.
Published: (2025)
Unsupervised Text Segmentation via Kernel Change-Point Detection on Sentence Embeddings
by: Jia, Mumin, et al.
Published: (2026)
by: Jia, Mumin, et al.
Published: (2026)
Scalable Parameter-Light Spectral Method for Clustering Short Text Embeddings with a Cohesion-Based Evaluation Metric
by: Neveditsin, Nikita, et al.
Published: (2025)
by: Neveditsin, Nikita, et al.
Published: (2025)
Entropy Centroids as Intrinsic Rewards for Test-Time Scaling
by: Zhao, Wenshuo, et al.
Published: (2026)
by: Zhao, Wenshuo, et al.
Published: (2026)
Dynamics of Spontaneous Topic Changes in Next Token Prediction with Self-Attention
by: Jia, Mumin, et al.
Published: (2025)
by: Jia, Mumin, et al.
Published: (2025)
Incremental Graph Construction Enables Robust Spectral Clustering of Texts
by: Pranjić, Marko, et al.
Published: (2026)
by: Pranjić, Marko, et al.
Published: (2026)
Vector Quantized Latent Concepts: A Scalable Alternative to Clustering-Based Concept Discovery
by: Yu, Xuemin, et al.
Published: (2026)
by: Yu, Xuemin, et al.
Published: (2026)
Adaptive Two Sided Laplace Transforms: A Learnable, Interpretable, and Scalable Replacement for Self-Attention
by: Kiruluta, Andrew
Published: (2025)
by: Kiruluta, Andrew
Published: (2025)
Fast Multipole Attention: A Scalable Multilevel Attention Mechanism for Text and Images
by: Kang, Yanming, et al.
Published: (2023)
by: Kang, Yanming, et al.
Published: (2023)
Beyond Tokens in Language Models: Interpreting Activations through Text Genre Chunks
by: Benito-Rodriguez, Éloïse, et al.
Published: (2025)
by: Benito-Rodriguez, Éloïse, et al.
Published: (2025)
Evaluating Text Summaries Generated by Large Language Models Using OpenAI's GPT
by: Shakil, Hassan, et al.
Published: (2024)
by: Shakil, Hassan, et al.
Published: (2024)
The Curved Spacetime of Transformer Architectures
by: Di Sipio, Riccardo, et al.
Published: (2025)
by: Di Sipio, Riccardo, et al.
Published: (2025)
Clustering Algorithms and RAG Enhancing Semi-Supervised Text Classification with Large LLMs
by: Zhong, Shan, et al.
Published: (2024)
by: Zhong, Shan, et al.
Published: (2024)
Improving Clustering on Occupational Text Data through Dimensionality Reduction
by: García, Iago Xabier Vázquez, et al.
Published: (2025)
by: García, Iago Xabier Vázquez, et al.
Published: (2025)
An Improved Deep Learning Model for Word Embeddings Based Clustering for Large Text Datasets
by: Sutrakar, Vijay Kumar, et al.
Published: (2025)
by: Sutrakar, Vijay Kumar, et al.
Published: (2025)
BiSparse-AAS: Bilinear Sparse Attention and Adaptive Spans Framework for Scalable and Efficient Text Summarization
by: Hagos, Desta Haileselassie, et al.
Published: (2025)
by: Hagos, Desta Haileselassie, et al.
Published: (2025)
Predictive Concept Decoders: Training Scalable End-to-End Interpretability Assistants
by: Huang, Vincent, et al.
Published: (2025)
by: Huang, Vincent, et al.
Published: (2025)
Graph Contrastive Learning via Cluster-refined Negative Sampling for Semi-supervised Text Classification
by: Ai, Wei, et al.
Published: (2024)
by: Ai, Wei, et al.
Published: (2024)
SIFiD: Reassess Summary Factual Inconsistency Detection with LLM
by: Yang, Jiuding, et al.
Published: (2024)
by: Yang, Jiuding, et al.
Published: (2024)
Interpretable Discriminative Text Representations via Agreement and Label Disentanglement
by: Wang, Tong, et al.
Published: (2026)
by: Wang, Tong, et al.
Published: (2026)
Comparing Feature Importance and Rule Extraction for Interpretability on Text Data
by: Lopardo, Gianluigi, et al.
Published: (2022)
by: Lopardo, Gianluigi, et al.
Published: (2022)
A General Framework for Producing Interpretable Semantic Text Embeddings
by: Sun, Yiqun, et al.
Published: (2024)
by: Sun, Yiqun, et al.
Published: (2024)
Interpretable Recognition of Cognitive Distortions in Natural Language Texts
by: Kolonin, Anton, et al.
Published: (2025)
by: Kolonin, Anton, et al.
Published: (2025)
Identifying Factual Inconsistencies in Summaries: Grounding LLM Inference via Task Taxonomy
by: Xu, Liyan, et al.
Published: (2024)
by: Xu, Liyan, et al.
Published: (2024)
Prediction of COPD Using Machine Learning, Clinical Summary Notes, and Vital Signs
by: Orangi-Fard, Negar
Published: (2024)
by: Orangi-Fard, Negar
Published: (2024)
MedDec: A Dataset for Extracting Medical Decisions from Discharge Summaries
by: Elgaar, Mohamed, et al.
Published: (2024)
by: Elgaar, Mohamed, et al.
Published: (2024)
LLM-Guided Semantic Bootstrapping for Interpretable Text Classification with Tsetlin Machines
by: Gao, Jiechao, et al.
Published: (2026)
by: Gao, Jiechao, et al.
Published: (2026)
Interpretable Predictability-Based AI Text Detection: A Replication Study
by: Skurla, Adam, et al.
Published: (2026)
by: Skurla, Adam, et al.
Published: (2026)
Proportional Fairness in Non-Centroid Clustering
by: Caragiannis, Ioannis, et al.
Published: (2024)
by: Caragiannis, Ioannis, et al.
Published: (2024)
Reasoning-Based Refinement of Unsupervised Text Clusters with LLMs
by: Islam, Tunazzina
Published: (2026)
by: Islam, Tunazzina
Published: (2026)
Semantic Token Reweighting for Interpretable and Controllable Text Embeddings in CLIP
by: Kim, Eunji, et al.
Published: (2024)
by: Kim, Eunji, et al.
Published: (2024)
Towards Scalable Meta-Learning of near-optimal Interpretable Models via Synthetic Model Generations
by: Myint, Kyaw Hpone, et al.
Published: (2025)
by: Myint, Kyaw Hpone, et al.
Published: (2025)
LIDS: LLM Summary Inference Under the Layered Lens
by: Park, Dylan, et al.
Published: (2026)
by: Park, Dylan, et al.
Published: (2026)
Fair Representation in Parliamentary Summaries: Measuring and Mitigating Inclusion Bias
by: Cunningham, Eoghan, et al.
Published: (2025)
by: Cunningham, Eoghan, et al.
Published: (2025)
Contrast-CAT: Contrasting Activations for Enhanced Interpretability in Transformer-based Text Classifiers
by: Han, Sungmin, et al.
Published: (2025)
by: Han, Sungmin, et al.
Published: (2025)
IPAD: Inverse Prompt for AI Detection - A Robust and Interpretable LLM-Generated Text Detector
by: Chen, Zheng, et al.
Published: (2025)
by: Chen, Zheng, et al.
Published: (2025)
Scalable Ensembling For Mitigating Reward Overoptimisation
by: Ahmed, Ahmed M., et al.
Published: (2024)
by: Ahmed, Ahmed M., et al.
Published: (2024)
Isotropy, Clusters, and Classifiers
by: Mickus, Timothee, et al.
Published: (2024)
by: Mickus, Timothee, et al.
Published: (2024)
TESS: Text-to-Text Self-Conditioned Simplex Diffusion
by: Mahabadi, Rabeeh Karimi, et al.
Published: (2023)
by: Mahabadi, Rabeeh Karimi, et al.
Published: (2023)
Morality is Contextual: Learning Interpretable Moral Contexts from Human Data with Probabilistic Clustering and Large Language Models
by: Morlat, Geoffroy, et al.
Published: (2025)
by: Morlat, Geoffroy, et al.
Published: (2025)
Similar Items
-
Consistent Kernel Change-Point Detection under m-Dependence for Text Segmentation
by: Diaz-Rodriguez, Jairo, et al.
Published: (2025) -
Unsupervised Text Segmentation via Kernel Change-Point Detection on Sentence Embeddings
by: Jia, Mumin, et al.
Published: (2026) -
Scalable Parameter-Light Spectral Method for Clustering Short Text Embeddings with a Cohesion-Based Evaluation Metric
by: Neveditsin, Nikita, et al.
Published: (2025) -
Entropy Centroids as Intrinsic Rewards for Test-Time Scaling
by: Zhao, Wenshuo, et al.
Published: (2026) -
Dynamics of Spontaneous Topic Changes in Next Token Prediction with Self-Attention
by: Jia, Mumin, et al.
Published: (2025)