LLMs Enable Bag-of-Texts Representations for Short-Text Clustering
Fuente:
arXiv
Saved in:
| Main Authors: | Lin, I-Fan, Hasibi, Faegheh, Verberne, Suzan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SPILL: Domain-Adaptive Intent Clustering based on Selection and Pooling with Large Language Models
by: Lin, I-Fan, et al.
Published: (2025)
by: Lin, I-Fan, et al.
Published: (2025)
LUMI: Unsupervised Intent Clustering with Multiple Pseudo-Labels
by: Lin, I-Fan, et al.
Published: (2025)
by: Lin, I-Fan, et al.
Published: (2025)
Generate then Refine: Data Augmentation for Zero-shot Intent Detection
by: Lin, I-Fan, et al.
Published: (2024)
by: Lin, I-Fan, et al.
Published: (2024)
The Impact of Quantization on Retrieval-Augmented Generation: An Analysis of Small LLMs
by: Yazan, Mert, et al.
Published: (2024)
by: Yazan, Mert, et al.
Published: (2024)
Undesirable Memorization in Large Language Models: A Survey
by: Satvaty, Ali, et al.
Published: (2024)
by: Satvaty, Ali, et al.
Published: (2024)
A Survey on Recent Advances in Conversational Data Generation
by: Soudani, Heydar, et al.
Published: (2024)
by: Soudani, Heydar, et al.
Published: (2024)
Improving RAG for Personalization with Author Features and Contrastive Examples
by: Yazan, Mert, et al.
Published: (2025)
by: Yazan, Mert, et al.
Published: (2025)
Bag of Tricks: Benchmarking of Jailbreak Attacks on LLMs
by: Xu, Zhao, et al.
Published: (2024)
by: Xu, Zhao, et al.
Published: (2024)
Bagging-Based Model Merging for Robust General Text Embeddings
by: Zhang, Hengran, et al.
Published: (2026)
by: Zhang, Hengran, et al.
Published: (2026)
PromptAug: Fine-grained Conflict Classification Using Data Augmentation
by: Warke, Oliver, et al.
Published: (2025)
by: Warke, Oliver, et al.
Published: (2025)
Text-ADBench: Text Anomaly Detection Benchmark based on LLMs Embedding
by: Xiao, Feng, et al.
Published: (2025)
by: Xiao, Feng, et al.
Published: (2025)
ReportLogic: Evaluating Logical Quality in Deep Research Reports
by: Zhao, Jujia, et al.
Published: (2026)
by: Zhao, Jujia, et al.
Published: (2026)
Skywork-Reward: Bag of Tricks for Reward Modeling in LLMs
by: Liu, Chris Yuhao, et al.
Published: (2024)
by: Liu, Chris Yuhao, et al.
Published: (2024)
PTD-SQL: Partitioning and Targeted Drilling with LLMs in Text-to-SQL
by: Luo, Ruilin, et al.
Published: (2024)
by: Luo, Ruilin, et al.
Published: (2024)
Reasoning-Based Refinement of Unsupervised Text Clusters with LLMs
by: Islam, Tunazzina
Published: (2026)
by: Islam, Tunazzina
Published: (2026)
Contrastive Learning Subspace for Text Clustering
by: Yong, Qian, et al.
Published: (2024)
by: Yong, Qian, et al.
Published: (2024)
German Text Embedding Clustering Benchmark
by: Wehrli, Silvan, et al.
Published: (2024)
by: Wehrli, Silvan, et al.
Published: (2024)
TextQuests: How Good are LLMs at Text-Based Video Games?
by: Phan, Long, et al.
Published: (2025)
by: Phan, Long, et al.
Published: (2025)
Soundwave: Less is More for Speech-Text Alignment in LLMs
by: Zhang, Yuhao, et al.
Published: (2025)
by: Zhang, Yuhao, et al.
Published: (2025)
ToBlend: Token-Level Blending With an Ensemble of LLMs to Attack AI-Generated Text Detection
by: Huang, Fan, et al.
Published: (2024)
by: Huang, Fan, et al.
Published: (2024)
Multi-Step Reasoning with Large Language Models, a Survey
by: Plaat, Aske, et al.
Published: (2024)
by: Plaat, Aske, et al.
Published: (2024)
Text or Pixels? It Takes Half: On the Token Efficiency of Visual Text Inputs in Multimodal LLMs
by: Li, Yanhong, et al.
Published: (2025)
by: Li, Yanhong, et al.
Published: (2025)
Conditioning LLMs to Generate Code-Switched Text
by: Heredia, Maite, et al.
Published: (2025)
by: Heredia, Maite, et al.
Published: (2025)
Disentangling Preference Representation and Text Generation for Efficient Individual Preference Alignment
by: Zhang, Jianfei, et al.
Published: (2024)
by: Zhang, Jianfei, et al.
Published: (2024)
EmbeddingGemma: Powerful and Lightweight Text Representations
by: Vera, Henrique Schechter, et al.
Published: (2025)
by: Vera, Henrique Schechter, et al.
Published: (2025)
Rep2Text: Decoding Full Text from a Single LLM Token Representation
by: Zhao, Haiyan, et al.
Published: (2025)
by: Zhao, Haiyan, et al.
Published: (2025)
FinSQL: Model-Agnostic LLMs-based Text-to-SQL Framework for Financial Analysis
by: Zhang, Chao, et al.
Published: (2024)
by: Zhang, Chao, et al.
Published: (2024)
Emotion Classification in Short English Texts using Deep Learning Techniques
by: Bhat, Siddhanth
Published: (2024)
by: Bhat, Siddhanth
Published: (2024)
Can LLMs Narrate Tabular Data? An Evaluation Framework for Natural Language Representations of Text-to-SQL System Outputs
by: Singh, Jyotika, et al.
Published: (2025)
by: Singh, Jyotika, et al.
Published: (2025)
Short-form Text Rewriting with Phi Silica
by: Tadimeti, Divya, et al.
Published: (2026)
by: Tadimeti, Divya, et al.
Published: (2026)
Estimating Causal Effects of Text Interventions Leveraging LLMs
by: Guo, Siyi, et al.
Published: (2024)
by: Guo, Siyi, et al.
Published: (2024)
Commonsense Knowledge Editing Based on Free-Text in LLMs
by: Huang, Xiusheng, et al.
Published: (2024)
by: Huang, Xiusheng, et al.
Published: (2024)
Difficulty Estimation and Simplification of French Text Using LLMs
by: Jamet, Henri, et al.
Published: (2024)
by: Jamet, Henri, et al.
Published: (2024)
Continuous Adversarial Text Representation Learning for Affective Recognition
by: Son, Seungah, et al.
Published: (2025)
by: Son, Seungah, et al.
Published: (2025)
ROME: Memorization Insights from Text, Logits and Representation
by: Li, Bo, et al.
Published: (2024)
by: Li, Bo, et al.
Published: (2024)
Fine Tuning vs. Retrieval Augmented Generation for Less Popular Knowledge
by: Soudani, Heydar, et al.
Published: (2024)
by: Soudani, Heydar, et al.
Published: (2024)
SparQLe: Speech Queries to Text Translation Through LLMs
by: Djanibekov, Amirbek, et al.
Published: (2025)
by: Djanibekov, Amirbek, et al.
Published: (2025)
Are LLMs Good Text Diacritizers? An Arabic and Yoruba Case Study
by: Toyin, Hawau Olamide, et al.
Published: (2025)
by: Toyin, Hawau Olamide, et al.
Published: (2025)
Automated Detection of Pre-training Text in Black-box LLMs
by: Hu, Ruihan, et al.
Published: (2025)
by: Hu, Ruihan, et al.
Published: (2025)
Comparing the Framing Effect in Humans and LLMs on Naturally Occurring Texts
by: Lior, Gili, et al.
Published: (2025)
by: Lior, Gili, et al.
Published: (2025)
Similar Items
-
SPILL: Domain-Adaptive Intent Clustering based on Selection and Pooling with Large Language Models
by: Lin, I-Fan, et al.
Published: (2025) -
LUMI: Unsupervised Intent Clustering with Multiple Pseudo-Labels
by: Lin, I-Fan, et al.
Published: (2025) -
Generate then Refine: Data Augmentation for Zero-shot Intent Detection
by: Lin, I-Fan, et al.
Published: (2024) -
The Impact of Quantization on Retrieval-Augmented Generation: An Analysis of Small LLMs
by: Yazan, Mert, et al.
Published: (2024) -
Undesirable Memorization in Large Language Models: A Survey
by: Satvaty, Ali, et al.
Published: (2024)