A comparative study of transformer-based embeddings for topic coherence
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ding, Alex, Rapaka, Tarun, Rodriguez, Willy, Yang, Jason |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multilingual transformer and BERTopic for short text topic modeling: The case of Serbian
von: Medvecki, Darija, et al.
Veröffentlicht: (2024)
von: Medvecki, Darija, et al.
Veröffentlicht: (2024)
A Human Word Association based model for topic detection in social networks
von: Khadivi, Mehrdad Ranjbar, et al.
Veröffentlicht: (2023)
von: Khadivi, Mehrdad Ranjbar, et al.
Veröffentlicht: (2023)
An Investigation of Linguistic Biases in LLM-Based Recommendations
von: Venkateswaran, Nitin, et al.
Veröffentlicht: (2026)
von: Venkateswaran, Nitin, et al.
Veröffentlicht: (2026)
Med-gte-hybrid: A contextual embedding transformer model for extracting actionable information from clinical texts
von: Kumar, Aditya, et al.
Veröffentlicht: (2025)
von: Kumar, Aditya, et al.
Veröffentlicht: (2025)
Generative AI for automatic topic labelling
von: Kozlowski, Diego, et al.
Veröffentlicht: (2024)
von: Kozlowski, Diego, et al.
Veröffentlicht: (2024)
From RLHF to Direct Alignment: A Theoretical Unification of Preference Learning for Large Language Models
von: Raheja, Tarun, et al.
Veröffentlicht: (2026)
von: Raheja, Tarun, et al.
Veröffentlicht: (2026)
Zero-shot prompt-based classification: topic labeling in times of foundation models in German Tweets
von: Münker, Simon, et al.
Veröffentlicht: (2024)
von: Münker, Simon, et al.
Veröffentlicht: (2024)
Manifold-based Verbalizer Space Re-embedding for Tuning-free Prompt-based Classification
von: Wang, Haochun, et al.
Veröffentlicht: (2023)
von: Wang, Haochun, et al.
Veröffentlicht: (2023)
WordAlchemy: A transformer-based Reverse Dictionary
von: Madaswar, Kanhaiya, et al.
Veröffentlicht: (2022)
von: Madaswar, Kanhaiya, et al.
Veröffentlicht: (2022)
What's in an embedding? Would a rose by any embedding smell as sweet?
von: Venkatasubramanian, Venkat
Veröffentlicht: (2024)
von: Venkatasubramanian, Venkat
Veröffentlicht: (2024)
NEU-ESC: A Comprehensive Vietnamese dataset for Educational Sentiment analysis and topic Classification toward multitask learning
von: Mai, Phan Quoc Hung, et al.
Veröffentlicht: (2025)
von: Mai, Phan Quoc Hung, et al.
Veröffentlicht: (2025)
Generalist embedding models are better at short-context clinical semantic search than specialized embedding models
von: Excoffier, Jean-Baptiste, et al.
Veröffentlicht: (2024)
von: Excoffier, Jean-Baptiste, et al.
Veröffentlicht: (2024)
Evaluating how LLM annotations represent diverse views on contentious topics
von: Brown, Megan A., et al.
Veröffentlicht: (2025)
von: Brown, Megan A., et al.
Veröffentlicht: (2025)
Quokka: Accelerating Program Verification with LLMs via Invariant Synthesis
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Modeling cognitive processes of natural reading with transformer-based Language Models
von: Bianchi, Bruno, et al.
Veröffentlicht: (2025)
von: Bianchi, Bruno, et al.
Veröffentlicht: (2025)
Improving embedding with contrastive fine-tuning on small datasets with expert-augmented scores
von: Lu, Jun, et al.
Veröffentlicht: (2024)
von: Lu, Jun, et al.
Veröffentlicht: (2024)
DIETA: A Decoder-only transformer-based model for Italian-English machine TrAnslation
von: Kasela, Pranav, et al.
Veröffentlicht: (2026)
von: Kasela, Pranav, et al.
Veröffentlicht: (2026)
Recent advancements in LLM Red-Teaming: Techniques, Defenses, and Ethical Considerations
von: Raheja, Tarun, et al.
Veröffentlicht: (2024)
von: Raheja, Tarun, et al.
Veröffentlicht: (2024)
Kallini et al. (2024) do not compare impossible languages with constituency-based ones
von: Hunter, Tim
Veröffentlicht: (2024)
von: Hunter, Tim
Veröffentlicht: (2024)
LLMs as mirrors of societal moral standards: reflection of cultural divergence and agreement across ethical topics
von: Meijer, Mijntje, et al.
Veröffentlicht: (2024)
von: Meijer, Mijntje, et al.
Veröffentlicht: (2024)
Human-Calibrated Automated Testing and Validation of Generative Language Models
von: Sudjianto, Agus, et al.
Veröffentlicht: (2024)
von: Sudjianto, Agus, et al.
Veröffentlicht: (2024)
Prompt-based Learning for Text Readability Assessment
von: Lee, Bruce W., et al.
Veröffentlicht: (2023)
von: Lee, Bruce W., et al.
Veröffentlicht: (2023)
Vocabulary embeddings organize linguistic structure early in language model training
von: Papadimitriou, Isabel, et al.
Veröffentlicht: (2025)
von: Papadimitriou, Isabel, et al.
Veröffentlicht: (2025)
Are language models rational? The case of coherence norms and belief revision
von: Hofweber, Thomas, et al.
Veröffentlicht: (2024)
von: Hofweber, Thomas, et al.
Veröffentlicht: (2024)
Simulated patient systems powered by large language model-based AI agents offer potential for transforming medical education
von: Yu, Huizi, et al.
Veröffentlicht: (2024)
von: Yu, Huizi, et al.
Veröffentlicht: (2024)
Epistemic Constitutionalism Or: how to avoid coherence bias
von: Loi, Michele
Veröffentlicht: (2026)
von: Loi, Michele
Veröffentlicht: (2026)
Team Fusion@ SU@ BC8 SympTEMIST track: transformer-based approach for symptom recognition and linking
von: Grazhdanski, Georgi, et al.
Veröffentlicht: (2026)
von: Grazhdanski, Georgi, et al.
Veröffentlicht: (2026)
Abusive text transformation using LLMs
von: Chandra, Rohitash, et al.
Veröffentlicht: (2025)
von: Chandra, Rohitash, et al.
Veröffentlicht: (2025)
Towards an automatic method for generating topical vocabulary test forms for specific reading passages
von: Flor, Michael, et al.
Veröffentlicht: (2025)
von: Flor, Michael, et al.
Veröffentlicht: (2025)
Navigating through the hidden embedding space: steering LLMs to improve mental health assessment
von: Ravenda, Federico, et al.
Veröffentlicht: (2025)
von: Ravenda, Federico, et al.
Veröffentlicht: (2025)
The Russian-focused embedders' exploration: ruMTEB benchmark and Russian embedding model design
von: Snegirev, Artem, et al.
Veröffentlicht: (2024)
von: Snegirev, Artem, et al.
Veröffentlicht: (2024)
SuperCoder: Assembly Program Superoptimization with Large Language Models
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Controlled LLM-based Reasoning for Clinical Trial Retrieval
von: Jullien, Mael, et al.
Veröffentlicht: (2024)
von: Jullien, Mael, et al.
Veröffentlicht: (2024)
Towards Safety Reasoning in LLMs: AI-agentic Deliberation for Policy-embedded CoT Data Creation
von: Kumarage, Tharindu, et al.
Veröffentlicht: (2025)
von: Kumarage, Tharindu, et al.
Veröffentlicht: (2025)
PRISM: A Transformer-based Language Model of Structured Clinical Event Data
von: Levine, Lionel, et al.
Veröffentlicht: (2025)
von: Levine, Lionel, et al.
Veröffentlicht: (2025)
Classification EM-PCA for clustering and embedding
von: Tighidet, Zineddine, et al.
Veröffentlicht: (2025)
von: Tighidet, Zineddine, et al.
Veröffentlicht: (2025)
Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey
von: Zhu, Jason, et al.
Veröffentlicht: (2025)
von: Zhu, Jason, et al.
Veröffentlicht: (2025)
A word association network methodology for evaluating implicit biases in LLMs compared to humans
von: Abramski, Katherine, et al.
Veröffentlicht: (2025)
von: Abramski, Katherine, et al.
Veröffentlicht: (2025)
Neural network embeddings recover value dimensions from psychometric survey items on par with human data
von: Pellert, Max, et al.
Veröffentlicht: (2025)
von: Pellert, Max, et al.
Veröffentlicht: (2025)
Referential ambiguity and clarification requests: comparing human and LLM behaviour
von: Madge, Chris, et al.
Veröffentlicht: (2025)
von: Madge, Chris, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Multilingual transformer and BERTopic for short text topic modeling: The case of Serbian
von: Medvecki, Darija, et al.
Veröffentlicht: (2024) -
A Human Word Association based model for topic detection in social networks
von: Khadivi, Mehrdad Ranjbar, et al.
Veröffentlicht: (2023) -
An Investigation of Linguistic Biases in LLM-Based Recommendations
von: Venkateswaran, Nitin, et al.
Veröffentlicht: (2026) -
Med-gte-hybrid: A contextual embedding transformer model for extracting actionable information from clinical texts
von: Kumar, Aditya, et al.
Veröffentlicht: (2025) -
Generative AI for automatic topic labelling
von: Kozlowski, Diego, et al.
Veröffentlicht: (2024)