Human-interpretable clustering of short-text using large language models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Miller, Justin K., Alexander, Tristram J. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards interpretable models for language proficiency assessment: Predicting the CEFR level of Estonian learner texts
von: Allkivi, Kais
Veröffentlicht: (2026)
von: Allkivi, Kais
Veröffentlicht: (2026)
State space models can express n-gram languages
von: Nandakumar, Vinoth, et al.
Veröffentlicht: (2023)
von: Nandakumar, Vinoth, et al.
Veröffentlicht: (2023)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023)
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023)
Human-like fleeting memory improves language learning but impairs reading time prediction in transformer language models
von: Thamma, Abishek, et al.
Veröffentlicht: (2025)
von: Thamma, Abishek, et al.
Veröffentlicht: (2025)
Inference acceleration for large language models using "stairs" assisted greedy generation
von: Grigaliūnas, Domas, et al.
Veröffentlicht: (2024)
von: Grigaliūnas, Domas, et al.
Veröffentlicht: (2024)
A Framework for Fine-Tuning LLMs using Heterogeneous Feedback
von: Aponte, Ryan, et al.
Veröffentlicht: (2024)
von: Aponte, Ryan, et al.
Veröffentlicht: (2024)
Context-Aware Clustering using Large Language Models
von: Tipirneni, Sindhu, et al.
Veröffentlicht: (2024)
von: Tipirneni, Sindhu, et al.
Veröffentlicht: (2024)
Enhancing Traffic Accident Classifications: Application of NLP Methods for City Safety
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
Digital Guardians: Can GPT-4, Perspective API, and Moderation API reliably detect hate speech in reader comments of German online newspapers?
von: Weber, Manuel, et al.
Veröffentlicht: (2025)
von: Weber, Manuel, et al.
Veröffentlicht: (2025)
Engineering A Large Language Model From Scratch
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2024)
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2024)
Large Language Model (LLM) Bias Index -- LLMBI
von: Oketunji, Abiodun Finbarrs, et al.
Veröffentlicht: (2023)
von: Oketunji, Abiodun Finbarrs, et al.
Veröffentlicht: (2023)
Mechanistic interpretability of large language models with applications to the financial services industry
von: Golgoon, Ashkan, et al.
Veröffentlicht: (2024)
von: Golgoon, Ashkan, et al.
Veröffentlicht: (2024)
Thread Detection and Response Generation using Transformers with Prompt Optimisation
von: T, Kevin Joshua, et al.
Veröffentlicht: (2024)
von: T, Kevin Joshua, et al.
Veröffentlicht: (2024)
Clinical information extraction for Low-resource languages with Few-shot learning using Pre-trained language models and Prompting
von: Richter-Pechanski, Phillip, et al.
Veröffentlicht: (2024)
von: Richter-Pechanski, Phillip, et al.
Veröffentlicht: (2024)
Plain language adaptations of biomedical text using LLMs: Comparision of evaluation metrics
von: Kocbek, Primoz, et al.
Veröffentlicht: (2025)
von: Kocbek, Primoz, et al.
Veröffentlicht: (2025)
Evaluating an evidence-guided reinforcement learning framework in aligning light-parameter large language models with decision-making cognition in psychiatric clinical reasoning
von: Lin, Xinxin, et al.
Veröffentlicht: (2026)
von: Lin, Xinxin, et al.
Veröffentlicht: (2026)
Unleashing the potential of prompt engineering for large language models
von: Chen, Banghao, et al.
Veröffentlicht: (2023)
von: Chen, Banghao, et al.
Veröffentlicht: (2023)
Determination of language families using deep learning
von: Lerner, Peter B.
Veröffentlicht: (2024)
von: Lerner, Peter B.
Veröffentlicht: (2024)
Are formal and functional linguistic mechanisms dissociated in language models?
von: Hanna, Michael, et al.
Veröffentlicht: (2025)
von: Hanna, Michael, et al.
Veröffentlicht: (2025)
Have Faith in Faithfulness: Going Beyond Circuit Overlap When Finding Model Mechanisms
von: Hanna, Michael, et al.
Veröffentlicht: (2024)
von: Hanna, Michael, et al.
Veröffentlicht: (2024)
Integrating Expert Labels into LLM-based Emission Goal Detection: Example Selection vs Automatic Prompt Design
von: Wrzalik, Marco, et al.
Veröffentlicht: (2024)
von: Wrzalik, Marco, et al.
Veröffentlicht: (2024)
Efficacy of ByT5 in Multilingual Translation of Biblical Texts for Underrepresented Languages
von: Aars, Corinne, et al.
Veröffentlicht: (2024)
von: Aars, Corinne, et al.
Veröffentlicht: (2024)
Scaling Laws for Forgetting When Fine-Tuning Large Language Models
von: Kalajdzievski, Damjan
Veröffentlicht: (2024)
von: Kalajdzievski, Damjan
Veröffentlicht: (2024)
The Role of Language Imbalance in Cross-lingual Generalisation: Insights from Cloned Language Experiments
von: Schäfer, Anton, et al.
Veröffentlicht: (2024)
von: Schäfer, Anton, et al.
Veröffentlicht: (2024)
Ensemble Language Models for Multilingual Sentiment Analysis
von: Hasan, Md Arid
Veröffentlicht: (2024)
von: Hasan, Md Arid
Veröffentlicht: (2024)
Linguistically-Informed Multilingual Instruction Tuning: Is There an Optimal Set of Languages to Tune?
von: Soykan, Gürkan, et al.
Veröffentlicht: (2024)
von: Soykan, Gürkan, et al.
Veröffentlicht: (2024)
MedFuzz: Exploring the Robustness of Large Language Models in Medical Question Answering
von: Ness, Robert Osazuwa, et al.
Veröffentlicht: (2024)
von: Ness, Robert Osazuwa, et al.
Veröffentlicht: (2024)
Sparsing Law: Towards Large Language Models with Greater Activation Sparsity
von: Luo, Yuqi, et al.
Veröffentlicht: (2024)
von: Luo, Yuqi, et al.
Veröffentlicht: (2024)
$S^3$ -- Semantic Signal Separation
von: Kardos, Márton, et al.
Veröffentlicht: (2024)
von: Kardos, Márton, et al.
Veröffentlicht: (2024)
MALoRA: Mixture of Asymmetric Low-Rank Adaptation for Enhanced Multi-Task Learning
von: Wang, Xujia, et al.
Veröffentlicht: (2024)
von: Wang, Xujia, et al.
Veröffentlicht: (2024)
Chain and Causal Attention for Efficient Entity Tracking
von: Fagnou, Erwan, et al.
Veröffentlicht: (2024)
von: Fagnou, Erwan, et al.
Veröffentlicht: (2024)
Efficient Solutions For An Intriguing Failure of LLMs: Long Context Window Does Not Mean LLMs Can Analyze Long Sequences Flawlessly
von: Hosseini, Peyman, et al.
Veröffentlicht: (2024)
von: Hosseini, Peyman, et al.
Veröffentlicht: (2024)
On the Effect of (Near) Duplicate Subwords in Language Modelling
von: Schäfer, Anton, et al.
Veröffentlicht: (2024)
von: Schäfer, Anton, et al.
Veröffentlicht: (2024)
BabyLlama-2: Ensemble-Distilled Models Consistently Outperform Teachers With Limited Data
von: Tastet, Jean-Loup, et al.
Veröffentlicht: (2024)
von: Tastet, Jean-Loup, et al.
Veröffentlicht: (2024)
Token Erasure as a Footprint of Implicit Vocabulary Items in LLMs
von: Feucht, Sheridan, et al.
Veröffentlicht: (2024)
von: Feucht, Sheridan, et al.
Veröffentlicht: (2024)
The Open Source Advantage in Large Language Models (LLMs)
von: Manchanda, Jiya, et al.
Veröffentlicht: (2024)
von: Manchanda, Jiya, et al.
Veröffentlicht: (2024)
MMSciBench: Benchmarking Language Models on Chinese Multimodal Scientific Problems
von: Ye, Xinwu, et al.
Veröffentlicht: (2025)
von: Ye, Xinwu, et al.
Veröffentlicht: (2025)
Influence-driven Curriculum Learning for Pre-training on Limited Data
von: Schoenegger, Loris, et al.
Veröffentlicht: (2025)
von: Schoenegger, Loris, et al.
Veröffentlicht: (2025)
Towards Latent Diffusion Suitable For Text
von: Midavaine, Nesta, et al.
Veröffentlicht: (2026)
von: Midavaine, Nesta, et al.
Veröffentlicht: (2026)
SpectralLoRA: Is Low-Frequency Structure Sufficient for LoRA Adaptation? A Spectral Analysis of Weight Updates
von: Singh, Rajveer
Veröffentlicht: (2026)
von: Singh, Rajveer
Veröffentlicht: (2026)
Ähnliche Einträge
-
Towards interpretable models for language proficiency assessment: Predicting the CEFR level of Estonian learner texts
von: Allkivi, Kais
Veröffentlicht: (2026) -
State space models can express n-gram languages
von: Nandakumar, Vinoth, et al.
Veröffentlicht: (2023) -
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023) -
Human-like fleeting memory improves language learning but impairs reading time prediction in transformer language models
von: Thamma, Abishek, et al.
Veröffentlicht: (2025) -
Inference acceleration for large language models using "stairs" assisted greedy generation
von: Grigaliūnas, Domas, et al.
Veröffentlicht: (2024)