A thorough benchmark of automatic text classification: From traditional approaches to large language models
Fuente:
arXiv
Guardado en:
| Autores principales: | Cunha, Washington, Rocha, Leonardo, Gonçalves, Marcos André |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CTDGSI: A comprehensive exploitation of instance selection methods for automatic text classification. VII Concurso de Teses, Dissertações e Trabalhos de Graduação em SI -- XXI Simpósio Brasileiro de Sistemas de Informação
por: Cunha, Washington, et al.
Publicado: (2025)
por: Cunha, Washington, et al.
Publicado: (2025)
Is 'Hope' a person or an idea? A pilot benchmark for NER: comparing traditional NLP tools and large language models on ambiguous entities
por: Latifi, Payam
Publicado: (2025)
por: Latifi, Payam
Publicado: (2025)
Fine-tuning of lightweight large language models for sentiment classification on heterogeneous financial textual data
por: Amorin, Alvaro Paredes, et al.
Publicado: (2025)
por: Amorin, Alvaro Paredes, et al.
Publicado: (2025)
Large language models struggle with ethnographic text annotation
por: Goodall, Leonardo S., et al.
Publicado: (2026)
por: Goodall, Leonardo S., et al.
Publicado: (2026)
ARC-Encoder: learning compressed text representations for large language models
por: Pilchen, Hippolyte, et al.
Publicado: (2025)
por: Pilchen, Hippolyte, et al.
Publicado: (2025)
Creativity Benchmark: A benchmark for marketing creativity for large language models
por: Bhat, Ninad, et al.
Publicado: (2025)
por: Bhat, Ninad, et al.
Publicado: (2025)
A dataset and benchmark for hospital course summarization with adapted large language models
por: Aali, Asad, et al.
Publicado: (2024)
por: Aali, Asad, et al.
Publicado: (2024)
An energy-based comparative analysis of common approaches to text classification in the Legal domain
por: Gultekin, Sinan, et al.
Publicado: (2023)
por: Gultekin, Sinan, et al.
Publicado: (2023)
Dissociating language and thought in large language models
por: Mahowald, Kyle, et al.
Publicado: (2023)
por: Mahowald, Kyle, et al.
Publicado: (2023)
ks-lit-3m: A 3.1 million word kashmiri text dataset for large language model pretraining
por: Malik, Haq Nawaz
Publicado: (2026)
por: Malik, Haq Nawaz
Publicado: (2026)
LongTail-Swap: benchmarking language models' abilities on rare words
por: Algayres, Robin, et al.
Publicado: (2025)
por: Algayres, Robin, et al.
Publicado: (2025)
Reshaping MOFs text mining with a dynamic multi-agents framework of large language model
por: Lin, Zuhong, et al.
Publicado: (2025)
por: Lin, Zuhong, et al.
Publicado: (2025)
TelcoLM: collecting data, adapting, and benchmarking language models for the telecommunication domain
por: Barboule, Camille, et al.
Publicado: (2024)
por: Barboule, Camille, et al.
Publicado: (2024)
On the attribution of confidence to large language models
por: Keeling, Geoff, et al.
Publicado: (2024)
por: Keeling, Geoff, et al.
Publicado: (2024)
Specialized text classification: an approach to classifying Open Banking transactions
por: TA, Duc Tuyen, et al.
Publicado: (2025)
por: TA, Duc Tuyen, et al.
Publicado: (2025)
Does language matter for spoken word classification? A multilingual generative meta-learning approach
por: Ziki, Batsirayi Mupamhi, et al.
Publicado: (2026)
por: Ziki, Batsirayi Mupamhi, et al.
Publicado: (2026)
A review on the use of large language models as virtual tutors
por: García-Méndez, Silvia, et al.
Publicado: (2024)
por: García-Méndez, Silvia, et al.
Publicado: (2024)
A survey of textual cyber abuse detection using cutting-edge language models and large language models
por: Diaz-Garcia, Jose A., et al.
Publicado: (2025)
por: Diaz-Garcia, Jose A., et al.
Publicado: (2025)
Zero-shot data citation function classification using transformer-based large language models (LLMs)
por: Byers, Neil, et al.
Publicado: (2025)
por: Byers, Neil, et al.
Publicado: (2025)
Quantifying non deterministic drift in large language models
por: Nicholson, Claire
Publicado: (2026)
por: Nicholson, Claire
Publicado: (2026)
Can large language models build causal graphs?
por: Long, Stephanie, et al.
Publicado: (2023)
por: Long, Stephanie, et al.
Publicado: (2023)
Multi-round jailbreak attack on large language models
por: Zhou, Yihua, et al.
Publicado: (2024)
por: Zhou, Yihua, et al.
Publicado: (2024)
Response: Emergent analogical reasoning in large language models
por: Hodel, Damian, et al.
Publicado: (2023)
por: Hodel, Damian, et al.
Publicado: (2023)
The 20 questions game to distinguish large language models
por: Richardeau, Gurvan, et al.
Publicado: (2024)
por: Richardeau, Gurvan, et al.
Publicado: (2024)
Representation in large language models
por: Yetman, Cameron
Publicado: (2025)
por: Yetman, Cameron
Publicado: (2025)
Failure of contextual invariance in large language models
por: Kumar, Sagar, et al.
Publicado: (2026)
por: Kumar, Sagar, et al.
Publicado: (2026)
A blind spot for large language models: Supradiegetic linguistic information
por: Zimmerman, Julia Witte, et al.
Publicado: (2023)
por: Zimmerman, Julia Witte, et al.
Publicado: (2023)
Hyacinth6B: A large language model for Traditional Chinese
por: Song, Chih-Wei, et al.
Publicado: (2024)
por: Song, Chih-Wei, et al.
Publicado: (2024)
Streamlining evidence based clinical recommendations with large language models
por: Li, Dubai, et al.
Publicado: (2025)
por: Li, Dubai, et al.
Publicado: (2025)
Re-evaluating Theory of Mind evaluation in large language models
por: Hu, Jennifer, et al.
Publicado: (2025)
por: Hu, Jennifer, et al.
Publicado: (2025)
Strong and weak alignment of large language models with human values
por: Khamassi, Mehdi, et al.
Publicado: (2024)
por: Khamassi, Mehdi, et al.
Publicado: (2024)
Correcting misinformation on social media with a large language model
por: Zhou, Xinyi, et al.
Publicado: (2024)
por: Zhou, Xinyi, et al.
Publicado: (2024)
Evaluating large language models in medical applications: a survey
por: Chen, Xiaolan, et al.
Publicado: (2024)
por: Chen, Xiaolan, et al.
Publicado: (2024)
Disentangling generalization and memorization in large language models using chess
por: Pleiss, Leonard S., et al.
Publicado: (2026)
por: Pleiss, Leonard S., et al.
Publicado: (2026)
MathDivide: Improved mathematical reasoning by large language models
por: Srivastava, Saksham Sahai, et al.
Publicado: (2024)
por: Srivastava, Saksham Sahai, et al.
Publicado: (2024)
Effects of term weighting approach with and without stop words removing on Arabic text classification
por: Alhenawi, Esra'a, et al.
Publicado: (2024)
por: Alhenawi, Esra'a, et al.
Publicado: (2024)
AI-AI Bias: large language models favor communications generated by large language models
por: Laurito, Walter, et al.
Publicado: (2024)
por: Laurito, Walter, et al.
Publicado: (2024)
Differentially-private text generation degrades output language quality
por: Çano, Erion, et al.
Publicado: (2025)
por: Çano, Erion, et al.
Publicado: (2025)
Optimizing watermarks for large language models
por: Wouters, Bram
Publicado: (2023)
por: Wouters, Bram
Publicado: (2023)
Alignment faking in large language models
por: Greenblatt, Ryan, et al.
Publicado: (2024)
por: Greenblatt, Ryan, et al.
Publicado: (2024)
Ejemplares similares
-
CTDGSI: A comprehensive exploitation of instance selection methods for automatic text classification. VII Concurso de Teses, Dissertações e Trabalhos de Graduação em SI -- XXI Simpósio Brasileiro de Sistemas de Informação
por: Cunha, Washington, et al.
Publicado: (2025) -
Is 'Hope' a person or an idea? A pilot benchmark for NER: comparing traditional NLP tools and large language models on ambiguous entities
por: Latifi, Payam
Publicado: (2025) -
Fine-tuning of lightweight large language models for sentiment classification on heterogeneous financial textual data
por: Amorin, Alvaro Paredes, et al.
Publicado: (2025) -
Large language models struggle with ethnographic text annotation
por: Goodall, Leonardo S., et al.
Publicado: (2026) -
ARC-Encoder: learning compressed text representations for large language models
por: Pilchen, Hippolyte, et al.
Publicado: (2025)