LLM Reading Tea Leaves: Automatically Evaluating Topic Models with Large Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Yang, Xiaohao, Zhao, He, Phung, Dinh, Buntine, Wray, Du, Lan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Neural Topic Modeling with Large Language Models in the Loop
por: Yang, Xiaohao, et al.
Publicado: (2024)
por: Yang, Xiaohao, et al.
Publicado: (2024)
Towards Generalising Neural Topical Representations
por: Yang, Xiaohao, et al.
Publicado: (2023)
por: Yang, Xiaohao, et al.
Publicado: (2023)
VerifiAgent: a Unified Verification Agent in Language Model Reasoning
por: Han, Jiuzhou, et al.
Publicado: (2025)
por: Han, Jiuzhou, et al.
Publicado: (2025)
A Survey on Out-of-Distribution Evaluation of Neural NLP Models
por: Li, Xinzhe, et al.
Publicado: (2023)
por: Li, Xinzhe, et al.
Publicado: (2023)
Towards Uncertainty-Aware Language Agent
por: Han, Jiuzhou, et al.
Publicado: (2024)
por: Han, Jiuzhou, et al.
Publicado: (2024)
Improving Symbolic Translation of Language Models for Logical Reasoning
por: Thatikonda, Ramya Keerthy, et al.
Publicado: (2026)
por: Thatikonda, Ramya Keerthy, et al.
Publicado: (2026)
Logical Reasoning with Outcome Reward Models for Test-Time Scaling
por: Thatikonda, Ramya Keerthy, et al.
Publicado: (2025)
por: Thatikonda, Ramya Keerthy, et al.
Publicado: (2025)
A Neural Topic Method Using a Large-Language-Model-in-the-Loop for Business Research
por: Ludwig, Stephan, et al.
Publicado: (2026)
por: Ludwig, Stephan, et al.
Publicado: (2026)
Reward Engineering for Generating Semi-structured Explanation
por: Han, Jiuzhou, et al.
Publicado: (2023)
por: Han, Jiuzhou, et al.
Publicado: (2023)
Discrete Diffusion Language Model for Efficient Text Summarization
por: Dat, Do Huu, et al.
Publicado: (2024)
por: Dat, Do Huu, et al.
Publicado: (2024)
LLM-XTM: Enhancing Cross-Lingual Topic Models with Large Language Models
por: Xuan, Minh Chu, et al.
Publicado: (2026)
por: Xuan, Minh Chu, et al.
Publicado: (2026)
Uncertainty-Based Methods for Automated Process Reward Data Construction and Output Aggregation in Mathematical Reasoning
por: Han, Jiuzhou, et al.
Publicado: (2025)
por: Han, Jiuzhou, et al.
Publicado: (2025)
Assessing the Sensitivity and Alignment of FOL Closeness Metrics
por: Thatikonda, Ramya Keerthy, et al.
Publicado: (2025)
por: Thatikonda, Ramya Keerthy, et al.
Publicado: (2025)
Extracting Consumer Insight from Text: A Large Language Model Approach to Emotion and Evaluation Measurement
por: Ludwig, Stephan, et al.
Publicado: (2026)
por: Ludwig, Stephan, et al.
Publicado: (2026)
PiVe: Prompting with Iterative Verification Improving Graph-based Generative Capability of LLMs
por: Han, Jiuzhou, et al.
Publicado: (2023)
por: Han, Jiuzhou, et al.
Publicado: (2023)
Automatic Generation and Evaluation of Reading Comprehension Test Items with Large Language Models
por: Säuberli, Andreas, et al.
Publicado: (2024)
por: Säuberli, Andreas, et al.
Publicado: (2024)
Comprehensive Evaluation of Large Language Models for Topic Modeling
por: Doi, Tomoki, et al.
Publicado: (2024)
por: Doi, Tomoki, et al.
Publicado: (2024)
Large Language Models for Automatic Detection of Sensitive Topics
por: Wen, Ruoyu, et al.
Publicado: (2024)
por: Wen, Ruoyu, et al.
Publicado: (2024)
Strategies for Improving NL-to-FOL Translation with LLMs: Data Generation, Incremental Fine-Tuning, and Verification
por: Thatikonda, Ramya Keerthy, et al.
Publicado: (2024)
por: Thatikonda, Ramya Keerthy, et al.
Publicado: (2024)
Multilingual LLM Prompting Strategies for Medical English-Vietnamese Machine Translation
por: Vo, Nhu, et al.
Publicado: (2025)
por: Vo, Nhu, et al.
Publicado: (2025)
MTP: A Dataset for Multi-Modal Turning Points in Casual Conversations
por: Ho, Gia-Bao Dinh, et al.
Publicado: (2024)
por: Ho, Gia-Bao Dinh, et al.
Publicado: (2024)
Addressing Topic Granularity and Hallucination in Large Language Models for Topic Modelling
por: Mu, Yida, et al.
Publicado: (2024)
por: Mu, Yida, et al.
Publicado: (2024)
ALScope: A Unified Toolkit for Deep Active Learning
por: Wu, Chenkai, et al.
Publicado: (2025)
por: Wu, Chenkai, et al.
Publicado: (2025)
ProtoMed-LLM: An Automatic Evaluation Framework for Large Language Models in Medical Protocol Formulation
por: Yi, Seungjun, et al.
Publicado: (2024)
por: Yi, Seungjun, et al.
Publicado: (2024)
Improving Vietnamese-English Medical Machine Translation
por: Vo, Nhu, et al.
Publicado: (2024)
por: Vo, Nhu, et al.
Publicado: (2024)
CVE-LLM : Ontology-Assisted Automatic Vulnerability Evaluation Using Large Language Models
por: Ghosh, Rikhiya, et al.
Publicado: (2025)
por: Ghosh, Rikhiya, et al.
Publicado: (2025)
Bridging the Evaluation Gap: Leveraging Large Language Models for Topic Model Evaluation
por: Tan, Zhiyin, et al.
Publicado: (2025)
por: Tan, Zhiyin, et al.
Publicado: (2025)
Your Large Language Models Are Leaving Fingerprints
por: McGovern, Hope, et al.
Publicado: (2024)
por: McGovern, Hope, et al.
Publicado: (2024)
The performances of the Chinese and U.S. Large Language Models on the Topic of Chinese Culture
por: Liu, Feiyan, et al.
Publicado: (2026)
por: Liu, Feiyan, et al.
Publicado: (2026)
XTRA: Cross-Lingual Topic Modeling with Topic and Representation Alignments
por: Nguyen, Tien Phat, et al.
Publicado: (2025)
por: Nguyen, Tien Phat, et al.
Publicado: (2025)
Evaluation of Automatic Speech Recognition Using Generative Large Language Models
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
por: Bañeras-Roux, Thibault, et al.
Publicado: (2026)
Knockout LLM Assessment: Using Large Language Models for Evaluations through Iterative Pairwise Comparisons
por: Sandan, Isik Baran, et al.
Publicado: (2025)
por: Sandan, Isik Baran, et al.
Publicado: (2025)
Large Language Models Offer an Alternative to the Traditional Approach of Topic Modelling
por: Mu, Yida, et al.
Publicado: (2024)
por: Mu, Yida, et al.
Publicado: (2024)
Toward Purpose-oriented Topic Model Evaluation enabled by Large Language Models
por: Tan, Zhiyin, et al.
Publicado: (2025)
por: Tan, Zhiyin, et al.
Publicado: (2025)
Large Language Models Struggle to Describe the Haystack without Human Help: Human-in-the-loop Evaluation of Topic Models
por: Li, Zongxia, et al.
Publicado: (2025)
por: Li, Zongxia, et al.
Publicado: (2025)
Do Language Models Enjoy Their Own Stories? Prompting Large Language Models for Automatic Story Evaluation
por: Chhun, Cyril, et al.
Publicado: (2024)
por: Chhun, Cyril, et al.
Publicado: (2024)
Automatic Instruction Evolving for Large Language Models
por: Zeng, Weihao, et al.
Publicado: (2024)
por: Zeng, Weihao, et al.
Publicado: (2024)
Multilingual Brain Surgeon: Large Language Models Can be Compressed Leaving No Language Behind
por: Zeng, Hongchuan, et al.
Publicado: (2024)
por: Zeng, Hongchuan, et al.
Publicado: (2024)
Reading Subtext: Evaluating Large Language Models on Short Story Summarization with Writers
por: Subbiah, Melanie, et al.
Publicado: (2024)
por: Subbiah, Melanie, et al.
Publicado: (2024)
RET-LLM: Towards a General Read-Write Memory for Large Language Models
por: Modarressi, Ali, et al.
Publicado: (2023)
por: Modarressi, Ali, et al.
Publicado: (2023)
Ejemplares similares
-
Neural Topic Modeling with Large Language Models in the Loop
por: Yang, Xiaohao, et al.
Publicado: (2024) -
Towards Generalising Neural Topical Representations
por: Yang, Xiaohao, et al.
Publicado: (2023) -
VerifiAgent: a Unified Verification Agent in Language Model Reasoning
por: Han, Jiuzhou, et al.
Publicado: (2025) -
A Survey on Out-of-Distribution Evaluation of Neural NLP Models
por: Li, Xinzhe, et al.
Publicado: (2023) -
Towards Uncertainty-Aware Language Agent
por: Han, Jiuzhou, et al.
Publicado: (2024)