Sonnet or Not, Bot? Poetry Evaluation for Large Models and Datasets
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Walsh, Melanie, Preus, Anna, Antoniak, Maria |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
so much depends / upon / a whitespace: Why Whitespace Matters for Poets and LLMs
von: Bhyravajjula, Sriharsh, et al.
Veröffentlicht: (2025)
von: Bhyravajjula, Sriharsh, et al.
Veröffentlicht: (2025)
Does ChatGPT Have a Poetic Style?
von: Walsh, Melanie, et al.
Veröffentlicht: (2024)
von: Walsh, Melanie, et al.
Veröffentlicht: (2024)
Trust No Bot: Discovering Personal Disclosures in Human-LLM Conversations in the Wild
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2024)
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2024)
The Afterlives of Shakespeare and Company in Online Social Readership
von: Antoniak, Maria, et al.
Veröffentlicht: (2024)
von: Antoniak, Maria, et al.
Veröffentlicht: (2024)
Capabilities and Evaluation Biases of Large Language Models in Classical Chinese Poetry Generation: A Case Study on Tang Poetry
von: Ma, Bolei, et al.
Veröffentlicht: (2025)
von: Ma, Bolei, et al.
Veröffentlicht: (2025)
LLM-Generated or Human-Written? Comparing Review and Non-Review Papers on ArXiv
von: Elazar, Yanai, et al.
Veröffentlicht: (2026)
von: Elazar, Yanai, et al.
Veröffentlicht: (2026)
From Ghazals to Sonnets: Decoding the Polysemous Expressions of Love Across Languages
von: Ali, Syed Mohammad Sualeh
Veröffentlicht: (2025)
von: Ali, Syed Mohammad Sualeh
Veröffentlicht: (2025)
Research Borderlands: Analysing Writing Across Research Cultures
von: Bhatt, Shaily, et al.
Veröffentlicht: (2025)
von: Bhatt, Shaily, et al.
Veröffentlicht: (2025)
What Does the Bot Say? Opportunities and Risks of Large Language Models in Social Media Bot Detection
von: Feng, Shangbin, et al.
Veröffentlicht: (2024)
von: Feng, Shangbin, et al.
Veröffentlicht: (2024)
Large Language Models for Classical Chinese Poetry Translation: Benchmarking, Evaluating, and Improving
von: Chen, Andong, et al.
Veröffentlicht: (2024)
von: Chen, Andong, et al.
Veröffentlicht: (2024)
Evaluating Diversity in Automatic Poetry Generation
von: Chen, Yanran, et al.
Veröffentlicht: (2024)
von: Chen, Yanran, et al.
Veröffentlicht: (2024)
WisdomBot: Tuning Large Language Models with Artificial Intelligence Knowledge
von: Chen, Jingyuan, et al.
Veröffentlicht: (2025)
von: Chen, Jingyuan, et al.
Veröffentlicht: (2025)
BotEval: Facilitating Interactive Human Evaluation
von: Cho, Hyundong, et al.
Veröffentlicht: (2024)
von: Cho, Hyundong, et al.
Veröffentlicht: (2024)
Information-Guided Identification of Training Data Imprint in (Proprietary) Large Language Models
von: Ravichander, Abhilasha, et al.
Veröffentlicht: (2025)
von: Ravichander, Abhilasha, et al.
Veröffentlicht: (2025)
Can Large Language Models Outperform Non-Experts in Poetry Evaluation? A Comparative Study Using the Consensual Assessment Technique
von: Sawicki, Piotr, et al.
Veröffentlicht: (2025)
von: Sawicki, Piotr, et al.
Veröffentlicht: (2025)
KPoEM: A Human-Annotated Dataset for Emotion Classification and RAG-Based Poetry Generation in Korean Modern Poetry
von: Lim, Iro, et al.
Veröffentlicht: (2025)
von: Lim, Iro, et al.
Veröffentlicht: (2025)
A Bolu: A Structured Dataset for the Computational Analysis of Sardinian Improvisational Poetry
von: Calderaro, Silvio, et al.
Veröffentlicht: (2026)
von: Calderaro, Silvio, et al.
Veröffentlicht: (2026)
Using Counterfactual Tasks to Evaluate the Generality of Analogical Reasoning in Large Language Models
von: Lewis, Martha, et al.
Veröffentlicht: (2024)
von: Lewis, Martha, et al.
Veröffentlicht: (2024)
Where Do People Tell Stories Online? Story Detection Across Online Communities
von: Antoniak, Maria, et al.
Veröffentlicht: (2023)
von: Antoniak, Maria, et al.
Veröffentlicht: (2023)
Evaluating the Robustness of Analogical Reasoning in Large Language Models
von: Lewis, Martha, et al.
Veröffentlicht: (2024)
von: Lewis, Martha, et al.
Veröffentlicht: (2024)
LAB: Large-Scale Alignment for ChatBots
von: Sudalairaj, Shivchander, et al.
Veröffentlicht: (2024)
von: Sudalairaj, Shivchander, et al.
Veröffentlicht: (2024)
Comparative Analysis of Document-Level Embedding Methods for Similarity Scoring on Shakespeare Sonnets and Taylor Swift Lyrics
von: Kramer, Klara
Veröffentlicht: (2024)
von: Kramer, Klara
Veröffentlicht: (2024)
A Dataset for Metaphor Detection in Early Medieval Hebrew Poetry
von: Toker, Michael, et al.
Veröffentlicht: (2024)
von: Toker, Michael, et al.
Veröffentlicht: (2024)
A Chinese Dataset for Evaluating the Safeguards in Large Language Models
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
von: Wang, Yuxia, et al.
Veröffentlicht: (2024)
Automated Evaluation of Meter and Rhyme in Russian Generative and Human-Authored Poetry
von: Koziev, Ilya
Veröffentlicht: (2025)
von: Koziev, Ilya
Veröffentlicht: (2025)
Reading Subtext: Evaluating Large Language Models on Short Story Summarization with Writers
von: Subbiah, Melanie, et al.
Veröffentlicht: (2024)
von: Subbiah, Melanie, et al.
Veröffentlicht: (2024)
CHBench: A Chinese Dataset for Evaluating Health in Large Language Models
von: Guo, Chenlu, et al.
Veröffentlicht: (2024)
von: Guo, Chenlu, et al.
Veröffentlicht: (2024)
Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry
von: Li, Jiang, et al.
Veröffentlicht: (2026)
von: Li, Jiang, et al.
Veröffentlicht: (2026)
Evaluating Text Creativity across Diverse Domains: A Dataset and Large Language Model Evaluator
von: Cao, Qian, et al.
Veröffentlicht: (2025)
von: Cao, Qian, et al.
Veröffentlicht: (2025)
Adversarial Poetry as a Universal Single-Turn Jailbreak Mechanism in Large Language Models
von: Bisconti, Piercosma, et al.
Veröffentlicht: (2025)
von: Bisconti, Piercosma, et al.
Veröffentlicht: (2025)
Extracting Training Dialogue Data from Large Language Model based Task Bots
von: Zhang, Shuo, et al.
Veröffentlicht: (2026)
von: Zhang, Shuo, et al.
Veröffentlicht: (2026)
Multi-Modal Framing Analysis of News
von: Arora, Arnav, et al.
Veröffentlicht: (2025)
von: Arora, Arnav, et al.
Veröffentlicht: (2025)
NLP for Maternal Healthcare: Perspectives and Guiding Principles in the Age of LLMs
von: Antoniak, Maria, et al.
Veröffentlicht: (2023)
von: Antoniak, Maria, et al.
Veröffentlicht: (2023)
Automating Dataset Updates Towards Reliable and Timely Evaluation of Large Language Models
von: Ying, Jiahao, et al.
Veröffentlicht: (2024)
von: Ying, Jiahao, et al.
Veröffentlicht: (2024)
AlbanianLLMSafety: A Safety Evaluation Dataset for Large Language Models in Albanian
von: Zaghouani, Wajdi, et al.
Veröffentlicht: (2026)
von: Zaghouani, Wajdi, et al.
Veröffentlicht: (2026)
Epistemic Diversity and Knowledge Collapse in Large Language Models
von: Wright, Dustin, et al.
Veröffentlicht: (2025)
von: Wright, Dustin, et al.
Veröffentlicht: (2025)
Quokka: An Open-source Large Language Model ChatBot for Material Science
von: Yang, Xianjun, et al.
Veröffentlicht: (2024)
von: Yang, Xianjun, et al.
Veröffentlicht: (2024)
Automatic Detection of Research Values from Scientific Abstracts Across Computer Science Subfields
von: Jiang, Hang, et al.
Veröffentlicht: (2025)
von: Jiang, Hang, et al.
Veröffentlicht: (2025)
Evaluation of Cultural Competence of Vision-Language Models
von: Yadav, Srishti, et al.
Veröffentlicht: (2025)
von: Yadav, Srishti, et al.
Veröffentlicht: (2025)
LFED: A Literary Fiction Evaluation Dataset for Large Language Models
von: Yu, Linhao, et al.
Veröffentlicht: (2024)
von: Yu, Linhao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
so much depends / upon / a whitespace: Why Whitespace Matters for Poets and LLMs
von: Bhyravajjula, Sriharsh, et al.
Veröffentlicht: (2025) -
Does ChatGPT Have a Poetic Style?
von: Walsh, Melanie, et al.
Veröffentlicht: (2024) -
Trust No Bot: Discovering Personal Disclosures in Human-LLM Conversations in the Wild
von: Mireshghallah, Niloofar, et al.
Veröffentlicht: (2024) -
The Afterlives of Shakespeare and Company in Online Social Readership
von: Antoniak, Maria, et al.
Veröffentlicht: (2024) -
Capabilities and Evaluation Biases of Large Language Models in Classical Chinese Poetry Generation: A Case Study on Tang Poetry
von: Ma, Bolei, et al.
Veröffentlicht: (2025)