BADGE: BADminton report Generation and Evaluation with LLM
Fuente:
arXiv
Saved in:
| Main Authors: | Chiang, Shang-Hsuan, Chao, Lin-Wei, Wang, Kuang-Da, Wang, Chih-Chuan, Peng, Wen-Chih |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Tree-of-Text: A Tree-based Prompting Framework for Table-to-Text Generation in the Sports Domain
by: Chiang, Shang-Hsuan, et al.
Published: (2026)
by: Chiang, Shang-Hsuan, et al.
Published: (2026)
Multi-agent KTO: Reinforcing Strategic Interactions of Large Language Model in Language Game
by: Ye, Rong, et al.
Published: (2025)
by: Ye, Rong, et al.
Published: (2025)
Creativity in LLM-based Multi-Agent Systems: A Survey
by: Lin, Yi-Cheng, et al.
Published: (2025)
by: Lin, Yi-Cheng, et al.
Published: (2025)
TICKing All the Boxes: Generated Checklists Improve LLM Evaluation and Generation
by: Cook, Jonathan, et al.
Published: (2024)
by: Cook, Jonathan, et al.
Published: (2024)
Do Students with Different Personality Traits Demonstrate Different Physiological Signals in Video-based Learning?
by: Tseng, Chun-Hsiung, et al.
Published: (2024)
by: Tseng, Chun-Hsiung, et al.
Published: (2024)
Evaluating LLM-Generated Lessons from the Language Learning Students' Perspective: A Short Case Study on Duolingo
by: Catalan, Carlos Rafael, et al.
Published: (2026)
by: Catalan, Carlos Rafael, et al.
Published: (2026)
LalaEval: A Holistic Human Evaluation Framework for Domain-Specific Large Language Models
by: Sun, Chongyan, et al.
Published: (2024)
by: Sun, Chongyan, et al.
Published: (2024)
Team Trifecta at Factify5WQA: Setting the Standard in Fact Verification with Fine-Tuning
by: Chiang, Shang-Hsuan, et al.
Published: (2024)
by: Chiang, Shang-Hsuan, et al.
Published: (2024)
Development and Evaluation of HopeBot: an LLM-based chatbot for structured and interactive PHQ-9 depression screening
by: Guo, Zhijun, et al.
Published: (2025)
by: Guo, Zhijun, et al.
Published: (2025)
Step Guided Reasoning: Improving Mathematical Reasoning using Guidance Generation and Step Reasoning
by: Cao, Lang, et al.
Published: (2024)
by: Cao, Lang, et al.
Published: (2024)
Do We Talk to Robots Like Therapists, and Do They Respond Accordingly? Language Alignment in AI Emotional Support
by: Chiang, Sophie, et al.
Published: (2025)
by: Chiang, Sophie, et al.
Published: (2025)
LLM Attributor: Interactive Visual Attribution for LLM Generation
by: Lee, Seongmin, et al.
Published: (2024)
by: Lee, Seongmin, et al.
Published: (2024)
Grading Scale Impact on LLM-as-a-Judge: Human-LLM Alignment Is Highest on 0-5 Grading Scale
by: Li, Weiyue, et al.
Published: (2026)
by: Li, Weiyue, et al.
Published: (2026)
Stories of Your Life as Others: A Round-Trip Evaluation of LLM-Generated Life Stories Conditioned on Rich Psychometric Profiles
by: Wigler, Ben, et al.
Published: (2026)
by: Wigler, Ben, et al.
Published: (2026)
Inclusion Arena: An Open Platform for Evaluating Large Foundation Models with Real-World Apps
by: Wang, Kangyu, et al.
Published: (2025)
by: Wang, Kangyu, et al.
Published: (2025)
SSKG Hub: An Expert-Guided Platform for LLM-Empowered Sustainability Standards Knowledge Graphs
by: He, Chaoyue, et al.
Published: (2026)
by: He, Chaoyue, et al.
Published: (2026)
Evaluating LLM Adaptation to Sociodemographic Factors: User Profile vs. Dialogue History
by: Zhong, Qishuai, et al.
Published: (2025)
by: Zhong, Qishuai, et al.
Published: (2025)
RNR: Teaching Large Language Models to Follow Roles and Rules
by: Wang, Kuan, et al.
Published: (2024)
by: Wang, Kuan, et al.
Published: (2024)
The Generative AI Paradox on Evaluation: What It Can Solve, It May Not Evaluate
by: Oh, Juhyun, et al.
Published: (2024)
by: Oh, Juhyun, et al.
Published: (2024)
Patchview: LLM-Powered Worldbuilding with Generative Dust and Magnet Visualization
by: Chung, John Joon Young, et al.
Published: (2024)
by: Chung, John Joon Young, et al.
Published: (2024)
Word Clouds as Common Voices: LLM-Assisted Visualization of Participant-Weighted Themes in Qualitative Interviews
by: Colonel, Joseph T., et al.
Published: (2025)
by: Colonel, Joseph T., et al.
Published: (2025)
Design Techniques for LLM-Powered Interactive Storytelling: A Case Study of the Dramamancer System
by: Wang, Tiffany, et al.
Published: (2026)
by: Wang, Tiffany, et al.
Published: (2026)
VeriLA: A Human-Centered Evaluation Framework for Interpretable Verification of LLM Agent Failures
by: Sung, Yoo Yeon, et al.
Published: (2025)
by: Sung, Yoo Yeon, et al.
Published: (2025)
Evaluating Behavioral Alignment in Conflict Dialogue: A Multi-Dimensional Comparison of LLM Agents and Humans
by: Kwon, Deuksin, et al.
Published: (2025)
by: Kwon, Deuksin, et al.
Published: (2025)
Benchmarking LLM Tool-Use in the Wild
by: Yu, Peijie, et al.
Published: (2026)
by: Yu, Peijie, et al.
Published: (2026)
Learning to Generate and Evaluate Fact-checking Explanations with Transformers
by: Feher, Darius, et al.
Published: (2024)
by: Feher, Darius, et al.
Published: (2024)
SouLLMate: An Application Enhancing Diverse Mental Health Support with Adaptive LLMs, Prompt Engineering, and RAG Techniques
by: Guo, Qiming, et al.
Published: (2024)
by: Guo, Qiming, et al.
Published: (2024)
Generative Echo Chamber? Effects of LLM-Powered Search Systems on Diverse Information Seeking
by: Sharma, Nikhil, et al.
Published: (2024)
by: Sharma, Nikhil, et al.
Published: (2024)
A Generalized LLM-Augmented BIM Framework: Application to a Speech-to-BIM system
by: Lee, Ghang, et al.
Published: (2024)
by: Lee, Ghang, et al.
Published: (2024)
AI-native Memory 2.0: Second Me
by: Wei, Jiale, et al.
Published: (2025)
by: Wei, Jiale, et al.
Published: (2025)
PsychBench: A comprehensive and professional benchmark for evaluating the performance of LLM-assisted psychiatric clinical practice
by: Liu, Shuyu, et al.
Published: (2025)
by: Liu, Shuyu, et al.
Published: (2025)
DailyLLM: Context-Aware Activity Log Generation Using Multi-Modal Sensors and LLMs
by: Tian, Ye, et al.
Published: (2025)
by: Tian, Ye, et al.
Published: (2025)
Pragmatics Meets Culture: Culturally-adapted Artwork Description Generation and Evaluation
by: Zhao, Lingjun, et al.
Published: (2026)
by: Zhao, Lingjun, et al.
Published: (2026)
EUDAIMONIA: Evaluating Undesirable Dynamics in AI
by: Huang, Jun Rui, et al.
Published: (2026)
by: Huang, Jun Rui, et al.
Published: (2026)
Through the Judge's Eyes: Inferred Thinking Traces Improve Reliability of LLM Raters
by: Zhang, Xingjian, et al.
Published: (2025)
by: Zhang, Xingjian, et al.
Published: (2025)
Crafting Hanzi as Narrative Bridges: An AI Co-Creation Workshop for Elderly Migrants
by: Zhan, Wen, et al.
Published: (2025)
by: Zhan, Wen, et al.
Published: (2025)
CHBench: A Cognitive Hierarchy Benchmark for Evaluating Strategic Reasoning Capability of LLMs
by: Liu, Hongtao, et al.
Published: (2025)
by: Liu, Hongtao, et al.
Published: (2025)
From Novice to Expert: LLM Agent Policy Optimization via Step-wise Reinforcement Learning
by: Deng, Zhirui, et al.
Published: (2024)
by: Deng, Zhirui, et al.
Published: (2024)
Understand User Opinions of Large Language Models via LLM-Powered In-the-Moment User Experience Interviews
by: Liu, Mengqiao, et al.
Published: (2025)
by: Liu, Mengqiao, et al.
Published: (2025)
Using Generative Text Models to Create Qualitative Codebooks for Student Evaluations of Teaching
by: Katz, Andrew, et al.
Published: (2024)
by: Katz, Andrew, et al.
Published: (2024)
Similar Items
-
Tree-of-Text: A Tree-based Prompting Framework for Table-to-Text Generation in the Sports Domain
by: Chiang, Shang-Hsuan, et al.
Published: (2026) -
Multi-agent KTO: Reinforcing Strategic Interactions of Large Language Model in Language Game
by: Ye, Rong, et al.
Published: (2025) -
Creativity in LLM-based Multi-Agent Systems: A Survey
by: Lin, Yi-Cheng, et al.
Published: (2025) -
TICKing All the Boxes: Generated Checklists Improve LLM Evaluation and Generation
by: Cook, Jonathan, et al.
Published: (2024) -
Do Students with Different Personality Traits Demonstrate Different Physiological Signals in Video-based Learning?
by: Tseng, Chun-Hsiung, et al.
Published: (2024)