ReflectSumm: A Benchmark for Course Reflection Summarization
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhong, Yang, Elaraby, Mohamed, Litman, Diane, Butt, Ahmed Ashraf, Menekse, Muhsin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ARC: Argument Representation and Coverage Analysis for Zero-Shot Long Document Summarization with Instruction Following LLMs
di: Elaraby, Mohamed, et al.
Pubblicazione: (2025)
di: Elaraby, Mohamed, et al.
Pubblicazione: (2025)
Discourse-Driven Evaluation: Unveiling Factual Inconsistency in Long Document Summarization
di: Zhong, Yang, et al.
Pubblicazione: (2025)
di: Zhong, Yang, et al.
Pubblicazione: (2025)
PreSumm: Predicting Summarization Performance Without Summarizing
di: Koniaev, Steven, et al.
Pubblicazione: (2025)
di: Koniaev, Steven, et al.
Pubblicazione: (2025)
Persuasiveness of Generated Free-Text Rationales in Subjective Decisions: A Case Study on Pairwise Argument Ranking
di: Elaraby, Mohamed, et al.
Pubblicazione: (2024)
di: Elaraby, Mohamed, et al.
Pubblicazione: (2024)
GlobeSumm: A Challenging Benchmark Towards Unifying Multi-lingual, Cross-lingual and Multi-document News Summarization
di: Ye, Yangfan, et al.
Pubblicazione: (2024)
di: Ye, Yangfan, et al.
Pubblicazione: (2024)
Efficient Layer-wise LLM Fine-tuning for Revision Intention Prediction
di: Liu, Zhexiong, et al.
Pubblicazione: (2025)
di: Liu, Zhexiong, et al.
Pubblicazione: (2025)
MedSumm: A Multimodal Approach to Summarizing Code-Mixed Hindi-English Clinical Queries
di: Ghosh, Akash, et al.
Pubblicazione: (2024)
di: Ghosh, Akash, et al.
Pubblicazione: (2024)
Medifact at PerAnsSumm 2025: Leveraging Lightweight Models for Perspective-Specific Summarization of Clinical Q&A Forums
di: Saeed, Nadia
Pubblicazione: (2025)
di: Saeed, Nadia
Pubblicazione: (2025)
UMB@PerAnsSumm 2025: Enhancing Perspective-Aware Summarization with Prompt Optimization and Supervised Fine-Tuning
di: Qi, Kristin, et al.
Pubblicazione: (2025)
di: Qi, Kristin, et al.
Pubblicazione: (2025)
ReFeed: Multi-dimensional Summarization Refinement with Reflective Reasoning on Feedback
di: Yun, Taewon, et al.
Pubblicazione: (2025)
di: Yun, Taewon, et al.
Pubblicazione: (2025)
CaseSumm: A Large-Scale Dataset for Long-Context Summarization from U.S. Supreme Court Opinions
di: Heddaya, Mourad, et al.
Pubblicazione: (2024)
di: Heddaya, Mourad, et al.
Pubblicazione: (2024)
Meta-Reflection: A Feedback-Free Reflection Learning Framework
di: Wang, Yaoke, et al.
Pubblicazione: (2024)
di: Wang, Yaoke, et al.
Pubblicazione: (2024)
Connecting the Dots: Benchmarking Reflective Memory in Long-Horizon Dialogue
di: Lin, Jingjie, et al.
Pubblicazione: (2026)
di: Lin, Jingjie, et al.
Pubblicazione: (2026)
MetaReflection: Learning Instructions for Language Agents using Past Reflections
di: Gupta, Priyanshu, et al.
Pubblicazione: (2024)
di: Gupta, Priyanshu, et al.
Pubblicazione: (2024)
MedReflect: Teaching Medical LLMs to Self-Improve via Reflective Correction
di: Huang, Yue, et al.
Pubblicazione: (2025)
di: Huang, Yue, et al.
Pubblicazione: (2025)
CAMEL: Confidence-Gated Reflection for Reward Modeling
di: Zhu, Zirui, et al.
Pubblicazione: (2026)
di: Zhu, Zirui, et al.
Pubblicazione: (2026)
ReflectionCoder: Learning from Reflection Sequence for Enhanced One-off Code Generation
di: Ren, Houxing, et al.
Pubblicazione: (2024)
di: Ren, Houxing, et al.
Pubblicazione: (2024)
ReflectivePrompt: Reflective evolution in autoprompting algorithms
di: Zhuravlev, Viktor N., et al.
Pubblicazione: (2025)
di: Zhuravlev, Viktor N., et al.
Pubblicazione: (2025)
Rethinking Reflection in Pre-Training
di: AI, Essential, et al.
Pubblicazione: (2025)
di: AI, Essential, et al.
Pubblicazione: (2025)
Stop Unnecessary Reflection: Training LRMs for Efficient Reasoning with Adaptive Reflection and Length Coordinated Penalty
di: Yu, Zewei, et al.
Pubblicazione: (2026)
di: Yu, Zewei, et al.
Pubblicazione: (2026)
eRevise+RF: A Writing Evaluation System for Assessing Student Essay Revisions and Providing Formative Feedback
di: Liu, Zhexiong, et al.
Pubblicazione: (2025)
di: Liu, Zhexiong, et al.
Pubblicazione: (2025)
Contextual ASR Error Handling with LLMs Augmentation for Goal-Oriented Conversational AI
di: Asano, Yuya, et al.
Pubblicazione: (2025)
di: Asano, Yuya, et al.
Pubblicazione: (2025)
MDSEval: A Meta-Evaluation Benchmark for Multimodal Dialogue Summarization
di: Liu, Yinhong, et al.
Pubblicazione: (2025)
di: Liu, Yinhong, et al.
Pubblicazione: (2025)
Reinforcement Learning from Reflective Feedback (RLRF): Aligning and Improving LLMs via Fine-Grained Self-Reflection
di: Lee, Kyungjae, et al.
Pubblicazione: (2024)
di: Lee, Kyungjae, et al.
Pubblicazione: (2024)
ReflectRM: Boosting Generative Reward Models via Self-Reflection within a Unified Judgment Framework
di: Qin, Kai, et al.
Pubblicazione: (2026)
di: Qin, Kai, et al.
Pubblicazione: (2026)
AugSumm: towards generalizable speech summarization using synthetic labels from large language model
di: Jung, Jee-weon, et al.
Pubblicazione: (2024)
di: Jung, Jee-weon, et al.
Pubblicazione: (2024)
FaithBench: A Diverse Hallucination Benchmark for Summarization by Modern LLMs
di: Bao, Forrest Sheng, et al.
Pubblicazione: (2024)
di: Bao, Forrest Sheng, et al.
Pubblicazione: (2024)
PCQPR: Proactive Conversational Question Planning with Reflection
di: Guo, Shasha, et al.
Pubblicazione: (2024)
di: Guo, Shasha, et al.
Pubblicazione: (2024)
Speaking of Language: Reflections on Metalanguage Research in NLP
di: Schneider, Nathan, et al.
Pubblicazione: (2026)
di: Schneider, Nathan, et al.
Pubblicazione: (2026)
Towards Better Chain-of-Thought: A Reflection on Effectiveness and Faithfulness
di: Li, Jiachun, et al.
Pubblicazione: (2024)
di: Li, Jiachun, et al.
Pubblicazione: (2024)
Large Language Models as Theory of Mind Aware Generative Agents with Counterfactual Reflection
di: Yang, Bo, et al.
Pubblicazione: (2025)
di: Yang, Bo, et al.
Pubblicazione: (2025)
Automatic Reflection Level Classification in Hungarian Student Essays
di: Csibi, Zsolt, et al.
Pubblicazione: (2026)
di: Csibi, Zsolt, et al.
Pubblicazione: (2026)
MyGO Multiplex CoT: A Method for Self-Reflection in Large Language Models via Double Chain of Thought Thinking
di: Ji, Shihao, et al.
Pubblicazione: (2025)
di: Ji, Shihao, et al.
Pubblicazione: (2025)
An Evaluation of Large Language Models on Text Summarization Tasks Using Prompt Engineering Techniques
di: Aly, Walid Mohamed, et al.
Pubblicazione: (2025)
di: Aly, Walid Mohamed, et al.
Pubblicazione: (2025)
ReflectDiffu:Reflect between Emotion-intent Contagion and Mimicry for Empathetic Response Generation via a RL-Diffusion Framework
di: Yuan, Jiahao, et al.
Pubblicazione: (2024)
di: Yuan, Jiahao, et al.
Pubblicazione: (2024)
CourtPressGER: A German Court Decision to Press Release Summarization Dataset
di: Nagl, Sebastian, et al.
Pubblicazione: (2025)
di: Nagl, Sebastian, et al.
Pubblicazione: (2025)
Reasoning: From Reflection to Solution
di: Li, Zixi
Pubblicazione: (2025)
di: Li, Zixi
Pubblicazione: (2025)
Positive Experience Reflection for Agents in Interactive Text Environments
di: Lippmann, Philip, et al.
Pubblicazione: (2024)
di: Lippmann, Philip, et al.
Pubblicazione: (2024)
SummExecEdit: A Factual Consistency Benchmark in Summarization with Executable Edits
di: Thorat, Onkar, et al.
Pubblicazione: (2024)
di: Thorat, Onkar, et al.
Pubblicazione: (2024)
American Sign Language Handshapes Reflect Pressures for Communicative Efficiency
di: Yin, Kayo, et al.
Pubblicazione: (2024)
di: Yin, Kayo, et al.
Pubblicazione: (2024)
Documenti analoghi
-
ARC: Argument Representation and Coverage Analysis for Zero-Shot Long Document Summarization with Instruction Following LLMs
di: Elaraby, Mohamed, et al.
Pubblicazione: (2025) -
Discourse-Driven Evaluation: Unveiling Factual Inconsistency in Long Document Summarization
di: Zhong, Yang, et al.
Pubblicazione: (2025) -
PreSumm: Predicting Summarization Performance Without Summarizing
di: Koniaev, Steven, et al.
Pubblicazione: (2025) -
Persuasiveness of Generated Free-Text Rationales in Subjective Decisions: A Case Study on Pairwise Argument Ranking
di: Elaraby, Mohamed, et al.
Pubblicazione: (2024) -
GlobeSumm: A Challenging Benchmark Towards Unifying Multi-lingual, Cross-lingual and Multi-document News Summarization
di: Ye, Yangfan, et al.
Pubblicazione: (2024)