CoKe: Customizable Fine-Grained Story Evaluation via Chain-of-Keyword Rationalization
Fuente:
arXiv
Salvato in:
| Autori principali: | Joshi, Brihi, Venkatapathy, Sriram, Bansal, Mohit, Peng, Nanyun, Chang, Haw-Shiuan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
REAL Sampling: Boosting Factuality and Diversity of Open-Ended Generation via Asymptotic Entropy
di: Chang, Haw-Shiuan, et al.
Pubblicazione: (2024)
di: Chang, Haw-Shiuan, et al.
Pubblicazione: (2024)
Explaining and Improving Contrastive Decoding by Extrapolating the Probabilities of a Huge and Hypothetical LM
di: Chang, Haw-Shiuan, et al.
Pubblicazione: (2024)
di: Chang, Haw-Shiuan, et al.
Pubblicazione: (2024)
Improving Language Model Personas via Rationalization with Psychological Scaffolds
di: Joshi, Brihi, et al.
Pubblicazione: (2025)
di: Joshi, Brihi, et al.
Pubblicazione: (2025)
LLM Self-Correction with DeCRIM: Decompose, Critique, and Refine for Enhanced Following of Instructions with Multiple Constraints
di: Ferraz, Thomas Palmeira, et al.
Pubblicazione: (2024)
di: Ferraz, Thomas Palmeira, et al.
Pubblicazione: (2024)
DreamRunner: Fine-Grained Compositional Story-to-Video Generation with Retrieval-Augmented Motion Adaptation
di: Wang, Zun, et al.
Pubblicazione: (2024)
di: Wang, Zun, et al.
Pubblicazione: (2024)
FLAMES: Improving LLM Math Reasoning via a Fine-Grained Analysis of the Data Synthesis Pipeline
di: Seegmiller, Parker, et al.
Pubblicazione: (2025)
di: Seegmiller, Parker, et al.
Pubblicazione: (2025)
Truncated Step-Level Sampling with Process Rewards for Retrieval-Augmented Reasoning
di: Samarinas, Chris, et al.
Pubblicazione: (2026)
di: Samarinas, Chris, et al.
Pubblicazione: (2026)
CS4: Measuring the Creativity of Large Language Models Automatically by Controlling the Number of Story-Writing Constraints
di: Atmakuru, Anirudh, et al.
Pubblicazione: (2024)
di: Atmakuru, Anirudh, et al.
Pubblicazione: (2024)
PrimeX: A Dataset of Worldview, Opinion, and Explanation
di: Koncel-Kedziorski, Rik, et al.
Pubblicazione: (2025)
di: Koncel-Kedziorski, Rik, et al.
Pubblicazione: (2025)
Latent Traits and Cross-Task Transfer: Deconstructing Dataset Interactions in LLM Fine-tuning
di: Krishna, Shambhavi, et al.
Pubblicazione: (2025)
di: Krishna, Shambhavi, et al.
Pubblicazione: (2025)
Tailoring Self-Rationalizers with Multi-Reward Distillation
di: Ramnath, Sahana, et al.
Pubblicazione: (2023)
di: Ramnath, Sahana, et al.
Pubblicazione: (2023)
Effective Reasoning Chains Reduce Intrinsic Dimensionality
di: Prasad, Archiki, et al.
Pubblicazione: (2026)
di: Prasad, Archiki, et al.
Pubblicazione: (2026)
PROMPT2BOX: Uncovering Entailment Structure among LLM Prompts
di: Bhuiya, Neeladri, et al.
Pubblicazione: (2026)
di: Bhuiya, Neeladri, et al.
Pubblicazione: (2026)
Video-Skill-CoT: Skill-based Chain-of-Thoughts for Domain-Adaptive Video Reasoning
di: Lee, Daeun, et al.
Pubblicazione: (2025)
di: Lee, Daeun, et al.
Pubblicazione: (2025)
Amulet: Putting Complex Multi-Turn Conversations on the Stand with LLM Juries
di: Ramnath, Sahana, et al.
Pubblicazione: (2025)
di: Ramnath, Sahana, et al.
Pubblicazione: (2025)
OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles
di: Deng, Yihe, et al.
Pubblicazione: (2025)
di: Deng, Yihe, et al.
Pubblicazione: (2025)
VISCO: Benchmarking Fine-Grained Critique and Correction Towards Self-Improvement in Visual Reasoning
di: Wu, Xueqing, et al.
Pubblicazione: (2024)
di: Wu, Xueqing, et al.
Pubblicazione: (2024)
QAPyramid: Fine-grained Evaluation of Content Selection for Text Summarization
di: Zhang, Shiyue, et al.
Pubblicazione: (2024)
di: Zhang, Shiyue, et al.
Pubblicazione: (2024)
Comparing Bad Apples to Good Oranges: Aligning Large Language Models via Joint Preference Optimization
di: Bansal, Hritik, et al.
Pubblicazione: (2024)
di: Bansal, Hritik, et al.
Pubblicazione: (2024)
ELI-Why: Evaluating the Pedagogical Utility of Language Model Explanations
di: Joshi, Brihi, et al.
Pubblicazione: (2025)
di: Joshi, Brihi, et al.
Pubblicazione: (2025)
Open-Domain Text Evaluation via Contrastive Distribution Methods
di: Lu, Sidi, et al.
Pubblicazione: (2023)
di: Lu, Sidi, et al.
Pubblicazione: (2023)
DiCoRe: Enhancing Zero-shot Event Detection via Divergent-Convergent LLM Reasoning
di: Parekh, Tanmay, et al.
Pubblicazione: (2025)
di: Parekh, Tanmay, et al.
Pubblicazione: (2025)
Fundamental Problems With Model Editing: How Should Rational Belief Revision Work in LLMs?
di: Hase, Peter, et al.
Pubblicazione: (2024)
di: Hase, Peter, et al.
Pubblicazione: (2024)
Believing without Seeing: Quality Scores for Contextualizing Vision-Language Model Explanations
di: He, Keyu, et al.
Pubblicazione: (2025)
di: He, Keyu, et al.
Pubblicazione: (2025)
PCOV-KWS: Multi-task Learning for Personalized Customizable Open Vocabulary Keyword Spotting
di: Pan, Jianan, et al.
Pubblicazione: (2026)
di: Pan, Jianan, et al.
Pubblicazione: (2026)
KPEval: Towards Fine-Grained Semantic-Based Keyphrase Evaluation
di: Wu, Di, et al.
Pubblicazione: (2023)
di: Wu, Di, et al.
Pubblicazione: (2023)
On the Loss of Context-awareness in General Instruction Fine-tuning
di: Wang, Yihan, et al.
Pubblicazione: (2024)
di: Wang, Yihan, et al.
Pubblicazione: (2024)
Control Large Language Models via Divide and Conquer
di: Li, Bingxuan, et al.
Pubblicazione: (2024)
di: Li, Bingxuan, et al.
Pubblicazione: (2024)
SkillVerse : Assessing and Enhancing LLMs with Tree Evaluation
di: Tian, Yufei, et al.
Pubblicazione: (2025)
di: Tian, Yufei, et al.
Pubblicazione: (2025)
ECoFLaP: Efficient Coarse-to-Fine Layer-Wise Pruning for Vision-Language Models
di: Sung, Yi-Lin, et al.
Pubblicazione: (2023)
di: Sung, Yi-Lin, et al.
Pubblicazione: (2023)
CAPTURe: Evaluating Spatial Reasoning in Vision Language Models via Occluded Object Counting
di: Pothiraj, Atin, et al.
Pubblicazione: (2025)
di: Pothiraj, Atin, et al.
Pubblicazione: (2025)
GenerationPrograms: Fine-grained Attribution with Executable Programs
di: Wan, David, et al.
Pubblicazione: (2025)
di: Wan, David, et al.
Pubblicazione: (2025)
StructFormer: Document Structure-based Masked Attention and its Impact on Language Model Pre-Training
di: Ponkshe, Kaustubh, et al.
Pubblicazione: (2024)
di: Ponkshe, Kaustubh, et al.
Pubblicazione: (2024)
CREMA: Generalizable and Efficient Video-Language Reasoning via Multimodal Modular Fusion
di: Yu, Shoubin, et al.
Pubblicazione: (2024)
di: Yu, Shoubin, et al.
Pubblicazione: (2024)
BRIEF: Bridging Retrieval and Inference for Multi-hop Reasoning via Compression
di: Li, Yuankai, et al.
Pubblicazione: (2024)
di: Li, Yuankai, et al.
Pubblicazione: (2024)
MAgICoRe: Multi-Agent, Iterative, Coarse-to-Fine Refinement for Reasoning
di: Chen, Justin Chih-Yao, et al.
Pubblicazione: (2024)
di: Chen, Justin Chih-Yao, et al.
Pubblicazione: (2024)
AMRFact: Enhancing Summarization Factuality Evaluation with AMR-Driven Negative Samples Generation
di: Qiu, Haoyi, et al.
Pubblicazione: (2023)
di: Qiu, Haoyi, et al.
Pubblicazione: (2023)
Evaluating Human Alignment and Model Faithfulness of LLM Rationale
di: Fayyaz, Mohsen, et al.
Pubblicazione: (2024)
di: Fayyaz, Mohsen, et al.
Pubblicazione: (2024)
Chaos with Keywords: Exposing Large Language Models Sycophantic Hallucination to Misleading Keywords and Evaluating Defense Strategies
di: RRV, Aswin, et al.
Pubblicazione: (2024)
di: RRV, Aswin, et al.
Pubblicazione: (2024)
Con-ReCall: Detecting Pre-training Data in LLMs via Contrastive Decoding
di: Wang, Cheng, et al.
Pubblicazione: (2024)
di: Wang, Cheng, et al.
Pubblicazione: (2024)
Documenti analoghi
-
REAL Sampling: Boosting Factuality and Diversity of Open-Ended Generation via Asymptotic Entropy
di: Chang, Haw-Shiuan, et al.
Pubblicazione: (2024) -
Explaining and Improving Contrastive Decoding by Extrapolating the Probabilities of a Huge and Hypothetical LM
di: Chang, Haw-Shiuan, et al.
Pubblicazione: (2024) -
Improving Language Model Personas via Rationalization with Psychological Scaffolds
di: Joshi, Brihi, et al.
Pubblicazione: (2025) -
LLM Self-Correction with DeCRIM: Decompose, Critique, and Refine for Enhanced Following of Instructions with Multiple Constraints
di: Ferraz, Thomas Palmeira, et al.
Pubblicazione: (2024) -
DreamRunner: Fine-Grained Compositional Story-to-Video Generation with Retrieval-Augmented Motion Adaptation
di: Wang, Zun, et al.
Pubblicazione: (2024)