What Makes a Good Story and How Can We Measure It? A Comprehensive Survey of Story Evaluation
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Dingyi, Jin, Qin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
StoryReasoning Dataset: Using Chain-of-Thought for Scene Understanding and Grounded Story Generation
di: Oliveira, Daniel A. P., et al.
Pubblicazione: (2025)
di: Oliveira, Daniel A. P., et al.
Pubblicazione: (2025)
PhysicsArena: The First Multimodal Physics Reasoning Benchmark Exploring Variable, Process, and Solution Dimensions
di: Dai, Song, et al.
Pubblicazione: (2025)
di: Dai, Song, et al.
Pubblicazione: (2025)
Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning
di: Tong, Jingqi, et al.
Pubblicazione: (2025)
di: Tong, Jingqi, et al.
Pubblicazione: (2025)
ICG: Improving Cover Image Generation via MLLM-based Prompting and Personalized Preference Alignment
di: Bian, Zhipeng, et al.
Pubblicazione: (2026)
di: Bian, Zhipeng, et al.
Pubblicazione: (2026)
Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning
di: Yang, Shan
Pubblicazione: (2026)
di: Yang, Shan
Pubblicazione: (2026)
Learning the meanings of function words from grounded language using a visual question answering model
di: Portelance, Eva, et al.
Pubblicazione: (2023)
di: Portelance, Eva, et al.
Pubblicazione: (2023)
Spatially-Aware Speaker for Vision-and-Language Navigation Instruction Generation
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
From Benchmarking to Reasoning: A Dual-Aspect, Large-Scale Evaluation of LLMs on Vietnamese Legal Text
di: Le, Van-Truong
Pubblicazione: (2026)
di: Le, Van-Truong
Pubblicazione: (2026)
StoryMovie: A Dataset for Semantic Alignment of Visual Stories with Movie Scripts and Subtitles
di: Oliveira, Daniel, et al.
Pubblicazione: (2026)
di: Oliveira, Daniel, et al.
Pubblicazione: (2026)
Story Generation from Visual Inputs: Techniques, Related Tasks, and Challenges
di: Oliveira, Daniel A. P., et al.
Pubblicazione: (2024)
di: Oliveira, Daniel A. P., et al.
Pubblicazione: (2024)
A Closer Look at Bias and Chain-of-Thought Faithfulness of Large (Vision) Language Models
di: Balasubramanian, Sriram, et al.
Pubblicazione: (2025)
di: Balasubramanian, Sriram, et al.
Pubblicazione: (2025)
Defending against Backdoor Attacks via Module Switching
di: Li, Weijun, et al.
Pubblicazione: (2025)
di: Li, Weijun, et al.
Pubblicazione: (2025)
More Than Meets the Eye: Measuring the Semiotic Gap in Vision-Language Models via Semantic Anchorage
di: He, Wei
Pubblicazione: (2026)
di: He, Wei
Pubblicazione: (2026)
K-MetBench: A Multi-Dimensional Benchmark for Fine-Grained Evaluation of Expert Reasoning, Locality, and Multimodality in Meteorology
di: Kim, Soyeon, et al.
Pubblicazione: (2026)
di: Kim, Soyeon, et al.
Pubblicazione: (2026)
VidNum-1.4K: A Comprehensive Benchmark for Video-based Numerical Reasoning
di: Cui, Shaoyang, et al.
Pubblicazione: (2026)
di: Cui, Shaoyang, et al.
Pubblicazione: (2026)
Reframing linguistic bootstrapping as joint inference using visually-grounded grammar induction models
di: Portelance, Eva, et al.
Pubblicazione: (2024)
di: Portelance, Eva, et al.
Pubblicazione: (2024)
GroundCap: A Visually Grounded Image Captioning Dataset
di: Oliveira, Daniel A. P., et al.
Pubblicazione: (2025)
di: Oliveira, Daniel A. P., et al.
Pubblicazione: (2025)
Relative Drawing Identification Complexity is Invariant to Modality in Vision-Language Models
di: Freitas, Diogo, et al.
Pubblicazione: (2025)
di: Freitas, Diogo, et al.
Pubblicazione: (2025)
Towards Explainable Fake Image Detection with Multi-Modal Large Language Models
di: Ji, Yikun, et al.
Pubblicazione: (2025)
di: Ji, Yikun, et al.
Pubblicazione: (2025)
Correspondence of high-dimensional emotion structures elicited by video clips between humans and Multimodal LLMs
di: Asanuma, Haruka, et al.
Pubblicazione: (2025)
di: Asanuma, Haruka, et al.
Pubblicazione: (2025)
OpenMap: Instruction Grounding via Open-Vocabulary Visual-Language Mapping
di: Li, Danyang, et al.
Pubblicazione: (2025)
di: Li, Danyang, et al.
Pubblicazione: (2025)
Evaluating Perspectival Biases in Cross-Modal Retrieval
di: Saengsukhiran, Teerapol, et al.
Pubblicazione: (2025)
di: Saengsukhiran, Teerapol, et al.
Pubblicazione: (2025)
CORDIAL: Can Multimodal Large Language Models Effectively Understand Coherence Relationships?
di: Ramakrishnan, Aashish Anantha, et al.
Pubblicazione: (2025)
di: Ramakrishnan, Aashish Anantha, et al.
Pubblicazione: (2025)
Taking Flight with Dialogue: Enabling Natural Language Control for PX4-based Drone Agent
di: Lim, Shoon Kit, et al.
Pubblicazione: (2025)
di: Lim, Shoon Kit, et al.
Pubblicazione: (2025)
VLEU: a Method for Automatic Evaluation for Generalizability of Text-to-Image Models
di: Cao, Jingtao, et al.
Pubblicazione: (2024)
di: Cao, Jingtao, et al.
Pubblicazione: (2024)
Survey Transfer Learning: Recycling Data with Silicon Responses
di: Amini, Ali
Pubblicazione: (2025)
di: Amini, Ali
Pubblicazione: (2025)
Bounding Hallucinations: Information-Theoretic Guarantees for RAG Systems via Merlin-Arthur Protocols
di: Deiseroth, Björn, et al.
Pubblicazione: (2025)
di: Deiseroth, Björn, et al.
Pubblicazione: (2025)
SpatialMath: Spatial Comprehension-Infused Symbolic Reasoning for Mathematical Problem-Solving
di: Bajpai, Ashutosh, et al.
Pubblicazione: (2026)
di: Bajpai, Ashutosh, et al.
Pubblicazione: (2026)
Universal Adversarial Attack on Aligned Multimodal LLMs
di: Rahmatullaev, Temurbek, et al.
Pubblicazione: (2025)
di: Rahmatullaev, Temurbek, et al.
Pubblicazione: (2025)
Memory-Efficient Differentially Private Training with Gradient Random Projection
di: Mulrooney, Alex, et al.
Pubblicazione: (2025)
di: Mulrooney, Alex, et al.
Pubblicazione: (2025)
PhysNote: Self-Knowledge Notes for Evolvable Physical Reasoning in Vision-Language Model
di: Zhang, Sinin, et al.
Pubblicazione: (2026)
di: Zhang, Sinin, et al.
Pubblicazione: (2026)
MM-SHAP: A Performance-agnostic Metric for Measuring Multimodal Contributions in Vision and Language Models & Tasks
di: Parcalabescu, Letitia, et al.
Pubblicazione: (2022)
di: Parcalabescu, Letitia, et al.
Pubblicazione: (2022)
StratXplore: Strategic Novelty-seeking and Instruction-aligned Exploration for Vision and Language Navigation
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
di: Gopinathan, Muraleekrishna, et al.
Pubblicazione: (2024)
Evaluating Voice Command Pipelines for Drone Control: From STT and LLM to Direct Classification and Siamese Networks
di: Simões, Lucca Emmanuel Pineli, et al.
Pubblicazione: (2024)
di: Simões, Lucca Emmanuel Pineli, et al.
Pubblicazione: (2024)
On Measuring Faithfulness or Self-consistency of Natural Language Explanations
di: Parcalabescu, Letitia, et al.
Pubblicazione: (2023)
di: Parcalabescu, Letitia, et al.
Pubblicazione: (2023)
MemeCraft: Contextual and Stance-Driven Multimodal Meme Generation
di: Wang, Han, et al.
Pubblicazione: (2024)
di: Wang, Han, et al.
Pubblicazione: (2024)
Emo3D: Metric and Benchmarking Dataset for 3D Facial Expression Generation from Emotion Description
di: Dehghani, Mahshid, et al.
Pubblicazione: (2024)
di: Dehghani, Mahshid, et al.
Pubblicazione: (2024)
WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation
di: Niu, Yuwei, et al.
Pubblicazione: (2025)
di: Niu, Yuwei, et al.
Pubblicazione: (2025)
Cinéaste: A Fine-grained Contextual Movie Question Answering Benchmark
di: Shah, Nisarg A., et al.
Pubblicazione: (2025)
di: Shah, Nisarg A., et al.
Pubblicazione: (2025)
ReSpace: Text-Driven Autoregressive 3D Indoor Scene Synthesis and Editing
di: Bucher, Martin JJ., et al.
Pubblicazione: (2025)
di: Bucher, Martin JJ., et al.
Pubblicazione: (2025)
Documenti analoghi
-
StoryReasoning Dataset: Using Chain-of-Thought for Scene Understanding and Grounded Story Generation
di: Oliveira, Daniel A. P., et al.
Pubblicazione: (2025) -
PhysicsArena: The First Multimodal Physics Reasoning Benchmark Exploring Variable, Process, and Solution Dimensions
di: Dai, Song, et al.
Pubblicazione: (2025) -
Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning
di: Tong, Jingqi, et al.
Pubblicazione: (2025) -
ICG: Improving Cover Image Generation via MLLM-based Prompting and Personalized Preference Alignment
di: Bian, Zhipeng, et al.
Pubblicazione: (2026) -
Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning
di: Yang, Shan
Pubblicazione: (2026)