A Modular Approach for Multimodal Summarization of TV Shows
Fuente:
arXiv
Saved in:
| Main Authors: | Mahon, Louis, Lapata, Mirella |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Context-Aware Hierarchical Merging for Long Document Summarization
by: Ou, Litu, et al.
Published: (2025)
by: Ou, Litu, et al.
Published: (2025)
Debating for Better Reasoning: An Unsupervised Multimodal Approach
by: Adhikari, Ashutosh, et al.
Published: (2025)
by: Adhikari, Ashutosh, et al.
Published: (2025)
Hierarchical Indexing for Retrieval-Augmented Opinion Summarization
by: Hosking, Tom, et al.
Published: (2024)
by: Hosking, Tom, et al.
Published: (2024)
Parameter-free Video Segmentation for Vision and Language Understanding
by: Mahon, Louis, et al.
Published: (2025)
by: Mahon, Louis, et al.
Published: (2025)
ScreenWriter: Automatic Screenplay Generation and Movie Summarisation
by: Mahon, Louis, et al.
Published: (2024)
by: Mahon, Louis, et al.
Published: (2024)
Explanatory Summarization with Discourse-Driven Planning
by: Liu, Dongqi, et al.
Published: (2025)
by: Liu, Dongqi, et al.
Published: (2025)
AMBROSIA: A Benchmark for Parsing Ambiguous Questions into Database Queries
by: Saparina, Irina, et al.
Published: (2024)
by: Saparina, Irina, et al.
Published: (2024)
What Is That Talk About? A Video-to-Text Summarization Dataset for Scientific Presentations
by: Liu, Dongqi, et al.
Published: (2025)
by: Liu, Dongqi, et al.
Published: (2025)
Decomposed Opinion Summarization with Verified Aspect-Aware Modules
by: Li, Miao, et al.
Published: (2025)
by: Li, Miao, et al.
Published: (2025)
CHIRON: Rich Character Representations in Long-Form Narratives
by: Gurung, Alexander, et al.
Published: (2024)
by: Gurung, Alexander, et al.
Published: (2024)
Integrating Large Language Models with Graph-based Reasoning for Conversational Question Answering
by: Jain, Parag, et al.
Published: (2024)
by: Jain, Parag, et al.
Published: (2024)
Improving Generalization in Semantic Parsing by Increasing Natural Language Variation
by: Saparina, Irina, et al.
Published: (2024)
by: Saparina, Irina, et al.
Published: (2024)
Learning to Reason for Long-Form Story Generation
by: Gurung, Alexander, et al.
Published: (2025)
by: Gurung, Alexander, et al.
Published: (2025)
mFACE: Multilingual Summarization with Factual Consistency Evaluation
by: Aharoni, Roee, et al.
Published: (2022)
by: Aharoni, Roee, et al.
Published: (2022)
Reasoning about Intent for Ambiguous Requests
by: Saparina, Irina, et al.
Published: (2025)
by: Saparina, Irina, et al.
Published: (2025)
Disambiguate First, Parse Later: Generating Interpretations for Ambiguity Resolution in Semantic Parsing
by: Saparina, Irina, et al.
Published: (2025)
by: Saparina, Irina, et al.
Published: (2025)
K*-Means: A Parameter-free Clustering Algorithm
by: Mahon, Louis, et al.
Published: (2025)
by: Mahon, Louis, et al.
Published: (2025)
Uncertainty Quantification in Retrieval Augmented Question Answering
by: Perez-Beltrachini, Laura, et al.
Published: (2025)
by: Perez-Beltrachini, Laura, et al.
Published: (2025)
PixT3: Pixel-based Table-To-Text Generation
by: Alonso, Iñigo, et al.
Published: (2023)
by: Alonso, Iñigo, et al.
Published: (2023)
Lightweight Latent Reasoning for Narrative Tasks
by: Gurung, Alexander, et al.
Published: (2025)
by: Gurung, Alexander, et al.
Published: (2025)
Archer: A Human-Labeled Text-to-SQL Dataset with Arithmetic, Commonsense and Hypothetical Reasoning
by: Zheng, Danna, et al.
Published: (2024)
by: Zheng, Danna, et al.
Published: (2024)
Low-Rank Adaptation for Multilingual Summarization: An Empirical Study
by: Whitehouse, Chenxi, et al.
Published: (2023)
by: Whitehouse, Chenxi, et al.
Published: (2023)
Little Red Riding Hood Goes Around the Globe:Crosslingual Story Planning and Generation with Large Language Models
by: Razumovskaia, Evgeniia, et al.
Published: (2022)
by: Razumovskaia, Evgeniia, et al.
Published: (2022)
Generating Visual Stories with Grounded and Coreferent Characters
by: Liu, Danyang, et al.
Published: (2024)
by: Liu, Danyang, et al.
Published: (2024)
$μ$PLAN: Summarizing using a Content Plan as Cross-Lingual Bridge
by: Huot, Fantine, et al.
Published: (2023)
by: Huot, Fantine, et al.
Published: (2023)
BookWorm: A Dataset for Character Description and Analysis
by: Papoudakis, Argyrios, et al.
Published: (2024)
by: Papoudakis, Argyrios, et al.
Published: (2024)
Compositional Generalisation for Explainable Hate Speech Detection
by: Calabrese, Agostina, et al.
Published: (2025)
by: Calabrese, Agostina, et al.
Published: (2025)
Help Me Write a Story: Evaluating LLMs' Ability to Generate Writing Feedback
by: Rashkin, Hannah, et al.
Published: (2025)
by: Rashkin, Hannah, et al.
Published: (2025)
How Reliable are LLMs as Knowledge Bases? Re-thinking Facutality and Consistency
by: Zheng, Danna, et al.
Published: (2024)
by: Zheng, Danna, et al.
Published: (2024)
Meta-Adaptive Prompt Distillation for Few-Shot Visual Question Answering
by: Gupta, Akash, et al.
Published: (2025)
by: Gupta, Akash, et al.
Published: (2025)
TABLET: A Large-Scale Dataset for Robust Visual Table Understanding
by: Alonso, Iñigo, et al.
Published: (2025)
by: Alonso, Iñigo, et al.
Published: (2025)
Think Before you Write: QA-Guided Reasoning for Character Descriptions in Books
by: Papoudakis, Argyrios, et al.
Published: (2026)
by: Papoudakis, Argyrios, et al.
Published: (2026)
TrustScore: Reference-Free Evaluation of LLM Response Trustworthiness
by: Zheng, Danna, et al.
Published: (2024)
by: Zheng, Danna, et al.
Published: (2024)
Masking in Multi-hop QA: An Analysis of How Language Models Perform with Context Permutation
by: Huang, Wenyu, et al.
Published: (2025)
by: Huang, Wenyu, et al.
Published: (2025)
Long-Form Information Alignment Evaluation Beyond Atomic Facts
by: Zheng, Danna, et al.
Published: (2025)
by: Zheng, Danna, et al.
Published: (2025)
SimLM: Can Language Models Infer Parameters of Physical Systems?
by: Memery, Sean, et al.
Published: (2023)
by: Memery, Sean, et al.
Published: (2023)
Evaluating LLMs for Targeted Concept Simplification for Domain-Specific Texts
by: Asthana, Sumit, et al.
Published: (2024)
by: Asthana, Sumit, et al.
Published: (2024)
Long-Context Reasoning Through Proxy-Based Chain-of-Thought Tuning
by: Li, Miao, et al.
Published: (2026)
by: Li, Miao, et al.
Published: (2026)
GraphLit: Learning Text-Enriched Dynamic Character Network Representations for Literary Study
by: Michel, Gaspard, et al.
Published: (2026)
by: Michel, Gaspard, et al.
Published: (2026)
MT-PingEval: Evaluating Multi-Turn Collaboration with Private Information Games
by: Eisenstein, Jacob, et al.
Published: (2026)
by: Eisenstein, Jacob, et al.
Published: (2026)
Similar Items
-
Context-Aware Hierarchical Merging for Long Document Summarization
by: Ou, Litu, et al.
Published: (2025) -
Debating for Better Reasoning: An Unsupervised Multimodal Approach
by: Adhikari, Ashutosh, et al.
Published: (2025) -
Hierarchical Indexing for Retrieval-Augmented Opinion Summarization
by: Hosking, Tom, et al.
Published: (2024) -
Parameter-free Video Segmentation for Vision and Language Understanding
by: Mahon, Louis, et al.
Published: (2025) -
ScreenWriter: Automatic Screenplay Generation and Movie Summarisation
by: Mahon, Louis, et al.
Published: (2024)