Finding the Right Moment: Human-Assisted Trailer Creation via Task Composition
Fuente:
arXiv
Saved in:
| Main Authors: | Papalampidi, Pinelopi, Keller, Frank, Lapata, Mirella |
|---|---|
| Format: | Preprint |
| Published: |
2021
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Parameter-free Video Segmentation for Vision and Language Understanding
by: Mahon, Louis, et al.
Published: (2025)
by: Mahon, Louis, et al.
Published: (2025)
Memory Consolidation Enables Long-Context Video Understanding
by: Balažević, Ivana, et al.
Published: (2024)
by: Balažević, Ivana, et al.
Published: (2024)
ScreenWriter: Automatic Screenplay Generation and Movie Summarisation
by: Mahon, Louis, et al.
Published: (2024)
by: Mahon, Louis, et al.
Published: (2024)
Dynamic Classifier-Free Diffusion Guidance via Online Feedback
by: Papalampidi, Pinelopi, et al.
Published: (2025)
by: Papalampidi, Pinelopi, et al.
Published: (2025)
Find the Cliffhanger: Multi-Modal Trailerness in Soap Operas
by: Bretti, Carlo, et al.
Published: (2024)
by: Bretti, Carlo, et al.
Published: (2024)
TABLET: A Large-Scale Dataset for Robust Visual Table Understanding
by: Alonso, Iñigo, et al.
Published: (2025)
by: Alonso, Iñigo, et al.
Published: (2025)
Meta-Adaptive Prompt Distillation for Few-Shot Visual Question Answering
by: Gupta, Akash, et al.
Published: (2025)
by: Gupta, Akash, et al.
Published: (2025)
Revisiting Text-to-Image Evaluation with Gecko: On Metrics, Prompts, and Human Ratings
by: Wiles, Olivia, et al.
Published: (2024)
by: Wiles, Olivia, et al.
Published: (2024)
A Simple Recipe for Contrastively Pre-training Video-First Encoders Beyond 16 Frames
by: Papalampidi, Pinelopi, et al.
Published: (2023)
by: Papalampidi, Pinelopi, et al.
Published: (2023)
Towards Automated Movie Trailer Generation
by: Argaw, Dawit Mureja, et al.
Published: (2024)
by: Argaw, Dawit Mureja, et al.
Published: (2024)
Self-Paced and Self-Corrective Masked Prediction for Movie Trailer Generation
by: Zhu, Sidan, et al.
Published: (2025)
by: Zhu, Sidan, et al.
Published: (2025)
BEAT: Rhythm-Elastic Alignment for Agentic Music-guided Movie Trailer Generation
by: Wang, Yutong, et al.
Published: (2026)
by: Wang, Yutong, et al.
Published: (2026)
Finding Visual Task Vectors
by: Hojel, Alberto, et al.
Published: (2024)
by: Hojel, Alberto, et al.
Published: (2024)
PosterOmni: Generalized Artistic Poster Creation via Task Distillation and Unified Reward Feedback
by: Chen, Sixiang, et al.
Published: (2026)
by: Chen, Sixiang, et al.
Published: (2026)
InfiniHuman: Infinite 3D Human Creation with Precise Control
by: Xue, Yuxuan, et al.
Published: (2025)
by: Xue, Yuxuan, et al.
Published: (2025)
How to Correctly Make Mistakes: A Framework for Constructing and Benchmarking Mistake Aware Egocentric Procedural Videos
by: Loginova, Olga, et al.
Published: (2026)
by: Loginova, Olga, et al.
Published: (2026)
Human-3Diffusion: Realistic Avatar Creation via Explicit 3D Consistent Diffusion Models
by: Xue, Yuxuan, et al.
Published: (2024)
by: Xue, Yuxuan, et al.
Published: (2024)
SemanticMoments: Training-Free Motion Similarity via Third Moment Features
by: Huberman, Saar, et al.
Published: (2026)
by: Huberman, Saar, et al.
Published: (2026)
MomentSeeker: A Task-Oriented Benchmark For Long-Video Moment Retrieval
by: Yuan, Huaying, et al.
Published: (2025)
by: Yuan, Huaying, et al.
Published: (2025)
Finding Optimal Video Moment without Training: Gaussian Boundary Optimization for Weakly Supervised Video Grounding
by: Kim, Sunoh, et al.
Published: (2026)
by: Kim, Sunoh, et al.
Published: (2026)
What Is That Talk About? A Video-to-Text Summarization Dataset for Scientific Presentations
by: Liu, Dongqi, et al.
Published: (2025)
by: Liu, Dongqi, et al.
Published: (2025)
ComboVerse: Compositional 3D Assets Creation Using Spatially-Aware Diffusion Guidance
by: Chen, Yongwei, et al.
Published: (2024)
by: Chen, Yongwei, et al.
Published: (2024)
PRISM: Perceptual Recognition for Identifying Standout Moments in Human-Centric Keyframe Extraction
by: Cakmak, Mert Can, et al.
Published: (2025)
by: Cakmak, Mert Can, et al.
Published: (2025)
Moment and Highlight Detection via MLLM Frame Segmentation
by: Jiwanta, I Putu Andika Bagas, et al.
Published: (2025)
by: Jiwanta, I Putu Andika Bagas, et al.
Published: (2025)
Background-aware Moment Detection for Video Moment Retrieval
by: Jung, Minjoon, et al.
Published: (2023)
by: Jung, Minjoon, et al.
Published: (2023)
Task-Oriented Human Grasp Synthesis via Context- and Task-Aware Diffusers
by: Liu, An-Lun, et al.
Published: (2025)
by: Liu, An-Lun, et al.
Published: (2025)
TR-DETR: Task-Reciprocal Transformer for Joint Moment Retrieval and Highlight Detection
by: Sun, Hao, et al.
Published: (2024)
by: Sun, Hao, et al.
Published: (2024)
Video Creation by Demonstration
by: Sun, Yihong, et al.
Published: (2024)
by: Sun, Yihong, et al.
Published: (2024)
Moment of Untruth: Dealing with Negative Queries in Video Moment Retrieval
by: Flanagan, Kevin, et al.
Published: (2025)
by: Flanagan, Kevin, et al.
Published: (2025)
Follow-Your-Creation: Empowering 4D Creation through Video Inpainting
by: Ma, Yue, et al.
Published: (2025)
by: Ma, Yue, et al.
Published: (2025)
VDOT: Efficient Unified Video Creation via Optimal Transport Distillation
by: Wang, Yutong, et al.
Published: (2025)
by: Wang, Yutong, et al.
Published: (2025)
Task-Driven Exploration: Decoupling and Inter-Task Feedback for Joint Moment Retrieval and Highlight Detection
by: Yang, Jin, et al.
Published: (2024)
by: Yang, Jin, et al.
Published: (2024)
MomentSeg: Moment-Centric Sampling for Enhanced Video Pixel Understanding
by: Dai, Ming, et al.
Published: (2025)
by: Dai, Ming, et al.
Published: (2025)
On the Robustness of Language Guidance for Low-Level Vision Tasks: Findings from Depth Estimation
by: Chatterjee, Agneet, et al.
Published: (2024)
by: Chatterjee, Agneet, et al.
Published: (2024)
CapHuman: Capture Your Moments in Parallel Universes
by: Liang, Chao, et al.
Published: (2024)
by: Liang, Chao, et al.
Published: (2024)
SpectralSplats: Robust Differentiable Tracking via Spectral Moment Supervision
by: Rimon, Avigail Cohen, et al.
Published: (2026)
by: Rimon, Avigail Cohen, et al.
Published: (2026)
MomentsNeRF: Leveraging Orthogonal Moments for Few-Shot Neural Rendering
by: AlMughrabi, Ahmad, et al.
Published: (2024)
by: AlMughrabi, Ahmad, et al.
Published: (2024)
MMTrail: A Multimodal Trailer Video Dataset with Language and Music Descriptions
by: Chi, Xiaowei, et al.
Published: (2024)
by: Chi, Xiaowei, et al.
Published: (2024)
Text-to-Edit: Controllable End-to-End Video Ad Creation via Multimodal LLMs
by: Cheng, Dabing, et al.
Published: (2025)
by: Cheng, Dabing, et al.
Published: (2025)
ACE++: Instruction-Based Image Creation and Editing via Context-Aware Content Filling
by: Mao, Chaojie, et al.
Published: (2025)
by: Mao, Chaojie, et al.
Published: (2025)
Similar Items
-
Parameter-free Video Segmentation for Vision and Language Understanding
by: Mahon, Louis, et al.
Published: (2025) -
Memory Consolidation Enables Long-Context Video Understanding
by: Balažević, Ivana, et al.
Published: (2024) -
ScreenWriter: Automatic Screenplay Generation and Movie Summarisation
by: Mahon, Louis, et al.
Published: (2024) -
Dynamic Classifier-Free Diffusion Guidance via Online Feedback
by: Papalampidi, Pinelopi, et al.
Published: (2025) -
Find the Cliffhanger: Multi-Modal Trailerness in Soap Operas
by: Bretti, Carlo, et al.
Published: (2024)