ScreenWriter: Automatic Screenplay Generation and Movie Summarisation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mahon, Louis, Lapata, Mirella |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Video Summarisation with Incident and Context Information using Generative AI
von: De Silva, Ulindu, et al.
Veröffentlicht: (2025)
von: De Silva, Ulindu, et al.
Veröffentlicht: (2025)
SoccerHigh: A Benchmark Dataset for Automatic Soccer Video Summarization
von: Díaz-Juan, Artur, et al.
Veröffentlicht: (2025)
von: Díaz-Juan, Artur, et al.
Veröffentlicht: (2025)
Intelligent Director: An Automatic Framework for Dynamic Visual Composition using ChatGPT
von: Zheng, Sixiao, et al.
Veröffentlicht: (2024)
von: Zheng, Sixiao, et al.
Veröffentlicht: (2024)
Automatic Recognition of Food Ingestion Environment from the AIM-2 Wearable Sensor
von: Huang, Yuning, et al.
Veröffentlicht: (2024)
von: Huang, Yuning, et al.
Veröffentlicht: (2024)
Parameter-free Video Segmentation for Vision and Language Understanding
von: Mahon, Louis, et al.
Veröffentlicht: (2025)
von: Mahon, Louis, et al.
Veröffentlicht: (2025)
Movie Trailer Genre Classification Using Multimodal Pretrained Features
von: Sulun, Serkan, et al.
Veröffentlicht: (2024)
von: Sulun, Serkan, et al.
Veröffentlicht: (2024)
A Survey on Generative AI and LLM for Video Generation, Understanding, and Streaming
von: Zhou, Pengyuan, et al.
Veröffentlicht: (2024)
von: Zhou, Pengyuan, et al.
Veröffentlicht: (2024)
Interactive Video Generation via Domain Adaptation
von: Rawal, Ishaan, et al.
Veröffentlicht: (2025)
von: Rawal, Ishaan, et al.
Veröffentlicht: (2025)
Kandinsky 3: Text-to-Image Synthesis for Multifunctional Generative Framework
von: Arkhipkin, Vladimir, et al.
Veröffentlicht: (2024)
von: Arkhipkin, Vladimir, et al.
Veröffentlicht: (2024)
ReCorD: Reasoning and Correcting Diffusion for HOI Generation
von: Jiang-Lin, Jian-Yu, et al.
Veröffentlicht: (2024)
von: Jiang-Lin, Jian-Yu, et al.
Veröffentlicht: (2024)
Revolutionizing Text-to-Image Retrieval as Autoregressive Token-to-Voken Generation
von: Li, Yongqi, et al.
Veröffentlicht: (2024)
von: Li, Yongqi, et al.
Veröffentlicht: (2024)
Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation
von: Yu, Lijun, et al.
Veröffentlicht: (2023)
von: Yu, Lijun, et al.
Veröffentlicht: (2023)
Distilling Generative-Discriminative Representations for Very Low-Resolution Face Recognition
von: Zhang, Junzheng, et al.
Veröffentlicht: (2024)
von: Zhang, Junzheng, et al.
Veröffentlicht: (2024)
Beyond Audio and Pose: A General-Purpose Framework for Video Synchronization
von: Shin, Yosub, et al.
Veröffentlicht: (2025)
von: Shin, Yosub, et al.
Veröffentlicht: (2025)
FakeParts: a New Family of AI-Generated DeepFakes
von: Liu, Ziyi, et al.
Veröffentlicht: (2025)
von: Liu, Ziyi, et al.
Veröffentlicht: (2025)
Relational Retrieval: Leveraging Known-Novel Interactions for Generalized Category Discovery
von: Xu, Yulin, et al.
Veröffentlicht: (2026)
von: Xu, Yulin, et al.
Veröffentlicht: (2026)
Controllable Audio-Visual Viewpoint Generation from 360° Spatial Information
von: Marinoni, Christian, et al.
Veröffentlicht: (2025)
von: Marinoni, Christian, et al.
Veröffentlicht: (2025)
EvAnimate: Event-conditioned Image-to-Video Generation for Human Animation
von: Qu, Qiang, et al.
Veröffentlicht: (2025)
von: Qu, Qiang, et al.
Veröffentlicht: (2025)
UniVid: Pyramid Diffusion Model for High Quality Video Generation
von: Xiao, Xinyu, et al.
Veröffentlicht: (2026)
von: Xiao, Xinyu, et al.
Veröffentlicht: (2026)
AlignDiT: Multimodal Aligned Diffusion Transformer for Synchronized Speech Generation
von: Choi, Jeongsoo, et al.
Veröffentlicht: (2025)
von: Choi, Jeongsoo, et al.
Veröffentlicht: (2025)
MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing
von: Zheng, Junjie, et al.
Veröffentlicht: (2025)
von: Zheng, Junjie, et al.
Veröffentlicht: (2025)
Terrain Diffusion Network: Climatic-Aware Terrain Generation with Geological Sketch Guidance
von: Hu, Zexin, et al.
Veröffentlicht: (2023)
von: Hu, Zexin, et al.
Veröffentlicht: (2023)
Moiré Video Authentication: A Physical Signature Against AI Video Generation
von: Qing, Yuan, et al.
Veröffentlicht: (2026)
von: Qing, Yuan, et al.
Veröffentlicht: (2026)
TIDE : Temporal-Aware Sparse Autoencoders for Interpretable Diffusion Transformers in Image Generation
von: Huang, Victor Shea-Jay, et al.
Veröffentlicht: (2025)
von: Huang, Victor Shea-Jay, et al.
Veröffentlicht: (2025)
MotionDreamer: One-to-Many Motion Synthesis with Localized Generative Masked Transformer
von: Wang, Yilin, et al.
Veröffentlicht: (2025)
von: Wang, Yilin, et al.
Veröffentlicht: (2025)
Autoregressive Image Generation with Linear Complexity: A Spatial-Aware Decay Perspective
von: Mao, Yuxin, et al.
Veröffentlicht: (2025)
von: Mao, Yuxin, et al.
Veröffentlicht: (2025)
Omni-Dish: Photorealistic and Faithful Image Generation and Editing for Arbitrary Chinese Dishes
von: Liu, Huijie, et al.
Veröffentlicht: (2025)
von: Liu, Huijie, et al.
Veröffentlicht: (2025)
AutoAWG: Adverse Weather Generation with Adaptive Multi-Controls for Automotive Videos
von: Hu, Jiagao, et al.
Veröffentlicht: (2026)
von: Hu, Jiagao, et al.
Veröffentlicht: (2026)
Fashion-RAG: Multimodal Fashion Image Editing via Retrieval-Augmented Generation
von: Sanguigni, Fulvio, et al.
Veröffentlicht: (2025)
von: Sanguigni, Fulvio, et al.
Veröffentlicht: (2025)
Official-NV: An LLM-Generated News Video Dataset for Multimodal Fake News Detection
von: Wang, Yihao, et al.
Veröffentlicht: (2024)
von: Wang, Yihao, et al.
Veröffentlicht: (2024)
A User-Friendly Framework for Generating Model-Preferred Prompts in Text-to-Image Synthesis
von: Hei, Nailei, et al.
Veröffentlicht: (2024)
von: Hei, Nailei, et al.
Veröffentlicht: (2024)
Q-Bench: A Benchmark for General-Purpose Foundation Models on Low-level Vision
von: Wu, Haoning, et al.
Veröffentlicht: (2023)
von: Wu, Haoning, et al.
Veröffentlicht: (2023)
Spotlighting Partially Visible Cinematic Language for Video-to-Audio Generation via Self-distillation
von: Huang, Feizhen, et al.
Veröffentlicht: (2025)
von: Huang, Feizhen, et al.
Veröffentlicht: (2025)
ELIQ: A Label-Free Framework for Quality Assessment of Evolving AI-Generated Images
von: Li, Xinyue, et al.
Veröffentlicht: (2026)
von: Li, Xinyue, et al.
Veröffentlicht: (2026)
OmniAvatar: Efficient Audio-Driven Avatar Video Generation with Adaptive Body Animation
von: Gan, Qijun, et al.
Veröffentlicht: (2025)
von: Gan, Qijun, et al.
Veröffentlicht: (2025)
Image is All You Need to Empower Large-scale Diffusion Models for In-Domain Generation
von: Cao, Pu, et al.
Veröffentlicht: (2023)
von: Cao, Pu, et al.
Veröffentlicht: (2023)
Test-Time Self-Adaptive Conditioning for Stable Audio-Driven Talking-Head Generation
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2026)
von: Zhang, Zhicheng, et al.
Veröffentlicht: (2026)
StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation
von: Wu, Yi, et al.
Veröffentlicht: (2025)
von: Wu, Yi, et al.
Veröffentlicht: (2025)
Long Video Diffusion Generation with Segmented Cross-Attention and Content-Rich Video Data Curation
von: Yan, Xin, et al.
Veröffentlicht: (2024)
von: Yan, Xin, et al.
Veröffentlicht: (2024)
MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance
von: Zhang, Yuang, et al.
Veröffentlicht: (2024)
von: Zhang, Yuang, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Video Summarisation with Incident and Context Information using Generative AI
von: De Silva, Ulindu, et al.
Veröffentlicht: (2025) -
SoccerHigh: A Benchmark Dataset for Automatic Soccer Video Summarization
von: Díaz-Juan, Artur, et al.
Veröffentlicht: (2025) -
Intelligent Director: An Automatic Framework for Dynamic Visual Composition using ChatGPT
von: Zheng, Sixiao, et al.
Veröffentlicht: (2024) -
Automatic Recognition of Food Ingestion Environment from the AIM-2 Wearable Sensor
von: Huang, Yuning, et al.
Veröffentlicht: (2024) -
Parameter-free Video Segmentation for Vision and Language Understanding
von: Mahon, Louis, et al.
Veröffentlicht: (2025)