Generating Narrated Lecture Videos from Slides with Synchronized Highlights
Fuente:
arXiv
Salvato in:
| Autore principale: | Holmberg, Alexander |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
AI-Generated Lecture Slides for Improving Slide Element Detection and Retrieval
di: Maniyar, Suyash, et al.
Pubblicazione: (2025)
di: Maniyar, Suyash, et al.
Pubblicazione: (2025)
Script-to-Slide Grounding: Grounding Script Sentences to Slide Objects for Automatic Instructional Video Generation
di: Suzuki, Rena, et al.
Pubblicazione: (2026)
di: Suzuki, Rena, et al.
Pubblicazione: (2026)
Unleash the Potential of CLIP for Video Highlight Detection
di: Han, Donghoon, et al.
Pubblicazione: (2024)
di: Han, Donghoon, et al.
Pubblicazione: (2024)
Controllable Contextualized Image Captioning: Directing the Visual Narrative through User-Defined Highlights
di: Mao, Shunqi, et al.
Pubblicazione: (2024)
di: Mao, Shunqi, et al.
Pubblicazione: (2024)
Video-MMLU: A Massive Multi-Discipline Lecture Understanding Benchmark
di: Song, Enxin, et al.
Pubblicazione: (2025)
di: Song, Enxin, et al.
Pubblicazione: (2025)
Learning Multi-modal Representations by Watching Hundreds of Surgical Video Lectures
di: Yuan, Kun, et al.
Pubblicazione: (2023)
di: Yuan, Kun, et al.
Pubblicazione: (2023)
HIPPO-Video: Simulating Watch Histories with Large Language Models for Personalized Video Highlighting
di: Lee, Jeongeun, et al.
Pubblicazione: (2025)
di: Lee, Jeongeun, et al.
Pubblicazione: (2025)
SlideCoder: Layout-aware RAG-enhanced Hierarchical Slide Generation from Design
di: Tang, Wenxin, et al.
Pubblicazione: (2025)
di: Tang, Wenxin, et al.
Pubblicazione: (2025)
SeqBench: Benchmarking Sequential Narrative Generation in Text-to-Video Models
di: Tang, Zhengxu, et al.
Pubblicazione: (2025)
di: Tang, Zhengxu, et al.
Pubblicazione: (2025)
SmartDirector: Keyframe-Conditioned Cinematic Video Generation with Narrative Pacing Control
di: Zhang, Zhida, et al.
Pubblicazione: (2026)
di: Zhang, Zhida, et al.
Pubblicazione: (2026)
Unsupervised Transcript-assisted Video Summarization and Highlight Detection
di: Barbakos, Spyros, et al.
Pubblicazione: (2025)
di: Barbakos, Spyros, et al.
Pubblicazione: (2025)
VideoLights: Feature Refinement and Cross-Task Alignment Transformer for Joint Video Highlight Detection and Moment Retrieval
di: Paul, Dhiman, et al.
Pubblicazione: (2024)
di: Paul, Dhiman, et al.
Pubblicazione: (2024)
Beyond Audio and Pose: A General-Purpose Framework for Video Synchronization
di: Shin, Yosub, et al.
Pubblicazione: (2025)
di: Shin, Yosub, et al.
Pubblicazione: (2025)
Watch Video, Catch Keyword: Context-aware Keyword Attention for Moment Retrieval and Highlight Detection
di: Um, Sung Jin, et al.
Pubblicazione: (2025)
di: Um, Sung Jin, et al.
Pubblicazione: (2025)
Automated Detection of Sport Highlights from Audio and Video Sources
di: Della Santa, Francesco, et al.
Pubblicazione: (2025)
di: Della Santa, Francesco, et al.
Pubblicazione: (2025)
GCAgent: Long-Video Understanding via Schematic and Narrative Episodic Memory
di: Yeo, Jeong Hun, et al.
Pubblicazione: (2025)
di: Yeo, Jeong Hun, et al.
Pubblicazione: (2025)
RACCooN: A Versatile Instructional Video Editing Framework with Auto-Generated Narratives
di: Yoon, Jaehong, et al.
Pubblicazione: (2024)
di: Yoon, Jaehong, et al.
Pubblicazione: (2024)
WSI-VQA: Interpreting Whole Slide Images by Generative Visual Question Answering
di: Chen, Pingyi, et al.
Pubblicazione: (2024)
di: Chen, Pingyi, et al.
Pubblicazione: (2024)
SlideChat: A Large Vision-Language Assistant for Whole-Slide Pathology Image Understanding
di: Chen, Ying, et al.
Pubblicazione: (2024)
di: Chen, Ying, et al.
Pubblicazione: (2024)
StochSync: Stochastic Diffusion Synchronization for Image Generation in Arbitrary Spaces
di: Yeo, Kyeongmin, et al.
Pubblicazione: (2025)
di: Yeo, Kyeongmin, et al.
Pubblicazione: (2025)
Tuning-Free Multi-Event Long Video Generation via Synchronized Coupled Sampling
di: Kim, Subin, et al.
Pubblicazione: (2025)
di: Kim, Subin, et al.
Pubblicazione: (2025)
QEVA: A Reference-Free Evaluation Metric for Narrative Video Summarization with Multimodal Question Answering
di: Jung, Woojun, et al.
Pubblicazione: (2026)
di: Jung, Woojun, et al.
Pubblicazione: (2026)
Masked Generative Video-to-Audio Transformers with Enhanced Synchronicity
di: Pascual, Santiago, et al.
Pubblicazione: (2024)
di: Pascual, Santiago, et al.
Pubblicazione: (2024)
VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos
di: Yu, Jiashuo, et al.
Pubblicazione: (2025)
di: Yu, Jiashuo, et al.
Pubblicazione: (2025)
Storynizor: Consistent Story Generation via Inter-Frame Synchronized and Shuffled ID Injection
di: Ma, Yuhang, et al.
Pubblicazione: (2024)
di: Ma, Yuhang, et al.
Pubblicazione: (2024)
Cross-Patient Pseudo Bags Generation and Curriculum Contrastive Learning for Imbalanced Multiclassification of Whole Slide Image
di: Wu, Yonghuang, et al.
Pubblicazione: (2024)
di: Wu, Yonghuang, et al.
Pubblicazione: (2024)
STORYANCHORS: Generating Consistent Multi-Scene Story Frames for Long-Form Narratives
di: Wang, Bo, et al.
Pubblicazione: (2025)
di: Wang, Bo, et al.
Pubblicazione: (2025)
Directing the Narrative: A Finetuning Method for Controlling Coherence and Style in Story Generation
di: Zhang, Jianzhang, et al.
Pubblicazione: (2026)
di: Zhang, Jianzhang, et al.
Pubblicazione: (2026)
Reangle-A-Video: 4D Video Generation as Video-to-Video Translation
di: Jeong, Hyeonho, et al.
Pubblicazione: (2025)
di: Jeong, Hyeonho, et al.
Pubblicazione: (2025)
PathNavigate: A Training-Free Pathology Agent with Surprise-Guided Scan and Shared Slide Memory for Whole-Slide Image VQA
di: Yang, Chunze, et al.
Pubblicazione: (2026)
di: Yang, Chunze, et al.
Pubblicazione: (2026)
Video-Infinity: Distributed Long Video Generation
di: Tan, Zhenxiong, et al.
Pubblicazione: (2024)
di: Tan, Zhenxiong, et al.
Pubblicazione: (2024)
Hypergraph Mamba for Efficient Whole Slide Image Understanding
di: Lu, Jiaxuan, et al.
Pubblicazione: (2025)
di: Lu, Jiaxuan, et al.
Pubblicazione: (2025)
Spatial Blindness in Whole-Slide Multiple Instance Learning
di: Li, Xiangyu, et al.
Pubblicazione: (2026)
di: Li, Xiangyu, et al.
Pubblicazione: (2026)
Transcriptomics-guided Slide Representation Learning in Computational Pathology
di: Jaume, Guillaume, et al.
Pubblicazione: (2024)
di: Jaume, Guillaume, et al.
Pubblicazione: (2024)
TARO: Timestep-Adaptive Representation Alignment with Onset-Aware Conditioning for Synchronized Video-to-Audio Synthesis
di: Ton, Tri, et al.
Pubblicazione: (2025)
di: Ton, Tri, et al.
Pubblicazione: (2025)
KPIs 2024 Challenge: Advancing Glomerular Segmentation from Patch- to Slide-Level
di: Deng, Ruining, et al.
Pubblicazione: (2025)
di: Deng, Ruining, et al.
Pubblicazione: (2025)
Video-As-Prompt: Unified Semantic Control for Video Generation
di: Bian, Yuxuan, et al.
Pubblicazione: (2025)
di: Bian, Yuxuan, et al.
Pubblicazione: (2025)
Conditional Video Generation for High-Efficiency Video Compression
di: Yi, Fangqiu, et al.
Pubblicazione: (2025)
di: Yi, Fangqiu, et al.
Pubblicazione: (2025)
Video-Bench: Human-Aligned Video Generation Benchmark
di: Han, Hui, et al.
Pubblicazione: (2025)
di: Han, Hui, et al.
Pubblicazione: (2025)
Video-T1: Test-Time Scaling for Video Generation
di: Liu, Fangfu, et al.
Pubblicazione: (2025)
di: Liu, Fangfu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
AI-Generated Lecture Slides for Improving Slide Element Detection and Retrieval
di: Maniyar, Suyash, et al.
Pubblicazione: (2025) -
Script-to-Slide Grounding: Grounding Script Sentences to Slide Objects for Automatic Instructional Video Generation
di: Suzuki, Rena, et al.
Pubblicazione: (2026) -
Unleash the Potential of CLIP for Video Highlight Detection
di: Han, Donghoon, et al.
Pubblicazione: (2024) -
Controllable Contextualized Image Captioning: Directing the Visual Narrative through User-Defined Highlights
di: Mao, Shunqi, et al.
Pubblicazione: (2024) -
Video-MMLU: A Massive Multi-Discipline Lecture Understanding Benchmark
di: Song, Enxin, et al.
Pubblicazione: (2025)