Prompt-Driven Agentic Video Editing System: Autonomous Comprehension of Long-Form, Story-Driven Media
Fuente:
arXiv
Saved in:
| Main Authors: | Ding, Zihan, Wang, Xinyi, Chen, Junlong, Kristensson, Per Ola, Shen, Junxiao |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generative AI for Accessible and Inclusive Extended Reality
by: Grubert, Jens, et al.
Published: (2024)
by: Grubert, Jens, et al.
Published: (2024)
Analyzing Multimodal Interaction Strategies for LLM-Assisted Manipulation of 3D Scenes
by: Chen, Junlong, et al.
Published: (2024)
by: Chen, Junlong, et al.
Published: (2024)
Large Language Model-assisted Speech and Pointing Benefits Multiple 3D Object Selection in Virtual Reality
by: Chen, Junlong, et al.
Published: (2024)
by: Chen, Junlong, et al.
Published: (2024)
Encode-Store-Retrieve: Augmenting Human Memory through Language-Encoded Egocentric Perception
by: Shen, Junxiao, et al.
Published: (2023)
by: Shen, Junxiao, et al.
Published: (2023)
ImageTalk: Designing a Multimodal AAC Text Generation System Driven by Image Recognition and Natural Language Generation
by: Yang, Boyin, et al.
Published: (2025)
by: Yang, Boyin, et al.
Published: (2025)
Evaluating the Usability of Microgestures for Text Editing Tasks in Virtual Reality
by: Li, Xiang, et al.
Published: (2025)
by: Li, Xiang, et al.
Published: (2025)
Unbounded: Object-Boundary Interaction in Mixed Reality
by: Lyu, Zhuoyue, et al.
Published: (2025)
by: Lyu, Zhuoyue, et al.
Published: (2025)
Bend It, Aim It, Tap It: Designing an On-Body Disambiguation Mechanism for Curve Selection in Mixed Reality
by: Li, Xiang, et al.
Published: (2025)
by: Li, Xiang, et al.
Published: (2025)
Optimizing Curve-Based Selection with On-Body Surfaces in Virtual Environments
by: Li, Xiang, et al.
Published: (2025)
by: Li, Xiang, et al.
Published: (2025)
Objestures: Everyday Objects Meet Mid-Air Gestures for Expressive Interaction
by: Lyu, Zhuoyue, et al.
Published: (2025)
by: Lyu, Zhuoyue, et al.
Published: (2025)
EnVisionVR: A Scene Interpretation Tool for Visual Accessibility in Virtual Reality
by: Chen, Junlong, et al.
Published: (2025)
by: Chen, Junlong, et al.
Published: (2025)
MyoInteract: A Framework for Fast Prototyping of Biomechanical HCI Tasks using Reinforcement Learning
by: Bhattarai, Ankit, et al.
Published: (2026)
by: Bhattarai, Ankit, et al.
Published: (2026)
How Do We Evaluate Experiences in Immersive Environments?
by: Li, Xiang, et al.
Published: (2026)
by: Li, Xiang, et al.
Published: (2026)
LocoScooter: Designing a Stationary Scooter-Based Locomotion System for Navigation in Virtual Reality
by: He, Wei, et al.
Published: (2026)
by: He, Wei, et al.
Published: (2026)
Optical Tag-Based Neuronavigation and Augmentation System for Non-Invasive Brain Stimulation
by: Hu, Xuyi, et al.
Published: (2026)
by: Hu, Xuyi, et al.
Published: (2026)
Handows: A Palm-Based Interactive Multi-Window Management System in Virtual Reality
by: Wang, Jindu, et al.
Published: (2025)
by: Wang, Jindu, et al.
Published: (2025)
A Multi-Camera Optical Tag Neuronavigation and AR Augmentation Framework for Non-Invasive Brain Stimulation
by: Hu, Xuyi, et al.
Published: (2026)
by: Hu, Xuyi, et al.
Published: (2026)
Lucia: A Temporal Computing Platform for Contextual Intelligence
by: Lin, Weizhe, et al.
Published: (2024)
by: Lin, Weizhe, et al.
Published: (2024)
Swarm manipulation: An efficient and accurate technique for multi-object manipulation in virtual reality
by: Li, Xiang, et al.
Published: (2024)
by: Li, Xiang, et al.
Published: (2024)
Cost-Aware Bayesian Optimization for Prototyping Interactive Devices
by: Langerak, Thomas, et al.
Published: (2026)
by: Langerak, Thomas, et al.
Published: (2026)
What Should We Engineer in Prompts? Training Humans in Requirement-Driven LLM Use
by: Ma, Qianou, et al.
Published: (2024)
by: Ma, Qianou, et al.
Published: (2024)
What Makes a Model Breathe? Understanding Reinforcement Learning Reward Function Design in Biomechanical User Simulation
by: Selder, Hannah, et al.
Published: (2025)
by: Selder, Hannah, et al.
Published: (2025)
Demystifying Reward Design in Reinforcement Learning for Upper Extremity Interaction: Practical Guidelines for Biomechanical Simulations in HCI
by: Selder, Hannah, et al.
Published: (2025)
by: Selder, Hannah, et al.
Published: (2025)
ReachVox: Clutter-free Reachability Visualization for Robot Motion Planning in Virtual Reality
by: Hauck, Steffen, et al.
Published: (2025)
by: Hauck, Steffen, et al.
Published: (2025)
Rewriting Video: Text-Driven Reauthoring of Video Footage
by: Wang, Sitong, et al.
Published: (2026)
by: Wang, Sitong, et al.
Published: (2026)
Accented Character Entry Using Physical Keyboards in Virtual Reality
by: Kalamkar, Snehanjali, et al.
Published: (2024)
by: Kalamkar, Snehanjali, et al.
Published: (2024)
AlphaPIG: The Nicest Way to Prolong Interactive Gestures in Extended Reality
by: Li, Yi, et al.
Published: (2025)
by: Li, Yi, et al.
Published: (2025)
The Stories We Tell About Data: Media Types for Data-Driven Storytelling
by: Zhao, Zhenpeng, et al.
Published: (2022)
by: Zhao, Zhenpeng, et al.
Published: (2022)
Investigating Creation Perspectives and Icon Placement Preferences for On-Body Menus in Virtual Reality
by: Li, Xiang, et al.
Published: (2024)
by: Li, Xiang, et al.
Published: (2024)
StoryNavi: On-Demand Narrative-Driven Reconstruction of Video Play With Generative AI
by: Xu, Alston Lantian, et al.
Published: (2024)
by: Xu, Alston Lantian, et al.
Published: (2024)
Hold Tight: Identifying Behavioral Patterns During Prolonged Work in VR through Video Analysis
by: Biener, Verena, et al.
Published: (2024)
by: Biener, Verena, et al.
Published: (2024)
A Crowdsourced Study of ChatBot Influence in Value-Driven Decision Making Scenarios
by: Wise, Anthony, et al.
Published: (2025)
by: Wise, Anthony, et al.
Published: (2025)
ChatHouseDiffusion: Prompt-Guided Generation and Editing of Floor Plans
by: Qin, Sizhong, et al.
Published: (2024)
by: Qin, Sizhong, et al.
Published: (2024)
StoryLensEdu: Personalized Learning Report Generation through Narrative-Driven Multi-Agent Systems
by: Shen, Leixian, et al.
Published: (2026)
by: Shen, Leixian, et al.
Published: (2026)
Prompt-with-Me: in-IDE Structured Prompt Management for LLM-Driven Software Engineering
by: Li, Ziyou, et al.
Published: (2025)
by: Li, Ziyou, et al.
Published: (2025)
Dynamic Prompt Middleware: Contextual Prompt Refinement Controls for Comprehension Tasks
by: Drosos, Ian, et al.
Published: (2024)
by: Drosos, Ian, et al.
Published: (2024)
Accessible Gesture-Driven Augmented Reality Interaction System
by: Wang, Yikan
Published: (2025)
by: Wang, Yikan
Published: (2025)
Creating Disability Story Videos with Generative AI: Motivation, Expression, and Sharing
by: Niu, Shuo, et al.
Published: (2026)
by: Niu, Shuo, et al.
Published: (2026)
ExpressEdit: Video Editing with Natural Language and Sketching
by: Tilekbay, Bekzat, et al.
Published: (2024)
by: Tilekbay, Bekzat, et al.
Published: (2024)
Vidmento: Creating Video Stories Through Context-Aware Expansion With Generative Video
by: Yeh, Catherine, et al.
Published: (2026)
by: Yeh, Catherine, et al.
Published: (2026)
Similar Items
-
Generative AI for Accessible and Inclusive Extended Reality
by: Grubert, Jens, et al.
Published: (2024) -
Analyzing Multimodal Interaction Strategies for LLM-Assisted Manipulation of 3D Scenes
by: Chen, Junlong, et al.
Published: (2024) -
Large Language Model-assisted Speech and Pointing Benefits Multiple 3D Object Selection in Virtual Reality
by: Chen, Junlong, et al.
Published: (2024) -
Encode-Store-Retrieve: Augmenting Human Memory through Language-Encoded Egocentric Perception
by: Shen, Junxiao, et al.
Published: (2023) -
ImageTalk: Designing a Multimodal AAC Text Generation System Driven by Image Recognition and Natural Language Generation
by: Yang, Boyin, et al.
Published: (2025)