Vidmento: Creating Video Stories Through Context-Aware Expansion With Generative Video
Fuente:
arXiv
Guardado en:
| Autores principales: | Yeh, Catherine, Truong, Anh, Dontcheva, Mira, Wang, Bryan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
VidTune: Creating Video Soundtracks with Generative Music and Contextual Thumbnails
por: Huh, Mina, et al.
Publicado: (2026)
por: Huh, Mina, et al.
Publicado: (2026)
A Text-Native Interface for Generative Video Authoring
por: Liu, Xingyu Bruce, et al.
Publicado: (2026)
por: Liu, Xingyu Bruce, et al.
Publicado: (2026)
Through the Looking-Glass: AI-Mediated Video Communication Reduces Interpersonal Trust and Confidence in Judgments
por: Fernández, Nelson Navajas, et al.
Publicado: (2026)
por: Fernández, Nelson Navajas, et al.
Publicado: (2026)
Node-Based Editing for Multimodal Generation of Text, Audio, Image, and Video
por: Kyaw, Alexander Htet, et al.
Publicado: (2025)
por: Kyaw, Alexander Htet, et al.
Publicado: (2025)
LAVE: LLM-Powered Agent Assistance and Language Augmentation for Video Editing
por: Wang, Bryan, et al.
Publicado: (2024)
por: Wang, Bryan, et al.
Publicado: (2024)
Focus360: Guiding User Attention in Immersive Videos for VR
por: Silva, Paulo Vitor S., et al.
Publicado: (2026)
por: Silva, Paulo Vitor S., et al.
Publicado: (2026)
DailyLLM: Context-Aware Activity Log Generation Using Multi-Modal Sensors and LLMs
por: Tian, Ye, et al.
Publicado: (2025)
por: Tian, Ye, et al.
Publicado: (2025)
IntentVLM: Open-Vocabulary Intention Recognition through Forward-Inverse Modeling with Video-Language Models
por: Rahimi, Hamed, et al.
Publicado: (2026)
por: Rahimi, Hamed, et al.
Publicado: (2026)
Livia: An Emotion-Aware AR Companion Powered by Modular AI Agents and Progressive Memory Compression
por: Xi, Rui, et al.
Publicado: (2025)
por: Xi, Rui, et al.
Publicado: (2025)
MetaDesigner: Advancing Artistic Typography Through AI-Driven, User-Centric, and Multilingual WordArt Synthesis
por: He, Jun-Yan, et al.
Publicado: (2024)
por: He, Jun-Yan, et al.
Publicado: (2024)
Secure & Personalized Music-to-Video Generation via CHARCHA
por: Agarwal, Mehul, et al.
Publicado: (2025)
por: Agarwal, Mehul, et al.
Publicado: (2025)
Human Aesthetic Preference-Based Large Text-to-Image Model Personalization: Kandinsky Generation as an Example
por: Zhou, Aven-Le, et al.
Publicado: (2024)
por: Zhou, Aven-Le, et al.
Publicado: (2024)
Archiving Body Movements: Collective Generation of Chinese Calligraphy
por: Zhou, Aven Le, et al.
Publicado: (2023)
por: Zhou, Aven Le, et al.
Publicado: (2023)
SimInterview: Transforming Business Education through Large Language Model-Based Simulated Multilingual Interview Training System
por: Nguyen, Truong Thanh Hung, et al.
Publicado: (2025)
por: Nguyen, Truong Thanh Hung, et al.
Publicado: (2025)
Leveraging LLMs to Create a Haptic Devices' Recommendation System
por: Liu, Yang, et al.
Publicado: (2025)
por: Liu, Yang, et al.
Publicado: (2025)
Simulacra Naturae: Generative Ecosystem driven by Agent-Based Simulations and Brain Organoid Collective Intelligence
por: Manoudaki, Nefeli, et al.
Publicado: (2025)
por: Manoudaki, Nefeli, et al.
Publicado: (2025)
Signals of Provenance: Practices & Challenges of Navigating Indicators in AI-Generated Media for Sighted and Blind Individuals
por: Ide, Ayae, et al.
Publicado: (2025)
por: Ide, Ayae, et al.
Publicado: (2025)
DreamLLM-3D: Affective Dream Reliving using Large Language Model and 3D Generative AI
por: Liu, Pinyao, et al.
Publicado: (2025)
por: Liu, Pinyao, et al.
Publicado: (2025)
Code2Video: A Code-centric Paradigm for Educational Video Generation
por: Chen, Yanzhe, et al.
Publicado: (2025)
por: Chen, Yanzhe, et al.
Publicado: (2025)
Creating Disability Story Videos with Generative AI: Motivation, Expression, and Sharing
por: Niu, Shuo, et al.
Publicado: (2026)
por: Niu, Shuo, et al.
Publicado: (2026)
Artic: AI-oriented Real-time Communication for MLLM Video Assistant
por: Wu, Jiangkai, et al.
Publicado: (2026)
por: Wu, Jiangkai, et al.
Publicado: (2026)
ANVIL: Analogies and Videos for Lecturers
por: Noviello, Yuri, et al.
Publicado: (2026)
por: Noviello, Yuri, et al.
Publicado: (2026)
Chain-of-Modality: Learning Manipulation Programs from Multimodal Human Videos with Vision-Language-Models
por: Wang, Chen, et al.
Publicado: (2025)
por: Wang, Chen, et al.
Publicado: (2025)
PodReels: Human-AI Co-Creation of Video Podcast Teasers
por: Wang, Sitong, et al.
Publicado: (2023)
por: Wang, Sitong, et al.
Publicado: (2023)
Proceedings of The third international workshop on eXplainable AI for the Arts (XAIxArts)
por: Ford, Corey, et al.
Publicado: (2025)
por: Ford, Corey, et al.
Publicado: (2025)
DiffMesh: A Motion-aware Diffusion Framework for Human Mesh Recovery from Videos
por: Zheng, Ce, et al.
Publicado: (2023)
por: Zheng, Ce, et al.
Publicado: (2023)
Chat with AI: The Surprising Turn of Real-time Video Communication from Human to AI
por: Wu, Jiangkai, et al.
Publicado: (2025)
por: Wu, Jiangkai, et al.
Publicado: (2025)
Self-supervised Spatio-Temporal Graph Mask-Passing Attention Network for Perceptual Importance Prediction of Multi-point Tactility
por: He, Dazhong, et al.
Publicado: (2024)
por: He, Dazhong, et al.
Publicado: (2024)
A Rhetorical Relations-Based Framework for Tailored Multimedia Document Summarization
por: Maredj, Azze-Eddine, et al.
Publicado: (2024)
por: Maredj, Azze-Eddine, et al.
Publicado: (2024)
An Empirical Evaluation of AI-Powered Non-Player Characters' Perceived Realism and Performance in Virtual Reality Environments
por: Korkiakoski, Mikko, et al.
Publicado: (2025)
por: Korkiakoski, Mikko, et al.
Publicado: (2025)
AMEX: Android Multi-annotation Expo Dataset for Mobile GUI Agents
por: Chai, Yuxiang, et al.
Publicado: (2024)
por: Chai, Yuxiang, et al.
Publicado: (2024)
Computational Analysis of Stress, Depression and Engagement in Mental Health: A Survey
por: Kumar, Puneet, et al.
Publicado: (2024)
por: Kumar, Puneet, et al.
Publicado: (2024)
Virtual Reality and Artificial Intelligence as Psychological Countermeasures in Space and Other Isolated and Confined Environments: A Scoping Review
por: Sharp, Jennifer, et al.
Publicado: (2025)
por: Sharp, Jennifer, et al.
Publicado: (2025)
A Multimedia Analytics Model for the Foundation Model Era
por: Worring, Marcel, et al.
Publicado: (2025)
por: Worring, Marcel, et al.
Publicado: (2025)
MuDoC: An Interactive Multimodal Document-grounded Conversational AI System
por: Taneja, Karan, et al.
Publicado: (2025)
por: Taneja, Karan, et al.
Publicado: (2025)
Rewriting Video: Text-Driven Reauthoring of Video Footage
por: Wang, Sitong, et al.
Publicado: (2026)
por: Wang, Sitong, et al.
Publicado: (2026)
MV-Crafter: An Intelligent System for Music-guided Video Generation
por: Chen, Chuer, et al.
Publicado: (2025)
por: Chen, Chuer, et al.
Publicado: (2025)
"I Can See Forever!": Evaluating Real-time VideoLLMs for Assisting Individuals with Visual Impairments
por: Zhang, Ziyi, et al.
Publicado: (2025)
por: Zhang, Ziyi, et al.
Publicado: (2025)
LLaVA-Scissor: Token Compression with Semantic Connected Components for Video LLMs
por: Sun, Boyuan, et al.
Publicado: (2025)
por: Sun, Boyuan, et al.
Publicado: (2025)
Color When It Counts: Grayscale-Guided Online Triggering for Always-On Streaming Video Sensing
por: Cai, Weitong, et al.
Publicado: (2026)
por: Cai, Weitong, et al.
Publicado: (2026)
Ejemplares similares
-
VidTune: Creating Video Soundtracks with Generative Music and Contextual Thumbnails
por: Huh, Mina, et al.
Publicado: (2026) -
A Text-Native Interface for Generative Video Authoring
por: Liu, Xingyu Bruce, et al.
Publicado: (2026) -
Through the Looking-Glass: AI-Mediated Video Communication Reduces Interpersonal Trust and Confidence in Judgments
por: Fernández, Nelson Navajas, et al.
Publicado: (2026) -
Node-Based Editing for Multimodal Generation of Text, Audio, Image, and Video
por: Kyaw, Alexander Htet, et al.
Publicado: (2025) -
LAVE: LLM-Powered Agent Assistance and Language Augmentation for Video Editing
por: Wang, Bryan, et al.
Publicado: (2024)