REGen: Multimodal Retrieval-Embedded Generation for Long-to-Short Video Editing
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Weihan, Ma, Yimeng, Huang, Jingyue, Li, Yang, Ma, Wenye, Berg-Kirkpatrick, Taylor, McAuley, Julian, Liang, Paul Pu, Dong, Hao-Wen |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TeaserGen: Generating Teasers for Long Documentaries
by: Xu, Weihan, et al.
Published: (2024)
by: Xu, Weihan, et al.
Published: (2024)
Generating Symbolic Music from Natural Language Prompts using an LLM-Enhanced Dataset
by: Xu, Weihan, et al.
Published: (2024)
by: Xu, Weihan, et al.
Published: (2024)
Video-Guided Text-to-Music Generation Using Public Domain Movie Collections
by: Kim, Haven, et al.
Published: (2025)
by: Kim, Haven, et al.
Published: (2025)
DITTO-2: Distilled Diffusion Inference-Time T-Optimization for Music Generation
by: Novack, Zachary, et al.
Published: (2024)
by: Novack, Zachary, et al.
Published: (2024)
WildFX: A DAW-Powered Pipeline for In-the-Wild Audio FX Graph Modeling
by: Yang, Qihui, et al.
Published: (2025)
by: Yang, Qihui, et al.
Published: (2025)
PDMX: A Large-Scale Public Domain MusicXML Dataset for Symbolic Music Processing
by: Long, Phillip, et al.
Published: (2024)
by: Long, Phillip, et al.
Published: (2024)
MuseTok: Symbolic Music Tokenization for Generation and Semantic Understanding
by: Huang, Jingyue, et al.
Published: (2025)
by: Huang, Jingyue, et al.
Published: (2025)
DITTO: Diffusion Inference-Time T-Optimization for Music Generation
by: Novack, Zachary, et al.
Published: (2024)
by: Novack, Zachary, et al.
Published: (2024)
Steering Autoregressive Music Generation with Recursive Feature Machines
by: Zhao, Daniel, et al.
Published: (2025)
by: Zhao, Daniel, et al.
Published: (2025)
Are you really listening? Boosting Perceptual Awareness in Music-QA Benchmarks
by: Zang, Yongyi, et al.
Published: (2025)
by: Zang, Yongyi, et al.
Published: (2025)
Composer Vector: Style-steering Symbolic Music Generation in a Latent Space
by: Jiang, Xunyi, et al.
Published: (2026)
by: Jiang, Xunyi, et al.
Published: (2026)
LVCHAT: Facilitating Long Video Comprehension
by: Wang, Yu, et al.
Published: (2024)
by: Wang, Yu, et al.
Published: (2024)
Presto! Distilling Steps and Layers for Accelerating Music Generation
by: Novack, Zachary, et al.
Published: (2024)
by: Novack, Zachary, et al.
Published: (2024)
Low-Resource Guidance for Controllable Latent Audio Diffusion
by: Novack, Zachary, et al.
Published: (2026)
by: Novack, Zachary, et al.
Published: (2026)
BiasEdit: Debiasing Stereotyped Language Models via Model Editing
by: Xu, Xin, et al.
Published: (2025)
by: Xu, Xin, et al.
Published: (2025)
Reddit2Deezer: A Scalable Dataset for Real-World Grounded Conversational Music Recommendation
by: Kim, Haven, et al.
Published: (2026)
by: Kim, Haven, et al.
Published: (2026)
Imagery as Inquiry: Exploring A Multimodal Dataset for Conversational Recommendation
by: Yoon, Se-eun, et al.
Published: (2024)
by: Yoon, Se-eun, et al.
Published: (2024)
Inductive Generative Recommendation via Retrieval-based Speculation
by: Ding, Yijie, et al.
Published: (2024)
by: Ding, Yijie, et al.
Published: (2024)
Purely Semantic Indexing for LLM-based Generative Recommendation and Retrieval
by: Zhang, Ruohan, et al.
Published: (2025)
by: Zhang, Ruohan, et al.
Published: (2025)
Embedding-Informed Adaptive Retrieval-Augmented Generation of Large Language Models
by: Huang, Chengkai, et al.
Published: (2024)
by: Huang, Chengkai, et al.
Published: (2024)
Train Once, Deploy Anywhere: Matryoshka Representation Learning for Multimodal Recommendation
by: Wang, Yueqi, et al.
Published: (2024)
by: Wang, Yueqi, et al.
Published: (2024)
Three central limit theorems for the unbounded excursion component of a Gaussian field
by: McAuley, Michael
Published: (2024)
by: McAuley, Michael
Published: (2024)
Children in Custody
by: McAuley, Mary
Published: (2022)
by: McAuley, Mary
Published: (2022)
Politics and the Soviet Union / Mary McAuley
by: McAuley, Mary
by: McAuley, Mary
Limit theorems for non-local functionals of smooth Gaussian fields via quasi-association
by: McAuley, Michael
Published: (2026)
by: McAuley, Michael
Published: (2026)
GSPRec: Temporal-Aware Graph Spectral Filtering for Recommendation
by: Rabiah, Ahmad Bin, et al.
Published: (2025)
by: Rabiah, Ahmad Bin, et al.
Published: (2025)
Extending Input Contexts of Language Models through Training on Segmented Sequences
by: Karypis, Petros, et al.
Published: (2023)
by: Karypis, Petros, et al.
Published: (2023)
On Generalization in Agentic Tool Calling: CoreThink Agentic Reasoner and MAVEN Dataset
by: Bhat, Vishvesh, et al.
Published: (2025)
by: Bhat, Vishvesh, et al.
Published: (2025)
FusID: Modality-Fused Semantic IDs for Generative Music Recommendation
by: Kim, Haven, et al.
Published: (2026)
by: Kim, Haven, et al.
Published: (2026)
Multi-Behavior Generative Recommendation
by: Liu, Zihan, et al.
Published: (2024)
by: Liu, Zihan, et al.
Published: (2024)
Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions
by: Hu, Yuanzhe, et al.
Published: (2025)
by: Hu, Yuanzhe, et al.
Published: (2025)
Retrieval Augmented Conversational Recommendation with Reinforcement Learning
by: Yue, Zhenrui, et al.
Published: (2026)
by: Yue, Zhenrui, et al.
Published: (2026)
Pctx: Tokenizing Personalized Context for Generative Recommendation
by: Zhong, Qiyong, et al.
Published: (2025)
by: Zhong, Qiyong, et al.
Published: (2025)
Fast Text-to-Audio Generation with Adversarial Post-Training
by: Novack, Zachary, et al.
Published: (2025)
by: Novack, Zachary, et al.
Published: (2025)
CoRAL: Collaborative Retrieval-Augmented Large Language Models Improve Long-tail Recommendation
by: Wu, Junda, et al.
Published: (2024)
by: Wu, Junda, et al.
Published: (2024)
Do Text Edits Generalize to Visual Generation? Benchmarking Cross-Modal Knowledge Editing in UMMs
by: Gao, Xin, et al.
Published: (2026)
by: Gao, Xin, et al.
Published: (2026)
CoreThink: A Symbolic Reasoning Layer to reason over Long Horizon Tasks with LLMs
by: Vaghasiya, Jay, et al.
Published: (2025)
by: Vaghasiya, Jay, et al.
Published: (2025)
Bridging Conversational and Collaborative Signals for Conversational Recommendation
by: Rabiah, Ahmad Bin, et al.
Published: (2024)
by: Rabiah, Ahmad Bin, et al.
Published: (2024)
Calibration-Disentangled Learning and Relevance-Prioritized Reranking for Calibrated Sequential Recommendation
by: Jeon, Hyunsik, et al.
Published: (2024)
by: Jeon, Hyunsik, et al.
Published: (2024)
StylePitcher: Generating Style-Following and Expressive Pitch Curves for Versatile Singing Tasks
by: Huang, Jingyue, et al.
Published: (2025)
by: Huang, Jingyue, et al.
Published: (2025)
Similar Items
-
TeaserGen: Generating Teasers for Long Documentaries
by: Xu, Weihan, et al.
Published: (2024) -
Generating Symbolic Music from Natural Language Prompts using an LLM-Enhanced Dataset
by: Xu, Weihan, et al.
Published: (2024) -
Video-Guided Text-to-Music Generation Using Public Domain Movie Collections
by: Kim, Haven, et al.
Published: (2025) -
DITTO-2: Distilled Diffusion Inference-Time T-Optimization for Music Generation
by: Novack, Zachary, et al.
Published: (2024) -
WildFX: A DAW-Powered Pipeline for In-the-Wild Audio FX Graph Modeling
by: Yang, Qihui, et al.
Published: (2025)