PipeFlow: Pipelined Processing and Motion-Aware Frame Selection for Long-Form Video Editing
Fuente:
arXiv
Saved in:
| Main Authors: | Munir, Mustafa, Rahman, Md Mostafijur, Bhardwaj, Kartikeya, Whatmough, Paul, Marculescu, Radu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RapidNet: Multi-Level Dilated Convolution Based Mobile Backbone
by: Munir, Mustafa, et al.
Published: (2024)
by: Munir, Mustafa, et al.
Published: (2024)
EMCAD: Efficient Multi-scale Convolutional Attention Decoding for Medical Image Segmentation
by: Rahman, Md Mostafijur, et al.
Published: (2024)
by: Rahman, Md Mostafijur, et al.
Published: (2024)
AdaptViG: Adaptive Vision GNN with Exponential Decay Gating
by: Munir, Mustafa, et al.
Published: (2025)
by: Munir, Mustafa, et al.
Published: (2025)
GreedyViG: Dynamic Axial Graph Construction for Efficient Vision GNNs
by: Munir, Mustafa, et al.
Published: (2024)
by: Munir, Mustafa, et al.
Published: (2024)
Ada-VE: Training-Free Consistent Video Editing Using Adaptive Motion Prior
by: Mahmud, Tanvir, et al.
Published: (2024)
by: Mahmud, Tanvir, et al.
Published: (2024)
MK-UNet: Multi-kernel Lightweight CNN for Medical Image Segmentation
by: Rahman, Md Mostafijur, et al.
Published: (2025)
by: Rahman, Md Mostafijur, et al.
Published: (2025)
LoMix: Learnable Weighted Multi-Scale Logits Mixing for Medical Image Segmentation
by: Rahman, Md Mostafijur, et al.
Published: (2025)
by: Rahman, Md Mostafijur, et al.
Published: (2025)
PP-SAM: Perturbed Prompts for Robust Adaptation of Segment Anything Model for Polyp Segmentation
by: Rahman, Md Mostafijur, et al.
Published: (2024)
by: Rahman, Md Mostafijur, et al.
Published: (2024)
ObjectAlign: Neuro-Symbolic Object Consistency Verification and Correction
by: Munir, Mustafa, et al.
Published: (2025)
by: Munir, Mustafa, et al.
Published: (2025)
Video Compression Meets Video Generation: Latent Inter-Frame Pruning with Attention Recovery
by: Menn, Dennis, et al.
Published: (2026)
by: Menn, Dennis, et al.
Published: (2026)
Scaling Graph Convolutions for Mobile Vision
by: Avery, William, et al.
Published: (2024)
by: Avery, William, et al.
Published: (2024)
VCMamba: Bridging Convolutions with Multi-Directional Mamba for Efficient Visual Representation
by: Munir, Mustafa, et al.
Published: (2025)
by: Munir, Mustafa, et al.
Published: (2025)
Multi-Scale High-Resolution Logarithmic Grapher Module for Efficient Vision GNNs
by: Munir, Mustafa, et al.
Published: (2025)
by: Munir, Mustafa, et al.
Published: (2025)
NeuS-QA: Grounding Long-Form Video Understanding in Temporal Logic and Neuro-Symbolic Reasoning
by: Shah, Sahil, et al.
Published: (2025)
by: Shah, Sahil, et al.
Published: (2025)
Fuel Gauge: Estimating Chain-of-Thought Length Ahead of Time in Large Multimodal Models
by: Yang, Yuedong, et al.
Published: (2026)
by: Yang, Yuedong, et al.
Published: (2026)
Zero-Shot Neural Architecture Search: Challenges, Solutions, and Opportunities
by: Li, Guihong, et al.
Published: (2023)
by: Li, Guihong, et al.
Published: (2023)
BanglaNet: Bangla Handwritten Character Recognition using Ensembling of Convolutional Neural Network
by: Saha, Chandrika, et al.
Published: (2024)
by: Saha, Chandrika, et al.
Published: (2024)
Mitigating Intra- and Inter-modal Forgetting in Continual Learning of Unified Multimodal Models
by: Wei, Xiwen, et al.
Published: (2025)
by: Wei, Xiwen, et al.
Published: (2025)
Motion-Aware Video Frame Interpolation
by: Han, Pengfei, et al.
Published: (2024)
by: Han, Pengfei, et al.
Published: (2024)
TRACE: Object Motion Editing in Videos with First-Frame Trajectory Guidance
by: Phung, Quynh, et al.
Published: (2026)
by: Phung, Quynh, et al.
Published: (2026)
Divide, then Ground: Adapting Frame Selection to Query Types for Long-Form Video Understanding
by: Li, Jialuo, et al.
Published: (2025)
by: Li, Jialuo, et al.
Published: (2025)
From Frames to Clips: Training-free Adaptive Key Clip Selection for Long-Form Video Understanding
by: Sun, Guangyu, et al.
Published: (2025)
by: Sun, Guangyu, et al.
Published: (2025)
Malaria Detection from Blood Cell Images Using XceptionNet
by: Nusrat, Warisa, et al.
Published: (2025)
by: Nusrat, Warisa, et al.
Published: (2025)
AttentionViG: Cross-Attention-Based Dynamic Neighbor Aggregation in Vision GNNs
by: Gedik, Hakan Emre, et al.
Published: (2025)
by: Gedik, Hakan Emre, et al.
Published: (2025)
Online-LoRA: Task-free Online Continual Learning via Low Rank Adaptation
by: Wei, Xiwen, et al.
Published: (2024)
by: Wei, Xiwen, et al.
Published: (2024)
AdaFlow: Efficient Long Video Editing via Adaptive Attention Slimming And Keyframe Selection
by: Zhang, Shuheng, et al.
Published: (2025)
by: Zhang, Shuheng, et al.
Published: (2025)
Video Reasoning without Training
by: Sridhar, Deepak, et al.
Published: (2025)
by: Sridhar, Deepak, et al.
Published: (2025)
MoVideo: Motion-Aware Video Generation with Diffusion Models
by: Liang, Jingyun, et al.
Published: (2023)
by: Liang, Jingyun, et al.
Published: (2023)
Event-Anchored Frame Selection for Effective Long-Video Understanding
by: Chen, Wang, et al.
Published: (2026)
by: Chen, Wang, et al.
Published: (2026)
MoCA-Video: Motion-Aware Concept Alignment for Consistent Video Editing
by: Zhang, Tong, et al.
Published: (2025)
by: Zhang, Tong, et al.
Published: (2025)
Adaptive Greedy Frame Selection for Long Video Understanding
by: Huang, Yuning, et al.
Published: (2026)
by: Huang, Yuning, et al.
Published: (2026)
Too Many Frames, Not All Useful: Efficient Strategies for Long-Form Video QA
by: Park, Jongwoo, et al.
Published: (2024)
by: Park, Jongwoo, et al.
Published: (2024)
Motion-Aware Generative Frame Interpolation
by: Zhang, Guozhen, et al.
Published: (2025)
by: Zhang, Guozhen, et al.
Published: (2025)
LumosFlow: Motion-Guided Long Video Generation
by: Chen, Jiahao, et al.
Published: (2025)
by: Chen, Jiahao, et al.
Published: (2025)
FRAG: Frame Selection Augmented Generation for Long Video and Long Document Understanding
by: Huang, De-An, et al.
Published: (2025)
by: Huang, De-An, et al.
Published: (2025)
Moment Sampling in Video LLMs for Long-Form Video QA
by: Chasmai, Mustafa, et al.
Published: (2025)
by: Chasmai, Mustafa, et al.
Published: (2025)
FlowNar: Scalable Streaming Narration for Long-Form Videos
by: Zhong, Zeyun, et al.
Published: (2026)
by: Zhong, Zeyun, et al.
Published: (2026)
CrediRAG: Network-Augmented Credibility-Based Retrieval for Misinformation Detection in Reddit
by: Ram, Ashwin, et al.
Published: (2024)
by: Ram, Ashwin, et al.
Published: (2024)
ExpertEdit: Learning Skill-Aware Motion Editing from Expert Videos
by: Somayazulu, Arjun, et al.
Published: (2026)
by: Somayazulu, Arjun, et al.
Published: (2026)
Frame by Familiar Frame: Understanding Replication in Video Diffusion Models
by: Rahman, Aimon, et al.
Published: (2024)
by: Rahman, Aimon, et al.
Published: (2024)
Similar Items
-
RapidNet: Multi-Level Dilated Convolution Based Mobile Backbone
by: Munir, Mustafa, et al.
Published: (2024) -
EMCAD: Efficient Multi-scale Convolutional Attention Decoding for Medical Image Segmentation
by: Rahman, Md Mostafijur, et al.
Published: (2024) -
AdaptViG: Adaptive Vision GNN with Exponential Decay Gating
by: Munir, Mustafa, et al.
Published: (2025) -
GreedyViG: Dynamic Axial Graph Construction for Efficient Vision GNNs
by: Munir, Mustafa, et al.
Published: (2024) -
Ada-VE: Training-Free Consistent Video Editing Using Adaptive Motion Prior
by: Mahmud, Tanvir, et al.
Published: (2024)