Joint Flow And Feature Refinement Using Attention For Video Restoration
Fuente:
arXiv
Saved in:
| Main Authors: | Merugu, Ranjith, Suhail, Mohammad Sameer, Sarashetti, Akshay P, Reddem, Venkata Bharath Reddy, Bajpai, Pankaj Kumar, Unde, Amit Satish |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CodeFormer++: Blind Face Restoration Using Deformable Registration and Deep Metric Learning
by: Reddem, Venkata Bharath Reddy, et al.
Published: (2025)
by: Reddem, Venkata Bharath Reddy, et al.
Published: (2025)
Joint Optimization of Buffer Delay and HARQ for Video Communications
by: Cheng, Baoping, et al.
Published: (2024)
by: Cheng, Baoping, et al.
Published: (2024)
SyncFlow: Toward Temporally Aligned Joint Audio-Video Generation from Text
by: Liu, Haohe, et al.
Published: (2024)
by: Liu, Haohe, et al.
Published: (2024)
ViFusion: In-Network Tensor Fusion for Scalable Video Feature Indexing
by: Wang, Yisu, et al.
Published: (2025)
by: Wang, Yisu, et al.
Published: (2025)
Reversing the Damage: A QP-Aware Transformer-Diffusion Approach for 8K Video Restoration under Codec Compression
by: Dehaghi, Ali Mollaahmadi, et al.
Published: (2024)
by: Dehaghi, Ali Mollaahmadi, et al.
Published: (2024)
Development of Immersive Virtual and Augmented Reality-Based Joint Attention Training Platform for Children with Autism
by: Samantaray, Ashirbad, et al.
Published: (2025)
by: Samantaray, Ashirbad, et al.
Published: (2025)
Attention of a Kiss: Exploring Attention Maps in Video Diffusion for XAIxArts
by: Cole, Adam, et al.
Published: (2025)
by: Cole, Adam, et al.
Published: (2025)
Balancing Semantic Relevance and Engagement in Related Video Recommendations
by: Jaspal, Amit, et al.
Published: (2025)
by: Jaspal, Amit, et al.
Published: (2025)
TextRefiner: Internal Visual Feature as Efficient Refiner for Vision-Language Models Prompt Tuning
by: Xie, Jingjing, et al.
Published: (2024)
by: Xie, Jingjing, et al.
Published: (2024)
Learning Quality from Complexity and Structure: A Feature-Fused XGBoost Model for Video Quality Assessment
by: Premkumar, Amritha, et al.
Published: (2025)
by: Premkumar, Amritha, et al.
Published: (2025)
BOLA360: Near-optimal View and Bitrate Adaptation for 360-degree Video Streaming
by: Zeynali, Ali, et al.
Published: (2023)
by: Zeynali, Ali, et al.
Published: (2023)
Transforming Video Subjective Testing with Training, Engagement, and Real-Time Feedback
by: Rahul, Kumar, et al.
Published: (2026)
by: Rahul, Kumar, et al.
Published: (2026)
Solving Copyright Infringement on Short Video Platforms: Novel Datasets and an Audio Restoration Deep Learning Pipeline
by: Oh, Minwoo, et al.
Published: (2025)
by: Oh, Minwoo, et al.
Published: (2025)
Seeing Further and Wider: Joint Spatio-Temporal Enlargement for Micro-Video Popularity Prediction
by: Wang, Dali, et al.
Published: (2026)
by: Wang, Dali, et al.
Published: (2026)
FlowVid: Taming Imperfect Optical Flows for Consistent Video-to-Video Synthesis
by: Liang, Feng, et al.
Published: (2023)
by: Liang, Feng, et al.
Published: (2023)
MMC: Iterative Refinement of VLM Reasoning via MCTS-based Multimodal Critique
by: Liu, Shuhang, et al.
Published: (2025)
by: Liu, Shuhang, et al.
Published: (2025)
Harnessing Multimodal Large Language Models for Personalized Product Search with Query-aware Refinement
by: Zhang, Beibei, et al.
Published: (2025)
by: Zhang, Beibei, et al.
Published: (2025)
Looking Backward: Streaming Video-to-Video Translation with Feature Banks
by: Liang, Feng, et al.
Published: (2024)
by: Liang, Feng, et al.
Published: (2024)
Synchronized Video Storytelling: Generating Video Narrations with Structured Storyline
by: Yang, Dingyi, et al.
Published: (2024)
by: Yang, Dingyi, et al.
Published: (2024)
Frozen LVLMs for Micro-Video Recommendation: A Systematic Study of Feature Extraction and Fusion
by: Sun, Huatuan, et al.
Published: (2025)
by: Sun, Huatuan, et al.
Published: (2025)
Quasi-Optimal Network Utility Maximization for Scalable Video Streaming
by: Talebi, Mohammad Sadegh, et al.
Published: (2011)
by: Talebi, Mohammad Sadegh, et al.
Published: (2011)
LASPA: Language Agnostic Speaker Disentanglement with Prefix-Tuned Cross-Attention
by: Menon, Aditya Srinivas, et al.
Published: (2025)
by: Menon, Aditya Srinivas, et al.
Published: (2025)
Where Does Vision Meet Language? Understanding and Refining Visual Fusion in MLLMs via Contrastive Attention
by: Song, Shezheng, et al.
Published: (2026)
by: Song, Shezheng, et al.
Published: (2026)
Hallucination Localization in Video Captioning
by: Nakada, Shota, et al.
Published: (2025)
by: Nakada, Shota, et al.
Published: (2025)
Music Grounding by Short Video
by: Xin, Zijie, et al.
Published: (2024)
by: Xin, Zijie, et al.
Published: (2024)
SVD: Spatial Video Dataset
by: Izadimehr, M. H., et al.
Published: (2025)
by: Izadimehr, M. H., et al.
Published: (2025)
A Clustering-Based Method for Automatic Educational Video Recommendation Using Deep Face-Features of Lecturers
by: Mendes, Paulo R. C., et al.
Published: (2020)
by: Mendes, Paulo R. C., et al.
Published: (2020)
Self-Attention and Hybrid Features for Replay and Deep-Fake Audio Detection
by: Huang, Lian, et al.
Published: (2024)
by: Huang, Lian, et al.
Published: (2024)
Deep Bi-directional Attention Network for Image Super-Resolution Quality Assessment
by: Li, Yixiao, et al.
Published: (2024)
by: Li, Yixiao, et al.
Published: (2024)
Multi-source Knowledge Enhanced Graph Attention Networks for Multimodal Fact Verification
by: Cao, Han, et al.
Published: (2024)
by: Cao, Han, et al.
Published: (2024)
diveXplore at the Video Browser Showdown 2024
by: Schoeffmann, Klaus, et al.
Published: (2025)
by: Schoeffmann, Klaus, et al.
Published: (2025)
Differentially Processed Optimized Collaborative Rich Text Editor
by: Jatana, Nishtha, et al.
Published: (2024)
by: Jatana, Nishtha, et al.
Published: (2024)
Cross-Attention Fusion of Visual and Geometric Features for Large Vocabulary Arabic Lipreading
by: Daou, Samar, et al.
Published: (2024)
by: Daou, Samar, et al.
Published: (2024)
HADUA: Hierarchical Attention and Dynamic Uniform Alignment for Robust Cross-Subject Emotion Recognition
by: Tang, Jiahao, et al.
Published: (2026)
by: Tang, Jiahao, et al.
Published: (2026)
Orthogonal Disentanglement with Projected Feature Alignment for Multimodal Emotion Recognition in Conversation
by: Che, Xinyi, et al.
Published: (2025)
by: Che, Xinyi, et al.
Published: (2025)
Latent Feature-Guided Conditional Diffusion for Generative Image Semantic Communication
by: Chen, Zehao, et al.
Published: (2025)
by: Chen, Zehao, et al.
Published: (2025)
Adaptive 3D Mesh Steganography Based on Feature-Preserving Distortion
by: Zhang, Yushu, et al.
Published: (2022)
by: Zhang, Yushu, et al.
Published: (2022)
Feature Coding in the Era of Large Models: Dataset, Test Conditions, and Benchmark
by: Gao, Changsheng, et al.
Published: (2024)
by: Gao, Changsheng, et al.
Published: (2024)
Feedback-Driven Rate Control for Learned Video Compression
by: Xu, Zhiheng, et al.
Published: (2026)
by: Xu, Zhiheng, et al.
Published: (2026)
Optimal Transcoding Preset Selection for Live Video Streaming
by: Nabizadeh, Zahra, et al.
Published: (2024)
by: Nabizadeh, Zahra, et al.
Published: (2024)
Similar Items
-
CodeFormer++: Blind Face Restoration Using Deformable Registration and Deep Metric Learning
by: Reddem, Venkata Bharath Reddy, et al.
Published: (2025) -
Joint Optimization of Buffer Delay and HARQ for Video Communications
by: Cheng, Baoping, et al.
Published: (2024) -
SyncFlow: Toward Temporally Aligned Joint Audio-Video Generation from Text
by: Liu, Haohe, et al.
Published: (2024) -
ViFusion: In-Network Tensor Fusion for Scalable Video Feature Indexing
by: Wang, Yisu, et al.
Published: (2025) -
Reversing the Damage: A QP-Aware Transformer-Diffusion Approach for 8K Video Restoration under Codec Compression
by: Dehaghi, Ali Mollaahmadi, et al.
Published: (2024)