Joint Flow And Feature Refinement Using Attention For Video Restoration
Fuente:
arXiv
Guardado en:
| Autores principales: | Merugu, Ranjith, Suhail, Mohammad Sameer, Sarashetti, Akshay P, Reddem, Venkata Bharath Reddy, Bajpai, Pankaj Kumar, Unde, Amit Satish |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
CodeFormer++: Blind Face Restoration Using Deformable Registration and Deep Metric Learning
por: Reddem, Venkata Bharath Reddy, et al.
Publicado: (2025)
por: Reddem, Venkata Bharath Reddy, et al.
Publicado: (2025)
Joint Optimization of Buffer Delay and HARQ for Video Communications
por: Cheng, Baoping, et al.
Publicado: (2024)
por: Cheng, Baoping, et al.
Publicado: (2024)
SyncFlow: Toward Temporally Aligned Joint Audio-Video Generation from Text
por: Liu, Haohe, et al.
Publicado: (2024)
por: Liu, Haohe, et al.
Publicado: (2024)
ViFusion: In-Network Tensor Fusion for Scalable Video Feature Indexing
por: Wang, Yisu, et al.
Publicado: (2025)
por: Wang, Yisu, et al.
Publicado: (2025)
Reversing the Damage: A QP-Aware Transformer-Diffusion Approach for 8K Video Restoration under Codec Compression
por: Dehaghi, Ali Mollaahmadi, et al.
Publicado: (2024)
por: Dehaghi, Ali Mollaahmadi, et al.
Publicado: (2024)
Development of Immersive Virtual and Augmented Reality-Based Joint Attention Training Platform for Children with Autism
por: Samantaray, Ashirbad, et al.
Publicado: (2025)
por: Samantaray, Ashirbad, et al.
Publicado: (2025)
Attention of a Kiss: Exploring Attention Maps in Video Diffusion for XAIxArts
por: Cole, Adam, et al.
Publicado: (2025)
por: Cole, Adam, et al.
Publicado: (2025)
Balancing Semantic Relevance and Engagement in Related Video Recommendations
por: Jaspal, Amit, et al.
Publicado: (2025)
por: Jaspal, Amit, et al.
Publicado: (2025)
TextRefiner: Internal Visual Feature as Efficient Refiner for Vision-Language Models Prompt Tuning
por: Xie, Jingjing, et al.
Publicado: (2024)
por: Xie, Jingjing, et al.
Publicado: (2024)
Learning Quality from Complexity and Structure: A Feature-Fused XGBoost Model for Video Quality Assessment
por: Premkumar, Amritha, et al.
Publicado: (2025)
por: Premkumar, Amritha, et al.
Publicado: (2025)
BOLA360: Near-optimal View and Bitrate Adaptation for 360-degree Video Streaming
por: Zeynali, Ali, et al.
Publicado: (2023)
por: Zeynali, Ali, et al.
Publicado: (2023)
Transforming Video Subjective Testing with Training, Engagement, and Real-Time Feedback
por: Rahul, Kumar, et al.
Publicado: (2026)
por: Rahul, Kumar, et al.
Publicado: (2026)
Solving Copyright Infringement on Short Video Platforms: Novel Datasets and an Audio Restoration Deep Learning Pipeline
por: Oh, Minwoo, et al.
Publicado: (2025)
por: Oh, Minwoo, et al.
Publicado: (2025)
Seeing Further and Wider: Joint Spatio-Temporal Enlargement for Micro-Video Popularity Prediction
por: Wang, Dali, et al.
Publicado: (2026)
por: Wang, Dali, et al.
Publicado: (2026)
FlowVid: Taming Imperfect Optical Flows for Consistent Video-to-Video Synthesis
por: Liang, Feng, et al.
Publicado: (2023)
por: Liang, Feng, et al.
Publicado: (2023)
MMC: Iterative Refinement of VLM Reasoning via MCTS-based Multimodal Critique
por: Liu, Shuhang, et al.
Publicado: (2025)
por: Liu, Shuhang, et al.
Publicado: (2025)
Harnessing Multimodal Large Language Models for Personalized Product Search with Query-aware Refinement
por: Zhang, Beibei, et al.
Publicado: (2025)
por: Zhang, Beibei, et al.
Publicado: (2025)
Looking Backward: Streaming Video-to-Video Translation with Feature Banks
por: Liang, Feng, et al.
Publicado: (2024)
por: Liang, Feng, et al.
Publicado: (2024)
Synchronized Video Storytelling: Generating Video Narrations with Structured Storyline
por: Yang, Dingyi, et al.
Publicado: (2024)
por: Yang, Dingyi, et al.
Publicado: (2024)
Frozen LVLMs for Micro-Video Recommendation: A Systematic Study of Feature Extraction and Fusion
por: Sun, Huatuan, et al.
Publicado: (2025)
por: Sun, Huatuan, et al.
Publicado: (2025)
Quasi-Optimal Network Utility Maximization for Scalable Video Streaming
por: Talebi, Mohammad Sadegh, et al.
Publicado: (2011)
por: Talebi, Mohammad Sadegh, et al.
Publicado: (2011)
LASPA: Language Agnostic Speaker Disentanglement with Prefix-Tuned Cross-Attention
por: Menon, Aditya Srinivas, et al.
Publicado: (2025)
por: Menon, Aditya Srinivas, et al.
Publicado: (2025)
Where Does Vision Meet Language? Understanding and Refining Visual Fusion in MLLMs via Contrastive Attention
por: Song, Shezheng, et al.
Publicado: (2026)
por: Song, Shezheng, et al.
Publicado: (2026)
Hallucination Localization in Video Captioning
por: Nakada, Shota, et al.
Publicado: (2025)
por: Nakada, Shota, et al.
Publicado: (2025)
Music Grounding by Short Video
por: Xin, Zijie, et al.
Publicado: (2024)
por: Xin, Zijie, et al.
Publicado: (2024)
SVD: Spatial Video Dataset
por: Izadimehr, M. H., et al.
Publicado: (2025)
por: Izadimehr, M. H., et al.
Publicado: (2025)
A Clustering-Based Method for Automatic Educational Video Recommendation Using Deep Face-Features of Lecturers
por: Mendes, Paulo R. C., et al.
Publicado: (2020)
por: Mendes, Paulo R. C., et al.
Publicado: (2020)
Self-Attention and Hybrid Features for Replay and Deep-Fake Audio Detection
por: Huang, Lian, et al.
Publicado: (2024)
por: Huang, Lian, et al.
Publicado: (2024)
Deep Bi-directional Attention Network for Image Super-Resolution Quality Assessment
por: Li, Yixiao, et al.
Publicado: (2024)
por: Li, Yixiao, et al.
Publicado: (2024)
Multi-source Knowledge Enhanced Graph Attention Networks for Multimodal Fact Verification
por: Cao, Han, et al.
Publicado: (2024)
por: Cao, Han, et al.
Publicado: (2024)
diveXplore at the Video Browser Showdown 2024
por: Schoeffmann, Klaus, et al.
Publicado: (2025)
por: Schoeffmann, Klaus, et al.
Publicado: (2025)
Differentially Processed Optimized Collaborative Rich Text Editor
por: Jatana, Nishtha, et al.
Publicado: (2024)
por: Jatana, Nishtha, et al.
Publicado: (2024)
Cross-Attention Fusion of Visual and Geometric Features for Large Vocabulary Arabic Lipreading
por: Daou, Samar, et al.
Publicado: (2024)
por: Daou, Samar, et al.
Publicado: (2024)
HADUA: Hierarchical Attention and Dynamic Uniform Alignment for Robust Cross-Subject Emotion Recognition
por: Tang, Jiahao, et al.
Publicado: (2026)
por: Tang, Jiahao, et al.
Publicado: (2026)
Orthogonal Disentanglement with Projected Feature Alignment for Multimodal Emotion Recognition in Conversation
por: Che, Xinyi, et al.
Publicado: (2025)
por: Che, Xinyi, et al.
Publicado: (2025)
Latent Feature-Guided Conditional Diffusion for Generative Image Semantic Communication
por: Chen, Zehao, et al.
Publicado: (2025)
por: Chen, Zehao, et al.
Publicado: (2025)
Adaptive 3D Mesh Steganography Based on Feature-Preserving Distortion
por: Zhang, Yushu, et al.
Publicado: (2022)
por: Zhang, Yushu, et al.
Publicado: (2022)
Feature Coding in the Era of Large Models: Dataset, Test Conditions, and Benchmark
por: Gao, Changsheng, et al.
Publicado: (2024)
por: Gao, Changsheng, et al.
Publicado: (2024)
Feedback-Driven Rate Control for Learned Video Compression
por: Xu, Zhiheng, et al.
Publicado: (2026)
por: Xu, Zhiheng, et al.
Publicado: (2026)
Optimal Transcoding Preset Selection for Live Video Streaming
por: Nabizadeh, Zahra, et al.
Publicado: (2024)
por: Nabizadeh, Zahra, et al.
Publicado: (2024)
Ejemplares similares
-
CodeFormer++: Blind Face Restoration Using Deformable Registration and Deep Metric Learning
por: Reddem, Venkata Bharath Reddy, et al.
Publicado: (2025) -
Joint Optimization of Buffer Delay and HARQ for Video Communications
por: Cheng, Baoping, et al.
Publicado: (2024) -
SyncFlow: Toward Temporally Aligned Joint Audio-Video Generation from Text
por: Liu, Haohe, et al.
Publicado: (2024) -
ViFusion: In-Network Tensor Fusion for Scalable Video Feature Indexing
por: Wang, Yisu, et al.
Publicado: (2025) -
Reversing the Damage: A QP-Aware Transformer-Diffusion Approach for 8K Video Restoration under Codec Compression
por: Dehaghi, Ali Mollaahmadi, et al.
Publicado: (2024)