Enhancing Space-time Video Super-resolution via Spatial-temporal Feature Interaction
Fuente:
arXiv
Saved in:
| Main Authors: | Yue, Zijie, Shi, Miaojing |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Weakly-Supervised Referring Video Object Segmentation through Text Supervision
by: Shi, Miaojing, et al.
Published: (2026)
by: Shi, Miaojing, et al.
Published: (2026)
LoSh: Long-Short Text Joint Prediction Network for Referring Video Object Segmentation
by: Yuan, Linfeng, et al.
Published: (2023)
by: Yuan, Linfeng, et al.
Published: (2023)
Text-promptable Object Counting via Quantity Awareness Enhancement
by: Shi, Miaojing, et al.
Published: (2025)
by: Shi, Miaojing, et al.
Published: (2025)
Bootstrapping MLLM for Weakly-Supervised Class-Agnostic Object Counting
by: Zhang, Xiaowen, et al.
Published: (2026)
by: Zhang, Xiaowen, et al.
Published: (2026)
Large Model driven Radiology Report Generation with Clinical Quality Reinforcement Learning
by: Zhou, Zijian, et al.
Published: (2024)
by: Zhou, Zijian, et al.
Published: (2024)
Bootstrapping Vision-language Models for Self-supervised Remote Physiological Measurement
by: Yue, Zijie, et al.
Published: (2024)
by: Yue, Zijie, et al.
Published: (2024)
Memory-guided Network with Uncertainty-based Feature Augmentation for Few-shot Semantic Segmentation
by: Chen, Xinyue, et al.
Published: (2024)
by: Chen, Xinyue, et al.
Published: (2024)
Space-Time Video Super-resolution with Neural Operator
by: Zhang, Yuantong, et al.
Published: (2024)
by: Zhang, Yuantong, et al.
Published: (2024)
FAAR: Efficient Frequency-Aware Multi-Task Fine-Tuning via Automatic Rank Selection
by: Fontana, Maxime, et al.
Published: (2026)
by: Fontana, Maxime, et al.
Published: (2026)
Enhancing Generalized Few-Shot Semantic Segmentation via Effective Knowledge Transfer
by: Chen, Xinyue, et al.
Published: (2024)
by: Chen, Xinyue, et al.
Published: (2024)
OpenPSG: Open-set Panoptic Scene Graph Generation via Large Multimodal Models
by: Zhou, Zijian, et al.
Published: (2024)
by: Zhou, Zijian, et al.
Published: (2024)
Boosting Object Detection with Zero-Shot Day-Night Domain Adaptation
by: Du, Zhipeng, et al.
Published: (2023)
by: Du, Zhipeng, et al.
Published: (2023)
Multitask Learning in Minimally Invasive Surgical Vision: A Review
by: Alabi, Oluwatosin, et al.
Published: (2024)
by: Alabi, Oluwatosin, et al.
Published: (2024)
VLPrompt: Vision-Language Prompting for Panoptic Scene Graph Generation
by: Zhou, Zijian, et al.
Published: (2023)
by: Zhou, Zijian, et al.
Published: (2023)
End-to-end Video Gaze Estimation via Capturing Head-face-eye Spatial-temporal Interaction Context
by: Guan, Yiran, et al.
Published: (2023)
by: Guan, Yiran, et al.
Published: (2023)
Real-time Spatial-temporal Traversability Assessment via Feature-based Sparse Gaussian Process
by: Hou, Zhenyu, et al.
Published: (2025)
by: Hou, Zhenyu, et al.
Published: (2025)
MPDrive: Improving Spatial Understanding with Marker-Based Prompt Learning for Autonomous Driving
by: Zhang, Zhiyuan, et al.
Published: (2025)
by: Zhang, Zhiyuan, et al.
Published: (2025)
STARS: Sparse Learning Correlation Filter with Spatio-temporal Regularization and Super-resolution Reconstruction for Thermal Infrared Target Tracking
by: Zhang, Shang, et al.
Published: (2025)
by: Zhang, Shang, et al.
Published: (2025)
UniPaint: Unified Space-time Video Inpainting via Mixture-of-Experts
by: Wan, Zhen, et al.
Published: (2024)
by: Wan, Zhen, et al.
Published: (2024)
Optimizing Dense Visual Predictions Through Multi-Task Coherence and Prioritization
by: Fontana, Maxime, et al.
Published: (2024)
by: Fontana, Maxime, et al.
Published: (2024)
Enhancing Video Super-Resolution via Implicit Resampling-based Alignment
by: Xu, Kai, et al.
Published: (2023)
by: Xu, Kai, et al.
Published: (2023)
Kalman-Inspired Feature Propagation for Video Face Super-Resolution
by: Feng, Ruicheng, et al.
Published: (2024)
by: Feng, Ruicheng, et al.
Published: (2024)
Exploiting Multimodal Spatial-temporal Patterns for Video Object Tracking
by: Hu, Xiantao, et al.
Published: (2024)
by: Hu, Xiantao, et al.
Published: (2024)
Test-time Training for Hyperspectral Image Super-resolution
by: Li, Ke, et al.
Published: (2024)
by: Li, Ke, et al.
Published: (2024)
First-order State Space Model for Lightweight Image Super-resolution
by: Zhu, Yujie, et al.
Published: (2025)
by: Zhu, Yujie, et al.
Published: (2025)
HSTR-Net: Reference Based Video Super-resolution with Dual Cameras
by: Suluhan, H. Umut, et al.
Published: (2023)
by: Suluhan, H. Umut, et al.
Published: (2023)
STAR: Spatial-Temporal Augmentation with Text-to-Video Models for Real-World Video Super-Resolution
by: Xie, Rui, et al.
Published: (2025)
by: Xie, Rui, et al.
Published: (2025)
MFSR: Multi-fractal Feature for Super-resolution Reconstruction with Fine Details Recovery
by: Yang, Lianping, et al.
Published: (2025)
by: Yang, Lianping, et al.
Published: (2025)
SurgPIS: Surgical-instrument-level Instances and Part-level Semantics for Weakly-supervised Part-aware Instance Segmentation
by: Wei, Meng, et al.
Published: (2025)
by: Wei, Meng, et al.
Published: (2025)
Text Promptable Surgical Instrument Segmentation with Vision-Language Models
by: Zhou, Zijian, et al.
Published: (2023)
by: Zhou, Zijian, et al.
Published: (2023)
Grounding Surgical Action Triplets with Instrument Instance Segmentation: A Dataset and Target-Aware Fusion Approach
by: Alabi, Oluwatosin, et al.
Published: (2025)
by: Alabi, Oluwatosin, et al.
Published: (2025)
Efficient Image Super-Resolution with Feature Interaction Weighted Hybrid Network
by: Li, Wenjie, et al.
Published: (2022)
by: Li, Wenjie, et al.
Published: (2022)
Practical Video Object Detection via Feature Selection and Aggregation
by: Shi, Yuheng, et al.
Published: (2024)
by: Shi, Yuheng, et al.
Published: (2024)
STDAN: Deformable Attention Network for Space-Time Video Super-Resolution
by: Wang, Hai, et al.
Published: (2022)
by: Wang, Hai, et al.
Published: (2022)
RefVSR++: Exploiting Reference Inputs for Reference-based Video Super-resolution
by: Zou, Han, et al.
Published: (2023)
by: Zou, Han, et al.
Published: (2023)
SpaceR: Reinforcing MLLMs in Video Spatial Reasoning
by: Ouyang, Kun, et al.
Published: (2025)
by: Ouyang, Kun, et al.
Published: (2025)
PAS-Mamba: Phase-Amplitude-Spatial State Space Model for MRI Reconstruction
by: Kui, Xiaoyan, et al.
Published: (2026)
by: Kui, Xiaoyan, et al.
Published: (2026)
Diffusion4D: Fast Spatial-temporal Consistent 4D Generation via Video Diffusion Models
by: Liang, Hanwen, et al.
Published: (2024)
by: Liang, Hanwen, et al.
Published: (2024)
CoInteract: Physically-Consistent Human-Object Interaction Video Synthesis via Spatially-Structured Co-Generation
by: Luo, Xiangyang, et al.
Published: (2026)
by: Luo, Xiangyang, et al.
Published: (2026)
STAR-Pose: Efficient Low-Resolution Video Human Pose Estimation via Spatial-Temporal Adaptive Super-Resolution
by: Jin, Yucheng, et al.
Published: (2025)
by: Jin, Yucheng, et al.
Published: (2025)
Similar Items
-
Weakly-Supervised Referring Video Object Segmentation through Text Supervision
by: Shi, Miaojing, et al.
Published: (2026) -
LoSh: Long-Short Text Joint Prediction Network for Referring Video Object Segmentation
by: Yuan, Linfeng, et al.
Published: (2023) -
Text-promptable Object Counting via Quantity Awareness Enhancement
by: Shi, Miaojing, et al.
Published: (2025) -
Bootstrapping MLLM for Weakly-Supervised Class-Agnostic Object Counting
by: Zhang, Xiaowen, et al.
Published: (2026) -
Large Model driven Radiology Report Generation with Clinical Quality Reinforcement Learning
by: Zhou, Zijian, et al.
Published: (2024)