Raformer: Redundancy-Aware Transformer for Video Wire Inpainting
Fuente:
arXiv
Salvato in:
| Autori principali: | Ji, Zhong, Su, Yimu, Zhang, Yan, Hou, Jiacheng, Pang, Yanwei, Han, Jungong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Fresh Look at Generalized Category Discovery through Non-negative Matrix Factorization
di: Ji, Zhong, et al.
Pubblicazione: (2024)
di: Ji, Zhong, et al.
Pubblicazione: (2024)
iEBAKER: Improved Remote Sensing Image-Text Retrieval Framework via Eliminate Before Align and Keyword Explicit Reasoning
di: Zhang, Yan, et al.
Pubblicazione: (2025)
di: Zhang, Yan, et al.
Pubblicazione: (2025)
DAOVI: Distortion-Aware Omnidirectional Video Inpainting
di: Seshimo, Ryosuke, et al.
Pubblicazione: (2025)
di: Seshimo, Ryosuke, et al.
Pubblicazione: (2025)
COCO-Inpaint: A Benchmark for Detecting and Localizing Inpainting-Based Image Manipulations
di: Yan, Haozhen, et al.
Pubblicazione: (2025)
di: Yan, Haozhen, et al.
Pubblicazione: (2025)
Relation-Aware Meta-Learning for Zero-shot Sketch-Based Image Retrieval
di: Liu, Yang, et al.
Pubblicazione: (2024)
di: Liu, Yang, et al.
Pubblicazione: (2024)
Wired Perspectives: Multi-View Wire Art Embraces Generative AI
di: Qu, Zhiyu, et al.
Pubblicazione: (2023)
di: Qu, Zhiyu, et al.
Pubblicazione: (2023)
Underlying Semantic Diffusion for Effective and Efficient In-Context Learning
di: Ji, Zhong, et al.
Pubblicazione: (2025)
di: Ji, Zhong, et al.
Pubblicazione: (2025)
Assessing Image Inpainting via Re-Inpainting Self-Consistency Evaluation
di: Chen, Tianyi, et al.
Pubblicazione: (2024)
di: Chen, Tianyi, et al.
Pubblicazione: (2024)
Interpretable Few-Shot Image Classification via Prototypical Concept-Guided Mixture of LoRA Experts
di: Ji, Zhong, et al.
Pubblicazione: (2025)
di: Ji, Zhong, et al.
Pubblicazione: (2025)
Multi-Stage Knowledge Integration of Vision-Language Models for Continual Learning
di: Zhang, Hongsheng, et al.
Pubblicazione: (2024)
di: Zhang, Hongsheng, et al.
Pubblicazione: (2024)
AdaTP: Attention-Debiased Token Pruning for Video Large Language Models
di: Sun, Fengyuan, et al.
Pubblicazione: (2025)
di: Sun, Fengyuan, et al.
Pubblicazione: (2025)
Tracking and Segmenting Anything in Any Modality
di: Zhang, Tianlu, et al.
Pubblicazione: (2025)
di: Zhang, Tianlu, et al.
Pubblicazione: (2025)
VideoPainter: Any-length Video Inpainting and Editing with Plug-and-Play Context Control
di: Bian, Yuxuan, et al.
Pubblicazione: (2025)
di: Bian, Yuxuan, et al.
Pubblicazione: (2025)
Diffree: Text-Guided Shape Free Object Inpainting with Diffusion Model
di: Zhao, Lirui, et al.
Pubblicazione: (2024)
di: Zhao, Lirui, et al.
Pubblicazione: (2024)
Knowing Your Target: Target-Aware Transformer Makes Better Spatio-Temporal Video Grounding
di: Gu, Xin, et al.
Pubblicazione: (2025)
di: Gu, Xin, et al.
Pubblicazione: (2025)
Unveiling Redundancy in Diffusion Transformers (DiTs): A Systematic Study
di: Sun, Xibo, et al.
Pubblicazione: (2024)
di: Sun, Xibo, et al.
Pubblicazione: (2024)
From Inheritance to Saturation: Disentangling the Evolution of Visual Redundancy for Architecture-Aware MLLM Inference Acceleration
di: Shi, Jiaqi, et al.
Pubblicazione: (2026)
di: Shi, Jiaqi, et al.
Pubblicazione: (2026)
Contact-Aware Amodal Completion for Human-Object Interaction via Multi-Regional Inpainting
di: Chi, Seunggeun, et al.
Pubblicazione: (2025)
di: Chi, Seunggeun, et al.
Pubblicazione: (2025)
Hybrid Discriminative Attribute-Object Embedding Network for Compositional Zero-Shot Learning
di: Liu, Yang, et al.
Pubblicazione: (2024)
di: Liu, Yang, et al.
Pubblicazione: (2024)
Semi-supervised Semantic Segmentation for Remote Sensing Images via Multi-scale Uncertainty Consistency and Cross-Teacher-Student Attention
di: Wang, Shanwen, et al.
Pubblicazione: (2025)
di: Wang, Shanwen, et al.
Pubblicazione: (2025)
FBPT: A Fully Binary Point Transformer
di: Hou, Zhixing, et al.
Pubblicazione: (2024)
di: Hou, Zhixing, et al.
Pubblicazione: (2024)
InpaintSLat: Inpainting Structured 3D Latents via Initial Noise Optimization
di: Chung, Jaeyoung, et al.
Pubblicazione: (2026)
di: Chung, Jaeyoung, et al.
Pubblicazione: (2026)
LF-ViT: Reducing Spatial Redundancy in Vision Transformer for Efficient Image Recognition
di: Hu, Youbing, et al.
Pubblicazione: (2024)
di: Hu, Youbing, et al.
Pubblicazione: (2024)
Compressed-Domain-Aware Online Video Super-Resolution
di: Wang, Yuhang, et al.
Pubblicazione: (2026)
di: Wang, Yuhang, et al.
Pubblicazione: (2026)
Focused Forcing: Content-Aware Per-Frame KV Selection for Efficient Autoregressive Video Diffusion
di: Cai, Peiliang, et al.
Pubblicazione: (2026)
di: Cai, Peiliang, et al.
Pubblicazione: (2026)
Motion-Aware Caching for Efficient Autoregressive Video Generation
di: Xu, Jing, et al.
Pubblicazione: (2026)
di: Xu, Jing, et al.
Pubblicazione: (2026)
RainFusion: Adaptive Video Generation Acceleration via Multi-Dimensional Visual Redundancy
di: Chen, Aiyue, et al.
Pubblicazione: (2025)
di: Chen, Aiyue, et al.
Pubblicazione: (2025)
UFO: Enhancing Diffusion-Based Video Generation with a Uniform Frame Organizer
di: Liu, Delong, et al.
Pubblicazione: (2024)
di: Liu, Delong, et al.
Pubblicazione: (2024)
Knowledge-Refined Dual Context-Aware Network for Partially Relevant Video Retrieval
di: Yang, Junkai, et al.
Pubblicazione: (2026)
di: Yang, Junkai, et al.
Pubblicazione: (2026)
Motion Blur Robust Wheat Pest Damage Detection with Dynamic Fuzzy Feature Fusion
di: Zhang, Han, et al.
Pubblicazione: (2026)
di: Zhang, Han, et al.
Pubblicazione: (2026)
FreeCond: Free Lunch in the Input Conditions of Text-Guided Inpainting
di: Hsiao, Teng-Fang, et al.
Pubblicazione: (2024)
di: Hsiao, Teng-Fang, et al.
Pubblicazione: (2024)
FacaDiffy: Inpainting Unseen Facade Parts Using Diffusion Models
di: Froech, Thomas, et al.
Pubblicazione: (2025)
di: Froech, Thomas, et al.
Pubblicazione: (2025)
BrushEdit: All-In-One Image Inpainting and Editing
di: Li, Yaowei, et al.
Pubblicazione: (2024)
di: Li, Yaowei, et al.
Pubblicazione: (2024)
Deep Learning for Video Anomaly Detection: A Review
di: Wu, Peng, et al.
Pubblicazione: (2024)
di: Wu, Peng, et al.
Pubblicazione: (2024)
MAGIC: Few-Shot Mask-Guided Anomaly Inpainting with Prompt Perturbation, Spatially Adaptive Guidance, and Context Awareness
di: Choi, JaeHyuck, et al.
Pubblicazione: (2025)
di: Choi, JaeHyuck, et al.
Pubblicazione: (2025)
LLMI3D: MLLM-based 3D Perception from a Single 2D Image
di: Yang, Fan, et al.
Pubblicazione: (2024)
di: Yang, Fan, et al.
Pubblicazione: (2024)
ConsDreamer: Advancing Multi-View Consistency for Zero-Shot Text-to-3D Generation
di: Zhou, Yuan, et al.
Pubblicazione: (2025)
di: Zhou, Yuan, et al.
Pubblicazione: (2025)
VMonarch: Efficient Video Diffusion Transformers with Structured Attention
di: Liang, Cheng, et al.
Pubblicazione: (2026)
di: Liang, Cheng, et al.
Pubblicazione: (2026)
Investigating Redundancy in Multimodal Large Language Models with Multiple Vision Encoders
di: Wang, Yizhou, et al.
Pubblicazione: (2025)
di: Wang, Yizhou, et al.
Pubblicazione: (2025)
CamViG: Camera Aware Image-to-Video Generation with Multimodal Transformers
di: Marmon, Andrew, et al.
Pubblicazione: (2024)
di: Marmon, Andrew, et al.
Pubblicazione: (2024)
Documenti analoghi
-
A Fresh Look at Generalized Category Discovery through Non-negative Matrix Factorization
di: Ji, Zhong, et al.
Pubblicazione: (2024) -
iEBAKER: Improved Remote Sensing Image-Text Retrieval Framework via Eliminate Before Align and Keyword Explicit Reasoning
di: Zhang, Yan, et al.
Pubblicazione: (2025) -
DAOVI: Distortion-Aware Omnidirectional Video Inpainting
di: Seshimo, Ryosuke, et al.
Pubblicazione: (2025) -
COCO-Inpaint: A Benchmark for Detecting and Localizing Inpainting-Based Image Manipulations
di: Yan, Haozhen, et al.
Pubblicazione: (2025) -
Relation-Aware Meta-Learning for Zero-shot Sketch-Based Image Retrieval
di: Liu, Yang, et al.
Pubblicazione: (2024)