LIPT: Latency-aware Image Processing Transformer
Fuente:
arXiv
Salvato in:
| Autori principali: | Qiao, Junbo, Li, Wei, Xie, Haizhen, Chen, Hanting, Zhou, Yunshuai, Tu, Zhijun, Hu, Jie, Lin, Shaohui |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Dynamic Contrastive Knowledge Distillation for Efficient Image Restoration
di: Zhou, Yunshuai, et al.
Pubblicazione: (2024)
di: Zhou, Yunshuai, et al.
Pubblicazione: (2024)
Autoregressive Image Generation with Vision Full-view Prompt
di: Cai, Miaomiao, et al.
Pubblicazione: (2025)
di: Cai, Miaomiao, et al.
Pubblicazione: (2025)
IPT-V2: Efficient Image Processing Transformer using Hierarchical Attentions
di: Tu, Zhijun, et al.
Pubblicazione: (2024)
di: Tu, Zhijun, et al.
Pubblicazione: (2024)
Data Upcycling Knowledge Distillation for Image Super-Resolution
di: Zhang, Yun, et al.
Pubblicazione: (2023)
di: Zhang, Yun, et al.
Pubblicazione: (2023)
U-DiTs: Downsample Tokens in U-Shaped Diffusion Transformers
di: Tian, Yuchuan, et al.
Pubblicazione: (2024)
di: Tian, Yuchuan, et al.
Pubblicazione: (2024)
EAM: Enhancing Anything with Diffusion Transformers for Blind Super-Resolution
di: Xie, Haizhen, et al.
Pubblicazione: (2025)
di: Xie, Haizhen, et al.
Pubblicazione: (2025)
Knowledge Distillation with Multi-granularity Mixture of Priors for Image Super-Resolution
di: Li, Simiao, et al.
Pubblicazione: (2024)
di: Li, Simiao, et al.
Pubblicazione: (2024)
RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought
di: Qiao, Junbo, et al.
Pubblicazione: (2025)
di: Qiao, Junbo, et al.
Pubblicazione: (2025)
Allo{SR}$^2$: Rectifying One-Step Super-Resolution to Stay Real via Allomorphic Generative Flows
di: Wang, Zihan, et al.
Pubblicazione: (2026)
di: Wang, Zihan, et al.
Pubblicazione: (2026)
OS-DiffVSR: Towards One-step Latent Diffusion Model for High-detailed Real-world Video Super-Resolution
di: Li, Hanting, et al.
Pubblicazione: (2025)
di: Li, Hanting, et al.
Pubblicazione: (2025)
Hi-Mamba: Hierarchical Mamba for Efficient Image Super-Resolution
di: Qiao, Junbo, et al.
Pubblicazione: (2024)
di: Qiao, Junbo, et al.
Pubblicazione: (2024)
Beyond Textual CoT: Interleaved Text-Image Chains with Deep Confidence Reasoning for Image Editing
di: Zou, Zhentao, et al.
Pubblicazione: (2025)
di: Zou, Zhentao, et al.
Pubblicazione: (2025)
Instruct-IPT: All-in-One Image Processing Transformer via Weight Modulation
di: Tian, Yuchuan, et al.
Pubblicazione: (2024)
di: Tian, Yuchuan, et al.
Pubblicazione: (2024)
ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices
di: Du, Kunpeng, et al.
Pubblicazione: (2026)
di: Du, Kunpeng, et al.
Pubblicazione: (2026)
DSPO: Direct Semantic Preference Optimization for Real-World Image Super-Resolution
di: Cai, Miaomiao, et al.
Pubblicazione: (2025)
di: Cai, Miaomiao, et al.
Pubblicazione: (2025)
Effective Diffusion Transformer Architecture for Image Super-Resolution
di: Cheng, Kun, et al.
Pubblicazione: (2024)
di: Cheng, Kun, et al.
Pubblicazione: (2024)
One-Step Diffusion-based Real-World Image Super-Resolution with Visual Perception Distillation
di: Wu, Xue, et al.
Pubblicazione: (2025)
di: Wu, Xue, et al.
Pubblicazione: (2025)
From Sequential to Spatial: Reordering Autoregression for Efficient Visual Generation
di: Wang, Siyang, et al.
Pubblicazione: (2025)
di: Wang, Siyang, et al.
Pubblicazione: (2025)
GIM: A Million-scale Benchmark for Generative Image Manipulation Detection and Localization
di: Chen, Yirui, et al.
Pubblicazione: (2024)
di: Chen, Yirui, et al.
Pubblicazione: (2024)
Latency-aware Unified Dynamic Networks for Efficient Image Recognition
di: Han, Yizeng, et al.
Pubblicazione: (2023)
di: Han, Yizeng, et al.
Pubblicazione: (2023)
One Step Diffusion-based Super-Resolution with Time-Aware Distillation
di: He, Xiao, et al.
Pubblicazione: (2024)
di: He, Xiao, et al.
Pubblicazione: (2024)
Segmenting and Understanding: Region-aware Semantic Attention for Fine-grained Image Quality Assessment with Large Language Models
di: Song, Chenyue, et al.
Pubblicazione: (2025)
di: Song, Chenyue, et al.
Pubblicazione: (2025)
Distilling Semantic Priors from SAM to Efficient Image Restoration Models
di: Zhang, Quan, et al.
Pubblicazione: (2024)
di: Zhang, Quan, et al.
Pubblicazione: (2024)
LF-ViT: Reducing Spatial Redundancy in Vision Transformer for Efficient Image Recognition
di: Hu, Youbing, et al.
Pubblicazione: (2024)
di: Hu, Youbing, et al.
Pubblicazione: (2024)
Region-aware Image-based Human Action Retrieval with Transformers
di: Wang, Hongsong, et al.
Pubblicazione: (2024)
di: Wang, Hongsong, et al.
Pubblicazione: (2024)
SAIGFormer: A Spatially-Adaptive Illumination-Guided Network for Low-Light Image Enhancement
di: Li, Hanting, et al.
Pubblicazione: (2025)
di: Li, Hanting, et al.
Pubblicazione: (2025)
Unleashing Vision Transformer Potential In Image Quality Assessment via Global-Local Adaptive Interaction
di: Li, Yu, et al.
Pubblicazione: (2026)
di: Li, Yu, et al.
Pubblicazione: (2026)
Mixture of Ranks with Degradation-Aware Routing for One-Step Real-World Image Super-Resolution
di: He, Xiao, et al.
Pubblicazione: (2025)
di: He, Xiao, et al.
Pubblicazione: (2025)
CompBench: Benchmarking Complex Instruction-guided Image Editing
di: Jia, Bohan, et al.
Pubblicazione: (2025)
di: Jia, Bohan, et al.
Pubblicazione: (2025)
Omni-Dimensional Frequency Learner for General Time Series Analysis
di: Chen, Xianing, et al.
Pubblicazione: (2024)
di: Chen, Xianing, et al.
Pubblicazione: (2024)
Knowledge Distillation via the Target-aware Transformer
di: Lin, Sihao, et al.
Pubblicazione: (2022)
di: Lin, Sihao, et al.
Pubblicazione: (2022)
Collaboration of Teachers for Semi-supervised Object Detection
di: Chen, Liyu, et al.
Pubblicazione: (2024)
di: Chen, Liyu, et al.
Pubblicazione: (2024)
IRBridge: Solving Image Restoration Bridge with Pre-trained Generative Diffusion Models
di: Wang, Hanting, et al.
Pubblicazione: (2025)
di: Wang, Hanting, et al.
Pubblicazione: (2025)
A boundary-aware point clustering approach in Euclidean and embedding spaces for roof plane segmentation
di: Li, Li, et al.
Pubblicazione: (2023)
di: Li, Li, et al.
Pubblicazione: (2023)
Frequency-Aware Transformer for Learned Image Compression
di: Li, Han, et al.
Pubblicazione: (2023)
di: Li, Han, et al.
Pubblicazione: (2023)
Latency-aware Road Anomaly Segmentation in Videos: A Photorealistic Dataset and New Metrics
di: Tian, Beiwen, et al.
Pubblicazione: (2024)
di: Tian, Beiwen, et al.
Pubblicazione: (2024)
sketch2symm: Symmetry-aware sketch-to-shape generation via semantic bridging
di: Zhou, Yan, et al.
Pubblicazione: (2025)
di: Zhou, Yan, et al.
Pubblicazione: (2025)
INTENT: Invariance and Discrimination-aware Noise Mitigation for Robust Composed Image Retrieval
di: Chen, Zhiwei, et al.
Pubblicazione: (2026)
di: Chen, Zhiwei, et al.
Pubblicazione: (2026)
Demystify Transformers & Convolutions in Modern Image Deep Networks
di: Hu, Xiaowei, et al.
Pubblicazione: (2022)
di: Hu, Xiaowei, et al.
Pubblicazione: (2022)
Depth-agnostic Single Image Dehazing
di: Xu, Honglei, et al.
Pubblicazione: (2024)
di: Xu, Honglei, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Dynamic Contrastive Knowledge Distillation for Efficient Image Restoration
di: Zhou, Yunshuai, et al.
Pubblicazione: (2024) -
Autoregressive Image Generation with Vision Full-view Prompt
di: Cai, Miaomiao, et al.
Pubblicazione: (2025) -
IPT-V2: Efficient Image Processing Transformer using Hierarchical Attentions
di: Tu, Zhijun, et al.
Pubblicazione: (2024) -
Data Upcycling Knowledge Distillation for Image Super-Resolution
di: Zhang, Yun, et al.
Pubblicazione: (2023) -
U-DiTs: Downsample Tokens in U-Shaped Diffusion Transformers
di: Tian, Yuchuan, et al.
Pubblicazione: (2024)