Improving Reconstruction of Representation Autoencoder
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Siyu, Qin, Chujie, Yin, Hubery, Yan, Qixin, Duan, Zheng-Peng, Li, Chen, Lyu, Jing, Guo, Chun-Le, Li, Chongyi |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
TimeMachine: Fine-Grained Facial Age Editing with Identity Preservation
di: Mi, Yilin, et al.
Pubblicazione: (2025)
di: Mi, Yilin, et al.
Pubblicazione: (2025)
Stand-In: A Lightweight and Plug-and-Play Identity Control for Video Generation
di: Xue, Bowen, et al.
Pubblicazione: (2025)
di: Xue, Bowen, et al.
Pubblicazione: (2025)
RAM++: Robust Representation Learning via Adaptive Mask for All-in-One Image Restoration
di: Zhang, Zilong, et al.
Pubblicazione: (2025)
di: Zhang, Zilong, et al.
Pubblicazione: (2025)
DRM: Diffusion-based Reward Model With Step-wise Guidance
di: Zhang, Jaxon, et al.
Pubblicazione: (2026)
di: Zhang, Jaxon, et al.
Pubblicazione: (2026)
A Diffusion-Based Framework for Occluded Object Movement
di: Duan, Zheng-Peng, et al.
Pubblicazione: (2025)
di: Duan, Zheng-Peng, et al.
Pubblicazione: (2025)
PerTouch: VLM-Driven Agent for Personalized and Semantic Image Retouching
di: Chang, Zewei, et al.
Pubblicazione: (2025)
di: Chang, Zewei, et al.
Pubblicazione: (2025)
FaceMe: Robust Blind Face Restoration with Personal Identification
di: Liu, Siyu, et al.
Pubblicazione: (2025)
di: Liu, Siyu, et al.
Pubblicazione: (2025)
Lighting Every Darkness with 3DGS: Fast Training and Real-Time Rendering for HDR View Synthesis
di: Jin, Xin, et al.
Pubblicazione: (2024)
di: Jin, Xin, et al.
Pubblicazione: (2024)
VQRAE: Representation Quantization Autoencoders for Multimodal Understanding, Generation and Reconstruction
di: Du, Sinan, et al.
Pubblicazione: (2025)
di: Du, Sinan, et al.
Pubblicazione: (2025)
Tackling the Singularities at the Endpoints of Time Intervals in Diffusion Models
di: Zhang, Pengze, et al.
Pubblicazione: (2024)
di: Zhang, Pengze, et al.
Pubblicazione: (2024)
DiT4SR: Taming Diffusion Transformer for Real-World Image Super-Resolution
di: Duan, Zheng-Peng, et al.
Pubblicazione: (2025)
di: Duan, Zheng-Peng, et al.
Pubblicazione: (2025)
Synergistic Multiscale Detail Refinement via Intrinsic Supervision for Underwater Image Enhancement
di: Zhang, Dehuan, et al.
Pubblicazione: (2023)
di: Zhang, Dehuan, et al.
Pubblicazione: (2023)
VTinker: Guided Flow Upsampling and Texture Mapping for High-Resolution Video Frame Interpolation
di: Wu, Chenyang, et al.
Pubblicazione: (2025)
di: Wu, Chenyang, et al.
Pubblicazione: (2025)
Iterative Predictor-Critic Code Decoding for Real-World Image Dehazing
di: Fu, Jiayi, et al.
Pubblicazione: (2025)
di: Fu, Jiayi, et al.
Pubblicazione: (2025)
ETC: training-free diffusion models acceleration with Error-aware Trend Consistency
di: Xie, Jiajian, et al.
Pubblicazione: (2025)
di: Xie, Jiajian, et al.
Pubblicazione: (2025)
DiffRetouch: Using Diffusion to Retouch on the Shoulder of Experts
di: Duan, Zheng-Peng, et al.
Pubblicazione: (2024)
di: Duan, Zheng-Peng, et al.
Pubblicazione: (2024)
Time-Aware One Step Diffusion Network for Real-World Image Super-Resolution
di: Zhang, Tianyi, et al.
Pubblicazione: (2025)
di: Zhang, Tianyi, et al.
Pubblicazione: (2025)
FlowSteer: Guiding Few-Step Image Synthesis with Authentic Trajectories
di: Ke, Lei, et al.
Pubblicazione: (2025)
di: Ke, Lei, et al.
Pubblicazione: (2025)
IG-Diff: Complex Night Scene Restoration with Illumination-Guided Diffusion Model
di: Chen, Yifan, et al.
Pubblicazione: (2026)
di: Chen, Yifan, et al.
Pubblicazione: (2026)
DecQ: Detail-Condensing Queries for Enhanced Reconstruction and Generation in Representation Autoencoders
di: Wang, Tianhang, et al.
Pubblicazione: (2026)
di: Wang, Tianhang, et al.
Pubblicazione: (2026)
UltraLED: Learning to See Everything in Ultra-High Dynamic Range Scenes
di: Meng, Yuang, et al.
Pubblicazione: (2025)
di: Meng, Yuang, et al.
Pubblicazione: (2025)
FlowConsist: Make Your Flow Consistent with Real Trajectory
di: Zhang, Tianyi, et al.
Pubblicazione: (2026)
di: Zhang, Tianyi, et al.
Pubblicazione: (2026)
NOVA: Sparse Control, Dense Synthesis for Pair-Free Video Editing
di: Pan, Tianlin, et al.
Pubblicazione: (2026)
di: Pan, Tianlin, et al.
Pubblicazione: (2026)
Restore Anything with Masks: Leveraging Mask Image Modeling for Blind All-in-One Image Restoration
di: Qin, Chu-Jie, et al.
Pubblicazione: (2024)
di: Qin, Chu-Jie, et al.
Pubblicazione: (2024)
YOSE: You Only Select Essential Tokens for Efficient DiT-based Video Object Removal
di: Wu, Chenyang, et al.
Pubblicazione: (2026)
di: Wu, Chenyang, et al.
Pubblicazione: (2026)
3D-UIR: 3D Gaussian for Underwater 3D Scene Reconstruction via Physics Based Appearance-Medium Decoupling
di: Yuan, Jieyu, et al.
Pubblicazione: (2025)
di: Yuan, Jieyu, et al.
Pubblicazione: (2025)
Improved Baselines with Representation Autoencoders
di: Singh, Jaskirat, et al.
Pubblicazione: (2026)
di: Singh, Jaskirat, et al.
Pubblicazione: (2026)
AMSP-UOD: When Vortex Convolution and Stochastic Perturbation Meet Underwater Object Detection
di: Zhou, Jingchun, et al.
Pubblicazione: (2023)
di: Zhou, Jingchun, et al.
Pubblicazione: (2023)
TC-AE: Unlocking Token Capacity for Deep Compression Autoencoders
di: Li, Teng, et al.
Pubblicazione: (2026)
di: Li, Teng, et al.
Pubblicazione: (2026)
Kalman-Inspired Feature Propagation for Video Face Super-Resolution
di: Feng, Ruicheng, et al.
Pubblicazione: (2024)
di: Feng, Ruicheng, et al.
Pubblicazione: (2024)
Scaling Text-to-Image Diffusion Transformers with Representation Autoencoders
di: Tong, Shengbang, et al.
Pubblicazione: (2026)
di: Tong, Shengbang, et al.
Pubblicazione: (2026)
Disentangled Learning Improves Implicit Neural Representations for Medical Reconstruction
di: Wu, Qing, et al.
Pubblicazione: (2026)
di: Wu, Qing, et al.
Pubblicazione: (2026)
SARMAE: Masked Autoencoder for SAR Representation Learning
di: Liu, Danxu, et al.
Pubblicazione: (2025)
di: Liu, Danxu, et al.
Pubblicazione: (2025)
Mixed Autoencoder for Self-supervised Visual Representation Learning
di: Chen, Kai, et al.
Pubblicazione: (2023)
di: Chen, Kai, et al.
Pubblicazione: (2023)
Audio-Omni: Extending Multi-modal Understanding to Versatile Audio Generation and Editing
di: Tian, Zeyue, et al.
Pubblicazione: (2026)
di: Tian, Zeyue, et al.
Pubblicazione: (2026)
DiverseAR: Boosting Diversity in Bitwise Autoregressive Image Generation
di: Yang, Ying, et al.
Pubblicazione: (2025)
di: Yang, Ying, et al.
Pubblicazione: (2025)
DiffPMAE: Diffusion Masked Autoencoders for Point Cloud Reconstruction
di: Li, Yanlong, et al.
Pubblicazione: (2023)
di: Li, Yanlong, et al.
Pubblicazione: (2023)
Diffusion Transformers with Representation Autoencoders
di: Zheng, Boyang, et al.
Pubblicazione: (2025)
di: Zheng, Boyang, et al.
Pubblicazione: (2025)
PSHuman: Photorealistic Single-image 3D Human Reconstruction using Cross-Scale Multiview Diffusion and Explicit Remeshing
di: Li, Peng, et al.
Pubblicazione: (2024)
di: Li, Peng, et al.
Pubblicazione: (2024)
PUGAN: Physical Model-Guided Underwater Image Enhancement Using GAN with Dual-Discriminators
di: Cong, Runmin, et al.
Pubblicazione: (2023)
di: Cong, Runmin, et al.
Pubblicazione: (2023)
Documenti analoghi
-
TimeMachine: Fine-Grained Facial Age Editing with Identity Preservation
di: Mi, Yilin, et al.
Pubblicazione: (2025) -
Stand-In: A Lightweight and Plug-and-Play Identity Control for Video Generation
di: Xue, Bowen, et al.
Pubblicazione: (2025) -
RAM++: Robust Representation Learning via Adaptive Mask for All-in-One Image Restoration
di: Zhang, Zilong, et al.
Pubblicazione: (2025) -
DRM: Diffusion-based Reward Model With Step-wise Guidance
di: Zhang, Jaxon, et al.
Pubblicazione: (2026) -
A Diffusion-Based Framework for Occluded Object Movement
di: Duan, Zheng-Peng, et al.
Pubblicazione: (2025)