RPiAE: A Representation-Pivoted Autoencoder Enhancing Both Image Generation and Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gong, Yue, Li, Hongyu, Liu, Shanyuan, Cheng, Bo, Ma, Yuhang, Wu, Liebucha, Wu, Xiaoyu, Zhang, Manyuan, Leng, Dawei, Yin, Yuhui, Zhang, Lijun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HiCo: Hierarchical Controllable Diffusion Model for Layout-to-image Generation
von: Cheng, Bo, et al.
Veröffentlicht: (2024)
von: Cheng, Bo, et al.
Veröffentlicht: (2024)
Bridge Diffusion Model: Bridge Chinese Text-to-Image Diffusion Model with English Communities
von: Liu, Shanyuan, et al.
Veröffentlicht: (2023)
von: Liu, Shanyuan, et al.
Veröffentlicht: (2023)
NAMI: Efficient Image Generation via Bridged Progressive Rectified Flow Transformers
von: Ma, Yuhang, et al.
Veröffentlicht: (2025)
von: Ma, Yuhang, et al.
Veröffentlicht: (2025)
CTA-Flux: Integrating Chinese Cultural Semantics into High-Quality English Text-to-Image Communities
von: Gong, Yue, et al.
Veröffentlicht: (2025)
von: Gong, Yue, et al.
Veröffentlicht: (2025)
NanoControl: A Lightweight Framework for Precise and Efficient Control in Diffusion Transformer
von: Liu, Shanyuan, et al.
Veröffentlicht: (2025)
von: Liu, Shanyuan, et al.
Veröffentlicht: (2025)
PlanGen: Towards Unified Layout Planning and Image Generation in Auto-Regressive Vision Language Models
von: He, Runze, et al.
Veröffentlicht: (2025)
von: He, Runze, et al.
Veröffentlicht: (2025)
FLUX-Makeup: High-Fidelity, Identity-Consistent, and Robust Makeup Transfer via Diffusion Transformer
von: Zhu, Jian, et al.
Veröffentlicht: (2025)
von: Zhu, Jian, et al.
Veröffentlicht: (2025)
RevealLayer: Disentangling Hidden and Visible Layers via Occlusion-Aware Image Decomposition
von: Wang, Binhao, et al.
Veröffentlicht: (2026)
von: Wang, Binhao, et al.
Veröffentlicht: (2026)
RefTon: Reference person shot assist virtual Try-on
von: Li, Liuzhuozheng, et al.
Veröffentlicht: (2025)
von: Li, Liuzhuozheng, et al.
Veröffentlicht: (2025)
U-StyDiT: Ultra-high Quality Artistic Style Transfer Using Diffusion Transformers
von: Zhang, Zhanjie, et al.
Veröffentlicht: (2025)
von: Zhang, Zhanjie, et al.
Veröffentlicht: (2025)
WISA: World Simulator Assistant for Physics-Aware Text-to-Video Generation
von: Wang, Jing, et al.
Veröffentlicht: (2025)
von: Wang, Jing, et al.
Veröffentlicht: (2025)
IAA: Inner-Adaptor Architecture Empowers Frozen Large Language Model with Multimodal Capabilities
von: Wang, Bin, et al.
Veröffentlicht: (2024)
von: Wang, Bin, et al.
Veröffentlicht: (2024)
ProteinAE: Protein Diffusion Autoencoders for Structure Encoding
von: Li, Shaoning, et al.
Veröffentlicht: (2025)
von: Li, Shaoning, et al.
Veröffentlicht: (2025)
RzenEmbed: Towards Comprehensive Multimodal Retrieval
von: Jian, Weijian, et al.
Veröffentlicht: (2025)
von: Jian, Weijian, et al.
Veröffentlicht: (2025)
Uni-Edit: Intelligent Editing Is A General Task For Unified Model Tuning
von: Zheng, Dian, et al.
Veröffentlicht: (2026)
von: Zheng, Dian, et al.
Veröffentlicht: (2026)
TC-AE: Unlocking Token Capacity for Deep Compression Autoencoders
von: Li, Teng, et al.
Veröffentlicht: (2026)
von: Li, Teng, et al.
Veröffentlicht: (2026)
SVD-AE: Simple Autoencoders for Collaborative Filtering
von: Hong, Seoyoung, et al.
Veröffentlicht: (2024)
von: Hong, Seoyoung, et al.
Veröffentlicht: (2024)
LMM-Det: Make Large Multimodal Models Excel in Object Detection
von: Li, Jincheng, et al.
Veröffentlicht: (2025)
von: Li, Jincheng, et al.
Veröffentlicht: (2025)
FG-CLIP: Fine-Grained Visual and Textual Alignment
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
KilonovAE: Exploring Kilonova Spectral Features with Autoencoders
von: Ford, N. M., et al.
Veröffentlicht: (2023)
von: Ford, N. M., et al.
Veröffentlicht: (2023)
Understanding Internal Representations of Recommendation Models with Sparse Autoencoders
von: Wang, Jiayin, et al.
Veröffentlicht: (2024)
von: Wang, Jiayin, et al.
Veröffentlicht: (2024)
Both Semantics and Reconstruction Matter: Making Representation Encoders Ready for Text-to-Image Generation and Editing
von: Zhang, Shilong, et al.
Veröffentlicht: (2025)
von: Zhang, Shilong, et al.
Veröffentlicht: (2025)
AE SemRL: Learning Semantic Association Rules with Autoencoders
von: Karabulut, Erkan, et al.
Veröffentlicht: (2024)
von: Karabulut, Erkan, et al.
Veröffentlicht: (2024)
Qihoo-T2X: An Efficient Proxy-Tokenized Diffusion Transformer for Text-to-Any-Task
von: Wang, Jing, et al.
Veröffentlicht: (2024)
von: Wang, Jing, et al.
Veröffentlicht: (2024)
Improved Baselines with Representation Autoencoders
von: Singh, Jaskirat, et al.
Veröffentlicht: (2026)
von: Singh, Jaskirat, et al.
Veröffentlicht: (2026)
RSAttAE: An Information-Aware Attention-based Autoencoder Recommender System
von: Taromi, Amirhossein Dadashzadeh, et al.
Veröffentlicht: (2025)
von: Taromi, Amirhossein Dadashzadeh, et al.
Veröffentlicht: (2025)
StrAE: Autoencoding for Pre-Trained Embeddings using Explicit Structure
von: Opper, Mattia, et al.
Veröffentlicht: (2023)
von: Opper, Mattia, et al.
Veröffentlicht: (2023)
FG-CLIP 2: A Bilingual Fine-grained Vision-Language Alignment Model
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
HaHeAE: Learning Generalisable Joint Representations of Human Hand and Head Movements in Extended Reality
von: Hu, Zhiming, et al.
Veröffentlicht: (2024)
von: Hu, Zhiming, et al.
Veröffentlicht: (2024)
Enhancing Text Authenticity: A Novel Hybrid Approach for AI-Generated Text Detection
von: Zhang, Ye, et al.
Veröffentlicht: (2024)
von: Zhang, Ye, et al.
Veröffentlicht: (2024)
Research on the Load Bearing and Impact Resistance of a Novel Structure Exhibiting Both Positive and Negative Poisson’s Ratios
von: Xidong Zhang, et al.
Veröffentlicht: (2024)
von: Xidong Zhang, et al.
Veröffentlicht: (2024)
Functional Autoencoder for Smoothing and Representation Learning
von: Wu, Sidi, et al.
Veröffentlicht: (2024)
von: Wu, Sidi, et al.
Veröffentlicht: (2024)
UniM$^2$AE: Multi-modal Masked Autoencoders with Unified 3D Representation for 3D Perception in Autonomous Driving
von: Zou, Jian, et al.
Veröffentlicht: (2023)
von: Zou, Jian, et al.
Veröffentlicht: (2023)
An injectable pH-responsive marine polysaccharide hydrogel (AE&LF@pOA) for sequential therapy of infected diabetic wounds.
von: Zhao, Meiyue, et al.
Veröffentlicht: (2026)
von: Zhao, Meiyue, et al.
Veröffentlicht: (2026)
TimeMAE: Self-Supervised Representations of Time Series with Decoupled Masked Autoencoders
von: Cheng, Mingyue, et al.
Veröffentlicht: (2023)
von: Cheng, Mingyue, et al.
Veröffentlicht: (2023)
AE-ViT: Token Enhancement for Vision Transformers via CNN-Based Autoencoder Ensembles
von: AIRCC
Veröffentlicht: (2025)
von: AIRCC
Veröffentlicht: (2025)
DNAEdit: Direct Noise Alignment for Text-Guided Rectified Flow Editing
von: Xie, Chenxi, et al.
Veröffentlicht: (2025)
von: Xie, Chenxi, et al.
Veröffentlicht: (2025)
threewater-dot/MvAE: MvAE
von: threewater-dot
Veröffentlicht: (2026)
von: threewater-dot
Veröffentlicht: (2026)
Improving Sparse Autoencoder with Dynamic Attention
von: Wang, Dongsheng, et al.
Veröffentlicht: (2026)
von: Wang, Dongsheng, et al.
Veröffentlicht: (2026)
DecQ: Detail-Condensing Queries for Enhanced Reconstruction and Generation in Representation Autoencoders
von: Wang, Tianhang, et al.
Veröffentlicht: (2026)
von: Wang, Tianhang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
HiCo: Hierarchical Controllable Diffusion Model for Layout-to-image Generation
von: Cheng, Bo, et al.
Veröffentlicht: (2024) -
Bridge Diffusion Model: Bridge Chinese Text-to-Image Diffusion Model with English Communities
von: Liu, Shanyuan, et al.
Veröffentlicht: (2023) -
NAMI: Efficient Image Generation via Bridged Progressive Rectified Flow Transformers
von: Ma, Yuhang, et al.
Veröffentlicht: (2025) -
CTA-Flux: Integrating Chinese Cultural Semantics into High-Quality English Text-to-Image Communities
von: Gong, Yue, et al.
Veröffentlicht: (2025) -
NanoControl: A Lightweight Framework for Precise and Efficient Control in Diffusion Transformer
von: Liu, Shanyuan, et al.
Veröffentlicht: (2025)