Semantics Lead the Way: Harmonizing Semantic and Texture Modeling with Asynchronous Latent Diffusion
Fuente:
arXiv
Saved in:
| Main Authors: | Pan, Yueming, Feng, Ruoyu, Dai, Qi, Wang, Yuqi, Lin, Wenfeng, Guo, Mingyu, Luo, Chong, Zheng, Nanning |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GlobalPaint: Spatiotemporal Coherent Video Outpainting with Global Feature Guidance
by: Pan, Yueming, et al.
Published: (2026)
by: Pan, Yueming, et al.
Published: (2026)
Bernini: Latent Semantic Planning for Video Diffusion
by: Bernini Team, et al.
Published: (2026)
by: Bernini Team, et al.
Published: (2026)
EVA: Mixture-of-Experts Semantic Variant Alignment for Compositional Zero-Shot Learning
by: Zhang, Xiao, et al.
Published: (2025)
by: Zhang, Xiao, et al.
Published: (2025)
Semantic Human Mesh Reconstruction with Textures
by: Zhan, Xiaoyu, et al.
Published: (2024)
by: Zhan, Xiaoyu, et al.
Published: (2024)
FlashPortrait: 6x Faster Infinite Portrait Animation with Adaptive Latent Prediction
by: Tu, Shuyuan, et al.
Published: (2025)
by: Tu, Shuyuan, et al.
Published: (2025)
VFM-VAE: Vision Foundation Models Can Be Good Tokenizers for Latent Diffusion Models
by: Bi, Tianci, et al.
Published: (2025)
by: Bi, Tianci, et al.
Published: (2025)
Open-Vocabulary Animal Keypoint Detection with Semantic-feature Matching
by: Zhang, Hao, et al.
Published: (2023)
by: Zhang, Hao, et al.
Published: (2023)
Depth-guided Texture Diffusion for Image Semantic Segmentation
by: Sun, Wei, et al.
Published: (2024)
by: Sun, Wei, et al.
Published: (2024)
SCASeg: Strip Cross-Attention for Efficient Semantic Segmentation
by: Xu, Guoan, et al.
Published: (2024)
by: Xu, Guoan, et al.
Published: (2024)
Mesh-Learner: Texturing Mesh with Spherical Harmonics
by: Wan, Yunfei, et al.
Published: (2025)
by: Wan, Yunfei, et al.
Published: (2025)
Unleashing the Potential of the Semantic Latent Space in Diffusion Models for Image Dehazing
by: Yang, Zizheng, et al.
Published: (2025)
by: Yang, Zizheng, et al.
Published: (2025)
StructVPR++: Distill Structural and Semantic Knowledge with Weighting Samples for Visual Place Recognition
by: Shen, Yanqing, et al.
Published: (2025)
by: Shen, Yanqing, et al.
Published: (2025)
Adversarially Domain-adaptive Latent Diffusion for Unsupervised Semantic Segmentation
by: Yu, Jongmin, et al.
Published: (2024)
by: Yu, Jongmin, et al.
Published: (2024)
Semantic-Enriched Latent Visual Reasoning
by: Xu, Tianrun, et al.
Published: (2026)
by: Xu, Tianrun, et al.
Published: (2026)
Dual-Representation Image Compression at Ultra-Low Bitrates via Explicit Semantics and Implicit Textures
by: Zhou, Chuqin, et al.
Published: (2026)
by: Zhou, Chuqin, et al.
Published: (2026)
Controllable Face Synthesis with Semantic Latent Diffusion Models
by: Ergasti, Alex, et al.
Published: (2024)
by: Ergasti, Alex, et al.
Published: (2024)
DiffHarmony: Latent Diffusion Model Meets Image Harmonization
by: Zhou, Pengfei, et al.
Published: (2024)
by: Zhou, Pengfei, et al.
Published: (2024)
NaTex: Seamless Texture Generation as Latent Color Diffusion
by: Lai, Zeqiang, et al.
Published: (2025)
by: Lai, Zeqiang, et al.
Published: (2025)
REGLUE Your Latents with Global and Local Semantics for Entangled Diffusion
by: Petsangourakis, Giorgos, et al.
Published: (2025)
by: Petsangourakis, Giorgos, et al.
Published: (2025)
SDiT: Semantic Region-Adaptive for Diffusion Transformers
by: Lin, Bowen, et al.
Published: (2026)
by: Lin, Bowen, et al.
Published: (2026)
ToProVAR: Efficient Visual Autoregressive Modeling via Tri-Dimensional Entropy-Aware Semantic Analysis and Sparsity Optimization
by: Chen, Jiayu, et al.
Published: (2026)
by: Chen, Jiayu, et al.
Published: (2026)
Diffusion in Diffusion: Cyclic One-Way Diffusion for Text-Vision-Conditioned Generation
by: Wang, Ruoyu, et al.
Published: (2023)
by: Wang, Ruoyu, et al.
Published: (2023)
CTGAN: Semantic-guided Conditional Texture Generator for 3D Shapes
by: Pan, Yi-Ting, et al.
Published: (2024)
by: Pan, Yi-Ting, et al.
Published: (2024)
FlowSSC: Universal Generative Monocular Semantic Scene Completion via One-Step Latent Diffusion
by: Xi, Zichen, et al.
Published: (2026)
by: Xi, Zichen, et al.
Published: (2026)
CCEdit: Creative and Controllable Video Editing via Diffusion Models
by: Feng, Ruoyu, et al.
Published: (2023)
by: Feng, Ruoyu, et al.
Published: (2023)
On the Influence of Shape, Texture and Color for Learning Semantic Segmentation
by: Mütze, Annika, et al.
Published: (2024)
by: Mütze, Annika, et al.
Published: (2024)
Structural and Statistical Texture Knowledge Distillation for Semantic Segmentation
by: Ji, Deyi, et al.
Published: (2023)
by: Ji, Deyi, et al.
Published: (2023)
Towards Self-Improvement of Diffusion Models via Group Preference Optimization
by: Chen, Renjie, et al.
Published: (2025)
by: Chen, Renjie, et al.
Published: (2025)
LaMD: Latent Motion Diffusion for Image-Conditional Video Generation
by: Hu, Yaosi, et al.
Published: (2023)
by: Hu, Yaosi, et al.
Published: (2023)
The Prism Hypothesis: Harmonizing Semantic and Pixel Representations via Unified Autoencoding
by: Fan, Weichen, et al.
Published: (2025)
by: Fan, Weichen, et al.
Published: (2025)
TNet: Terrace Convolutional Decoder Network for Remote Sensing Image Semantic Segmentation
by: Dai, Chengqian, et al.
Published: (2025)
by: Dai, Chengqian, et al.
Published: (2025)
Generalizing WiFi Gesture Recognition via Large-Model-Aware Semantic Distillation and Alignment
by: Cui, Feng-Qi, et al.
Published: (2025)
by: Cui, Feng-Qi, et al.
Published: (2025)
ContentV: Efficient Training of Video Generation Models with Limited Compute
by: Lin, Wenfeng, et al.
Published: (2025)
by: Lin, Wenfeng, et al.
Published: (2025)
Discovering Interpretable Directions in the Semantic Latent Space of Diffusion Models
by: Haas, René, et al.
Published: (2023)
by: Haas, René, et al.
Published: (2023)
VEOcc: Voxel-Centric Online Semantic Occupancy Prediction For Embodied Scene Understanding
by: Wang, Ruoyu, et al.
Published: (2026)
by: Wang, Ruoyu, et al.
Published: (2026)
DSNet: A Novel Way to Use Atrous Convolutions in Semantic Segmentation
by: Guo, Zilu, et al.
Published: (2024)
by: Guo, Zilu, et al.
Published: (2024)
Semantically Structured Image Compression via Irregular Group-Based Decoupling
by: Feng, Ruoyu, et al.
Published: (2023)
by: Feng, Ruoyu, et al.
Published: (2023)
SpecEdit: Training-Free Acceleration for Diffusion based Image Editing via Semantic Locking
by: Yan, Zhengan, et al.
Published: (2026)
by: Yan, Zhengan, et al.
Published: (2026)
SparseOcc: Rethinking Sparse Latent Representation for Vision-Based Semantic Occupancy Prediction
by: Tang, Pin, et al.
Published: (2024)
by: Tang, Pin, et al.
Published: (2024)
Towards Robust Semantic Correspondence: A Benchmark and Insights
by: Chong, Wenyue
Published: (2025)
by: Chong, Wenyue
Published: (2025)
Similar Items
-
GlobalPaint: Spatiotemporal Coherent Video Outpainting with Global Feature Guidance
by: Pan, Yueming, et al.
Published: (2026) -
Bernini: Latent Semantic Planning for Video Diffusion
by: Bernini Team, et al.
Published: (2026) -
EVA: Mixture-of-Experts Semantic Variant Alignment for Compositional Zero-Shot Learning
by: Zhang, Xiao, et al.
Published: (2025) -
Semantic Human Mesh Reconstruction with Textures
by: Zhan, Xiaoyu, et al.
Published: (2024) -
FlashPortrait: 6x Faster Infinite Portrait Animation with Adaptive Latent Prediction
by: Tu, Shuyuan, et al.
Published: (2025)