Structured 3D Latents Are Surprisingly Powerful: Unleashing Generalizable Style with 2D Diffusion
Fuente:
arXiv
Saved in:
| Main Authors: | Qiao, Yiran, Lu, Yiren, Zhou, Yunlai, Liu, Disheng, Hou, Linlin, Yang, Rui, Yin, Yu, Ma, Jing |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DefenseSplat: Enhancing the Robustness of 3D Gaussian Splatting via Frequency-Aware Filtering
by: Qiao, Yiran, et al.
Published: (2026)
by: Qiao, Yiran, et al.
Published: (2026)
CAUSAL3D: A Comprehensive Benchmark for Causal Learning from Visual Data
by: Liu, Disheng, et al.
Published: (2025)
by: Liu, Disheng, et al.
Published: (2025)
AdvSplat: Adversarial Attacks on Feed-Forward Gaussian Splatting Models
by: Qiao, Yiran, et al.
Published: (2026)
by: Qiao, Yiran, et al.
Published: (2026)
GSMem: 3D Gaussian Splatting as Persistent Spatial Memory for Zero-Shot Embodied Exploration and Reasoning
by: Lu, Yiren, et al.
Published: (2026)
by: Lu, Yiren, et al.
Published: (2026)
BARD-GS: Blur-Aware Reconstruction of Dynamic Scenes via Gaussian Splatting
by: Lu, Yiren, et al.
Published: (2025)
by: Lu, Yiren, et al.
Published: (2025)
Segment then Splat: Unified 3D Open-Vocabulary Segmentation via Gaussian Splatting
by: Lu, Yiren, et al.
Published: (2025)
by: Lu, Yiren, et al.
Published: (2025)
Counterfactual Visual Explanation via Causally-Guided Adversarial Steering
by: Qiao, Yiran, et al.
Published: (2025)
by: Qiao, Yiran, et al.
Published: (2025)
When 'YES' Meets 'BUT': Can Large Models Comprehend Contradictory Humor Through Comparative Reasoning?
by: Liang, Tuo, et al.
Published: (2025)
by: Liang, Tuo, et al.
Published: (2025)
Cracking the Code of Juxtaposition: Can AI Models Understand the Humorous Contradictions
by: Hu, Zhe, et al.
Published: (2024)
by: Hu, Zhe, et al.
Published: (2024)
Certified Causal Defense with Generalizable Robustness
by: Qiao, Yiran, et al.
Published: (2024)
by: Qiao, Yiran, et al.
Published: (2024)
MorphAny3D: Unleashing the Power of Structured Latent in 3D Morphing
by: Sun, Xiaokun, et al.
Published: (2026)
by: Sun, Xiaokun, et al.
Published: (2026)
View-consistent Object Removal in Radiance Fields
by: Lu, Yiren, et al.
Published: (2024)
by: Lu, Yiren, et al.
Published: (2024)
Controllable 3D Face Generation with Conditional Style Code Diffusion
by: Shen, Xiaolong, et al.
Published: (2023)
by: Shen, Xiaolong, et al.
Published: (2023)
Unleashing the potential: AI empowered advanced metasurface research
by: Yunlai Fu, et al.
Published: (2024)
by: Yunlai Fu, et al.
Published: (2024)
DiffuseST: Unleashing the Capability of the Diffusion Model for Style Transfer
by: Hu, Ying, et al.
Published: (2024)
by: Hu, Ying, et al.
Published: (2024)
Distractor-free Generalizable 3D Gaussian Splatting
by: Bao, Yanqi, et al.
Published: (2024)
by: Bao, Yanqi, et al.
Published: (2024)
Unleashing the Potential of the Semantic Latent Space in Diffusion Models for Image Dehazing
by: Yang, Zizheng, et al.
Published: (2025)
by: Yang, Zizheng, et al.
Published: (2025)
GeoWizard: Unleashing the Diffusion Priors for 3D Geometry Estimation from a Single Image
by: Fu, Xiao, et al.
Published: (2024)
by: Fu, Xiao, et al.
Published: (2024)
Native and Compact Structured Latents for 3D Generation
by: Xiang, Jianfeng, et al.
Published: (2025)
by: Xiang, Jianfeng, et al.
Published: (2025)
Topology-Aware Latent Diffusion for 3D Shape Generation
by: Hu, Jiangbei, et al.
Published: (2024)
by: Hu, Jiangbei, et al.
Published: (2024)
StyleDecoupler: Generalizable Artistic Style Disentanglement
by: Jia, Zexi, et al.
Published: (2026)
by: Jia, Zexi, et al.
Published: (2026)
HuGDiffusion: Generalizable Single-Image Human Rendering via 3D Gaussian Diffusion
by: Tang, Yingzhi, et al.
Published: (2025)
by: Tang, Yingzhi, et al.
Published: (2025)
DAG: Unleash the Potential of Diffusion Model for Open-Vocabulary 3D Affordance Grounding
by: Wang, Hanqing, et al.
Published: (2025)
by: Wang, Hanqing, et al.
Published: (2025)
Structured 3D Latents for Scalable and Versatile 3D Generation
by: Xiang, Jianfeng, et al.
Published: (2024)
by: Xiang, Jianfeng, et al.
Published: (2024)
Unleashing the Power of Chain-of-Prediction for Monocular 3D Object Detection
by: Zhang, Zhihao, et al.
Published: (2025)
by: Zhang, Zhihao, et al.
Published: (2025)
SSGaussian: Semantic-Aware and Structure-Preserving 3D Style Transfer
by: Xu, Jimin, et al.
Published: (2025)
by: Xu, Jimin, et al.
Published: (2025)
VISTA: Triplet-Supervised Video Style Transfer with Diffusion Transformers
by: Song, Yiren, et al.
Published: (2026)
by: Song, Yiren, et al.
Published: (2026)
GeoSAM2: Unleashing the Power of SAM2 for 3D Part Segmentation
by: Deng, Ken, et al.
Published: (2025)
by: Deng, Ken, et al.
Published: (2025)
VP3D: Unleashing 2D Visual Prompt for Text-to-3D Generation
by: Chen, Yang, et al.
Published: (2024)
by: Chen, Yang, et al.
Published: (2024)
GraspLDP: Towards Generalizable Grasping Policy via Latent Diffusion
by: Xiang, Enda, et al.
Published: (2026)
by: Xiang, Enda, et al.
Published: (2026)
Generalizable Humanoid Manipulation with 3D Diffusion Policies
by: Ze, Yanjie, et al.
Published: (2024)
by: Ze, Yanjie, et al.
Published: (2024)
StructLDM: Structured Latent Diffusion for 3D Human Generation
by: Hu, Tao, et al.
Published: (2024)
by: Hu, Tao, et al.
Published: (2024)
SD-DiT: Unleashing the Power of Self-supervised Discrimination in Diffusion Transformer
by: Zhu, Rui, et al.
Published: (2024)
by: Zhu, Rui, et al.
Published: (2024)
GenDP: 3D Semantic Fields for Category-Level Generalizable Diffusion Policy
by: Wang, Yixuan, et al.
Published: (2024)
by: Wang, Yixuan, et al.
Published: (2024)
Unleashing the Potential of Neighbors: Diffusion-based Latent Neighbor Generation for Session-based Recommendation
by: Yang, Yuhan, et al.
Published: (2026)
by: Yang, Yuhan, et al.
Published: (2026)
EEG-Driven 3D Object Reconstruction with Style Consistency and Diffusion Prior
by: Xiang, Xin, et al.
Published: (2024)
by: Xiang, Xin, et al.
Published: (2024)
Unleashing the Potential of Pre-Trained Diffusion Models for Generalizable Person Re-Identification
by: Li, Jiachen, et al.
Published: (2025)
by: Li, Jiachen, et al.
Published: (2025)
Tune-Your-Style: Intensity-tunable 3D Style Transfer with Gaussian Splatting
by: Zhao, Yian, et al.
Published: (2026)
by: Zhao, Yian, et al.
Published: (2026)
WMAdapter: Adding WaterMark Control to Latent Diffusion Models
by: Ci, Hai, et al.
Published: (2024)
by: Ci, Hai, et al.
Published: (2024)
Reconstruction Matters: Learning Geometry-Aligned BEV Representation through 3D Gaussian Splatting
by: Lu, Yiren, et al.
Published: (2026)
by: Lu, Yiren, et al.
Published: (2026)
Similar Items
-
DefenseSplat: Enhancing the Robustness of 3D Gaussian Splatting via Frequency-Aware Filtering
by: Qiao, Yiran, et al.
Published: (2026) -
CAUSAL3D: A Comprehensive Benchmark for Causal Learning from Visual Data
by: Liu, Disheng, et al.
Published: (2025) -
AdvSplat: Adversarial Attacks on Feed-Forward Gaussian Splatting Models
by: Qiao, Yiran, et al.
Published: (2026) -
GSMem: 3D Gaussian Splatting as Persistent Spatial Memory for Zero-Shot Embodied Exploration and Reasoning
by: Lu, Yiren, et al.
Published: (2026) -
BARD-GS: Blur-Aware Reconstruction of Dynamic Scenes via Gaussian Splatting
by: Lu, Yiren, et al.
Published: (2025)