Saved in:
| Main Author: | Xu, Philip |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2512.22294 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GAP3D: Generative Alignment of VLM Latents to Patch-Level Embeddings for 3D Generation
by: Gkotsi, Polytimi Anna, et al.
Published: (2026)
by: Gkotsi, Polytimi Anna, et al.
Published: (2026)
PartRAG: Retrieval-Augmented Part-Level 3D Generation and Editing
by: Li, Peize, et al.
Published: (2026)
by: Li, Peize, et al.
Published: (2026)
Controllable Generation of Large-Scale 3D Urban Layouts with Semantic and Structural Guidance
by: Niu, Mengyuan, et al.
Published: (2025)
by: Niu, Mengyuan, et al.
Published: (2025)
REPARO: Compositional 3D Assets Generation with Differentiable 3D Layout Alignment
by: Han, Haonan, et al.
Published: (2024)
by: Han, Haonan, et al.
Published: (2024)
DreamControl: Control-Based Text-to-3D Generation with 3D Self-Prior
by: Huang, Tianyu, et al.
Published: (2023)
by: Huang, Tianyu, et al.
Published: (2023)
Hyperbolic Hierarchical Alignment Reasoning Network for Text-3D Retrieval
by: Li, Wenrui, et al.
Published: (2025)
by: Li, Wenrui, et al.
Published: (2025)
3DAlign-DAER: Dynamic Attention Policy and Efficient Retrieval Strategy for Fine-grained 3D-Text Alignment at Scale
by: Fan, Yijia, et al.
Published: (2025)
by: Fan, Yijia, et al.
Published: (2025)
PLA4D: Pixel-Level Alignments for Text-to-4D Gaussian Splatting
by: Miao, Qiaowei, et al.
Published: (2024)
by: Miao, Qiaowei, et al.
Published: (2024)
DAC: 2D-3D Retrieval with Noisy Labels via Divide-and-Conquer Alignment and Correction
by: Gan, Chaofan, et al.
Published: (2024)
by: Gan, Chaofan, et al.
Published: (2024)
InterAct: Advancing Large-Scale Versatile 3D Human-Object Interaction Generation
by: Xu, Sirui, et al.
Published: (2025)
by: Xu, Sirui, et al.
Published: (2025)
One4D: Unified 4D Generation and Reconstruction via Decoupled LoRA Control
by: Mi, Zhenxing, et al.
Published: (2025)
by: Mi, Zhenxing, et al.
Published: (2025)
IL3D: A Large-Scale Indoor Layout Dataset for LLM-Driven 3D Scene Generation
by: Zhou, Wenxu, et al.
Published: (2025)
by: Zhou, Wenxu, et al.
Published: (2025)
MCA: 2D-3D Retrieval with Noisy Labels via Multi-level Adaptive Correction and Alignment
by: Zou, Gui, et al.
Published: (2025)
by: Zou, Gui, et al.
Published: (2025)
CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation
by: Wang, Qinghe, et al.
Published: (2025)
by: Wang, Qinghe, et al.
Published: (2025)
Leveling3D: Leveling Up 3D Reconstruction with Feed-Forward 3D Gaussian Splatting and Geometry-Aware Generation
by: Huang, Yiming, et al.
Published: (2026)
by: Huang, Yiming, et al.
Published: (2026)
BrightDreamer: Generic 3D Gaussian Generative Framework for Fast Text-to-3D Synthesis
by: Jiang, Lutao, et al.
Published: (2024)
by: Jiang, Lutao, et al.
Published: (2024)
COMOGen: A Controllable Text-to-3D Multi-object Generation Framework
by: Sun, Shaorong, et al.
Published: (2024)
by: Sun, Shaorong, et al.
Published: (2024)
4DWorldBench: A Comprehensive Evaluation Framework for 3D/4D World Generation Models
by: Lu, Yiting, et al.
Published: (2025)
by: Lu, Yiting, et al.
Published: (2025)
VGGHeads: 3D Multi Head Alignment with a Large-Scale Synthetic Dataset
by: Kupyn, Orest, et al.
Published: (2024)
by: Kupyn, Orest, et al.
Published: (2024)
PointAlign: Feature-Level Alignment Regularization for 3D Vision-Language Models
by: Su, Yuanhao, et al.
Published: (2026)
by: Su, Yuanhao, et al.
Published: (2026)
Direct3D: Scalable Image-to-3D Generation via 3D Latent Diffusion Transformer
by: Wu, Shuang, et al.
Published: (2024)
by: Wu, Shuang, et al.
Published: (2024)
3D Vision and Language Pretraining with Large-Scale Synthetic Data
by: Yang, Dejie, et al.
Published: (2024)
by: Yang, Dejie, et al.
Published: (2024)
CLAY: A Controllable Large-scale Generative Model for Creating High-quality 3D Assets
by: Zhang, Longwen, et al.
Published: (2024)
by: Zhang, Longwen, et al.
Published: (2024)
LASER: Layer-wise Scale Alignment for Training-Free Streaming 4D Reconstruction
by: Ding, Tianye, et al.
Published: (2025)
by: Ding, Tianye, et al.
Published: (2025)
SesaHand: Enhancing 3D Hand Reconstruction via Controllable Generation with Semantic and Structural Alignment
by: Zhao, Zhuoran, et al.
Published: (2026)
by: Zhao, Zhuoran, et al.
Published: (2026)
Hunyuan3D-Omni: A Unified Framework for Controllable Generation of 3D Assets
by: Hunyuan3D, Team, et al.
Published: (2025)
by: Hunyuan3D, Team, et al.
Published: (2025)
3D-LSPTM: An Automatic Framework with 3D-Large-Scale Pretrained Model for Laryngeal Cancer Detection Using Laryngoscopic Videos
by: Qiu, Meiyu, et al.
Published: (2024)
by: Qiu, Meiyu, et al.
Published: (2024)
3DWG: 3D Weakly Supervised Visual Grounding via Category and Instance-Level Alignment
by: Li, Xiaoqi, et al.
Published: (2025)
by: Li, Xiaoqi, et al.
Published: (2025)
Pixel-to-4D: Camera-Controlled Image-to-Video Generation with Dynamic 3D Gaussians
by: de Almeida, Melonie, et al.
Published: (2026)
by: de Almeida, Melonie, et al.
Published: (2026)
Looking 3D: Anomaly Detection with 2D-3D Alignment
by: Bhunia, Ankan, et al.
Published: (2024)
by: Bhunia, Ankan, et al.
Published: (2024)
iControl3D: An Interactive System for Controllable 3D Scene Generation
by: Li, Xingyi, et al.
Published: (2024)
by: Li, Xingyi, et al.
Published: (2024)
HOTS3D: Hyper-Spherical Optimal Transport for Semantic Alignment of Text-to-3D Generation
by: Li, Zezeng, et al.
Published: (2024)
by: Li, Zezeng, et al.
Published: (2024)
A3D: Does Diffusion Dream about 3D Alignment?
by: Ignatyev, Savva, et al.
Published: (2024)
by: Ignatyev, Savva, et al.
Published: (2024)
LPA3D: 3D Room-Level Scene Generation from In-the-Wild Images
by: Yang, Ming-Jia, et al.
Published: (2025)
by: Yang, Ming-Jia, et al.
Published: (2025)
ImageNet3D: Towards General-Purpose Object-Level 3D Understanding
by: Ma, Wufei, et al.
Published: (2024)
by: Ma, Wufei, et al.
Published: (2024)
RAISECity: A Multimodal Agent Framework for Reality-Aligned 3D World Generation at City-Scale
by: Wang, Shengyuan, et al.
Published: (2025)
by: Wang, Shengyuan, et al.
Published: (2025)
SplatFont3D: Structure-Aware Text-to-3D Artistic Font Generation with Part-Level Style Control
by: Gan, Ji, et al.
Published: (2025)
by: Gan, Ji, et al.
Published: (2025)
SC4D: Sparse-Controlled Video-to-4D Generation and Motion Transfer
by: Wu, Zijie, et al.
Published: (2024)
by: Wu, Zijie, et al.
Published: (2024)
Learning Segment Similarity and Alignment in Large-Scale Content Based Video Retrieval
by: Jiang, Chen, et al.
Published: (2023)
by: Jiang, Chen, et al.
Published: (2023)
LAM3D: Large Image-Point-Cloud Alignment Model for 3D Reconstruction from Single Image
by: Cui, Ruikai, et al.
Published: (2024)
by: Cui, Ruikai, et al.
Published: (2024)
Similar Items
-
GAP3D: Generative Alignment of VLM Latents to Patch-Level Embeddings for 3D Generation
by: Gkotsi, Polytimi Anna, et al.
Published: (2026) -
PartRAG: Retrieval-Augmented Part-Level 3D Generation and Editing
by: Li, Peize, et al.
Published: (2026) -
Controllable Generation of Large-Scale 3D Urban Layouts with Semantic and Structural Guidance
by: Niu, Mengyuan, et al.
Published: (2025) -
REPARO: Compositional 3D Assets Generation with Differentiable 3D Layout Alignment
by: Han, Haonan, et al.
Published: (2024) -
DreamControl: Control-Based Text-to-3D Generation with 3D Self-Prior
by: Huang, Tianyu, et al.
Published: (2023)