Text-Image Conditioned 3D Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Cen, Jiazhong, Fang, Jiemin, Li, Sikuang, Wu, Guanjun, Yang, Chen, Yi, Taoran, Zhou, Zanwei, Bao, Zhikuan, Xie, Lingxi, Shen, Wei, Tian, Qi |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
WorldGrow: Generating Infinite 3D World
by: Li, Sikuang, et al.
Published: (2025)
by: Li, Sikuang, et al.
Published: (2025)
UniLat3D: Geometry-Appearance Unified Latents for Single-Stage 3D Generation
by: Wu, Guanjun, et al.
Published: (2025)
by: Wu, Guanjun, et al.
Published: (2025)
Segment Anything in 3D with Radiance Fields
by: Cen, Jiazhong, et al.
Published: (2023)
by: Cen, Jiazhong, et al.
Published: (2023)
Few-step Flow for 3D Generation via Marginal-Data Transport Distillation
by: Zhou, Zanwei, et al.
Published: (2025)
by: Zhou, Zanwei, et al.
Published: (2025)
GaussianDreamerPro: Text to Manipulable 3D Gaussians with Highly Enhanced Quality
by: Yi, Taoran, et al.
Published: (2024)
by: Yi, Taoran, et al.
Published: (2024)
Segment Any 4D Gaussians
by: Ji, Shengxiang, et al.
Published: (2024)
by: Ji, Shengxiang, et al.
Published: (2024)
Segment Any 3D Gaussians
by: Cen, Jiazhong, et al.
Published: (2023)
by: Cen, Jiazhong, et al.
Published: (2023)
GaussianDreamer: Fast Generation from Text to 3D Gaussians by Bridging 2D and 3D Diffusion Models
by: Yi, Taoran, et al.
Published: (2023)
by: Yi, Taoran, et al.
Published: (2023)
Tackling View-Dependent Semantics in 3D Language Gaussian Splatting
by: Cen, Jiazhong, et al.
Published: (2025)
by: Cen, Jiazhong, et al.
Published: (2025)
4D Gaussian Splatting for Real-Time Dynamic Scene Rendering
by: Wu, Guanjun, et al.
Published: (2023)
by: Wu, Guanjun, et al.
Published: (2023)
GaussianObject: High-Quality 3D Object Reconstruction from Four Views with Gaussian Splatting
by: Yang, Chen, et al.
Published: (2024)
by: Yang, Chen, et al.
Published: (2024)
Dereflection Any Image with Diffusion Priors and Diversified Data
by: Hu, Jichen, et al.
Published: (2025)
by: Hu, Jichen, et al.
Published: (2025)
WorldAct: Activating Monolithic 3D Worlds into Interactive-Ready Object-Centric Scenes
by: Hu, Jichen, et al.
Published: (2026)
by: Hu, Jichen, et al.
Published: (2026)
GaussianEditor: Editing 3D Gaussians Delicately with Text Instructions
by: Wang, Junjie, et al.
Published: (2023)
by: Wang, Junjie, et al.
Published: (2023)
Fast High Dynamic Range Radiance Fields for Dynamic Scenes
by: Wu, Guanjun, et al.
Published: (2024)
by: Wu, Guanjun, et al.
Published: (2024)
LiftImage3D: Lifting Any Single Image to 3D Gaussians with Video Generation Priors
by: Chen, Yabo, et al.
Published: (2024)
by: Chen, Yabo, et al.
Published: (2024)
Cascade-Zero123: One Image to Highly Consistent 3D with Self-Prompted Nearby Views
by: Chen, Yabo, et al.
Published: (2023)
by: Chen, Yabo, et al.
Published: (2023)
Snap-Snap: Taking Two Images to Reconstruct 3D Human Gaussians in Milliseconds
by: Lu, Jia, et al.
Published: (2025)
by: Lu, Jia, et al.
Published: (2025)
Text-Animator: Controllable Visual Text Video Generation
by: Liu, Lin, et al.
Published: (2024)
by: Liu, Lin, et al.
Published: (2024)
Generating Human Motion in 3D Scenes from Text Descriptions
by: Cen, Zhi, et al.
Published: (2024)
by: Cen, Zhi, et al.
Published: (2024)
Incorporating Visual Experts to Resolve the Information Loss in Multimodal Large Language Models
by: He, Xin, et al.
Published: (2024)
by: He, Xin, et al.
Published: (2024)
Real-time Reflectance Generation for UAV Multispectral Imagery using an Onboard Downwelling Spectrometer in Varied Weather Conditions
by: Xie, Jiayang, et al.
Published: (2024)
by: Xie, Jiayang, et al.
Published: (2024)
BideDPO: Conditional Image Generation with Simultaneous Text and Condition Alignment
by: Zhou, Dewei, et al.
Published: (2025)
by: Zhou, Dewei, et al.
Published: (2025)
Parameter Efficient Fine-tuning via Cross Block Orchestration for Segment Anything Model
by: Peng, Zelin, et al.
Published: (2023)
by: Peng, Zelin, et al.
Published: (2023)
EMMA: Efficient Multimodal Understanding, Generation, and Editing with a Unified Architecture
by: He, Xin, et al.
Published: (2025)
by: He, Xin, et al.
Published: (2025)
Hypothesis Generation via LLM-Automated Language Bias for ILP
by: Yang, Yang, et al.
Published: (2025)
by: Yang, Yang, et al.
Published: (2025)
AI Decodes Historical Chinese Archives to Reveal Lost Climate History
by: He, Sida, et al.
Published: (2026)
by: He, Sida, et al.
Published: (2026)
Enhancing Electric‐Gas–Integrated Energy Systems: Optimal Coupling Strategies for Mitigating Voltage Sag Effects
by: Wei Zhao, et al.
Published: (2025)
by: Wei Zhao, et al.
Published: (2025)
Mixpert: Mitigating Multimodal Learning Conflicts with Efficient Mixture-of-Vision-Experts
by: He, Xin, et al.
Published: (2025)
by: He, Xin, et al.
Published: (2025)
RepEval: Effective Text Evaluation with LLM Representation
by: Sheng, Shuqian, et al.
Published: (2024)
by: Sheng, Shuqian, et al.
Published: (2024)
MotionDuet: Dual-Conditioned 3D Human Motion Generation with Video-Regularized Text Learning
by: Zhang, Yi-Yang, et al.
Published: (2025)
by: Zhang, Yi-Yang, et al.
Published: (2025)
On the Cost and Benefits of Training Context with Utterance or Full Conversation Training: A Comparative Stud
by: Liu, Hyouin, et al.
Published: (2025)
by: Liu, Hyouin, et al.
Published: (2025)
3DIS: Depth-Driven Decoupled Instance Synthesis for Text-to-Image Generation
by: Zhou, Dewei, et al.
Published: (2024)
by: Zhou, Dewei, et al.
Published: (2024)
EndoGSLAM: Real-Time Dense Reconstruction and Tracking in Endoscopic Surgeries using Gaussian Splatting
by: Wang, Kailing, et al.
Published: (2024)
by: Wang, Kailing, et al.
Published: (2024)
Lifespan of the Non-resistive Hall-MHD System with Small Magnetic Gradient
by: Yang, Linbin, et al.
Published: (2025)
by: Yang, Linbin, et al.
Published: (2025)
The topology and isochronicity on complex Hamiltonian systems with homogeneous nonlinearities
by: Dong, Guangfeng, et al.
Published: (2023)
by: Dong, Guangfeng, et al.
Published: (2023)
EMGauss: Continuous Slice-to-3D Reconstruction via Dynamic Gaussian Modeling in Volume Electron Microscopy
by: He, Yumeng, et al.
Published: (2025)
by: He, Yumeng, et al.
Published: (2025)
Parameter Identification and Frequency Adjustment for Achieving Constant Voltage Output in Series‐None Compensated Wireless Charging System Under Offset Conditions
by: Yiming Zhang, et al.
Published: (2026)
by: Yiming Zhang, et al.
Published: (2026)
Text2Weight: Bridging Natural Language and Neural Network Weight Spaces
by: Tian, Bowen, et al.
Published: (2025)
by: Tian, Bowen, et al.
Published: (2025)
LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding
by: Qiu, Jihao, et al.
Published: (2026)
by: Qiu, Jihao, et al.
Published: (2026)
Similar Items
-
WorldGrow: Generating Infinite 3D World
by: Li, Sikuang, et al.
Published: (2025) -
UniLat3D: Geometry-Appearance Unified Latents for Single-Stage 3D Generation
by: Wu, Guanjun, et al.
Published: (2025) -
Segment Anything in 3D with Radiance Fields
by: Cen, Jiazhong, et al.
Published: (2023) -
Few-step Flow for 3D Generation via Marginal-Data Transport Distillation
by: Zhou, Zanwei, et al.
Published: (2025) -
GaussianDreamerPro: Text to Manipulable 3D Gaussians with Highly Enhanced Quality
by: Yi, Taoran, et al.
Published: (2024)