Guardado en:
| Autores principales: | Zhang, Jinzhi, Xiong, Feng, Xu, Mu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2412.02202 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
G3PT: Unleash the power of Autoregressive Modeling in 3D Generation via Cross-scale Querying Transformer
por: Zhang, Jinzhi, et al.
Publicado: (2024)
por: Zhang, Jinzhi, et al.
Publicado: (2024)
MVPainter: Accurate and Detailed 3D Texture Generation via Multi-View Diffusion with Geometric Control
por: Shao, Mingqi, et al.
Publicado: (2025)
por: Shao, Mingqi, et al.
Publicado: (2025)
Flow caching for autoregressive video generation
por: Ma, Yuexiao, et al.
Publicado: (2026)
por: Ma, Yuexiao, et al.
Publicado: (2026)
HumanRig: Learning Automatic Rigging for Humanoid Character in a Large Scale Dataset
por: Chu, Zedong, et al.
Publicado: (2024)
por: Chu, Zedong, et al.
Publicado: (2024)
Predicting 3D representations for Dynamic Scenes
por: Qi, Di, et al.
Publicado: (2025)
por: Qi, Di, et al.
Publicado: (2025)
I2V3D: Controllable image-to-video generation with 3D guidance
por: Zhang, Zhiyuan, et al.
Publicado: (2025)
por: Zhang, Zhiyuan, et al.
Publicado: (2025)
Dora: Sampling and Benchmarking for 3D Shape Variational Auto-Encoders
por: Chen, Rui, et al.
Publicado: (2024)
por: Chen, Rui, et al.
Publicado: (2024)
Event-boosted Deformable 3D Gaussians for Dynamic Scene Reconstruction
por: Xu, Wenhao, et al.
Publicado: (2024)
por: Xu, Wenhao, et al.
Publicado: (2024)
Not all tokens contribute equally to diffusion learning
por: Zhang, Guoqing, et al.
Publicado: (2026)
por: Zhang, Guoqing, et al.
Publicado: (2026)
VarGes: Improving Variation in Co-Speech 3D Gesture Generation via StyleCLIPS
por: Meng, Ming, et al.
Publicado: (2025)
por: Meng, Ming, et al.
Publicado: (2025)
ODGS: 3D Scene Reconstruction from Omnidirectional Images with 3D Gaussian Splattings
por: Lee, Suyoung, et al.
Publicado: (2024)
por: Lee, Suyoung, et al.
Publicado: (2024)
NOVA-3D: Non-overlapped Views for 3D Anime Character Reconstruction
por: Wang, Hongsheng, et al.
Publicado: (2024)
por: Wang, Hongsheng, et al.
Publicado: (2024)
RCGDet3D: Rethinking 4D Radar-Camera Fusion-based 3D Object Detection with Enhanced Radar Feature Encoding
por: Xiong, Weiyi, et al.
Publicado: (2026)
por: Xiong, Weiyi, et al.
Publicado: (2026)
ARM3D: Attention-based relation module for indoor 3D object detection
por: Lan, Yuqing, et al.
Publicado: (2022)
por: Lan, Yuqing, et al.
Publicado: (2022)
Byte-level generative predictions for forensics multimedia carving
por: Lee, Jaewon, et al.
Publicado: (2026)
por: Lee, Jaewon, et al.
Publicado: (2026)
InstructLayout: Instruction-Driven 2D and 3D Layout Synthesis with Semantic Graph Prior
por: Lin, Chenguo, et al.
Publicado: (2024)
por: Lin, Chenguo, et al.
Publicado: (2024)
SCA3D: Enhancing Cross-modal 3D Retrieval via 3D Shape and Caption Paired Data Augmentation
por: Ren, Junlong, et al.
Publicado: (2025)
por: Ren, Junlong, et al.
Publicado: (2025)
OmniPhysGS: 3D Constitutive Gaussians for General Physics-Based Dynamics Generation
por: Lin, Yuchen, et al.
Publicado: (2025)
por: Lin, Yuchen, et al.
Publicado: (2025)
CEI-3D: Collaborative Explicit-Implicit 3D Reconstruction for Realistic and Fine-Grained Object Editing
por: Shi, Yue, et al.
Publicado: (2026)
por: Shi, Yue, et al.
Publicado: (2026)
Group Critical-token Policy Optimization for Autoregressive Image Generation
por: Zhang, Guohui, et al.
Publicado: (2025)
por: Zhang, Guohui, et al.
Publicado: (2025)
VEDAL: Variational Error-Driven Asynchronous Learning for 3D Gaussian Splatting Pruning
por: Li, Aoduo, et al.
Publicado: (2026)
por: Li, Aoduo, et al.
Publicado: (2026)
CRAG: Can 3D Generative Models Help 3D Assembly?
por: Jiang, Zeyu, et al.
Publicado: (2026)
por: Jiang, Zeyu, et al.
Publicado: (2026)
RadarGaussianDet3D: Gaussian Representation-based Real-time 3D Object Detection with 4D Automotive Radars
por: Xiong, Weiyi, et al.
Publicado: (2025)
por: Xiong, Weiyi, et al.
Publicado: (2025)
StereoDETR: Stereo-based Transformer for 3D Object Detection
por: Mu, Shiyi, et al.
Publicado: (2025)
por: Mu, Shiyi, et al.
Publicado: (2025)
SR3D: Unleashing Single-view 3D Reconstruction for Transparent and Specular Object Grasping
por: Zhang, Mingxu, et al.
Publicado: (2025)
por: Zhang, Mingxu, et al.
Publicado: (2025)
Open-Vocabulary High-Resolution 3D (OVHR3D) Data Segmentation and Annotation Framework
por: Xu, Jiuyi, et al.
Publicado: (2024)
por: Xu, Jiuyi, et al.
Publicado: (2024)
CO^3: Cooperative Unsupervised 3D Representation Learning for Autonomous Driving
por: Chen, Runjian, et al.
Publicado: (2022)
por: Chen, Runjian, et al.
Publicado: (2022)
COM3D: Leveraging Cross-View Correspondence and Cross-Modal Mining for 3D Retrieval
por: Wu, Hao, et al.
Publicado: (2024)
por: Wu, Hao, et al.
Publicado: (2024)
When Worse is Better: Navigating the compression-generation tradeoff in visual tokenization
por: Ramanujan, Vivek, et al.
Publicado: (2024)
por: Ramanujan, Vivek, et al.
Publicado: (2024)
Rein3D: Reinforced 3D Indoor Scene Generation with Panoramic Video Diffusion Models
por: Wang, Dehui, et al.
Publicado: (2026)
por: Wang, Dehui, et al.
Publicado: (2026)
InstructScene: Instruction-Driven 3D Indoor Scene Synthesis with Semantic Graph Prior
por: Lin, Chenguo, et al.
Publicado: (2024)
por: Lin, Chenguo, et al.
Publicado: (2024)
Gaussian Variation Field Diffusion for High-fidelity Video-to-4D Synthesis
por: Zhang, Bowen, et al.
Publicado: (2025)
por: Zhang, Bowen, et al.
Publicado: (2025)
FantasyWorld: Geometry-Consistent World Modeling via Unified Video and 3D Prediction
por: Dai, Yixiang, et al.
Publicado: (2025)
por: Dai, Yixiang, et al.
Publicado: (2025)
Visual enhancement and 3D representation for underwater scenes: a review
por: Huang, Guoxi, et al.
Publicado: (2025)
por: Huang, Guoxi, et al.
Publicado: (2025)
Hyper3D: Efficient 3D Representation via Hybrid Triplane and Octree Feature for Enhanced 3D Shape Variational Auto-Encoders
por: Guo, Jingyu, et al.
Publicado: (2025)
por: Guo, Jingyu, et al.
Publicado: (2025)
B2N3D: Progressive Learning from Binary to N-ary Relationships for 3D Object Grounding
por: Xiao, Feng, et al.
Publicado: (2025)
por: Xiao, Feng, et al.
Publicado: (2025)
Resolving compositional and conformational heterogeneity in cryo-EM with deformable 3D Gaussian representations
por: He, Bintao, et al.
Publicado: (2025)
por: He, Bintao, et al.
Publicado: (2025)
Uncertainty-Aware AB3DMOT by Variational 3D Object Detection
por: Oleksiienko, Illia, et al.
Publicado: (2023)
por: Oleksiienko, Illia, et al.
Publicado: (2023)
PhysAlign: Physics-Coherent Image-to-Video Generation through Feature and 3D Representation Alignment
por: Xiong, Zhexiao, et al.
Publicado: (2026)
por: Xiong, Zhexiao, et al.
Publicado: (2026)
SDesc3D: Towards Layout-Aware 3D Indoor Scene Generation from Short Descriptions
por: Feng, Jie, et al.
Publicado: (2026)
por: Feng, Jie, et al.
Publicado: (2026)
Ejemplares similares
-
G3PT: Unleash the power of Autoregressive Modeling in 3D Generation via Cross-scale Querying Transformer
por: Zhang, Jinzhi, et al.
Publicado: (2024) -
MVPainter: Accurate and Detailed 3D Texture Generation via Multi-View Diffusion with Geometric Control
por: Shao, Mingqi, et al.
Publicado: (2025) -
Flow caching for autoregressive video generation
por: Ma, Yuexiao, et al.
Publicado: (2026) -
HumanRig: Learning Automatic Rigging for Humanoid Character in a Large Scale Dataset
por: Chu, Zedong, et al.
Publicado: (2024) -
Predicting 3D representations for Dynamic Scenes
por: Qi, Di, et al.
Publicado: (2025)