Learning Coherent Matrixized Representation in Latent Space for Volumetric 4D Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Qitong, Feng, Mingtao, Wu, Zijie, Sun, Shijie, Dong, Weisheng, Wang, Yaonan, Mian, Ajmal |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
External Knowledge Enhanced 3D Scene Generation from Sketch
by: Wu, Zijie, et al.
Published: (2024)
by: Wu, Zijie, et al.
Published: (2024)
3D Object Detection from Point Cloud via Voting Step Diffusion
by: Hou, Haoran, et al.
Published: (2024)
by: Hou, Haoran, et al.
Published: (2024)
Multiview Point Cloud Registration Based on Minimum Potential Energy for Free-Form Blade Measurement
by: Wu, Zijie, et al.
Published: (2025)
by: Wu, Zijie, et al.
Published: (2025)
FLaTEC: Frequency-Disentangled Latent Triplanes for Efficient Compression of LiDAR Point Clouds
by: Zhang, Xiaoge, et al.
Published: (2025)
by: Zhang, Xiaoge, et al.
Published: (2025)
Diffusion-Driven Self-Supervised Learning for Shape Reconstruction and Pose Estimation
by: Sun, Jingtao, et al.
Published: (2024)
by: Sun, Jingtao, et al.
Published: (2024)
DiffCom: Decoupled Sparse Priors Guided Diffusion Compression for Point Clouds
by: Zhang, Xiaoge, et al.
Published: (2024)
by: Zhang, Xiaoge, et al.
Published: (2024)
Referring Human Pose and Mask Estimation in the Wild
by: Miao, Bo, et al.
Published: (2024)
by: Miao, Bo, et al.
Published: (2024)
Controlled Data Rebalancing in Multi-Task Learning for Real-World Image Super-Resolution
by: Lin, Shuchen, et al.
Published: (2025)
by: Lin, Shuchen, et al.
Published: (2025)
Class-Partitioned VQ-VAE and Latent Flow Matching for Point Cloud Scene Generation
by: Edirimuni, Dasith de Silva, et al.
Published: (2026)
by: Edirimuni, Dasith de Silva, et al.
Published: (2026)
SDesc3D: Towards Layout-Aware 3D Indoor Scene Generation from Short Descriptions
by: Feng, Jie, et al.
Published: (2026)
by: Feng, Jie, et al.
Published: (2026)
Implicit Neural Representation-Based Continuous Single Image Super-Resolution: An Empirical Benchmark
by: Nasir, Tayyab, et al.
Published: (2026)
by: Nasir, Tayyab, et al.
Published: (2026)
Simultaneous Multiple Object Detection and Pose Estimation using 3D Model Infusion with Monocular Vision
by: Li, Congliang, et al.
Published: (2022)
by: Li, Congliang, et al.
Published: (2022)
Soft Masked Transformer for Point Cloud Processing with Skip Attention-Based Upsampling
by: He, Yong, et al.
Published: (2024)
by: He, Yong, et al.
Published: (2024)
Deep Learning Based 3D Segmentation: A Survey
by: He, Yong, et al.
Published: (2021)
by: He, Yong, et al.
Published: (2021)
Safety Without Semantic Disruptions: Editing-free Safe Image Generation via Context-preserving Dual Latent Reconstruction
by: Vice, Jordan, et al.
Published: (2024)
by: Vice, Jordan, et al.
Published: (2024)
Back2Color: Domain-Adaptive Synthetic-to-Real Monocular Depth Estimation for Dynamic Traffic Scenes
by: Zhu, Yufan, et al.
Published: (2024)
by: Zhu, Yufan, et al.
Published: (2024)
SeqAffordSplat: Scene-level Sequential Affordance Reasoning on 3D Gaussian Splatting
by: Li, Di, et al.
Published: (2025)
by: Li, Di, et al.
Published: (2025)
Occlusion-aware Text-Image-Point Cloud Pretraining for Open-World 3D Object Recognition
by: Nguyen, Khanh, et al.
Published: (2025)
by: Nguyen, Khanh, et al.
Published: (2025)
PointDiffuse: A Dual-Conditional Diffusion Model for Enhanced Point Cloud Semantic Segmentation
by: He, Yong, et al.
Published: (2025)
by: He, Yong, et al.
Published: (2025)
ARMFlow: AutoRegressive MeanFlow for Online 3D Human Reaction Generation
by: Geng, Zichen, et al.
Published: (2025)
by: Geng, Zichen, et al.
Published: (2025)
Exploiting Semantic and Pixel Representations for Ultra-Low Bitrate Image Compression
by: Wei, Hao, et al.
Published: (2026)
by: Wei, Hao, et al.
Published: (2026)
Occlusion-Aware 3D Hand-Object Pose Estimation with Masked AutoEncoders
by: Yang, Hui, et al.
Published: (2025)
by: Yang, Hui, et al.
Published: (2025)
FastBO: Fast HPO and NAS with Adaptive Fidelity Identification
by: Jiang, Jiantong, et al.
Published: (2024)
by: Jiang, Jiantong, et al.
Published: (2024)
Auto-Regressive Diffusion for Generating 3D Human-Object Interactions
by: Geng, Zichen, et al.
Published: (2025)
by: Geng, Zichen, et al.
Published: (2025)
GVGEN: Text-to-3D Generation with Volumetric Representation
by: He, Xianglong, et al.
Published: (2024)
by: He, Xianglong, et al.
Published: (2024)
Domain-invariant Prototypes for Semantic Segmentation
by: Yang, Zhengeng, et al.
Published: (2022)
by: Yang, Zhengeng, et al.
Published: (2022)
Mantis: Mamba-native Tuning is Efficient for 3D Point Cloud Foundation Models
by: Guo, Zihao, et al.
Published: (2026)
by: Guo, Zihao, et al.
Published: (2026)
D3Seg: Dependency-Aware Diffusion for Brain Tumor Segmentation with Missing Modalities
by: Ali, Danish, et al.
Published: (2026)
by: Ali, Danish, et al.
Published: (2026)
Manipulating and Mitigating Generative Model Biases without Retraining
by: Vice, Jordan, et al.
Published: (2024)
by: Vice, Jordan, et al.
Published: (2024)
Exploring Bias in over 100 Text-to-Image Generative Models
by: Vice, Jordan, et al.
Published: (2025)
by: Vice, Jordan, et al.
Published: (2025)
CymbaDiff: Structured Spatial Diffusion for Sketch-based 3D Semantic Urban Scene Generation
by: Liang, Li, et al.
Published: (2025)
by: Liang, Li, et al.
Published: (2025)
MonoDiff9D: Monocular Category-Level 9D Object Pose Estimation via Diffusion Model
by: Liu, Jian, et al.
Published: (2025)
by: Liu, Jian, et al.
Published: (2025)
Modeling Human Skeleton Joint Dynamics for Fall Detection
by: Zahan, Sania, et al.
Published: (2025)
by: Zahan, Sania, et al.
Published: (2025)
SDFA: Structure Aware Discriminative Feature Aggregation for Efficient Human Fall Detection in Video
by: Zahan, Sania, et al.
Published: (2025)
by: Zahan, Sania, et al.
Published: (2025)
Real-time Photorealistic Dynamic Scene Representation and Rendering with 4D Gaussian Splatting
by: Yang, Zeyu, et al.
Published: (2023)
by: Yang, Zeyu, et al.
Published: (2023)
Diff9D: Diffusion-Based Domain-Generalized Category-Level 9-DoF Object Pose Estimation
by: Liu, Jian, et al.
Published: (2025)
by: Liu, Jian, et al.
Published: (2025)
Sparse Points to Dense Clouds: Enhancing 3D Detection with Limited LiDAR Data
by: Kumar, Aakash, et al.
Published: (2024)
by: Kumar, Aakash, et al.
Published: (2024)
Disentangled Hierarchical VAE for 3D Human-Human Interaction Generation
by: Geng, Zichen, et al.
Published: (2026)
by: Geng, Zichen, et al.
Published: (2026)
Align then Adapt: Rethinking Parameter-Efficient Transfer Learning in 4D Perception
by: Sun, Yiding, et al.
Published: (2026)
by: Sun, Yiding, et al.
Published: (2026)
UniHM: Universal Human Motion Generation with Object Interactions in Indoor Scenes
by: Geng, Zichen, et al.
Published: (2025)
by: Geng, Zichen, et al.
Published: (2025)
Similar Items
-
External Knowledge Enhanced 3D Scene Generation from Sketch
by: Wu, Zijie, et al.
Published: (2024) -
3D Object Detection from Point Cloud via Voting Step Diffusion
by: Hou, Haoran, et al.
Published: (2024) -
Multiview Point Cloud Registration Based on Minimum Potential Energy for Free-Form Blade Measurement
by: Wu, Zijie, et al.
Published: (2025) -
FLaTEC: Frequency-Disentangled Latent Triplanes for Efficient Compression of LiDAR Point Clouds
by: Zhang, Xiaoge, et al.
Published: (2025) -
Diffusion-Driven Self-Supervised Learning for Shape Reconstruction and Pose Estimation
by: Sun, Jingtao, et al.
Published: (2024)