A Unified Framework for Human-centric Point Cloud Video Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Yiteng, Ye, Kecheng, Han, Xiao, Ren, Yiming, Zhu, Xinge, Ma, Yuexin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Practical Human Motion Prediction with LiDAR Point Clouds
by: Han, Xiao, et al.
Published: (2024)
by: Han, Xiao, et al.
Published: (2024)
ReMoGen: Real-time Human Interaction-to-Reaction Generation via Modular Learning from Diverse Data
by: Ye, Yaoqin, et al.
Published: (2026)
by: Ye, Yaoqin, et al.
Published: (2026)
Sparkle: A Robust and Versatile Representation for Point Cloud based Human Motion Capture
by: Ren, Yiming, et al.
Published: (2026)
by: Ren, Yiming, et al.
Published: (2026)
Registration between Point Cloud Streams and Sequential Bounding Boxes via Gradient Descent
by: Li, Xuesong, et al.
Published: (2024)
by: Li, Xuesong, et al.
Published: (2024)
FreeCap: Hybrid Calibration-Free Motion Capture in Open Environments
by: Xue, Aoru, et al.
Published: (2024)
by: Xue, Aoru, et al.
Published: (2024)
HUNTER: Unsupervised Human-centric 3D Detection via Transferring Knowledge from Synthetic Instances to Real Scenes
by: Yao, Yichen, et al.
Published: (2024)
by: Yao, Yichen, et al.
Published: (2024)
Learning to Adapt SAM for Segmenting Cross-domain Point Clouds
by: Peng, Xidong, et al.
Published: (2023)
by: Peng, Xidong, et al.
Published: (2023)
LaserHuman: Language-guided Scene-aware Human Motion Generation in Free Environment
by: Cong, Peishan, et al.
Published: (2024)
by: Cong, Peishan, et al.
Published: (2024)
OctreeOcc: Efficient and Multi-Granularity Occupancy Prediction Using Octree Queries
by: Lu, Yuhang, et al.
Published: (2023)
by: Lu, Yuhang, et al.
Published: (2023)
SocialMirror: Reconstructing 3D Human Interaction Behaviors from Monocular Videos with Semantic and Geometric Guidance
by: Xia, Qi, et al.
Published: (2026)
by: Xia, Qi, et al.
Published: (2026)
LiveHPS: LiDAR-based Scene-level Human Pose and Shape Estimation in Free Environment
by: Ren, Yiming, et al.
Published: (2024)
by: Ren, Yiming, et al.
Published: (2024)
HUMOF: Human Motion Forecasting in Interactive Social Scenes
by: Sun, Caiyi, et al.
Published: (2025)
by: Sun, Caiyi, et al.
Published: (2025)
EvolvingGrasp: Evolutionary Grasp Generation via Efficient Preference Alignment
by: Zhu, Yufei, et al.
Published: (2025)
by: Zhu, Yufei, et al.
Published: (2025)
UniBioTransfer: A Unified Framework for Multiple Biometrics Transfer
by: Sun, Caiyi, et al.
Published: (2026)
by: Sun, Caiyi, et al.
Published: (2026)
MoE3D: Mixture of Experts meets Multi-Modal 3D Understanding
by: Li, Yu, et al.
Published: (2025)
by: Li, Yu, et al.
Published: (2025)
Dynamic 3D Point Cloud Sequences as 2D Videos
by: Zeng, Yiming, et al.
Published: (2024)
by: Zeng, Yiming, et al.
Published: (2024)
LiveHPS++: Robust and Coherent Motion Capture in Dynamic Free Environment
by: Ren, Yiming, et al.
Published: (2024)
by: Ren, Yiming, et al.
Published: (2024)
Gait Recognition in Large-scale Free Environment via Single LiDAR
by: Han, Xiao, et al.
Published: (2022)
by: Han, Xiao, et al.
Published: (2022)
Can LVLMs Obtain a Driver's License? A Benchmark Towards Reliable AGI for Autonomous Driving
by: Lu, Yuhang, et al.
Published: (2024)
by: Lu, Yuhang, et al.
Published: (2024)
GPT4Point: A Unified Framework for Point-Language Understanding and Generation
by: Qi, Zhangyang, et al.
Published: (2023)
by: Qi, Zhangyang, et al.
Published: (2023)
Uni4D: A Unified Self-Supervised Learning Framework for Point Cloud Videos
by: Zuo, Zhi, et al.
Published: (2025)
by: Zuo, Zhi, et al.
Published: (2025)
HumanSAM: Classifying Human-centric Forgery Videos in Human Spatial, Appearance, and Motion Anomaly
by: Liu, Chang, et al.
Published: (2025)
by: Liu, Chang, et al.
Published: (2025)
Point Cloud Quantization through Multimodal Prompting for 3D Understanding
by: Li, Hongxuan, et al.
Published: (2025)
by: Li, Hongxuan, et al.
Published: (2025)
PoinTramba: A Hybrid Transformer-Mamba Framework for Point Cloud Analysis
by: Wang, Zicheng, et al.
Published: (2024)
by: Wang, Zicheng, et al.
Published: (2024)
ZigzagPointMamba: Spatial-Semantic Mamba for Point Cloud Understanding
by: Diao, Linshuang, et al.
Published: (2025)
by: Diao, Linshuang, et al.
Published: (2025)
FreqPolicy: Frequency Autoregressive Visuomotor Policy with Continuous Tokens
by: Zhong, Yiming, et al.
Published: (2025)
by: Zhong, Yiming, et al.
Published: (2025)
PointMamba: A Simple State Space Model for Point Cloud Analysis
by: Liang, Dingkang, et al.
Published: (2024)
by: Liang, Dingkang, et al.
Published: (2024)
Driving with A Thousand Faces: A Benchmark for Closed-Loop Personalized End-to-End Autonomous Driving
by: Dong, Xiaoru, et al.
Published: (2026)
by: Dong, Xiaoru, et al.
Published: (2026)
Geometrically-driven Aggregation for Zero-shot 3D Point Cloud Understanding
by: Mei, Guofeng, et al.
Published: (2023)
by: Mei, Guofeng, et al.
Published: (2023)
On Exploring PDE Modeling for Point Cloud Video Representation Learning
by: Huang, Zhuoxu, et al.
Published: (2024)
by: Huang, Zhuoxu, et al.
Published: (2024)
Point-In-Context: Understanding Point Cloud via In-Context Learning
by: Liu, Mengyuan, et al.
Published: (2024)
by: Liu, Mengyuan, et al.
Published: (2024)
CLIP-based Point Cloud Classification via Point Cloud to Image Translation
by: Ghose, Shuvozit, et al.
Published: (2024)
by: Ghose, Shuvozit, et al.
Published: (2024)
EgoExo-Gen: Ego-centric Video Prediction by Watching Exo-centric Videos
by: Xu, Jilan, et al.
Published: (2025)
by: Xu, Jilan, et al.
Published: (2025)
VideoEval-Pro: Robust and Realistic Long Video Understanding Evaluation
by: Ma, Wentao, et al.
Published: (2025)
by: Ma, Wentao, et al.
Published: (2025)
KPM-Bench: A Kinematic Parsing Motion Benchmark for Fine-grained Motion-centric Video Understanding
by: Lin, Boda, et al.
Published: (2026)
by: Lin, Boda, et al.
Published: (2026)
Situational Scene Graph for Structured Human-centric Situation Understanding
by: Sugandhika, Chinthani, et al.
Published: (2024)
by: Sugandhika, Chinthani, et al.
Published: (2024)
ORV: 4D Occupancy-centric Robot Video Generation
by: Yang, Xiuyu, et al.
Published: (2025)
by: Yang, Xiuyu, et al.
Published: (2025)
STAGE: A Stream-Centric Generative World Model for Long-Horizon Driving-Scene Simulation
by: Wang, Jiamin, et al.
Published: (2025)
by: Wang, Jiamin, et al.
Published: (2025)
Hyper-Bagel: A Unified Acceleration Framework for Multimodal Understanding and Generation
by: Lu, Yanzuo, et al.
Published: (2025)
by: Lu, Yanzuo, et al.
Published: (2025)
UniVideo: Unified Understanding, Generation, and Editing for Videos
by: Wei, Cong, et al.
Published: (2025)
by: Wei, Cong, et al.
Published: (2025)
Similar Items
-
Towards Practical Human Motion Prediction with LiDAR Point Clouds
by: Han, Xiao, et al.
Published: (2024) -
ReMoGen: Real-time Human Interaction-to-Reaction Generation via Modular Learning from Diverse Data
by: Ye, Yaoqin, et al.
Published: (2026) -
Sparkle: A Robust and Versatile Representation for Point Cloud based Human Motion Capture
by: Ren, Yiming, et al.
Published: (2026) -
Registration between Point Cloud Streams and Sequential Bounding Boxes via Gradient Descent
by: Li, Xuesong, et al.
Published: (2024) -
FreeCap: Hybrid Calibration-Free Motion Capture in Open Environments
by: Xue, Aoru, et al.
Published: (2024)