4DPC$^2$hat: Towards Dynamic Point Cloud Understanding with Failure-Aware Bootstrapping
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Xindan, Yan, Weilong, Shi, Yufei, Qiu, Xuerui, He, Tao, Li, Ying, Li, Ming, Fan, Hehe |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SciEducator: Scientific Video Understanding and Educating via Deming-Cycle Multi-Agent System
by: Xu, Zhiyu, et al.
Published: (2025)
by: Xu, Zhiyu, et al.
Published: (2025)
FaVChat: Hierarchical Prompt-Query Guided Facial Video Understanding with Data-Efficient GRPO
by: Zhao, Fufangchen, et al.
Published: (2025)
by: Zhao, Fufangchen, et al.
Published: (2025)
UniF$^2$ace: A Unified Fine-grained Face Understanding and Generation Model
by: Li, Junzhe, et al.
Published: (2025)
by: Li, Junzhe, et al.
Published: (2025)
Scaling Video Understanding via Compact Latent Multi-Agent Collaboration
by: Chen, Kerui, et al.
Published: (2026)
by: Chen, Kerui, et al.
Published: (2026)
Prompt-Aware Adapter: Towards Learning Adaptive Visual Tokens for Multimodal Large Language Models
by: Zhang, Yue, et al.
Published: (2024)
by: Zhang, Yue, et al.
Published: (2024)
One Sentence, One Drama: Personalized Short-Form Drama Generation via Multi-Agent Systems
by: Shi, Yufei, et al.
Published: (2026)
by: Shi, Yufei, et al.
Published: (2026)
ProtChatGPT: Towards Understanding Proteins with Large Language Models
by: Wang, Chao, et al.
Published: (2024)
by: Wang, Chao, et al.
Published: (2024)
MA-Bench: Towards Fine-grained Micro-Action Understanding
by: Li, Kun, et al.
Published: (2026)
by: Li, Kun, et al.
Published: (2026)
DPC: Dual-Prompt Collaboration for Tuning Vision-Language Models
by: Li, Haoyang, et al.
Published: (2025)
by: Li, Haoyang, et al.
Published: (2025)
PMA: Towards Parameter-Efficient Point Cloud Understanding via Point Mamba Adapter
by: Zha, Yaohua, et al.
Published: (2025)
by: Zha, Yaohua, et al.
Published: (2025)
Adversarial Unsupervised Domain Adaptation for 3D Semantic Segmentation with 2D Image Fusion of Dense Depth
by: Xindan Zhang, et al.
Published: (2024)
by: Xindan Zhang, et al.
Published: (2024)
Representation Learning for Point Cloud Understanding
by: Yan, Siming
Published: (2025)
by: Yan, Siming
Published: (2025)
Text-Scene: A Scene-to-Language Parsing Framework for 3D Scene Understanding
by: Li, Haoyuan, et al.
Published: (2025)
by: Li, Haoyuan, et al.
Published: (2025)
IT-DPC-SRI: A Cloud-Optimized Archive of Italian Radar Precipitation (2010-2025)
by: Franch, Gabriele, et al.
Published: (2026)
by: Franch, Gabriele, et al.
Published: (2026)
TV-Dialogue: Crafting Theme-Aware Video Dialogues with Immersive Interaction
by: Wang, Sai, et al.
Published: (2025)
by: Wang, Sai, et al.
Published: (2025)
Let's Reward Step-by-Step: Step-Aware Contrastive Alignment for Vision-Language Navigation in Continuous Environments
by: Li, Haoyuan, et al.
Published: (2026)
by: Li, Haoyuan, et al.
Published: (2026)
Understanding Dynamic Scenes in Ego Centric 4D Point Clouds
by: Huang, Junsheng, et al.
Published: (2025)
by: Huang, Junsheng, et al.
Published: (2025)
Why Do DiT Editors Drift? Plug-and-Play Low Frequency Alignment in VAE Latent Space
by: Wang, Xiaoce, et al.
Published: (2026)
by: Wang, Xiaoce, et al.
Published: (2026)
Exploiting GPT-4 Vision for Zero-shot Point Cloud Understanding
by: Sun, Qi, et al.
Published: (2024)
by: Sun, Qi, et al.
Published: (2024)
Point-In-Context: Understanding Point Cloud via In-Context Learning
by: Liu, Mengyuan, et al.
Published: (2024)
by: Liu, Mengyuan, et al.
Published: (2024)
IGASA: Integrated Geometry-Aware and Skip-Attention Modules for Enhanced Point Cloud Registration
by: Zhang, Dongxu, et al.
Published: (2026)
by: Zhang, Dongxu, et al.
Published: (2026)
ROI-Guided Point Cloud Geometry Compression Towards Human and Machine Vision
by: Liang, Xie, et al.
Published: (2025)
by: Liang, Xie, et al.
Published: (2025)
DeMoGen: Towards Decompositional Human Motion Generation with Energy-Based Diffusion Models
by: Zhang, Jianrong, et al.
Published: (2025)
by: Zhang, Jianrong, et al.
Published: (2025)
TechCoach: Towards Technical-Point-Aware Descriptive Action Coaching
by: Li, Yuan-Ming, et al.
Published: (2024)
by: Li, Yuan-Ming, et al.
Published: (2024)
TOPA: Extending Large Language Models for Video Understanding via Text-Only Pre-Alignment
by: Li, Wei, et al.
Published: (2024)
by: Li, Wei, et al.
Published: (2024)
Point Cloud Mamba: Point Cloud Learning via State Space Model
by: Zhang, Tao, et al.
Published: (2024)
by: Zhang, Tao, et al.
Published: (2024)
Bootstrap Dynamic-Aware 3D Visual Representation for Scalable Robot Learning
by: Liang, Qiwei, et al.
Published: (2025)
by: Liang, Qiwei, et al.
Published: (2025)
TransNormal: Dense Visual Semantics for Diffusion-based Transparent Object Normal Estimation
by: Li, Mingwei, et al.
Published: (2026)
by: Li, Mingwei, et al.
Published: (2026)
Imperceptible Adversarial Attacks on Point Clouds Guided by Point-to-Surface Field
by: Tang, Keke, et al.
Published: (2024)
by: Tang, Keke, et al.
Published: (2024)
Study on the Anti‐Vasculogenic Mimicry Effect of Duchesnea indica (Andr.) Focke Acidic Polysaccharide in Colon Cancer
by: Qiuying Shi, et al.
Published: (2025)
by: Qiuying Shi, et al.
Published: (2025)
Ground Awareness in Deep Learning for Large Outdoor Point Cloud Segmentation
by: Qiu, Kevin, et al.
Published: (2025)
by: Qiu, Kevin, et al.
Published: (2025)
Prompt-Aware Controllable Shadow Removal
by: Chen, Kerui, et al.
Published: (2025)
by: Chen, Kerui, et al.
Published: (2025)
Utonia: Toward One Encoder for All Point Clouds
by: Zhang, Yujia, et al.
Published: (2026)
by: Zhang, Yujia, et al.
Published: (2026)
Towards Foundation Models for 3D Scene Understanding: Instance-Aware Self-Supervised Learning for Point Clouds
by: Yang, Bin, et al.
Published: (2026)
by: Yang, Bin, et al.
Published: (2026)
PointCaM: Cut-and-Mix for Open-Set Point Cloud Learning
by: Hong, Jie, et al.
Published: (2022)
by: Hong, Jie, et al.
Published: (2022)
Mamba Learns in Context: Structure-Aware Domain Generalization for Multi-Task Point Cloud Understanding
by: Jiang, Jincen, et al.
Published: (2026)
by: Jiang, Jincen, et al.
Published: (2026)
Deformation-based In-Context Learning for Point Cloud Understanding
by: Lin, Chengxing, et al.
Published: (2026)
by: Lin, Chengxing, et al.
Published: (2026)
ML-SemReg: Boosting Point Cloud Registration with Multi-level Semantic Consistency
by: Yan, Shaocheng, et al.
Published: (2024)
by: Yan, Shaocheng, et al.
Published: (2024)
Transferable and Undefendable Point Cloud Attacks via Medial Axis Transform
by: Tang, Keke, et al.
Published: (2025)
by: Tang, Keke, et al.
Published: (2025)
KAN or MLP? Point Cloud Shows the Way Forward
by: Shi, Yan, et al.
Published: (2025)
by: Shi, Yan, et al.
Published: (2025)
Similar Items
-
SciEducator: Scientific Video Understanding and Educating via Deming-Cycle Multi-Agent System
by: Xu, Zhiyu, et al.
Published: (2025) -
FaVChat: Hierarchical Prompt-Query Guided Facial Video Understanding with Data-Efficient GRPO
by: Zhao, Fufangchen, et al.
Published: (2025) -
UniF$^2$ace: A Unified Fine-grained Face Understanding and Generation Model
by: Li, Junzhe, et al.
Published: (2025) -
Scaling Video Understanding via Compact Latent Multi-Agent Collaboration
by: Chen, Kerui, et al.
Published: (2026) -
Prompt-Aware Adapter: Towards Learning Adaptive Visual Tokens for Multimodal Large Language Models
by: Zhang, Yue, et al.
Published: (2024)