Collaborative Learning for 3D Hand-Object Reconstruction and Compositional Action Recognition from Egocentric RGB Videos Using Superquadrics
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Tse, Tze Ho Elden, Feng, Runyang, Zheng, Linfang, Park, Jiho, Gao, Yixing, Kim, Jihie, Leonardis, Ales, Chang, Hyung Jin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
High-Resolution Spatiotemporal Modeling with Global-Local State Space Models for Video-Based Human Pose Estimation
von: Feng, Runyang, et al.
Veröffentlicht: (2025)
von: Feng, Runyang, et al.
Veröffentlicht: (2025)
GeoReF: Geometric Alignment Across Shape Variation for Category-level Object Pose Refinement
von: Zheng, Linfang, et al.
Veröffentlicht: (2024)
von: Zheng, Linfang, et al.
Veröffentlicht: (2024)
Leveraging RGB Images for Pre-Training of Event-Based Hand Pose Estimation
von: Liu, Ruicong, et al.
Veröffentlicht: (2025)
von: Liu, Ruicong, et al.
Veröffentlicht: (2025)
TIGeR: Text-Instructed Generation and Refinement for Template-Free Hand-Object Interaction
von: Huang, Yiyao, et al.
Veröffentlicht: (2025)
von: Huang, Yiyao, et al.
Veröffentlicht: (2025)
Visual Intention Grounding for Egocentric Assistants
von: Sun, Pengzhan, et al.
Veröffentlicht: (2025)
von: Sun, Pengzhan, et al.
Veröffentlicht: (2025)
DAS3R: Dynamics-Aware Gaussian Splatting for Static Scene Reconstruction
von: Xu, Kai, et al.
Veröffentlicht: (2024)
von: Xu, Kai, et al.
Veröffentlicht: (2024)
NCRF: Neural Contact Radiance Fields for Free-Viewpoint Rendering of Hand-Object Interaction
von: Zhang, Zhongqun, et al.
Veröffentlicht: (2024)
von: Zhang, Zhongqun, et al.
Veröffentlicht: (2024)
Improving Human Motion Plausibility with Body Momentum
von: Nguyen, Ha Linh, et al.
Veröffentlicht: (2025)
von: Nguyen, Ha Linh, et al.
Veröffentlicht: (2025)
Reconstructing Objects along Hand Interaction Timelines in Egocentric Video
von: Zhu, Zhifan, et al.
Veröffentlicht: (2025)
von: Zhu, Zhifan, et al.
Veröffentlicht: (2025)
SA-GS: Semantic-Aware Gaussian Splatting for Large Scene Reconstruction with Geometry Constrain
von: Xiong, Butian, et al.
Veröffentlicht: (2024)
von: Xiong, Butian, et al.
Veröffentlicht: (2024)
A Constrained Optimization Approach for Gaussian Splatting from Coarsely-posed Images and Noisy Lidar Point Clouds
von: Peng, Jizong, et al.
Veröffentlicht: (2025)
von: Peng, Jizong, et al.
Veröffentlicht: (2025)
Get a Grip: Reconstructing Hand-Object Stable Grasps in Egocentric Videos
von: Zhu, Zhifan, et al.
Veröffentlicht: (2023)
von: Zhu, Zhifan, et al.
Veröffentlicht: (2023)
Cross-Modal Action Recognition in Egocentric Video Using Mamba: Integrating RGB and Hand Skeleton Streams via CLS Token Fusion Strategies
von: Gorostegui, Juan Ignacio Bustos, et al.
Veröffentlicht: (2026)
von: Gorostegui, Juan Ignacio Bustos, et al.
Veröffentlicht: (2026)
Force-Aware 3D Contact Modeling for Stable Grasp Generation
von: Chen, Zhuo, et al.
Veröffentlicht: (2025)
von: Chen, Zhuo, et al.
Veröffentlicht: (2025)
Humans as Checkerboards: Calibrating Camera Motion Scale for World-Coordinate Human Mesh Recovery
von: Yang, Fengyuan, et al.
Veröffentlicht: (2024)
von: Yang, Fengyuan, et al.
Veröffentlicht: (2024)
StableSketcher: Enhancing Diffusion Model for Pixel-based Sketch Generation via Visual Question Answering Feedback
von: Park, Jiho, et al.
Veröffentlicht: (2025)
von: Park, Jiho, et al.
Veröffentlicht: (2025)
Articulation in Motion: Prior-free Part Mobility Analysis for Articulated Objects By Dynamic-Static Disentanglement
von: Ai, Hao, et al.
Veröffentlicht: (2026)
von: Ai, Hao, et al.
Veröffentlicht: (2026)
In My Perspective, In My Hands: Accurate Egocentric 2D Hand Pose and Action Recognition
von: Mucha, Wiktor, et al.
Veröffentlicht: (2024)
von: Mucha, Wiktor, et al.
Veröffentlicht: (2024)
HandNeRF: Learning to Reconstruct Hand-Object Interaction Scene from a Single RGB Image
von: Choi, Hongsuk, et al.
Veröffentlicht: (2023)
von: Choi, Hongsuk, et al.
Veröffentlicht: (2023)
Improving Object Detection via Local-global Contrastive Learning
von: Triantafyllidou, Danai, et al.
Veröffentlicht: (2024)
von: Triantafyllidou, Danai, et al.
Veröffentlicht: (2024)
Benchmarks and Challenges in Pose Estimation for Egocentric Hand Interactions with Objects
von: Fan, Zicong, et al.
Veröffentlicht: (2024)
von: Fan, Zicong, et al.
Veröffentlicht: (2024)
Object Aware Egocentric Online Action Detection
von: An, Joungbin, et al.
Veröffentlicht: (2024)
von: An, Joungbin, et al.
Veröffentlicht: (2024)
WHOLE: World-Grounded Hand-Object Lifted from Egocentric Videos
von: Ye, Yufei, et al.
Veröffentlicht: (2026)
von: Ye, Yufei, et al.
Veröffentlicht: (2026)
Do Egocentric Video-Language Models Truly Understand Hand-Object Interactions?
von: Xu, Boshen, et al.
Veröffentlicht: (2024)
von: Xu, Boshen, et al.
Veröffentlicht: (2024)
Modeling Fine-Grained Hand-Object Dynamics for Egocentric Video Representation Learning
von: Pei, Baoqi, et al.
Veröffentlicht: (2025)
von: Pei, Baoqi, et al.
Veröffentlicht: (2025)
GHOST: Fast Category-agnostic Hand-Object Interaction Reconstruction from RGB Videos using Gaussian Splatting
von: Aboukhadra, Ahmed Tawfik, et al.
Veröffentlicht: (2026)
von: Aboukhadra, Ahmed Tawfik, et al.
Veröffentlicht: (2026)
Dexterous Manipulation Policies from RGB Human Videos via 3D Hand-Object Trajectory Reconstruction
von: Chen, Hongyi, et al.
Veröffentlicht: (2026)
von: Chen, Hongyi, et al.
Veröffentlicht: (2026)
Label-Efficient Object Detection via Region Proposal Network Pre-Training
von: Dong, Nanqing, et al.
Veröffentlicht: (2022)
von: Dong, Nanqing, et al.
Veröffentlicht: (2022)
MEgoHand: Multimodal Egocentric Hand-Object Interaction Motion Generation
von: Zhou, Bohan, et al.
Veröffentlicht: (2025)
von: Zhou, Bohan, et al.
Veröffentlicht: (2025)
EgoGrasp: World-Space Hand-Object Interaction Estimation from Egocentric Videos
von: Fu, Hongming, et al.
Veröffentlicht: (2026)
von: Fu, Hongming, et al.
Veröffentlicht: (2026)
HaWoR: World-Space Hand Motion Reconstruction from Egocentric Videos
von: Zhang, Jinglei, et al.
Veröffentlicht: (2025)
von: Zhang, Jinglei, et al.
Veröffentlicht: (2025)
Masked Video and Body-worn IMU Autoencoder for Egocentric Action Recognition
von: Zhang, Mingfang, et al.
Veröffentlicht: (2024)
von: Zhang, Mingfang, et al.
Veröffentlicht: (2024)
Exploiting Spatial-Temporal Context for Interacting Hand Reconstruction on Monocular RGB Video
von: Zhao, Weichao, et al.
Veröffentlicht: (2023)
von: Zhao, Weichao, et al.
Veröffentlicht: (2023)
Heatmap Pooling Network for Action Recognition from RGB Videos
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025)
von: Liu, Mengyuan, et al.
Veröffentlicht: (2025)
EgoHandICL: Egocentric 3D Hand Reconstruction with In-Context Learning
von: Xie, Binzhu, et al.
Veröffentlicht: (2026)
von: Xie, Binzhu, et al.
Veröffentlicht: (2026)
EgoX: Egocentric Video Generation from a Single Exocentric Video
von: Kang, Taewoong, et al.
Veröffentlicht: (2025)
von: Kang, Taewoong, et al.
Veröffentlicht: (2025)
SHARP: Segmentation of Hands and Arms by Range using Pseudo-Depth for Enhanced Egocentric 3D Hand Pose Estimation and Action Recognition
von: Mucha, Wiktor, et al.
Veröffentlicht: (2024)
von: Mucha, Wiktor, et al.
Veröffentlicht: (2024)
Diff-IP2D: Diffusion-Based Hand-Object Interaction Prediction on Egocentric Videos
von: Ma, Junyi, et al.
Veröffentlicht: (2024)
von: Ma, Junyi, et al.
Veröffentlicht: (2024)
SEA: Evaluating Sketch Abstraction Efficiency via Element-level Commonsense Visual Question Answering
von: Park, Jiho, et al.
Veröffentlicht: (2026)
von: Park, Jiho, et al.
Veröffentlicht: (2026)
Superquadric Motion and Superquadric Hyperbolic Split Quaternion Algebra Via Gielis Formula
von: Zehra Özdemir, et al.
Veröffentlicht: (2025)
von: Zehra Özdemir, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
High-Resolution Spatiotemporal Modeling with Global-Local State Space Models for Video-Based Human Pose Estimation
von: Feng, Runyang, et al.
Veröffentlicht: (2025) -
GeoReF: Geometric Alignment Across Shape Variation for Category-level Object Pose Refinement
von: Zheng, Linfang, et al.
Veröffentlicht: (2024) -
Leveraging RGB Images for Pre-Training of Event-Based Hand Pose Estimation
von: Liu, Ruicong, et al.
Veröffentlicht: (2025) -
TIGeR: Text-Instructed Generation and Refinement for Template-Free Hand-Object Interaction
von: Huang, Yiyao, et al.
Veröffentlicht: (2025) -
Visual Intention Grounding for Egocentric Assistants
von: Sun, Pengzhan, et al.
Veröffentlicht: (2025)