HVG-3D: Bridging Real and Simulation Domains for 3D-Conditional Hand-Object Interaction Video Synthesis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chen, Mingjin, Chen, Junhao, Fan, Zhaoxin, Lee, Yujian, Dang, Zichen, Wang, Lili, Cui, Yawen, Chau, Lap-Pui, Wang, Yi |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
3DGeoDet: General-purpose Geometry-aware Image-based 3D Object Detection
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
Interaction-aware Representation Modeling with Co-occurrence Consistency for Egocentric Hand-Object Parsing
von: Su, Yuejiao, et al.
Veröffentlicht: (2026)
von: Su, Yuejiao, et al.
Veröffentlicht: (2026)
LaSSM: Efficient Semantic-Spatial Query Decoding via Local Aggregation and State Space Models for 3D Instance Segmentation
von: Yao, Lei, et al.
Veröffentlicht: (2026)
von: Yao, Lei, et al.
Veröffentlicht: (2026)
GVSynergy-Det: Synergistic Gaussian-Voxel Representations for Multi-View 3D Object Detection
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
von: Zhang, Yi, et al.
Veröffentlicht: (2025)
CaRe-Ego: Contact-aware Relationship Modeling for Egocentric Interactive Hand-object Segmentation
von: Su, Yuejiao, et al.
Veröffentlicht: (2024)
von: Su, Yuejiao, et al.
Veröffentlicht: (2024)
SGIFormer: Semantic-guided and Geometric-enhanced Interleaving Transformer for 3D Instance Segmentation
von: Yao, Lei, et al.
Veröffentlicht: (2024)
von: Yao, Lei, et al.
Veröffentlicht: (2024)
Egocentric Human-Object Interaction Detection: A New Benchmark and Method
von: Deng, Kunyuan, et al.
Veröffentlicht: (2025)
von: Deng, Kunyuan, et al.
Veröffentlicht: (2025)
EVA02-AT: Egocentric Video-Language Understanding with Spatial-Temporal Rotary Positional Embeddings and Symmetric Optimization
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2025)
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2025)
OccProphet: Pushing Efficiency Frontier of Camera-Only 4D Occupancy Forecasting with Observer-Forecaster-Refiner Framework
von: Chen, Junliang, et al.
Veröffentlicht: (2025)
von: Chen, Junliang, et al.
Veröffentlicht: (2025)
HAMMER: Harnessing MLLM via Cross-Modal Integration for Intention-Driven 3D Affordance Grounding
von: Yao, Lei, et al.
Veröffentlicht: (2026)
von: Yao, Lei, et al.
Veröffentlicht: (2026)
GaussianCross: Cross-modal Self-supervised 3D Representation Learning via Gaussian Splatting
von: Yao, Lei, et al.
Veröffentlicht: (2025)
von: Yao, Lei, et al.
Veröffentlicht: (2025)
Ultraman: Single Image 3D Human Reconstruction with Ultra Speed and Detail
von: Chen, Mingjin, et al.
Veröffentlicht: (2024)
von: Chen, Mingjin, et al.
Veröffentlicht: (2024)
Fuzzy-aware Loss for Source-free Domain Adaptation in Visual Emotion Recognition
von: Zheng, Ying, et al.
Veröffentlicht: (2025)
von: Zheng, Ying, et al.
Veröffentlicht: (2025)
ProCal: Probability Calibration for Neighborhood-Guided Source-Free Domain Adaptation
von: Zheng, Ying, et al.
Veröffentlicht: (2026)
von: Zheng, Ying, et al.
Veröffentlicht: (2026)
Symmetric Multi-Similarity Loss for EPIC-KITCHENS-100 Multi-Instance Retrieval Challenge 2024
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoqi, et al.
Veröffentlicht: (2024)
Weakly-supervised Part-Attention and Mentored Networks for Vehicle Re-Identification
von: Tang, Lisha, et al.
Veröffentlicht: (2021)
von: Tang, Lisha, et al.
Veröffentlicht: (2021)
SP-MoMamba: Superpixel-driven Mixture of State Space Experts for Efficient Image Super-Resolution
von: Zou, Wenbin, et al.
Veröffentlicht: (2026)
von: Zou, Wenbin, et al.
Veröffentlicht: (2026)
PADetBench: Towards Benchmarking Physical Attacks against Object Detection
von: Lian, Jiawei, et al.
Veröffentlicht: (2024)
von: Lian, Jiawei, et al.
Veröffentlicht: (2024)
Evolution-Inspired Sample Competition for Deep Neural Network Optimization
von: Zheng, Ying, et al.
Veröffentlicht: (2026)
von: Zheng, Ying, et al.
Veröffentlicht: (2026)
GUI-C$^2$: Coarse-to-Fine GUI Grounding via Difficulty-Aware Reinforcement Learning
von: Li, Junlong, et al.
Veröffentlicht: (2026)
von: Li, Junlong, et al.
Veröffentlicht: (2026)
Towards Blind Bitstream-corrupted Video Recovery via a Visual Foundation Model-driven Framework
von: Liu, Tianyi, et al.
Veröffentlicht: (2025)
von: Liu, Tianyi, et al.
Veröffentlicht: (2025)
EDVD-LLaMA: Explainable Deepfake Video Detection via Multimodal Large Language Model Reasoning
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
A Survey on Occupancy Perception for Autonomous Driving: The Information Fusion Perspective
von: Xu, Huaiyuan, et al.
Veröffentlicht: (2024)
von: Xu, Huaiyuan, et al.
Veröffentlicht: (2024)
HandBooster: Boosting 3D Hand-Mesh Reconstruction by Conditional Synthesis and Sampling of Hand-Object Interactions
von: Xu, Hao, et al.
Veröffentlicht: (2024)
von: Xu, Hao, et al.
Veröffentlicht: (2024)
MASS: Mesh-inellipse Aligned Deformable Surfel Splatting for Hand Reconstruction and Rendering from Egocentric Monocular Video
von: Zhu, Haoyu, et al.
Veröffentlicht: (2026)
von: Zhu, Haoyu, et al.
Veröffentlicht: (2026)
3D Reconstruction of Objects in Hands without Real World 3D Supervision
von: Prakash, Aditya, et al.
Veröffentlicht: (2023)
von: Prakash, Aditya, et al.
Veröffentlicht: (2023)
PEM: Perception Error Model for Virtual Testing of Autonomous Vehicles
von: Piazzoni, Andrea, et al.
Veröffentlicht: (2023)
von: Piazzoni, Andrea, et al.
Veröffentlicht: (2023)
HOT3D: Hand and Object Tracking in 3D from Egocentric Multi-View Videos
von: Banerjee, Prithviraj, et al.
Veröffentlicht: (2024)
von: Banerjee, Prithviraj, et al.
Veröffentlicht: (2024)
ANNEXE: Unified Analyzing, Answering, and Pixel Grounding for Egocentric Interaction
von: Su, Yuejiao, et al.
Veröffentlicht: (2025)
von: Su, Yuejiao, et al.
Veröffentlicht: (2025)
RoboFlow4D: A Lightweight Flow World Model Toward Real-Time Flow-Guided Robotic Manipulation
von: Lin, Sixu, et al.
Veröffentlicht: (2026)
von: Lin, Sixu, et al.
Veröffentlicht: (2026)
A Survey of Embodied Learning for Object-Centric Robotic Manipulation
von: Zheng, Ying, et al.
Veröffentlicht: (2024)
von: Zheng, Ying, et al.
Veröffentlicht: (2024)
Reconstructing Hand-Held Objects in 3D from Images and Videos
von: Wu, Jane, et al.
Veröffentlicht: (2024)
von: Wu, Jane, et al.
Veröffentlicht: (2024)
MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model
von: Tong, Jinguang, et al.
Veröffentlicht: (2026)
von: Tong, Jinguang, et al.
Veröffentlicht: (2026)
Idea23D: Collaborative LMM Agents Enable 3D Model Generation from Interleaved Multimodal Inputs
von: Chen, Junhao, et al.
Veröffentlicht: (2024)
von: Chen, Junhao, et al.
Veröffentlicht: (2024)
SIGHT: Synthesizing Image-Text Conditioned and Geometry-Guided 3D Hand-Object Trajectories
von: Gavryushin, Alexey, et al.
Veröffentlicht: (2025)
von: Gavryushin, Alexey, et al.
Veröffentlicht: (2025)
ForeHOI: Feed-forward 3D Object Reconstruction from Daily Hand-Object Interaction Videos
von: Chen, Yuantao, et al.
Veröffentlicht: (2026)
von: Chen, Yuantao, et al.
Veröffentlicht: (2026)
Accelerating Diffusion-based Video Editing via Heterogeneous Caching: Beyond Full Computing at Sampled Denoising Timestep
von: Liu, Tianyi, et al.
Veröffentlicht: (2026)
von: Liu, Tianyi, et al.
Veröffentlicht: (2026)
MLPHand: Real Time Multi-View 3D Hand Mesh Reconstruction via MLP Modeling
von: Yang, Jian, et al.
Veröffentlicht: (2024)
von: Yang, Jian, et al.
Veröffentlicht: (2024)
Semantic Representation Attack against Aligned Large Language Models
von: Lian, Jiawei, et al.
Veröffentlicht: (2025)
von: Lian, Jiawei, et al.
Veröffentlicht: (2025)
Revealing the Intrinsic Ethical Vulnerability of Aligned Large Language Models
von: Lian, Jiawei, et al.
Veröffentlicht: (2025)
von: Lian, Jiawei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
3DGeoDet: General-purpose Geometry-aware Image-based 3D Object Detection
von: Zhang, Yi, et al.
Veröffentlicht: (2025) -
Interaction-aware Representation Modeling with Co-occurrence Consistency for Egocentric Hand-Object Parsing
von: Su, Yuejiao, et al.
Veröffentlicht: (2026) -
LaSSM: Efficient Semantic-Spatial Query Decoding via Local Aggregation and State Space Models for 3D Instance Segmentation
von: Yao, Lei, et al.
Veröffentlicht: (2026) -
GVSynergy-Det: Synergistic Gaussian-Voxel Representations for Multi-View 3D Object Detection
von: Zhang, Yi, et al.
Veröffentlicht: (2025) -
CaRe-Ego: Contact-aware Relationship Modeling for Egocentric Interactive Hand-object Segmentation
von: Su, Yuejiao, et al.
Veröffentlicht: (2024)