Perceiving and Acting in First-Person: A Dataset and Benchmark for Egocentric Human-Object-Human Interactions
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Liang, Yang, Chengqun, Lin, Zili, Xu, Fei, Liu, Yifan, Xu, Congsheng, Zhang, Yiyi, Qin, Jie, Sheng, Xingdong, Liu, Yunhui, Jin, Xin, Yan, Yichao, Zeng, Wenjun, Yang, Xiaokang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
HIMO: A New Benchmark for Full-Body Human Interacting with Multiple Objects
by: Lv, Xintao, et al.
Published: (2024)
by: Lv, Xintao, et al.
Published: (2024)
$E^{3}$Gen: Efficient, Expressive and Editable Avatars Generation
by: Zhang, Weitian, et al.
Published: (2024)
by: Zhang, Weitian, et al.
Published: (2024)
IPAD: Industrial Process Anomaly Detection Dataset
by: Liu, Jinfan, et al.
Published: (2024)
by: Liu, Jinfan, et al.
Published: (2024)
MotionBank: A Large-scale Video Motion Benchmark with Disentangled Rule-based Annotations
by: Xu, Liang, et al.
Published: (2024)
by: Xu, Liang, et al.
Published: (2024)
Rethinking Clothes Changing Person ReID: Conflicts, Synthesis, and Optimization
by: Li, Junjie, et al.
Published: (2024)
by: Li, Junjie, et al.
Published: (2024)
ReGenNet: Towards Human Action-Reaction Synthesis
by: Xu, Liang, et al.
Published: (2024)
by: Xu, Liang, et al.
Published: (2024)
POLAR: A Portrait OLAT Dataset and Generative Framework for Illumination-Aware Face Modeling
by: Chen, Zhuo, et al.
Published: (2025)
by: Chen, Zhuo, et al.
Published: (2025)
LOME: Learning Human-Object Manipulation with Action-Conditioned Egocentric World Model
by: Gao, Quankai, et al.
Published: (2026)
by: Gao, Quankai, et al.
Published: (2026)
SingingBot: An Avatar-Driven System for Robotic Face Singing Performance
by: Xu, Zhuoxiong, et al.
Published: (2026)
by: Xu, Zhuoxiong, et al.
Published: (2026)
PairHuman: A High-Fidelity Photographic Dataset for Customized Dual-Person Generation
by: Pan, Ting, et al.
Published: (2025)
by: Pan, Ting, et al.
Published: (2025)
EMHI: A Multimodal Egocentric Human Motion Dataset with HMD and Body-Worn IMUs
by: Fan, Zhen, et al.
Published: (2024)
by: Fan, Zhen, et al.
Published: (2024)
Egocentric Human-Object Interaction Detection: A New Benchmark and Method
by: Deng, Kunyuan, et al.
Published: (2025)
by: Deng, Kunyuan, et al.
Published: (2025)
HOI4D: A 4D Egocentric Dataset for Category-Level Human-Object Interaction
by: Liu, Yunze, et al.
Published: (2022)
by: Liu, Yunze, et al.
Published: (2022)
MCOD: The First Challenging Benchmark for Multispectral Camouflaged Object Detection
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
Fine-Grained Human Pose Editing Assessment via Layer-Selective MLLMs
by: Sun, Ningyu, et al.
Published: (2026)
by: Sun, Ningyu, et al.
Published: (2026)
EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World
by: Punamiya, Ryan, et al.
Published: (2026)
by: Punamiya, Ryan, et al.
Published: (2026)
Interact-Custom: Customized Human Object Interaction Image Generation
by: Xu, Zhu, et al.
Published: (2025)
by: Xu, Zhu, et al.
Published: (2025)
Cellulose Acetate Nanofiber and Sodium Alginate‐Based Conductive Aerogel for Human Motion Monitoring
by: Mengyang Bao, et al.
Published: (2025)
by: Mengyang Bao, et al.
Published: (2025)
$π_0$-EqM: Equilibrium Matching for Closed-Loop Vision-Language-Action Control
by: Liu, Huanming, et al.
Published: (2026)
by: Liu, Huanming, et al.
Published: (2026)
DexHiL: A Human-in-the-Loop Framework for Vision-Language-Action Model Post-Training in Dexterous Manipulation
by: Han, Yifan, et al.
Published: (2026)
by: Han, Yifan, et al.
Published: (2026)
Human-centered In-building Embodied Delivery Benchmark
by: Xu, Zhuoqun, et al.
Published: (2024)
by: Xu, Zhuoqun, et al.
Published: (2024)
Hyperspectral Remote Sensing Images Salient Object Detection: The First Benchmark Dataset and Baseline
by: Liu, Peifu, et al.
Published: (2025)
by: Liu, Peifu, et al.
Published: (2025)
Do Egocentric Video-Language Models Truly Understand Hand-Object Interactions?
by: Xu, Boshen, et al.
Published: (2024)
by: Xu, Boshen, et al.
Published: (2024)
Human-Object Interaction from Human-Level Instructions
by: Wu, Zhen, et al.
Published: (2024)
by: Wu, Zhen, et al.
Published: (2024)
TAL: Two-stream Adaptive Learning for Generalizable Person Re-identification
by: Yan, Yichao, et al.
Published: (2021)
by: Yan, Yichao, et al.
Published: (2021)
Person Identification from Egocentric Human-Object Interactions using 3D Hand Pose
by: Hamza, Muhammad, et al.
Published: (2025)
by: Hamza, Muhammad, et al.
Published: (2025)
CORE4D: A 4D Human-Object-Human Interaction Dataset for Collaborative Object REarrangement
by: Liu, Yun, et al.
Published: (2024)
by: Liu, Yun, et al.
Published: (2024)
MMOT: The First Challenging Benchmark for Drone-based Multispectral Multi-Object Tracking
by: Li, Tianhao, et al.
Published: (2025)
by: Li, Tianhao, et al.
Published: (2025)
EgoChoir: Capturing 3D Human-Object Interaction Regions from Egocentric Views
by: Yang, Yuhang, et al.
Published: (2024)
by: Yang, Yuhang, et al.
Published: (2024)
FCMBench: The First Large-scale Financial Credit Multimodal Benchmark for Real-world Applications
by: Yang, Yehui, et al.
Published: (2026)
by: Yang, Yehui, et al.
Published: (2026)
POV: Prompt-Oriented View-Agnostic Learning for Egocentric Hand-Object Interaction in the Multi-View World
by: Xu, Boshen, et al.
Published: (2024)
by: Xu, Boshen, et al.
Published: (2024)
Towards Domain-Generalized Open-Vocabulary Object Detection: A Progressive Domain-invariant Cross-modal Alignment Method
by: Xu, Xiaoran, et al.
Published: (2026)
by: Xu, Xiaoran, et al.
Published: (2026)
MODA: The First Challenging Benchmark for Multispectral Object Detection in Aerial Images
by: Han, Shuaihao, et al.
Published: (2025)
by: Han, Shuaihao, et al.
Published: (2025)
Human Body Restoration with One-Step Diffusion Model and A New Benchmark
by: Gong, Jue, et al.
Published: (2025)
by: Gong, Jue, et al.
Published: (2025)
Do You Guys Want to Dance: Zero-Shot Compositional Human Dance Generation with Multiple Persons
by: Xu, Zhe, et al.
Published: (2024)
by: Xu, Zhe, et al.
Published: (2024)
Visible-Thermal Tiny Object Detection: A Benchmark Dataset and Baselines
by: Ying, Xinyi, et al.
Published: (2024)
by: Ying, Xinyi, et al.
Published: (2024)
Benchmarks and Challenges in Pose Estimation for Egocentric Hand Interactions with Objects
by: Fan, Zicong, et al.
Published: (2024)
by: Fan, Zicong, et al.
Published: (2024)
Transferable Human Mobility Network Reconstruction with neuroGravity
by: Yang, Jinming, et al.
Published: (2026)
by: Yang, Jinming, et al.
Published: (2026)
Human-controllable AI: Meaningful Human Control
by: Liu, Chengke, et al.
Published: (2025)
by: Liu, Chengke, et al.
Published: (2025)
Oriented Tiny Object Detection: A Dataset, Benchmark, and Dynamic Unbiased Learning
by: Xu, Chang, et al.
Published: (2024)
by: Xu, Chang, et al.
Published: (2024)
Similar Items
-
HIMO: A New Benchmark for Full-Body Human Interacting with Multiple Objects
by: Lv, Xintao, et al.
Published: (2024) -
$E^{3}$Gen: Efficient, Expressive and Editable Avatars Generation
by: Zhang, Weitian, et al.
Published: (2024) -
IPAD: Industrial Process Anomaly Detection Dataset
by: Liu, Jinfan, et al.
Published: (2024) -
MotionBank: A Large-scale Video Motion Benchmark with Disentangled Rule-based Annotations
by: Xu, Liang, et al.
Published: (2024) -
Rethinking Clothes Changing Person ReID: Conflicts, Synthesis, and Optimization
by: Li, Junjie, et al.
Published: (2024)