RoHOI: Robustness Benchmark for Human-Object Interaction Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wen, Di, Peng, Kunyu, Yang, Kailun, Chen, Yufan, Liu, Ruiping, Zheng, Junwei, Roitberg, Alina, Paudel, Danda Pani, Van Gool, Luc, Stiefelhagen, Rainer |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RefAtomNet++: Advancing Referring Atomic Video Action Recognition using Semantic Retrieval based Multi-Trajectory Mamba
von: Peng, Kunyu, et al.
Veröffentlicht: (2025)
von: Peng, Kunyu, et al.
Veröffentlicht: (2025)
Elevating Skeleton-Based Action Recognition with Efficient Multi-Modality Self-Supervision
von: Wei, Yiping, et al.
Veröffentlicht: (2023)
von: Wei, Yiping, et al.
Veröffentlicht: (2023)
Skeleton-Based Human Action Recognition with Noisy Labels
von: Xu, Yi, et al.
Veröffentlicht: (2024)
von: Xu, Yi, et al.
Veröffentlicht: (2024)
Exploring Self-supervised Skeleton-based Action Recognition in Occluded Environments
von: Chen, Yifei, et al.
Veröffentlicht: (2023)
von: Chen, Yifei, et al.
Veröffentlicht: (2023)
Referring Atomic Video Action Recognition
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
TransKD: Transformer Knowledge Distillation for Efficient Semantic Segmentation
von: Liu, Ruiping, et al.
Veröffentlicht: (2022)
von: Liu, Ruiping, et al.
Veröffentlicht: (2022)
InterEdit: Navigating Text-Guided Multi-Human 3D Motion Editing
von: Yang, Yebin, et al.
Veröffentlicht: (2026)
von: Yang, Yebin, et al.
Veröffentlicht: (2026)
ProOOD: Prototype-Guided Out-of-Distribution 3D Occupancy Prediction
von: Zhang, Yuheng, et al.
Veröffentlicht: (2026)
von: Zhang, Yuheng, et al.
Veröffentlicht: (2026)
Mitigating Label Noise using Prompt-Based Hyperbolic Meta-Learning in Open-Set Domain Generalization
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
RoDLA: Benchmarking the Robustness of Document Layout Analysis Models
von: Chen, Yufan, et al.
Veröffentlicht: (2024)
von: Chen, Yufan, et al.
Veröffentlicht: (2024)
Fourier Prompt Tuning for Modality-Incomplete Scene Segmentation
von: Liu, Ruiping, et al.
Veröffentlicht: (2024)
von: Liu, Ruiping, et al.
Veröffentlicht: (2024)
EReLiFM: Evidential Reliability-Aware Residual Flow Meta-Learning for Open-Set Domain Generalization under Noisy Labels
von: Peng, Kunyu, et al.
Veröffentlicht: (2025)
von: Peng, Kunyu, et al.
Veröffentlicht: (2025)
Exploring Few-Shot Adaptation for Activity Recognition on Diverse Domains
von: Peng, Kunyu, et al.
Veröffentlicht: (2023)
von: Peng, Kunyu, et al.
Veröffentlicht: (2023)
Lego: Learning to Disentangle and Invert Personalized Concepts Beyond Object Appearance in Text-to-Image Diffusion Models
von: Motamed, Saman, et al.
Veröffentlicht: (2023)
von: Motamed, Saman, et al.
Veröffentlicht: (2023)
Exploring Video-Based Driver Activity Recognition under Noisy Labels
von: Fan, Linjuan, et al.
Veröffentlicht: (2025)
von: Fan, Linjuan, et al.
Veröffentlicht: (2025)
Towards Activated Muscle Group Estimation in the Wild
von: Peng, Kunyu, et al.
Veröffentlicht: (2023)
von: Peng, Kunyu, et al.
Veröffentlicht: (2023)
$M^2$-Occ: Resilient 3D Semantic Occupancy Prediction for Autonomous Driving with Incomplete Camera Inputs
von: Lin, Kaixin, et al.
Veröffentlicht: (2026)
von: Lin, Kaixin, et al.
Veröffentlicht: (2026)
RICO: Two Realistic Benchmarks and an In-Depth Analysis for Incremental Learning in Object Detection
von: Neuwirth-Trapp, Matthias, et al.
Veröffentlicht: (2025)
von: Neuwirth-Trapp, Matthias, et al.
Veröffentlicht: (2025)
Seeing Together: Multi-Robot Cooperative Egocentric Spatial Reasoning with Multimodal Large Language Models
von: Peng, Kunyu, et al.
Veröffentlicht: (2026)
von: Peng, Kunyu, et al.
Veröffentlicht: (2026)
EvenNICER-SLAM: Event-based Neural Implicit Encoding SLAM
von: Chen, Shi, et al.
Veröffentlicht: (2024)
von: Chen, Shi, et al.
Veröffentlicht: (2024)
HopaDIFF: Holistic-Partial Aware Fourier Conditioned Diffusion for Referring Human Action Segmentation in Multi-Person Scenarios
von: Peng, Kunyu, et al.
Veröffentlicht: (2025)
von: Peng, Kunyu, et al.
Veröffentlicht: (2025)
Incremental Object Detection with Prompt-based Methods
von: Neuwirth-Trapp, Matthias, et al.
Veröffentlicht: (2025)
von: Neuwirth-Trapp, Matthias, et al.
Veröffentlicht: (2025)
OAFuser: Towards Omni-Aperture Fusion for Light Field Semantic Segmentation
von: Teng, Fei, et al.
Veröffentlicht: (2023)
von: Teng, Fei, et al.
Veröffentlicht: (2023)
EchoTrack: Auditory Referring Multi-Object Tracking for Autonomous Driving
von: Lin, Jiacheng, et al.
Veröffentlicht: (2024)
von: Lin, Jiacheng, et al.
Veröffentlicht: (2024)
Segment-to-Act: Label-Noise-Robust Action-Prompted Video Segmentation Towards Embodied Intelligence
von: Li, Wenxin, et al.
Veröffentlicht: (2025)
von: Li, Wenxin, et al.
Veröffentlicht: (2025)
Open Panoramic Segmentation
von: Zheng, Junwei, et al.
Veröffentlicht: (2024)
von: Zheng, Junwei, et al.
Veröffentlicht: (2024)
Implicit-Zoo: A Large-Scale Dataset of Neural Implicit Functions for 2D Images and 3D Scenes
von: Ma, Qi, et al.
Veröffentlicht: (2024)
von: Ma, Qi, et al.
Veröffentlicht: (2024)
Continuous Pose for Monocular Cameras in Neural Implicit Representation
von: Ma, Qi, et al.
Veröffentlicht: (2023)
von: Ma, Qi, et al.
Veröffentlicht: (2023)
From Synchrony to Sequence: Exo-to-Ego Generation via Interpolation
von: Mahdi, Mohammad, et al.
Veröffentlicht: (2026)
von: Mahdi, Mohammad, et al.
Veröffentlicht: (2026)
Vision encoders should be image size agnostic and task driven
von: Prisadnikov, Nedyalko, et al.
Veröffentlicht: (2025)
von: Prisadnikov, Nedyalko, et al.
Veröffentlicht: (2025)
Self-supervised pretraining for an iterative image size agnostic vision transformer
von: Prisadnikov, Nedyalko, et al.
Veröffentlicht: (2026)
von: Prisadnikov, Nedyalko, et al.
Veröffentlicht: (2026)
Taming CLIP for Fine-grained and Structured Visual Understanding of Museum Exhibits
von: Balauca, Ada-Astrid, et al.
Veröffentlicht: (2024)
von: Balauca, Ada-Astrid, et al.
Veröffentlicht: (2024)
A Simple and Generalist Approach for Panoptic Segmentation
von: Prisadnikov, Nedyalko, et al.
Veröffentlicht: (2024)
von: Prisadnikov, Nedyalko, et al.
Veröffentlicht: (2024)
Graph-based Document Structure Analysis
von: Chen, Yufan, et al.
Veröffentlicht: (2025)
von: Chen, Yufan, et al.
Veröffentlicht: (2025)
Advancing Open-Set Domain Generalization Using Evidential Bi-Level Hardest Domain Scheduler
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)
Occlusion-Aware Seamless Segmentation
von: Cao, Yihong, et al.
Veröffentlicht: (2024)
von: Cao, Yihong, et al.
Veröffentlicht: (2024)
RHO: Robust Holistic OSM-Based Metric Cross-View Geo-Localization
von: Zheng, Junwei, et al.
Veröffentlicht: (2026)
von: Zheng, Junwei, et al.
Veröffentlicht: (2026)
Towards Multi-Source Domain Generalization for Sleep Staging with Noisy Labels
von: Wang, Kening, et al.
Veröffentlicht: (2026)
von: Wang, Kening, et al.
Veröffentlicht: (2026)
AdaptiveClick: Clicks-aware Transformer with Adaptive Focal Loss for Interactive Image Segmentation
von: Lin, Jiacheng, et al.
Veröffentlicht: (2023)
von: Lin, Jiacheng, et al.
Veröffentlicht: (2023)
IMPACT-HOI: Supervisory Control for Onset-Anchored Partial HOI Event Construction
von: Zhang, Haoshen, et al.
Veröffentlicht: (2026)
von: Zhang, Haoshen, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
RefAtomNet++: Advancing Referring Atomic Video Action Recognition using Semantic Retrieval based Multi-Trajectory Mamba
von: Peng, Kunyu, et al.
Veröffentlicht: (2025) -
Elevating Skeleton-Based Action Recognition with Efficient Multi-Modality Self-Supervision
von: Wei, Yiping, et al.
Veröffentlicht: (2023) -
Skeleton-Based Human Action Recognition with Noisy Labels
von: Xu, Yi, et al.
Veröffentlicht: (2024) -
Exploring Self-supervised Skeleton-based Action Recognition in Occluded Environments
von: Chen, Yifei, et al.
Veröffentlicht: (2023) -
Referring Atomic Video Action Recognition
von: Peng, Kunyu, et al.
Veröffentlicht: (2024)