Referring Atomic Video Action Recognition
Fuente:
arXiv
Guardado en:
| Autores principales: | Peng, Kunyu, Fu, Jia, Yang, Kailun, Wen, Di, Chen, Yufan, Liu, Ruiping, Zheng, Junwei, Zhang, Jiaming, Sarfraz, M. Saquib, Stiefelhagen, Rainer, Roitberg, Alina |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Skeleton-Based Human Action Recognition with Noisy Labels
por: Xu, Yi, et al.
Publicado: (2024)
por: Xu, Yi, et al.
Publicado: (2024)
Elevating Skeleton-Based Action Recognition with Efficient Multi-Modality Self-Supervision
por: Wei, Yiping, et al.
Publicado: (2023)
por: Wei, Yiping, et al.
Publicado: (2023)
Exploring Few-Shot Adaptation for Activity Recognition on Diverse Domains
por: Peng, Kunyu, et al.
Publicado: (2023)
por: Peng, Kunyu, et al.
Publicado: (2023)
Exploring Self-supervised Skeleton-based Action Recognition in Occluded Environments
por: Chen, Yifei, et al.
Publicado: (2023)
por: Chen, Yifei, et al.
Publicado: (2023)
RefAtomNet++: Advancing Referring Atomic Video Action Recognition using Semantic Retrieval based Multi-Trajectory Mamba
por: Peng, Kunyu, et al.
Publicado: (2025)
por: Peng, Kunyu, et al.
Publicado: (2025)
Mitigating Label Noise using Prompt-Based Hyperbolic Meta-Learning in Open-Set Domain Generalization
por: Peng, Kunyu, et al.
Publicado: (2024)
por: Peng, Kunyu, et al.
Publicado: (2024)
Fourier Prompt Tuning for Modality-Incomplete Scene Segmentation
por: Liu, Ruiping, et al.
Publicado: (2024)
por: Liu, Ruiping, et al.
Publicado: (2024)
Towards Activated Muscle Group Estimation in the Wild
por: Peng, Kunyu, et al.
Publicado: (2023)
por: Peng, Kunyu, et al.
Publicado: (2023)
TransKD: Transformer Knowledge Distillation for Efficient Semantic Segmentation
por: Liu, Ruiping, et al.
Publicado: (2022)
por: Liu, Ruiping, et al.
Publicado: (2022)
Exploring Video-Based Driver Activity Recognition under Noisy Labels
por: Fan, Linjuan, et al.
Publicado: (2025)
por: Fan, Linjuan, et al.
Publicado: (2025)
RoHOI: Robustness Benchmark for Human-Object Interaction Detection
por: Wen, Di, et al.
Publicado: (2025)
por: Wen, Di, et al.
Publicado: (2025)
HopaDIFF: Holistic-Partial Aware Fourier Conditioned Diffusion for Referring Human Action Segmentation in Multi-Person Scenarios
por: Peng, Kunyu, et al.
Publicado: (2025)
por: Peng, Kunyu, et al.
Publicado: (2025)
$M^2$-Occ: Resilient 3D Semantic Occupancy Prediction for Autonomous Driving with Incomplete Camera Inputs
por: Lin, Kaixin, et al.
Publicado: (2026)
por: Lin, Kaixin, et al.
Publicado: (2026)
OAFuser: Towards Omni-Aperture Fusion for Light Field Semantic Segmentation
por: Teng, Fei, et al.
Publicado: (2023)
por: Teng, Fei, et al.
Publicado: (2023)
Segment-to-Act: Label-Noise-Robust Action-Prompted Video Segmentation Towards Embodied Intelligence
por: Li, Wenxin, et al.
Publicado: (2025)
por: Li, Wenxin, et al.
Publicado: (2025)
EchoTrack: Auditory Referring Multi-Object Tracking for Autonomous Driving
por: Lin, Jiacheng, et al.
Publicado: (2024)
por: Lin, Jiacheng, et al.
Publicado: (2024)
Occlusion-Aware Seamless Segmentation
por: Cao, Yihong, et al.
Publicado: (2024)
por: Cao, Yihong, et al.
Publicado: (2024)
Advancing Open-Set Domain Generalization Using Evidential Bi-Level Hardest Domain Scheduler
por: Peng, Kunyu, et al.
Publicado: (2024)
por: Peng, Kunyu, et al.
Publicado: (2024)
InterEdit: Navigating Text-Guided Multi-Human 3D Motion Editing
por: Yang, Yebin, et al.
Publicado: (2026)
por: Yang, Yebin, et al.
Publicado: (2026)
Behind Every Domain There is a Shift: Adapting Distortion-aware Vision Transformers for Panoramic Semantic Segmentation
por: Zhang, Jiaming, et al.
Publicado: (2022)
por: Zhang, Jiaming, et al.
Publicado: (2022)
Can we Trust Unreliable Voxels? Exploring 3D Semantic Occupancy Prediction under Label Noise
por: Li, Wenxin, et al.
Publicado: (2026)
por: Li, Wenxin, et al.
Publicado: (2026)
Seeing Beyond: Extrapolative Domain Adaptive Panoramic Segmentation
por: Zheng, Yuanfan, et al.
Publicado: (2026)
por: Zheng, Yuanfan, et al.
Publicado: (2026)
Unlocking Constraints: Source-Free Occlusion-Aware Seamless Segmentation
por: Cao, Yihong, et al.
Publicado: (2025)
por: Cao, Yihong, et al.
Publicado: (2025)
LF Tracy: A Unified Single-Pipeline Approach for Salient Object Detection in Light Field Cameras
por: Teng, Fei, et al.
Publicado: (2024)
por: Teng, Fei, et al.
Publicado: (2024)
MateRobot: Material Recognition in Wearable Robotics for People with Visual Impairments
por: Zheng, Junwei, et al.
Publicado: (2023)
por: Zheng, Junwei, et al.
Publicado: (2023)
Out-of-Distribution Semantic Occupancy Prediction
por: Zhang, Yuheng, et al.
Publicado: (2025)
por: Zhang, Yuheng, et al.
Publicado: (2025)
CoBEV: Elevating Roadside 3D Object Detection with Depth and Height Complementarity
por: Shi, Hao, et al.
Publicado: (2023)
por: Shi, Hao, et al.
Publicado: (2023)
AdaptiveClick: Clicks-aware Transformer with Adaptive Focal Loss for Interactive Image Segmentation
por: Lin, Jiacheng, et al.
Publicado: (2023)
por: Lin, Jiacheng, et al.
Publicado: (2023)
Go Beyond Earth: Understanding Human Actions and Scenes in Microgravity Environments
por: Wen, Di, et al.
Publicado: (2025)
por: Wen, Di, et al.
Publicado: (2025)
Graph-based Document Structure Analysis
por: Chen, Yufan, et al.
Publicado: (2025)
por: Chen, Yufan, et al.
Publicado: (2025)
LFX: Towards Unified Light Field Dense Semantic Segmentation and Salient Object Detection
por: Teng, Fei, et al.
Publicado: (2025)
por: Teng, Fei, et al.
Publicado: (2025)
DriveXQA: Cross-modal Visual Question Answering for Adverse Driving Scene Understanding
por: Tao, Mingzhe, et al.
Publicado: (2026)
por: Tao, Mingzhe, et al.
Publicado: (2026)
EgoExoMem: Cross-View Memory Reasoning over Synchronized Egocentric and Exocentric Videos
por: Liu, Ruiping, et al.
Publicado: (2026)
por: Liu, Ruiping, et al.
Publicado: (2026)
Hallucinating 360°: Panoramic Street-View Generation via Local Scenes Diffusion and Probabilistic Prompting
por: Teng, Fei, et al.
Publicado: (2025)
por: Teng, Fei, et al.
Publicado: (2025)
Open Panoramic Segmentation
por: Zheng, Junwei, et al.
Publicado: (2024)
por: Zheng, Junwei, et al.
Publicado: (2024)
EReLiFM: Evidential Reliability-Aware Residual Flow Meta-Learning for Open-Set Domain Generalization under Noisy Labels
por: Peng, Kunyu, et al.
Publicado: (2025)
por: Peng, Kunyu, et al.
Publicado: (2025)
P2U-SLAM: A Monocular Wide-FoV SLAM System Based on Point Uncertainty and Pose Uncertainty
por: Zhang, Yufan, et al.
Publicado: (2024)
por: Zhang, Yufan, et al.
Publicado: (2024)
Expression Prompt Collaboration Transformer for Universal Referring Video Object Segmentation
por: Chen, Jiajun, et al.
Publicado: (2023)
por: Chen, Jiajun, et al.
Publicado: (2023)
Data Diet: Can Trimming PET/CT Datasets Enhance Lesion Segmentation?
por: Jaus, Alexander, et al.
Publicado: (2024)
por: Jaus, Alexander, et al.
Publicado: (2024)
OmniTrack++: Omnidirectional Multi-Object Tracking by Learning Large-FoV Trajectory Feedback
por: Luo, Kai, et al.
Publicado: (2025)
por: Luo, Kai, et al.
Publicado: (2025)
Ejemplares similares
-
Skeleton-Based Human Action Recognition with Noisy Labels
por: Xu, Yi, et al.
Publicado: (2024) -
Elevating Skeleton-Based Action Recognition with Efficient Multi-Modality Self-Supervision
por: Wei, Yiping, et al.
Publicado: (2023) -
Exploring Few-Shot Adaptation for Activity Recognition on Diverse Domains
por: Peng, Kunyu, et al.
Publicado: (2023) -
Exploring Self-supervised Skeleton-based Action Recognition in Occluded Environments
por: Chen, Yifei, et al.
Publicado: (2023) -
RefAtomNet++: Advancing Referring Atomic Video Action Recognition using Semantic Retrieval based Multi-Trajectory Mamba
por: Peng, Kunyu, et al.
Publicado: (2025)