Human Action Recognition without Human
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kataoka, Hirokatsu, Hara, Kensho, Satoh, Yutaka |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2016
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Formula-Supervised Visual-Geometric Pre-training
von: Yamada, Ryosuke, et al.
Veröffentlicht: (2024)
von: Yamada, Ryosuke, et al.
Veröffentlicht: (2024)
Can masking background and object reduce static bias for zero-shot action recognition?
von: Fukuzawa, Takumi, et al.
Veröffentlicht: (2025)
von: Fukuzawa, Takumi, et al.
Veröffentlicht: (2025)
Noise-Tolerant Learning for Audio-Visual Action Recognition
von: Han, Haochen, et al.
Veröffentlicht: (2022)
von: Han, Haochen, et al.
Veröffentlicht: (2022)
InstructHumans: Editing Animated 3D Human Textures with Instructions
von: Zhu, Jiayin, et al.
Veröffentlicht: (2024)
von: Zhu, Jiayin, et al.
Veröffentlicht: (2024)
Hierarchical Action Recognition: A Contrastive Video-Language Approach with Hierarchical Interactions
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
von: Zhang, Rui, et al.
Veröffentlicht: (2024)
Wavelet-Decoupling Contrastive Enhancement Network for Fine-Grained Skeleton-Based Action Recognition
von: Chang, Haochen, et al.
Veröffentlicht: (2024)
von: Chang, Haochen, et al.
Veröffentlicht: (2024)
SkeFi: Cross-Modal Knowledge Transfer for Wireless Skeleton-Based Action Recognition
von: Huang, Shunyu, et al.
Veröffentlicht: (2026)
von: Huang, Shunyu, et al.
Veröffentlicht: (2026)
Human Motion Video Generation: A Survey
von: Xue, Haiwei, et al.
Veröffentlicht: (2025)
von: Xue, Haiwei, et al.
Veröffentlicht: (2025)
Spatial-Temporal Human-Object Interaction Detection
von: Sun, Xu, et al.
Veröffentlicht: (2025)
von: Sun, Xu, et al.
Veröffentlicht: (2025)
HAIC: Improving Human Action Understanding and Generation with Better Captions for Multi-modal Large Language Models
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
von: Wang, Xiao, et al.
Veröffentlicht: (2025)
Generating Attribute-Aware Human Motions from Textual Prompt
von: Wang, Xinghan, et al.
Veröffentlicht: (2025)
von: Wang, Xinghan, et al.
Veröffentlicht: (2025)
Memory-Guided View Refinement for Dynamic Human-in-the-loop EQA
von: Lu, Xin, et al.
Veröffentlicht: (2026)
von: Lu, Xin, et al.
Veröffentlicht: (2026)
On the Robustness of Human-Object Interaction Detection against Distribution Shift
von: Xie, Chi, et al.
Veröffentlicht: (2025)
von: Xie, Chi, et al.
Veröffentlicht: (2025)
OneHOI: Unifying Human-Object Interaction Generation and Editing
von: Hoe, Jiun Tian, et al.
Veröffentlicht: (2026)
von: Hoe, Jiun Tian, et al.
Veröffentlicht: (2026)
PP-Motion: Physical-Perceptual Fidelity Evaluation for Human Motion Generation
von: Zhao, Sihan, et al.
Veröffentlicht: (2025)
von: Zhao, Sihan, et al.
Veröffentlicht: (2025)
MoRAG -- Multi-Fusion Retrieval Augmented Generation for Human Motion
von: Kalakonda, Sai Shashank, et al.
Veröffentlicht: (2024)
von: Kalakonda, Sai Shashank, et al.
Veröffentlicht: (2024)
Scalable Image Coding for Humans and Machines Using Feature Fusion Network
von: Shindo, Takahiro, et al.
Veröffentlicht: (2024)
von: Shindo, Takahiro, et al.
Veröffentlicht: (2024)
MSVBench: Towards Human-Level Evaluation of Multi-Shot Video Generation
von: Shi, Haoyuan, et al.
Veröffentlicht: (2026)
von: Shi, Haoyuan, et al.
Veröffentlicht: (2026)
ROI-Guided Point Cloud Geometry Compression Towards Human and Machine Vision
von: Liang, Xie, et al.
Veröffentlicht: (2025)
von: Liang, Xie, et al.
Veröffentlicht: (2025)
Text-controlled Motion Mamba: Text-Instructed Temporal Grounding of Human Motion
von: Wang, Xinghan, et al.
Veröffentlicht: (2024)
von: Wang, Xinghan, et al.
Veröffentlicht: (2024)
Interpretable Concept-based Deep Learning Framework for Multimodal Human Behavior Modeling
von: Li, Xinyu, et al.
Veröffentlicht: (2025)
von: Li, Xinyu, et al.
Veröffentlicht: (2025)
Exploring Mutual Cross-Modal Attention for Context-Aware Human Affordance Generation
von: Roy, Prasun, et al.
Veröffentlicht: (2025)
von: Roy, Prasun, et al.
Veröffentlicht: (2025)
Discriminative-Generative Synergy for Occlusion Robust 3D Human Mesh Recovery
von: Liu, Yang, et al.
Veröffentlicht: (2026)
von: Liu, Yang, et al.
Veröffentlicht: (2026)
Controllable Complex Human Motion Video Generation via Text-to-Skeleton Cascades
von: Taghipour, Ashkan, et al.
Veröffentlicht: (2026)
von: Taghipour, Ashkan, et al.
Veröffentlicht: (2026)
Unified Coding for Both Human Perception and Generalized Machine Analytics with CLIP Supervision
von: Yin, Kangsheng, et al.
Veröffentlicht: (2025)
von: Yin, Kangsheng, et al.
Veröffentlicht: (2025)
Text-guided Synthetic Geometric Augmentation for Zero-shot 3D Understanding
von: Torimi, Kohei, et al.
Veröffentlicht: (2025)
von: Torimi, Kohei, et al.
Veröffentlicht: (2025)
Radio Frequency Signal based Human Silhouette Segmentation: A Sequential Diffusion Approach
von: Wen, Penghui, et al.
Veröffentlicht: (2024)
von: Wen, Penghui, et al.
Veröffentlicht: (2024)
Multimodal Large Language Model is a Human-Aligned Annotator for Text-to-Image Generation
von: Wu, Xun, et al.
Veröffentlicht: (2024)
von: Wu, Xun, et al.
Veröffentlicht: (2024)
HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning
von: Chen, Liyang, et al.
Veröffentlicht: (2025)
von: Chen, Liyang, et al.
Veröffentlicht: (2025)
UniAnimate-DiT: Human Image Animation with Large-Scale Video Diffusion Transformer
von: Wang, Xiang, et al.
Veröffentlicht: (2025)
von: Wang, Xiang, et al.
Veröffentlicht: (2025)
SOSControl: Enhancing Human Motion Generation through Saliency-Aware Symbolic Orientation and Timing Control
von: Au, Ho Yin, et al.
Veröffentlicht: (2025)
von: Au, Ho Yin, et al.
Veröffentlicht: (2025)
A Human-Annotated Video Dataset for Training and Evaluation of 360-Degree Video Summarization Methods
von: Kontostathis, Ioannis, et al.
Veröffentlicht: (2024)
von: Kontostathis, Ioannis, et al.
Veröffentlicht: (2024)
EditHF-1M: A Million-Scale Rich Human Preference Feedback for Image Editing
von: Xu, Zitong, et al.
Veröffentlicht: (2026)
von: Xu, Zitong, et al.
Veröffentlicht: (2026)
HDiffTG: A Lightweight Hybrid Diffusion-Transformer-GCN Architecture for 3D Human Pose Estimation
von: Fu, Yajie, et al.
Veröffentlicht: (2025)
von: Fu, Yajie, et al.
Veröffentlicht: (2025)
LongInsightBench: A Comprehensive Benchmark for Evaluating Omni-Modal Models on Human-Centric Long-Video Understanding
von: Han, ZhaoYang, et al.
Veröffentlicht: (2025)
von: Han, ZhaoYang, et al.
Veröffentlicht: (2025)
Dynamic Resolution Guidance for Facial Expression Recognition
von: Wang, Songpan, et al.
Veröffentlicht: (2024)
von: Wang, Songpan, et al.
Veröffentlicht: (2024)
InteractEdit: Zero-Shot Editing of Human-Object Interactions in Images
von: Hoe, Jiun Tian, et al.
Veröffentlicht: (2025)
von: Hoe, Jiun Tian, et al.
Veröffentlicht: (2025)
Optimized Learned Image Compression for Facial Expression Recognition
von: Li, Xiumei, et al.
Veröffentlicht: (2025)
von: Li, Xiumei, et al.
Veröffentlicht: (2025)
Image Referenced Sketch Colorization Based on Animation Creation Workflow
von: Yan, Dingkun, et al.
Veröffentlicht: (2025)
von: Yan, Dingkun, et al.
Veröffentlicht: (2025)
Probabilistic Temporal Masked Attention for Cross-view Online Action Detection
von: Xie, Liping, et al.
Veröffentlicht: (2025)
von: Xie, Liping, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Formula-Supervised Visual-Geometric Pre-training
von: Yamada, Ryosuke, et al.
Veröffentlicht: (2024) -
Can masking background and object reduce static bias for zero-shot action recognition?
von: Fukuzawa, Takumi, et al.
Veröffentlicht: (2025) -
Noise-Tolerant Learning for Audio-Visual Action Recognition
von: Han, Haochen, et al.
Veröffentlicht: (2022) -
InstructHumans: Editing Animated 3D Human Textures with Instructions
von: Zhu, Jiayin, et al.
Veröffentlicht: (2024) -
Hierarchical Action Recognition: A Contrastive Video-Language Approach with Hierarchical Interactions
von: Zhang, Rui, et al.
Veröffentlicht: (2024)