End-to-End Shared Attention Estimation via Group Detection with Feedback Refinement
Fuente:
arXiv
Saved in:
| Main Authors: | Nakatani, Chihiro, Ukita, Norimichi, Odobez, Jean-Marc |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Dynamic Group Detection using VLM-augmented Temporal Groupness Graph
by: Yokoyama, Kaname, et al.
Published: (2025)
by: Yokoyama, Kaname, et al.
Published: (2025)
Group-DINOmics: Incorporating People Dynamics into DINO for Self-supervised Group Activity Feature Learning
by: Tezuka, Ryuki, et al.
Published: (2026)
by: Tezuka, Ryuki, et al.
Published: (2026)
Learning Group Activity Features Through Person Attribute Prediction
by: Nakatani, Chihiro, et al.
Published: (2024)
by: Nakatani, Chihiro, et al.
Published: (2024)
Human-in-the-loop Adaptation in Group Activity Feature Learning for Team Sports Video Retrieval
by: Nakatani, Chihiro, et al.
Published: (2026)
by: Nakatani, Chihiro, et al.
Published: (2026)
Size-Variable Virtual Try-On with Physical Clothes Size
by: Yamashita, Yohei, et al.
Published: (2024)
by: Yamashita, Yohei, et al.
Published: (2024)
Human Motion Prediction via Test-domain-aware Adaptation with Easily-available Human Motions Estimated from Videos
by: Shimbo, Katsuki, et al.
Published: (2025)
by: Shimbo, Katsuki, et al.
Published: (2025)
SAMIDARE: Advanced Tracking-by-Segmentation for Dense Scenarios
by: Hirano, Shozaburo, et al.
Published: (2026)
by: Hirano, Shozaburo, et al.
Published: (2026)
Efficient Cost-and-Quality Controllable Arbitrary-scale Super-resolution with Fourier Constraints
by: Akita, Kazutoshi, et al.
Published: (2025)
by: Akita, Kazutoshi, et al.
Published: (2025)
Time-series Initialization and Conditioning for Video-agnostic Stabilization of Video Super-Resolution using Recurrent Networks
by: Mori, Hiroshi, et al.
Published: (2024)
by: Mori, Hiroshi, et al.
Published: (2024)
Inpainting-Driven Mask Optimization for Object Removal
by: Shimosato, Kodai, et al.
Published: (2024)
by: Shimosato, Kodai, et al.
Published: (2024)
Depth Estimation fusing Image and Radar Measurements with Uncertain Directions
by: Kotani, Masaya, et al.
Published: (2024)
by: Kotani, Masaya, et al.
Published: (2024)
Multi-Person Pose Estimation Evaluation Using Optimal Transportation and Improved Pose Matching
by: Moriki, Takato, et al.
Published: (2026)
by: Moriki, Takato, et al.
Published: (2026)
Enhancing 3D Gaze Estimation in the Wild using Weak Supervision with Gaze Following Labels
by: Vuillecard, Pierre, et al.
Published: (2025)
by: Vuillecard, Pierre, et al.
Published: (2025)
Test-time Cost-and-Quality Controllable Arbitrary-Scale Super-Resolution with Variable Fourier Components
by: Akita, Kazutoshi, et al.
Published: (2024)
by: Akita, Kazutoshi, et al.
Published: (2024)
Data-Driven Stochastic Motion Evaluation and Optimization with Image by Spatially-Aligned Temporal Encoding
by: Oba, Takeru, et al.
Published: (2023)
by: Oba, Takeru, et al.
Published: (2023)
MMCM: Multimodality-aware Metric using Clustering-based Modes for Probabilistic Human Motion Prediction
by: Tokoro, Kyotaro, et al.
Published: (2025)
by: Tokoro, Kyotaro, et al.
Published: (2025)
Burst Super-Resolution with Diffusion Models for Improving Perceptual Quality
by: Tokoro, Kyotaro, et al.
Published: (2024)
by: Tokoro, Kyotaro, et al.
Published: (2024)
Selective Social-Interaction via Individual Importance for Fast Human Trajectory Prediction
by: Urano, Yota, et al.
Published: (2025)
by: Urano, Yota, et al.
Published: (2025)
Joint Learning of Blind Super-Resolution and Crack Segmentation for Realistic Degraded Images
by: Kondo, Yuki, et al.
Published: (2023)
by: Kondo, Yuki, et al.
Published: (2023)
CacheFlow: Fast Human Motion Prediction by Cached Normalizing Flow
by: Maeda, Takahiro, et al.
Published: (2025)
by: Maeda, Takahiro, et al.
Published: (2025)
Physical Plausibility-aware Trajectory Prediction via Locomotion Embodiment
by: Taketsugu, Hiromu, et al.
Published: (2025)
by: Taketsugu, Hiromu, et al.
Published: (2025)
Multimodal Active Measurement for Human Mesh Recovery in Close Proximity
by: Maeda, Takahiro, et al.
Published: (2023)
by: Maeda, Takahiro, et al.
Published: (2023)
MoGA: Mixture-of-Groups Attention for End-to-End Long Video Generation
by: Jia, Weinan, et al.
Published: (2025)
by: Jia, Weinan, et al.
Published: (2025)
Pseudo-Expert Regularized Offline RL for End-to-End Autonomous Driving in Photorealistic Closed-Loop Environments
by: Noguchi, Chihiro, et al.
Published: (2025)
by: Noguchi, Chihiro, et al.
Published: (2025)
ChildPlay-Hand: A Dataset of Hand Manipulations in the Wild
by: Farkhondeh, Arya, et al.
Published: (2024)
by: Farkhondeh, Arya, et al.
Published: (2024)
Efficient Burst Super-Resolution with One-step Diffusion
by: Kawai, Kento, et al.
Published: (2025)
by: Kawai, Kento, et al.
Published: (2025)
NTIRE 2023 Image Shadow Removal Challenge Technical Report: Team IIM_TTI
by: Kondo, Yuki, et al.
Published: (2024)
by: Kondo, Yuki, et al.
Published: (2024)
Micro-Expression-Aware Avatar Fingerprinting via Inter-Frame Feature Differencing
by: Chapariniya, Masoumeh, et al.
Published: (2026)
by: Chapariniya, Masoumeh, et al.
Published: (2026)
End-to-End Visual Autonomous Parking via Control-Aided Attention
by: Chen, Chao, et al.
Published: (2025)
by: Chen, Chao, et al.
Published: (2025)
DiffRefiner: Coarse to Fine Trajectory Planning via Diffusion Refinement with Semantic Interaction for End to End Autonomous Driving
by: Yin, Liuhan, et al.
Published: (2025)
by: Yin, Liuhan, et al.
Published: (2025)
Investigating Identity Signals in Conversational Facial Dynamics via Disentangled Expression Features
by: Chapariniya, Masoumeh, et al.
Published: (2025)
by: Chapariniya, Masoumeh, et al.
Published: (2025)
Eyes on VLM: Benchmarking Gaze Following and Social Gaze Prediction in Vision Language Models
by: Wang, Hengfei, et al.
Published: (2026)
by: Wang, Hengfei, et al.
Published: (2026)
Exploring the Zero-Shot Capabilities of Vision-Language Models for Improving Gaze Following
by: Gupta, Anshul, et al.
Published: (2024)
by: Gupta, Anshul, et al.
Published: (2024)
End-to-End LiDAR optimization for 3D point cloud registration
by: Katyan, Siddhant, et al.
Published: (2026)
by: Katyan, Siddhant, et al.
Published: (2026)
M2SVid: End-to-End Inpainting and Refinement for Monocular-to-Stereo Video Conversion
by: Shvetsova, Nina, et al.
Published: (2025)
by: Shvetsova, Nina, et al.
Published: (2025)
Differentiable NMS via Sinkhorn Matching for End-to-End Fabric Defect Detection
by: Lu, Zhengyang, et al.
Published: (2025)
by: Lu, Zhengyang, et al.
Published: (2025)
Guiding Attention in End-to-End Driving Models
by: Porres, Diego, et al.
Published: (2024)
by: Porres, Diego, et al.
Published: (2024)
An End-to-End Framework for Video Multi-Person Pose Estimation
by: Wei, Zhihong
Published: (2025)
by: Wei, Zhihong
Published: (2025)
Collision Risk Estimation via Loss Prediction in End-to-End Autonomous Driving
by: Xiong, Ziliang, et al.
Published: (2025)
by: Xiong, Ziliang, et al.
Published: (2025)
DEYO: DETR with YOLO for End-to-End Object Detection
by: Ouyang, Haodong
Published: (2024)
by: Ouyang, Haodong
Published: (2024)
Similar Items
-
Dynamic Group Detection using VLM-augmented Temporal Groupness Graph
by: Yokoyama, Kaname, et al.
Published: (2025) -
Group-DINOmics: Incorporating People Dynamics into DINO for Self-supervised Group Activity Feature Learning
by: Tezuka, Ryuki, et al.
Published: (2026) -
Learning Group Activity Features Through Person Attribute Prediction
by: Nakatani, Chihiro, et al.
Published: (2024) -
Human-in-the-loop Adaptation in Group Activity Feature Learning for Team Sports Video Retrieval
by: Nakatani, Chihiro, et al.
Published: (2026) -
Size-Variable Virtual Try-On with Physical Clothes Size
by: Yamashita, Yohei, et al.
Published: (2024)