Beyond the Individual: Introducing Group Intention Forecasting with SHOT Dataset
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Ruixu, Wang, Yuran, Hu, Xinyi, Mai, Chaoyu, Liu, Wenxuan, Xu, Danni, Zhong, Xian, Wang, Zheng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SPAN: Continuous Modeling of Suspicion Progression for Temporal Intention Localization
por: Hu, Xinyi, et al.
Publicado: (2025)
por: Hu, Xinyi, et al.
Publicado: (2025)
Anomize: Better Open Vocabulary Video Anomaly Detection
por: Li, Fei, et al.
Publicado: (2025)
por: Li, Fei, et al.
Publicado: (2025)
Beyond Literal Descriptions: Understanding and Locating Open-World Objects Aligned with Human Intentions
por: Wang, Wenxuan, et al.
Publicado: (2024)
por: Wang, Wenxuan, et al.
Publicado: (2024)
Beyond Single Models: Mitigating Multimodal Hallucinations via Adaptive Token Ensemble Decoding
por: Li, Jinlin, et al.
Publicado: (2025)
por: Li, Jinlin, et al.
Publicado: (2025)
Uncertainty-Aware Token Importance Estimation in Spiking Transformers
por: Liu, Wenxuan, et al.
Publicado: (2026)
por: Liu, Wenxuan, et al.
Publicado: (2026)
Beyond the Horizon: Decoupling Multi-View UAV Action Recognition via Partial Order Transfer
por: Liu, Wenxuan, et al.
Publicado: (2025)
por: Liu, Wenxuan, et al.
Publicado: (2025)
Vision Transformer based Random Walk for Group Re-Identification
por: Zhang, Guoqing, et al.
Publicado: (2024)
por: Zhang, Guoqing, et al.
Publicado: (2024)
RobuSTereo: Robust Zero-Shot Stereo Matching under Adverse Weather
por: Wang, Yuran, et al.
Publicado: (2025)
por: Wang, Yuran, et al.
Publicado: (2025)
Towards Dense and Accurate Radar Perception Via Efficient Cross-Modal Diffusion Model
por: Zhang, Ruibin, et al.
Publicado: (2024)
por: Zhang, Ruibin, et al.
Publicado: (2024)
SpikeDerain: Unveiling Clear Videos from Rainy Sequences Using Color Spike Streams
por: Liang, Hanwen, et al.
Publicado: (2025)
por: Liang, Hanwen, et al.
Publicado: (2025)
360VFI: A Dataset and Benchmark for Omnidirectional Video Frame Interpolation
por: Lu, Wenxuan, et al.
Publicado: (2024)
por: Lu, Wenxuan, et al.
Publicado: (2024)
Towards Low-latency Event-based Visual Recognition with Hybrid Step-wise Distillation Spiking Neural Networks
por: Zhong, Xian, et al.
Publicado: (2024)
por: Zhong, Xian, et al.
Publicado: (2024)
ASM-UNet: Adaptive Scan Mamba Integrating Group Commonalities and Individual Variations for Fine-Grained Segmentation
por: Wang, Bo, et al.
Publicado: (2025)
por: Wang, Bo, et al.
Publicado: (2025)
ONE-SHOT: Compositional Human-Environment Video Synthesis via Spatial-Decoupled Motion Injection and Hybrid Context Integration
por: Yang, Fengyuan, et al.
Publicado: (2026)
por: Yang, Fengyuan, et al.
Publicado: (2026)
Boosting Zero-shot Stereo Matching using Large-scale Mixed Images Sources in the Real World
por: Wang, Yuran, et al.
Publicado: (2025)
por: Wang, Yuran, et al.
Publicado: (2025)
HAD: Hierarchical Asymmetric Distillation to Bridge Spatio-Temporal Gaps in Event-Based Object Tracking
por: Deng, Yao, et al.
Publicado: (2025)
por: Deng, Yao, et al.
Publicado: (2025)
Brain-Inspired Multimodal Spiking Neural Network for Image-Text Retrieval
por: Zong, Xintao, et al.
Publicado: (2026)
por: Zong, Xintao, et al.
Publicado: (2026)
SOTA: Spike-Navigated Optimal TrAnsport Saliency Region Detection in Composite-bias Videos
por: Liu, Wenxuan, et al.
Publicado: (2025)
por: Liu, Wenxuan, et al.
Publicado: (2025)
Intention-Conditioned Long-Term Human Egocentric Action Forecasting
por: Mascaro, Esteve Valls, et al.
Publicado: (2022)
por: Mascaro, Esteve Valls, et al.
Publicado: (2022)
Unveiling Parts Beyond Objects:Towards Finer-Granularity Referring Expression Segmentation
por: Wang, Wenxuan, et al.
Publicado: (2023)
por: Wang, Wenxuan, et al.
Publicado: (2023)
Devil is in Details: Locality-Aware 3D Abdominal CT Volume Generation for Self-Supervised Organ Segmentation
por: Wang, Yuran, et al.
Publicado: (2024)
por: Wang, Yuran, et al.
Publicado: (2024)
OccludeNet: A Causal Journey into Mixed-View Actor-Centric Video Action Recognition under Occlusions
por: Zhou, Guanyu, et al.
Publicado: (2024)
por: Zhou, Guanyu, et al.
Publicado: (2024)
IFNet: Deep Imaging and Focusing for Handheld SAR with Millimeter-wave Signals
por: Li, Yadong, et al.
Publicado: (2024)
por: Li, Yadong, et al.
Publicado: (2024)
Towards Comprehensive Multimodal Perception: Introducing the Touch-Language-Vision Dataset
por: Cheng, Ning, et al.
Publicado: (2024)
por: Cheng, Ning, et al.
Publicado: (2024)
Introducing SDICE: An Index for Assessing Diversity of Synthetic Medical Datasets
por: Alam, Mohammed Talha, et al.
Publicado: (2024)
por: Alam, Mohammed Talha, et al.
Publicado: (2024)
Introducing HOT3D: An Egocentric Dataset for 3D Hand and Object Tracking
por: Banerjee, Prithviraj, et al.
Publicado: (2024)
por: Banerjee, Prithviraj, et al.
Publicado: (2024)
Mono2Stereo: Monocular Knowledge Transfer for Enhanced Stereo Matching
por: Wang, Yuran, et al.
Publicado: (2024)
por: Wang, Yuran, et al.
Publicado: (2024)
Can Video Diffusion Model Reconstruct 4D Geometry?
por: Mai, Jinjie, et al.
Publicado: (2025)
por: Mai, Jinjie, et al.
Publicado: (2025)
Video-Zero: Self-Evolution Video Understanding
por: Zhang, Ruixu, et al.
Publicado: (2026)
por: Zhang, Ruixu, et al.
Publicado: (2026)
Beyond Dataset Distillation: Lossless Dataset Concentration via Diffusion-Assisted Distribution Alignment
por: Liu, Tongfei, et al.
Publicado: (2026)
por: Liu, Tongfei, et al.
Publicado: (2026)
SU-YOLO: Spiking Neural Network for Efficient Underwater Object Detection
por: Li, Chenyang, et al.
Publicado: (2025)
por: Li, Chenyang, et al.
Publicado: (2025)
Learning Group Interactions and Semantic Intentions for Multi-Object Trajectory Prediction
por: Qi, Mengshi, et al.
Publicado: (2024)
por: Qi, Mengshi, et al.
Publicado: (2024)
DeMo: Decoupling Motion Forecasting into Directional Intentions and Dynamic States
por: Zhang, Bozhou, et al.
Publicado: (2024)
por: Zhang, Bozhou, et al.
Publicado: (2024)
Wildfire Smoke Detection System: Model Architecture, Training Mechanism, and Dataset
por: Wang, Chong, et al.
Publicado: (2023)
por: Wang, Chong, et al.
Publicado: (2023)
Expanding Zero-Shot Object Counting with Rich Prompts
por: Zhu, Huilin, et al.
Publicado: (2025)
por: Zhu, Huilin, et al.
Publicado: (2025)
Beyond Average: Individualized Visual Scanpath Prediction
por: Chen, Xianyu, et al.
Publicado: (2024)
por: Chen, Xianyu, et al.
Publicado: (2024)
Does Semantic Noise Initialization Transfer from Images to Videos? A Paired Diagnostic Study
por: Jing, Yixiao, et al.
Publicado: (2026)
por: Jing, Yixiao, et al.
Publicado: (2026)
Scone: Bridging Composition and Distinction in Subject-Driven Image Generation via Unified Understanding-Generation Modeling
por: Wang, Yuran, et al.
Publicado: (2025)
por: Wang, Yuran, et al.
Publicado: (2025)
Passive Non-Line-of-Sight Imaging with Light Transport Modulation
por: Zhang, Jiarui, et al.
Publicado: (2023)
por: Zhang, Jiarui, et al.
Publicado: (2023)
Hyperspectral Image Dataset for Individual Penguin Identification
por: Noboru, Youta, et al.
Publicado: (2024)
por: Noboru, Youta, et al.
Publicado: (2024)
Ejemplares similares
-
SPAN: Continuous Modeling of Suspicion Progression for Temporal Intention Localization
por: Hu, Xinyi, et al.
Publicado: (2025) -
Anomize: Better Open Vocabulary Video Anomaly Detection
por: Li, Fei, et al.
Publicado: (2025) -
Beyond Literal Descriptions: Understanding and Locating Open-World Objects Aligned with Human Intentions
por: Wang, Wenxuan, et al.
Publicado: (2024) -
Beyond Single Models: Mitigating Multimodal Hallucinations via Adaptive Token Ensemble Decoding
por: Li, Jinlin, et al.
Publicado: (2025) -
Uncertainty-Aware Token Importance Estimation in Spiking Transformers
por: Liu, Wenxuan, et al.
Publicado: (2026)