Crowded Video Individual Counting Informed by Social Grouping and Spatial-Temporal Displacement Priors
Fuente:
arXiv
Saved in:
| Main Authors: | Lu, Hao, Zhu, Xuhui, Zhang, Wenjing, Li, Yanan, Bai, Xiang |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Video Individual Counting With Implicit One-to-Many Matching
by: Zhu, Xuhui, et al.
Published: (2025)
by: Zhu, Xuhui, et al.
Published: (2025)
Foggy Crowd Counting: Combining Physical Priors and KAN-Graph
by: Wang, Yuhao, et al.
Published: (2025)
by: Wang, Yuhao, et al.
Published: (2025)
FGENet: Fine-Grained Extraction Network for Congested Crowd Counting
by: Ma, Hao-Yuan, et al.
Published: (2024)
by: Ma, Hao-Yuan, et al.
Published: (2024)
Embodied Crowd Counting
by: Long, Runling, et al.
Published: (2025)
by: Long, Runling, et al.
Published: (2025)
VMambaCC: A Visual State Space Model for Crowd Counting
by: Ma, Hao-Yuan, et al.
Published: (2024)
by: Ma, Hao-Yuan, et al.
Published: (2024)
VCBench: A Streaming Counting Benchmark for Spatial-Temporal State Maintenance in Long Videos
by: Liu, Pengyiang, et al.
Published: (2026)
by: Liu, Pengyiang, et al.
Published: (2026)
EnvSocial-Diff: A Diffusion-Based Crowd Simulation Model with Environmental Conditioning and Individual-Group Interaction
by: Zhao, Bingxue, et al.
Published: (2026)
by: Zhao, Bingxue, et al.
Published: (2026)
RACANet: Reliability-Aware Crowd Anchor Network for RGB-T Crowd Counting
by: Shi, Jinghao, et al.
Published: (2026)
by: Shi, Jinghao, et al.
Published: (2026)
CountFormer: Multi-View Crowd Counting Transformer
by: Mo, Hong, et al.
Published: (2024)
by: Mo, Hong, et al.
Published: (2024)
The Effectiveness of a Simplified Model Structure for Crowd Counting
by: Chen, Lei, et al.
Published: (2024)
by: Chen, Lei, et al.
Published: (2024)
Density Estimation and Crowd Counting
by: Sunil, Balachandra Devarangadi, et al.
Published: (2025)
by: Sunil, Balachandra Devarangadi, et al.
Published: (2025)
FitControler: Toward Fit-Aware Virtual Try-On
by: Yang, Lu, et al.
Published: (2025)
by: Yang, Lu, et al.
Published: (2025)
Structured Video-Language Modeling with Temporal Grouping and Spatial Grounding
by: Xiong, Yuanhao, et al.
Published: (2023)
by: Xiong, Yuanhao, et al.
Published: (2023)
Rethinking Global Context in Crowd Counting
by: Sun, Guolei, et al.
Published: (2021)
by: Sun, Guolei, et al.
Published: (2021)
Single Domain Generalization for Crowd Counting
by: Peng, Zhuoxuan, et al.
Published: (2024)
by: Peng, Zhuoxuan, et al.
Published: (2024)
Learning Discriminative Features for Crowd Counting
by: Chen, Yuehai, et al.
Published: (2023)
by: Chen, Yuehai, et al.
Published: (2023)
A Dual-Modulation Framework for RGB-T Crowd Counting via Spatially Modulated Attention and Adaptive Fusion
by: Feng, Yuhong, et al.
Published: (2025)
by: Feng, Yuhong, et al.
Published: (2025)
Temporal-Visual Semantic Alignment: A Unified Architecture for Transferring Spatial Priors from Vision Models to Zero-Shot Temporal Tasks
by: Ma, Xiangkai, et al.
Published: (2025)
by: Ma, Xiangkai, et al.
Published: (2025)
CrowdVLM-R1: Expanding R1 Ability to Vision Language Model for Crowd Counting using Fuzzy Group Relative Policy Reward
by: Wang, Zhiqiang, et al.
Published: (2025)
by: Wang, Zhiqiang, et al.
Published: (2025)
Video Individual Counting for Moving Drones
by: Fan, Yaowu, et al.
Published: (2025)
by: Fan, Yaowu, et al.
Published: (2025)
One-Shot Crowd Counting With Density Guidance For Scene Adaptation
by: Chen, Jiwei, et al.
Published: (2026)
by: Chen, Jiwei, et al.
Published: (2026)
Multi-View Crowd Counting With Self-Supervised Learning
by: Mo, Hong, et al.
Published: (2025)
by: Mo, Hong, et al.
Published: (2025)
Transformer-Based Dual-Optical Attention Fusion Crowd Head Point Counting and Localization Network
by: Zhou, Fei, et al.
Published: (2025)
by: Zhou, Fei, et al.
Published: (2025)
Regressor-Segmenter Mutual Prompt Learning for Crowd Counting
by: Guo, Mingyue, et al.
Published: (2023)
by: Guo, Mingyue, et al.
Published: (2023)
WSCF-MVCC: Weakly-supervised Calibration-free Multi-view Crowd Counting
by: Li, Bin, et al.
Published: (2025)
by: Li, Bin, et al.
Published: (2025)
DyCrowd: Towards Dynamic Crowd Reconstruction from a Large-scene Video
by: Wen, Hao, et al.
Published: (2025)
by: Wen, Hao, et al.
Published: (2025)
Semi-Supervised Crowd Counting with Contextual Modeling: Facilitating Holistic Understanding of Crowd Scenes
by: Qian, Yifei, et al.
Published: (2023)
by: Qian, Yifei, et al.
Published: (2023)
Counting Fish with Temporal Representations of Sonar Video
by: Van Brunt, Kai, et al.
Published: (2025)
by: Van Brunt, Kai, et al.
Published: (2025)
L2HCount:Generalizing Crowd Counting from Low to High Crowd Density via Density Simulation
by: Xu, Guoliang, et al.
Published: (2025)
by: Xu, Guoliang, et al.
Published: (2025)
Curriculum for Crowd Counting -- Is it Worthy?
by: Khan, Muhammad Asif, et al.
Published: (2024)
by: Khan, Muhammad Asif, et al.
Published: (2024)
ProgRoCC: A Progressive Approach to Rough Crowd Counting
by: Jiang, Shengqin, et al.
Published: (2025)
by: Jiang, Shengqin, et al.
Published: (2025)
EHNet: An Efficient Hybrid Network for Crowd Counting and Localization
by: Yan, Yuqing, et al.
Published: (2025)
by: Yan, Yuqing, et al.
Published: (2025)
Local Information Matters: A Rethink of Crowd Counting
by: Pan, Tianhang, et al.
Published: (2025)
by: Pan, Tianhang, et al.
Published: (2025)
Multi-modal Crowd Counting via Modal Emulation
by: Wang, Chenhao, et al.
Published: (2024)
by: Wang, Chenhao, et al.
Published: (2024)
Learning Temporally Consistent Video Depth from Video Diffusion Priors
by: Shao, Jiahao, et al.
Published: (2024)
by: Shao, Jiahao, et al.
Published: (2024)
Autoregressive Video Autoencoder with Decoupled Temporal and Spatial Context
by: Shen, Cuifeng, et al.
Published: (2025)
by: Shen, Cuifeng, et al.
Published: (2025)
Spatial-Temporal Graph Mamba for Music-Guided Dance Video Synthesis
by: Tang, Hao, et al.
Published: (2025)
by: Tang, Hao, et al.
Published: (2025)
Active View Selection for Scene-level Multi-view Crowd Counting and Localization with Limited Labels
by: Zhang, Qi, et al.
Published: (2025)
by: Zhang, Qi, et al.
Published: (2025)
VLIPP: Towards Physically Plausible Video Generation with Vision and Language Informed Physical Prior
by: Yang, Xindi, et al.
Published: (2025)
by: Yang, Xindi, et al.
Published: (2025)
DVFace: Spatio-Temporal Dual-Prior Diffusion for Video Face Restoration
by: Chen, Zheng, et al.
Published: (2026)
by: Chen, Zheng, et al.
Published: (2026)
Similar Items
-
Video Individual Counting With Implicit One-to-Many Matching
by: Zhu, Xuhui, et al.
Published: (2025) -
Foggy Crowd Counting: Combining Physical Priors and KAN-Graph
by: Wang, Yuhao, et al.
Published: (2025) -
FGENet: Fine-Grained Extraction Network for Congested Crowd Counting
by: Ma, Hao-Yuan, et al.
Published: (2024) -
Embodied Crowd Counting
by: Long, Runling, et al.
Published: (2025) -
VMambaCC: A Visual State Space Model for Crowd Counting
by: Ma, Hao-Yuan, et al.
Published: (2024)