Video Individual Counting With Implicit One-to-Many Matching
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhu, Xuhui, Xu, Jing, Wang, Bingjie, Dai, Huikang, Lu, Hao |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Crowded Video Individual Counting Informed by Social Grouping and Spatial-Temporal Displacement Priors
par: Lu, Hao, et autres
Publié: (2026)
par: Lu, Hao, et autres
Publié: (2026)
Differentiable NMS via Sinkhorn Matching for End-to-End Fabric Defect Detection
par: Lu, Zhengyang, et autres
Publié: (2025)
par: Lu, Zhengyang, et autres
Publié: (2025)
Video Individual Counting for Moving Drones
par: Fan, Yaowu, et autres
Publié: (2025)
par: Fan, Yaowu, et autres
Publié: (2025)
TokenBinder: Text-Video Retrieval with One-to-Many Alignment Paradigm
par: Zhang, Bingqing, et autres
Publié: (2024)
par: Zhang, Bingqing, et autres
Publié: (2024)
Deep Understanding of Soccer Match Videos
par: Xu, Shikun, et autres
Publié: (2024)
par: Xu, Shikun, et autres
Publié: (2024)
One-Shot Crowd Counting With Density Guidance For Scene Adaptation
par: Chen, Jiwei, et autres
Publié: (2026)
par: Chen, Jiwei, et autres
Publié: (2026)
Match4Annotate: Propagating Sparse Video Annotations via Implicit Neural Feature Matching
par: Zhang, Zhuorui, et autres
Publié: (2026)
par: Zhang, Zhuorui, et autres
Publié: (2026)
Video Individual Counting and Tracking from Moving Drones: A Benchmark and Methods
par: Fan, Yaowu, et autres
Publié: (2026)
par: Fan, Yaowu, et autres
Publié: (2026)
EmoLat: Text-driven Image Sentiment Transfer via Emotion Latent Space
par: Zhang, Jing, et autres
Publié: (2026)
par: Zhang, Jing, et autres
Publié: (2026)
EmoKGEdit: Training-free Affective Injection via Visual Cue Transformation
par: Zhang, Jing, et autres
Publié: (2026)
par: Zhang, Jing, et autres
Publié: (2026)
CausalSR: Structural Causal Model-Driven Super-Resolution with Counterfactual Inference
par: Lu, Zhengyang, et autres
Publié: (2025)
par: Lu, Zhengyang, et autres
Publié: (2025)
Geometry-Aware Implicit Memory for Video World Models
par: Wei, Zhengxuan, et autres
Publié: (2026)
par: Wei, Zhengxuan, et autres
Publié: (2026)
Deforming Videos to Masks: Flow Matching for Referring Video Segmentation
par: Wang, Zanyi, et autres
Publié: (2025)
par: Wang, Zanyi, et autres
Publié: (2025)
Matcher: Segment Anything with One Shot Using All-Purpose Feature Matching
par: Liu, Yang, et autres
Publié: (2023)
par: Liu, Yang, et autres
Publié: (2023)
Efficient Masked AutoEncoder for Video Object Counting and A Large-Scale Benchmark
par: Cao, Bing, et autres
Publié: (2024)
par: Cao, Bing, et autres
Publié: (2024)
One Identity, Many Roles: Multimodal Entity Coreference for Enhanced Video Situation Recognition
par: Darur, Balaji, et autres
Publié: (2026)
par: Darur, Balaji, et autres
Publié: (2026)
Match-Stereo-Videos: Bidirectional Alignment for Consistent Dynamic Stereo Matching
par: Jing, Junpeng, et autres
Publié: (2024)
par: Jing, Junpeng, et autres
Publié: (2024)
Fast Encoding and Decoding for Implicit Video Representation
par: Chen, Hao, et autres
Publié: (2024)
par: Chen, Hao, et autres
Publié: (2024)
Reward Forcing: Efficient Streaming Video Generation with Rewarded Distribution Matching Distillation
par: Lu, Yunhong, et autres
Publié: (2025)
par: Lu, Yunhong, et autres
Publié: (2025)
Blind Video Super-Resolution based on Implicit Kernels
par: Zhu, Qiang, et autres
Publié: (2025)
par: Zhu, Qiang, et autres
Publié: (2025)
From One-to-One to Many-to-Many: Dynamic Cross-Layer Injection for Deep Vision-Language Fusion
par: Chen, Cheng, et autres
Publié: (2026)
par: Chen, Cheng, et autres
Publié: (2026)
One View, Many Worlds: Single-Image to 3D Object Meets Generative Domain Randomization for One-Shot 6D Pose Estimation
par: Geng, Zheng, et autres
Publié: (2025)
par: Geng, Zheng, et autres
Publié: (2025)
NeRV-Diffusion: Diffuse Implicit Neural Representations for Video Synthesis
par: Ren, Yixuan, et autres
Publié: (2025)
par: Ren, Yixuan, et autres
Publié: (2025)
Revisit Event Generation Model: Self-Supervised Learning of Event-to-Video Reconstruction with Implicit Neural Representations
par: Wang, Zipeng, et autres
Publié: (2024)
par: Wang, Zipeng, et autres
Publié: (2024)
InstaRevive: One-Step Image Enhancement via Dynamic Score Matching
par: Zhu, Yixuan, et autres
Publié: (2025)
par: Zhu, Yixuan, et autres
Publié: (2025)
Match Stereo Videos via Bidirectional Alignment
par: Jing, Junpeng, et autres
Publié: (2024)
par: Jing, Junpeng, et autres
Publié: (2024)
GeoDM: Geometry-aware Distribution Matching for Dataset Distillation
par: Li, Xuhui, et autres
Publié: (2025)
par: Li, Xuhui, et autres
Publié: (2025)
Learning Sewing Patterns via Latent Flow Matching of Implicit Fields
par: Cao, Cong, et autres
Publié: (2026)
par: Cao, Cong, et autres
Publié: (2026)
Depth-Aware Super-Resolution via Distance-Adaptive Variational Formulation
par: Guo, Tianhao, et autres
Publié: (2025)
par: Guo, Tianhao, et autres
Publié: (2025)
From Transparent to Opaque: Rethinking Neural Implicit Surfaces with $α$-NeuS
par: Zhang, Haoran, et autres
Publié: (2024)
par: Zhang, Haoran, et autres
Publié: (2024)
Enhancing Video Super-Resolution via Implicit Resampling-based Alignment
par: Xu, Kai, et autres
Publié: (2023)
par: Xu, Kai, et autres
Publié: (2023)
FMVP: Masked Flow Matching for Adversarial Video Purification
par: Tang, Duoxun, et autres
Publié: (2026)
par: Tang, Duoxun, et autres
Publié: (2026)
Stereo Any Video: Temporally Consistent Stereo Matching
par: Jing, Junpeng, et autres
Publié: (2025)
par: Jing, Junpeng, et autres
Publié: (2025)
Plant Taxonomy Meets Plant Counting: A Fine-Grained, Taxonomic Dataset for Counting Hundreds of Plant Species
par: Xu, Jinyu, et autres
Publié: (2026)
par: Xu, Jinyu, et autres
Publié: (2026)
One-Step Diffusion Distillation through Score Implicit Matching
par: Luo, Weijian, et autres
Publié: (2024)
par: Luo, Weijian, et autres
Publié: (2024)
MultiCounter: Multiple Action Agnostic Repetition Counting in Untrimmed Videos
par: Tang, Yin, et autres
Publié: (2024)
par: Tang, Yin, et autres
Publié: (2024)
Many-for-Many: Unify the Training of Multiple Video and Image Generation and Manipulation Tasks
par: Li, Ruibin, et autres
Publié: (2025)
par: Li, Ruibin, et autres
Publié: (2025)
Video2LoRA: Unified Semantic-Controlled Video Generation via Per-Reference-Video LoRA
par: Wu, Zexi, et autres
Publié: (2026)
par: Wu, Zexi, et autres
Publié: (2026)
Memories are One-to-Many Mapping Alleviators in Talking Face Generation
par: Tang, Anni, et autres
Publié: (2022)
par: Tang, Anni, et autres
Publié: (2022)
Impact of Sunglasses on One-to-Many Facial Identification Accuracy
par: Tian, Sicong, et autres
Publié: (2024)
par: Tian, Sicong, et autres
Publié: (2024)
Documents similaires
-
Crowded Video Individual Counting Informed by Social Grouping and Spatial-Temporal Displacement Priors
par: Lu, Hao, et autres
Publié: (2026) -
Differentiable NMS via Sinkhorn Matching for End-to-End Fabric Defect Detection
par: Lu, Zhengyang, et autres
Publié: (2025) -
Video Individual Counting for Moving Drones
par: Fan, Yaowu, et autres
Publié: (2025) -
TokenBinder: Text-Video Retrieval with One-to-Many Alignment Paradigm
par: Zhang, Bingqing, et autres
Publié: (2024) -
Deep Understanding of Soccer Match Videos
par: Xu, Shikun, et autres
Publié: (2024)