View-Centric Multi-Object Tracking with Homographic Matching in Moving UAV
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Ji, Deyi, Zhu, Lanyun, Gao, Siqi, Zhu, Qi, Zhao, Yiru, Xu, Peng, Ding, Yue, Lu, Hongtao, Ye, Jieping, Wu, Feng, Zhao, Feng |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Tree-of-Table: Unleashing the Power of LLMs for Enhanced Large-Scale Table Understanding
par: Ji, Deyi, et autres
Publié: (2024)
par: Ji, Deyi, et autres
Publié: (2024)
Discrete Latent Perspective Learning for Segmentation and Detection
par: Ji, Deyi, et autres
Publié: (2024)
par: Ji, Deyi, et autres
Publié: (2024)
Structural and Statistical Texture Knowledge Distillation and Learning for Segmentation
par: Ji, Deyi, et autres
Publié: (2025)
par: Ji, Deyi, et autres
Publié: (2025)
PPTFormer: Pseudo Multi-Perspective Transformer for UAV Segmentation
par: Ji, Deyi, et autres
Publié: (2024)
par: Ji, Deyi, et autres
Publié: (2024)
Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification
par: Zhu, Lanyun, et autres
Publié: (2025)
par: Zhu, Lanyun, et autres
Publié: (2025)
LLaFS: When Large Language Models Meet Few-Shot Segmentation
par: Zhu, Lanyun, et autres
Publié: (2023)
par: Zhu, Lanyun, et autres
Publié: (2023)
IBD: Alleviating Hallucinations in Large Vision-Language Models via Image-Biased Decoding
par: Zhu, Lanyun, et autres
Publié: (2024)
par: Zhu, Lanyun, et autres
Publié: (2024)
ChangeNet: Multi-Temporal Asymmetric Change Detection Dataset
par: Ji, Deyi, et autres
Publié: (2023)
par: Ji, Deyi, et autres
Publié: (2023)
Self-signals Driven Multi-LLM Debate for Efficient and Accurate Reasoning
par: Chen, Xuhang, et autres
Publié: (2025)
par: Chen, Xuhang, et autres
Publié: (2025)
Retrv-R1: A Reasoning-Driven MLLM Framework for Universal and Efficient Multimodal Retrieval
par: Zhu, Lanyun, et autres
Publié: (2025)
par: Zhu, Lanyun, et autres
Publié: (2025)
Multi Player Tracking in Ice Hockey with Homographic Projections
par: Prakash, Harish, et autres
Publié: (2024)
par: Prakash, Harish, et autres
Publié: (2024)
SAM3-Adapter: Efficient Adaptation of Segment Anything 3 for Camouflage Object Segmentation, Shadow Detection, and Medical Image Segmentation
par: Chen, Tianrun, et autres
Publié: (2025)
par: Chen, Tianrun, et autres
Publié: (2025)
SAMITE: Position Prompted SAM2 with Calibrated Memory for Visual Object Tracking
par: Xu, Qianxiong, et autres
Publié: (2025)
par: Xu, Qianxiong, et autres
Publié: (2025)
StreamSense: Streaming Social Task Detection with Selective Vision-Language Model Routing
par: Wang, Han, et autres
Publié: (2026)
par: Wang, Han, et autres
Publié: (2026)
RAVEN: Robust Advertisement Video Violation Temporal Grounding via Reinforcement Reasoning
par: Ji, Deyi, et autres
Publié: (2025)
par: Ji, Deyi, et autres
Publié: (2025)
Retrospective Matching Network‐Based One‐Shot Multi‐Object Tracking Method for UAV
par: Hui Zhao, et autres
Publié: (2025)
par: Hui Zhao, et autres
Publié: (2025)
Breaking the Box: Enhancing Remote Sensing Image Segmentation with Freehand Sketches
par: Zang, Ying, et autres
Publié: (2025)
par: Zang, Ying, et autres
Publié: (2025)
xLSTM-UNet can be an Effective 2D & 3D Medical Image Segmentation Backbone with Vision-LSTM (ViL) better than its Mamba Counterpart
par: Chen, Tianrun, et autres
Publié: (2024)
par: Chen, Tianrun, et autres
Publié: (2024)
CamGeo: Sparse Camera-Conditioned Image-to-Video Generation with 3D Geometry Priors
par: Liu, Xuanyi, et autres
Publié: (2026)
par: Liu, Xuanyi, et autres
Publié: (2026)
StreamCacheVGGT: Streaming Visual Geometry Transformers with Robust Scoring and Hybrid Cache Compression
par: Liu, Xuanyi, et autres
Publié: (2026)
par: Liu, Xuanyi, et autres
Publié: (2026)
Homograph Attacks on Maghreb Sentiment Analyzers
par: Qachfar, Fatima Zahra, et autres
Publié: (2024)
par: Qachfar, Fatima Zahra, et autres
Publié: (2024)
Towards Autonomous UAV Visual Object Search in City Space: Benchmark and Agentic Methodology
par: Ji, Yatai, et autres
Publié: (2025)
par: Ji, Yatai, et autres
Publié: (2025)
Video-Zero: Self-Evolution Video Understanding
par: Zhang, Ruixu, et autres
Publié: (2026)
par: Zhang, Ruixu, et autres
Publié: (2026)
WaterWave: Bridging Underwater Image Enhancement into Video Streams via Wavelet-based Temporal Consistency Field
par: Zhu, Qi, et autres
Publié: (2025)
par: Zhu, Qi, et autres
Publié: (2025)
Seeing the Unseen: Mask-Driven Positional Encoding and Strip-Convolution Context Modeling for Cross-View Object Geo-Localization
par: Hu, Shuhan, et autres
Publié: (2025)
par: Hu, Shuhan, et autres
Publié: (2025)
RaTrack: Moving Object Detection and Tracking with 4D Radar Point Cloud
par: Pan, Zhijun, et autres
Publié: (2023)
par: Pan, Zhijun, et autres
Publié: (2023)
Robust 4D Visual Geometry Transformer with Uncertainty-Aware Priors
par: Zang, Ying, et autres
Publié: (2026)
par: Zang, Ying, et autres
Publié: (2026)
Aligning LLM Uncertainty with Human Disagreement in Subjectivity Analysis
par: Lu, Junyu, et autres
Publié: (2026)
par: Lu, Junyu, et autres
Publié: (2026)
Object-Centric Instruction Augmentation for Robotic Manipulation
par: Wen, Junjie, et autres
Publié: (2024)
par: Wen, Junjie, et autres
Publié: (2024)
SAM2-Adapter: Evaluating & Adapting Segment Anything 2 in Downstream Tasks: Camouflage, Shadow, Medical Image Segmentation, and More
par: Chen, Tianrun, et autres
Publié: (2024)
par: Chen, Tianrun, et autres
Publié: (2024)
Segment Anything with Motion, Geometry, and Semantic Adaptation for Complex Nonlinear Visual Object Tracking
par: Zhu, Deyi, et autres
Publié: (2026)
par: Zhu, Deyi, et autres
Publié: (2026)
In Defense and Revival of Bayesian Filtering for Thermal Infrared Object Tracking
par: Gao, Peng, et autres
Publié: (2024)
par: Gao, Peng, et autres
Publié: (2024)
Multi-Agent VLMs Guided Self-Training with PNU Loss for Low-Resource Offensive Content Detection
par: Wang, Han, et autres
Publié: (2025)
par: Wang, Han, et autres
Publié: (2025)
ARGUS: Policy-Adaptive Ad Governance via Evolving Reinforcement with Adversarial Umpiring
par: Ji, Deyi, et autres
Publié: (2026)
par: Ji, Deyi, et autres
Publié: (2026)
Multiple Object Tracking as ID Prediction
par: Gao, Ruopeng, et autres
Publié: (2024)
par: Gao, Ruopeng, et autres
Publié: (2024)
4DVGGT-D: 4D Visual Geometry Transformer with Improved Dynamic Depth Estimation
par: Zang, Ying, et autres
Publié: (2026)
par: Zang, Ying, et autres
Publié: (2026)
Learning an Adaptive and View-Invariant Vision Transformer for Real-Time UAV Tracking
par: Wu, You, et autres
Publié: (2024)
par: Wu, You, et autres
Publié: (2024)
AerialMind: Towards Referring Multi-Object Tracking in UAV Scenarios
par: Chen, Chenglizhao, et autres
Publié: (2025)
par: Chen, Chenglizhao, et autres
Publié: (2025)
FuxiMT: Sparsifying Large Language Models for Chinese-Centric Multilingual Machine Translation
par: Zhu, Shaolin, et autres
Publié: (2025)
par: Zhu, Shaolin, et autres
Publié: (2025)
CST Anti-UAV: A Thermal Infrared Benchmark for Tiny UAV Tracking in Complex Scenes
par: Xie, Bin, et autres
Publié: (2025)
par: Xie, Bin, et autres
Publié: (2025)
Documents similaires
-
Tree-of-Table: Unleashing the Power of LLMs for Enhanced Large-Scale Table Understanding
par: Ji, Deyi, et autres
Publié: (2024) -
Discrete Latent Perspective Learning for Segmentation and Detection
par: Ji, Deyi, et autres
Publié: (2024) -
Structural and Statistical Texture Knowledge Distillation and Learning for Segmentation
par: Ji, Deyi, et autres
Publié: (2025) -
PPTFormer: Pseudo Multi-Perspective Transformer for UAV Segmentation
par: Ji, Deyi, et autres
Publié: (2024) -
Not Every Patch is Needed: Towards a More Efficient and Effective Backbone for Video-based Person Re-identification
par: Zhu, Lanyun, et autres
Publié: (2025)