Perception-to-Pursuit: Track-Centric Temporal Reasoning for Open-World Drone Detection and Autonomous Chasing
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Oruganti, Venkatakrishna Reddy |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SGLoc: Semantic Localization System for Camera Pose Estimation from 3D Gaussian Splatting Representation
von: Xu, Beining, et al.
Veröffentlicht: (2025)
von: Xu, Beining, et al.
Veröffentlicht: (2025)
Adaptive Thresholding for Visual Place Recognition using Negative Gaussian Mixture Statistics
von: Trinh, Nick, et al.
Veröffentlicht: (2025)
von: Trinh, Nick, et al.
Veröffentlicht: (2025)
Rethinking Camera Choice: An Empirical Study on Fisheye Camera Properties in Robotic Manipulation
von: Xue, Han, et al.
Veröffentlicht: (2026)
von: Xue, Han, et al.
Veröffentlicht: (2026)
Can VLMs Unlock Semantic Anomaly Detection? A Framework for Structured Reasoning
von: Brusnicki, Roberto, et al.
Veröffentlicht: (2025)
von: Brusnicki, Roberto, et al.
Veröffentlicht: (2025)
Large Language Models and 3D Vision for Intelligent Robotic Perception and Autonomy
von: Mehta, Vinit, et al.
Veröffentlicht: (2025)
von: Mehta, Vinit, et al.
Veröffentlicht: (2025)
Temporally Consistent Object 6D Pose Estimation for Robot Control
von: Zorina, Kateryna, et al.
Veröffentlicht: (2026)
von: Zorina, Kateryna, et al.
Veröffentlicht: (2026)
Systematic Comparison of Projection Methods for Monocular 3D Human Pose Estimation on Fisheye Images
von: Käs, Stephanie, et al.
Veröffentlicht: (2025)
von: Käs, Stephanie, et al.
Veröffentlicht: (2025)
From Photons to Physics: Autonomous Indoor Drones and the Future of Objective Property Assessment
von: Teikari, Petteri, et al.
Veröffentlicht: (2025)
von: Teikari, Petteri, et al.
Veröffentlicht: (2025)
Unveiling the Potential of iMarkers: Invisible Fiducial Markers for Advanced Robotics
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
SuperPoint-SLAM3: Augmenting ORB-SLAM3 with Deep Features, Adaptive NMS, and Learning-Based Loop Closure
von: Syed, Shahram Najam, et al.
Veröffentlicht: (2025)
von: Syed, Shahram Najam, et al.
Veröffentlicht: (2025)
vS-Graphs: Tightly Coupling Visual SLAM and 3D Scene Graphs Exploiting Hierarchical Scene Understanding
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
Evaluating the Impact of Synthetic Data on Object Detection Tasks in Autonomous Driving
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
CARScenes: Semantic VLM Dataset for Safe Autonomous Driving
von: He, Yuankai, et al.
Veröffentlicht: (2025)
von: He, Yuankai, et al.
Veröffentlicht: (2025)
Shaken, Not Stirred: A Novel Dataset for Visual Understanding of Glasses in Human-Robot Bartending Tasks
von: Gajdošech, Lukáš, et al.
Veröffentlicht: (2025)
von: Gajdošech, Lukáš, et al.
Veröffentlicht: (2025)
Botany Meets Robotics in Alpine Scree Monitoring
von: De Benedittis, Davide, et al.
Veröffentlicht: (2025)
von: De Benedittis, Davide, et al.
Veröffentlicht: (2025)
Have We Mastered Scale in Deep Monocular Visual SLAM? The ScaleMaster Dataset and Benchmark
von: Ju, Hyoseok, et al.
Veröffentlicht: (2026)
von: Ju, Hyoseok, et al.
Veröffentlicht: (2026)
BEVTraj: Map-Free End-to-End Trajectory Prediction in Bird's-Eye View with Deformable Attention and Sparse Goal Proposals
von: Kong, Minsang, et al.
Veröffentlicht: (2025)
von: Kong, Minsang, et al.
Veröffentlicht: (2025)
Is Single-View Mesh Reconstruction Ready for Robotics?
von: Nolte, Frederik, et al.
Veröffentlicht: (2025)
von: Nolte, Frederik, et al.
Veröffentlicht: (2025)
StemVLA:An Open-Source Vision-Language-Action Model with Future 3D Spatial Geometry Knowledge and 4D Historical Representation
von: Xiao, Jiasong, et al.
Veröffentlicht: (2026)
von: Xiao, Jiasong, et al.
Veröffentlicht: (2026)
Verifiable Obstacle Detection
von: Bansal, Ayoosh, et al.
Veröffentlicht: (2022)
von: Bansal, Ayoosh, et al.
Veröffentlicht: (2022)
Human-Centric Perception for Child Sexual Abuse Imagery
von: Laranjeira, Camila, et al.
Veröffentlicht: (2026)
von: Laranjeira, Camila, et al.
Veröffentlicht: (2026)
Thermal and RGB Images Work Better Together in Wind Turbine Damage Detection
von: Svystun, Serhii, et al.
Veröffentlicht: (2024)
von: Svystun, Serhii, et al.
Veröffentlicht: (2024)
YOLO-APD: Enhancing YOLOv8 for Robust Pedestrian Detection on Complex Road Geometries
von: Joctum, Aquino, et al.
Veröffentlicht: (2025)
von: Joctum, Aquino, et al.
Veröffentlicht: (2025)
MARS: Multi-Agent Robotic System with Multimodal Large Language Models for Assistive Intelligence
von: Gao, Renjun
Veröffentlicht: (2025)
von: Gao, Renjun
Veröffentlicht: (2025)
Infrastructure-Centric World Models: Bridging Temporal Depth and Spatial Breadth for Roadside Perception
von: Meng, Siyuan, et al.
Veröffentlicht: (2026)
von: Meng, Siyuan, et al.
Veröffentlicht: (2026)
TinyNav: End-to-End TinyML for Real-Time Autonomous Navigation on Microcontrollers
von: Roy, Pooria, et al.
Veröffentlicht: (2026)
von: Roy, Pooria, et al.
Veröffentlicht: (2026)
CODEI: Resource-Efficient Task-Driven Co-Design of Perception and Decision Making for Mobile Robots Applied to Autonomous Vehicles
von: Milojevic, Dejan, et al.
Veröffentlicht: (2025)
von: Milojevic, Dejan, et al.
Veröffentlicht: (2025)
Seeing Roads Through Words: A Language-Guided Framework for RGB-T Driving Scene Segmentation
von: Reddy, Ruturaj, et al.
Veröffentlicht: (2026)
von: Reddy, Ruturaj, et al.
Veröffentlicht: (2026)
Single-Shot Metric Depth from Focused Plenoptic Cameras
von: Lasheras-Hernandez, Blanca, et al.
Veröffentlicht: (2024)
von: Lasheras-Hernandez, Blanca, et al.
Veröffentlicht: (2024)
Distant Object Localisation from Noisy Image Segmentation Sequences
von: Pesonen, Julius, et al.
Veröffentlicht: (2025)
von: Pesonen, Julius, et al.
Veröffentlicht: (2025)
A Multi-purpose Tracking Framework for Salmon Welfare Monitoring in Challenging Environments
von: Høgstedt, Espen Uri, et al.
Veröffentlicht: (2025)
von: Høgstedt, Espen Uri, et al.
Veröffentlicht: (2025)
Locate 3D: Real-World Object Localization via Self-Supervised Learning in 3D
von: Arnaud, Sergio, et al.
Veröffentlicht: (2025)
von: Arnaud, Sergio, et al.
Veröffentlicht: (2025)
Traffic Scene Small Target Detection Method Based on YOLOv8n-SPTS Model for Autonomous Driving
von: Wu, Songhan
Veröffentlicht: (2025)
von: Wu, Songhan
Veröffentlicht: (2025)
How do Foundation Models Compare to Skeleton-Based Approaches for Gesture Recognition in Human-Robot Interaction?
von: Käs, Stephanie, et al.
Veröffentlicht: (2025)
von: Käs, Stephanie, et al.
Veröffentlicht: (2025)
Decoupled Sensitivity-Consistency Learning for Weakly Supervised Video Anomaly Detection
von: Zheng, Hantao, et al.
Veröffentlicht: (2026)
von: Zheng, Hantao, et al.
Veröffentlicht: (2026)
FCBV-Net: Category-Level Robotic Garment Smoothing via Feature-Conditioned Bimanual Value Prediction
von: Daba, Mohammed, et al.
Veröffentlicht: (2025)
von: Daba, Mohammed, et al.
Veröffentlicht: (2025)
FACT: Multinomial Misalignment Classification for Point Cloud Registration
von: Dillén, Ludvig, et al.
Veröffentlicht: (2025)
von: Dillén, Ludvig, et al.
Veröffentlicht: (2025)
Autonomous Underwater Cognitive System for Adaptive Navigation: A SLAM-Integrated Cognitive Architecture
von: Jayarathne, K. A. I. N, et al.
Veröffentlicht: (2025)
von: Jayarathne, K. A. I. N, et al.
Veröffentlicht: (2025)
Edge-Enabled Collaborative Object Detection for Real-Time Multi-Vehicle Perception
von: Richards, Everett, et al.
Veröffentlicht: (2025)
von: Richards, Everett, et al.
Veröffentlicht: (2025)
Pedestrian motion prediction evaluation for urban autonomous driving
von: Zabolotnii, Dmytro, et al.
Veröffentlicht: (2024)
von: Zabolotnii, Dmytro, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SGLoc: Semantic Localization System for Camera Pose Estimation from 3D Gaussian Splatting Representation
von: Xu, Beining, et al.
Veröffentlicht: (2025) -
Adaptive Thresholding for Visual Place Recognition using Negative Gaussian Mixture Statistics
von: Trinh, Nick, et al.
Veröffentlicht: (2025) -
Rethinking Camera Choice: An Empirical Study on Fisheye Camera Properties in Robotic Manipulation
von: Xue, Han, et al.
Veröffentlicht: (2026) -
Can VLMs Unlock Semantic Anomaly Detection? A Framework for Structured Reasoning
von: Brusnicki, Roberto, et al.
Veröffentlicht: (2025) -
Large Language Models and 3D Vision for Intelligent Robotic Perception and Autonomy
von: Mehta, Vinit, et al.
Veröffentlicht: (2025)