YOLO-APD: Enhancing YOLOv8 for Robust Pedestrian Detection on Complex Road Geometries
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Joctum, Aquino, Kandiri, John |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FCBV-Net: Category-Level Robotic Garment Smoothing via Feature-Conditioned Bimanual Value Prediction
von: Daba, Mohammed, et al.
Veröffentlicht: (2025)
von: Daba, Mohammed, et al.
Veröffentlicht: (2025)
MARS: Multi-Agent Robotic System with Multimodal Large Language Models for Assistive Intelligence
von: Gao, Renjun
Veröffentlicht: (2025)
von: Gao, Renjun
Veröffentlicht: (2025)
Locate 3D: Real-World Object Localization via Self-Supervised Learning in 3D
von: Arnaud, Sergio, et al.
Veröffentlicht: (2025)
von: Arnaud, Sergio, et al.
Veröffentlicht: (2025)
TinyNav: End-to-End TinyML for Real-Time Autonomous Navigation on Microcontrollers
von: Roy, Pooria, et al.
Veröffentlicht: (2026)
von: Roy, Pooria, et al.
Veröffentlicht: (2026)
CC-SGG: Corner Case Scenario Generation using Learned Scene Graphs
von: Drayson, George, et al.
Veröffentlicht: (2023)
von: Drayson, George, et al.
Veröffentlicht: (2023)
MaSC: A Masked Similarity Metric for Evaluating Concept-Driven Generation
von: Bartkowiak, Patryk, et al.
Veröffentlicht: (2026)
von: Bartkowiak, Patryk, et al.
Veröffentlicht: (2026)
SmartDate: AI-Driven Precision Sorting and Quality Control in Date Fruits
von: Eskaf, Khaled
Veröffentlicht: (2025)
von: Eskaf, Khaled
Veröffentlicht: (2025)
Botany Meets Robotics in Alpine Scree Monitoring
von: De Benedittis, Davide, et al.
Veröffentlicht: (2025)
von: De Benedittis, Davide, et al.
Veröffentlicht: (2025)
Have We Mastered Scale in Deep Monocular Visual SLAM? The ScaleMaster Dataset and Benchmark
von: Ju, Hyoseok, et al.
Veröffentlicht: (2026)
von: Ju, Hyoseok, et al.
Veröffentlicht: (2026)
NV3D: Leveraging Spatial Shape Through Normal Vector-based 3D Object Detection
von: Chaowakarn, Krittin, et al.
Veröffentlicht: (2025)
von: Chaowakarn, Krittin, et al.
Veröffentlicht: (2025)
Evaluating the Impact of Synthetic Data on Object Detection Tasks in Autonomous Driving
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
Thermal and RGB Images Work Better Together in Wind Turbine Damage Detection
von: Svystun, Serhii, et al.
Veröffentlicht: (2024)
von: Svystun, Serhii, et al.
Veröffentlicht: (2024)
Selective Progress-Aware Querying for Human-in-the-Loop Reinforcement Learning
von: Muraleedharan, Anujith, et al.
Veröffentlicht: (2025)
von: Muraleedharan, Anujith, et al.
Veröffentlicht: (2025)
SemanticFeels: Semantic Labeling during In-Hand Manipulation
von: Khalil, Anas Al Shikh, et al.
Veröffentlicht: (2026)
von: Khalil, Anas Al Shikh, et al.
Veröffentlicht: (2026)
Seeing Roads Through Words: A Language-Guided Framework for RGB-T Driving Scene Segmentation
von: Reddy, Ruturaj, et al.
Veröffentlicht: (2026)
von: Reddy, Ruturaj, et al.
Veröffentlicht: (2026)
Systematic Comparison of Projection Methods for Monocular 3D Human Pose Estimation on Fisheye Images
von: Käs, Stephanie, et al.
Veröffentlicht: (2025)
von: Käs, Stephanie, et al.
Veröffentlicht: (2025)
BEVTraj: Map-Free End-to-End Trajectory Prediction in Bird's-Eye View with Deformable Attention and Sparse Goal Proposals
von: Kong, Minsang, et al.
Veröffentlicht: (2025)
von: Kong, Minsang, et al.
Veröffentlicht: (2025)
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
von: Chahine, Makram, et al.
Veröffentlicht: (2024)
von: Chahine, Makram, et al.
Veröffentlicht: (2024)
Lifelong Learning in Vision-Language Models: Enhanced EWC with Cross-Modal Knowledge Retention
von: Durrani, Hamza Ahmed, et al.
Veröffentlicht: (2026)
von: Durrani, Hamza Ahmed, et al.
Veröffentlicht: (2026)
Single-Shot Metric Depth from Focused Plenoptic Cameras
von: Lasheras-Hernandez, Blanca, et al.
Veröffentlicht: (2024)
von: Lasheras-Hernandez, Blanca, et al.
Veröffentlicht: (2024)
Perception-to-Pursuit: Track-Centric Temporal Reasoning for Open-World Drone Detection and Autonomous Chasing
von: Oruganti, Venkatakrishna Reddy
Veröffentlicht: (2026)
von: Oruganti, Venkatakrishna Reddy
Veröffentlicht: (2026)
Visual Categorization Across Minds and Models: Cognitive Analysis of Human Labeling and Neuro-Symbolic Integration
von: Kabgere, Chethana Prasad
Veröffentlicht: (2025)
von: Kabgere, Chethana Prasad
Veröffentlicht: (2025)
Cyclic Contrastive Knowledge Transfer for Open-Vocabulary Object Detection
von: Zhang, Chuhan, et al.
Veröffentlicht: (2025)
von: Zhang, Chuhan, et al.
Veröffentlicht: (2025)
Cooperative Perception: A Resource-Efficient Framework for Multi-Drone 3D Scene Reconstruction Using Federated Diffusion and NeRF
von: Pourmandi, Massoud
Veröffentlicht: (2025)
von: Pourmandi, Massoud
Veröffentlicht: (2025)
StemVLA:An Open-Source Vision-Language-Action Model with Future 3D Spatial Geometry Knowledge and 4D Historical Representation
von: Xiao, Jiasong, et al.
Veröffentlicht: (2026)
von: Xiao, Jiasong, et al.
Veröffentlicht: (2026)
Unveiling the Potential of iMarkers: Invisible Fiducial Markers for Advanced Robotics
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
SuperPoint-SLAM3: Augmenting ORB-SLAM3 with Deep Features, Adaptive NMS, and Learning-Based Loop Closure
von: Syed, Shahram Najam, et al.
Veröffentlicht: (2025)
von: Syed, Shahram Najam, et al.
Veröffentlicht: (2025)
Large Language Models and 3D Vision for Intelligent Robotic Perception and Autonomy
von: Mehta, Vinit, et al.
Veröffentlicht: (2025)
von: Mehta, Vinit, et al.
Veröffentlicht: (2025)
vS-Graphs: Tightly Coupling Visual SLAM and 3D Scene Graphs Exploiting Hierarchical Scene Understanding
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
Temporally Consistent Object 6D Pose Estimation for Robot Control
von: Zorina, Kateryna, et al.
Veröffentlicht: (2026)
von: Zorina, Kateryna, et al.
Veröffentlicht: (2026)
An Analysis of Layer-Freezing Strategies for Enhanced Transfer Learning in YOLO Architectures
von: Dobrzycki, Andrzej D., et al.
Veröffentlicht: (2025)
von: Dobrzycki, Andrzej D., et al.
Veröffentlicht: (2025)
SGLoc: Semantic Localization System for Camera Pose Estimation from 3D Gaussian Splatting Representation
von: Xu, Beining, et al.
Veröffentlicht: (2025)
von: Xu, Beining, et al.
Veröffentlicht: (2025)
Adaptive Thresholding for Visual Place Recognition using Negative Gaussian Mixture Statistics
von: Trinh, Nick, et al.
Veröffentlicht: (2025)
von: Trinh, Nick, et al.
Veröffentlicht: (2025)
Rethinking Camera Choice: An Empirical Study on Fisheye Camera Properties in Robotic Manipulation
von: Xue, Han, et al.
Veröffentlicht: (2026)
von: Xue, Han, et al.
Veröffentlicht: (2026)
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
von: Shahin, Nada, et al.
Veröffentlicht: (2025)
von: Shahin, Nada, et al.
Veröffentlicht: (2025)
VLM-VPI: A Vision-Language Reasoning Framework for Improving Automated Vehicle-Pedestrian Interactions
von: Pu, Qingwen, et al.
Veröffentlicht: (2026)
von: Pu, Qingwen, et al.
Veröffentlicht: (2026)
FeedbackSTS-Det: Sparse Frames-Based Spatio-Temporal Semantic Feedback Network for Moving Infrared Small Target Detection
von: Huang, Yian, et al.
Veröffentlicht: (2026)
von: Huang, Yian, et al.
Veröffentlicht: (2026)
Method of UAV Inspection of Photovoltaic Modules Using Thermal and RGB Data Fusion
von: Lysyi, Andrii, et al.
Veröffentlicht: (2025)
von: Lysyi, Andrii, et al.
Veröffentlicht: (2025)
Machine Learning Based Path Planning for Improved Rover Navigation (Pre-Print Version)
von: Abcouwer, Neil, et al.
Veröffentlicht: (2020)
von: Abcouwer, Neil, et al.
Veröffentlicht: (2020)
Is Single-View Mesh Reconstruction Ready for Robotics?
von: Nolte, Frederik, et al.
Veröffentlicht: (2025)
von: Nolte, Frederik, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FCBV-Net: Category-Level Robotic Garment Smoothing via Feature-Conditioned Bimanual Value Prediction
von: Daba, Mohammed, et al.
Veröffentlicht: (2025) -
MARS: Multi-Agent Robotic System with Multimodal Large Language Models for Assistive Intelligence
von: Gao, Renjun
Veröffentlicht: (2025) -
Locate 3D: Real-World Object Localization via Self-Supervised Learning in 3D
von: Arnaud, Sergio, et al.
Veröffentlicht: (2025) -
TinyNav: End-to-End TinyML for Real-Time Autonomous Navigation on Microcontrollers
von: Roy, Pooria, et al.
Veröffentlicht: (2026) -
CC-SGG: Corner Case Scenario Generation using Learned Scene Graphs
von: Drayson, George, et al.
Veröffentlicht: (2023)