Vision Transformers for End-to-End Vision-Based Quadrotor Obstacle Avoidance
Fuente:
arXiv
Saved in:
| Main Authors: | Bhattacharya, Anish, Rao, Nishanth, Parikh, Dhruv, Kunapuli, Pratik, Wu, Yuwei, Tao, Yuezhan, Matni, Nikolai, Kumar, Vijay |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Monocular Event-Based Vision for Obstacle Avoidance with a Quadrotor
by: Bhattacharya, Anish, et al.
Published: (2024)
by: Bhattacharya, Anish, et al.
Published: (2024)
Vision-Proprioception Fusion with Mamba2 in End-to-End Reinforcement Learning for Motion Control
by: Tao, Xiaowen, et al.
Published: (2025)
by: Tao, Xiaowen, et al.
Published: (2025)
LocoMamba: Vision-Driven Locomotion via End-to-End Deep Reinforcement Learning with Mamba
by: Wang, Yinuo, et al.
Published: (2025)
by: Wang, Yinuo, et al.
Published: (2025)
Force-EvT: A Closer Look at Robotic Gripper Force Measurement with Event-based Vision Transformer
by: Guo, Qianyu, et al.
Published: (2024)
by: Guo, Qianyu, et al.
Published: (2024)
The Era of End-to-End Autonomy: Transitioning from Rule-Based Driving to Large Driving Models
by: Nebot, Eduardo, et al.
Published: (2026)
by: Nebot, Eduardo, et al.
Published: (2026)
Safe Interval Motion Planning for Quadrotors in Dynamic Environments
by: Huang, Songhao, et al.
Published: (2024)
by: Huang, Songhao, et al.
Published: (2024)
MonoSIM: An open source SIL framework for Ackermann Vehicular Systems with Monocular Vision
by: Rahman, Shantanu, et al.
Published: (2026)
by: Rahman, Shantanu, et al.
Published: (2026)
QuadKAN: KAN-Enhanced Quadruped Motion Control via End-to-End Reinforcement Learning
by: Wang, Yinuo, et al.
Published: (2025)
by: Wang, Yinuo, et al.
Published: (2025)
Vision Controlled Orthotic Hand Exoskeleton
by: Blais, Connor, et al.
Published: (2025)
by: Blais, Connor, et al.
Published: (2025)
Leveraging Symmetry to Accelerate Learning of Trajectory Tracking Controllers for Free-Flying Robotic Systems
by: Welde, Jake, et al.
Published: (2024)
by: Welde, Jake, et al.
Published: (2024)
Leveling the Playing Field: Carefully Comparing Classical and Learned Controllers for Quadrotor Trajectory Tracking
by: Kunapuli, Pratik, et al.
Published: (2025)
by: Kunapuli, Pratik, et al.
Published: (2025)
Traffic Regulation-aware Path Planning with Regulation Databases and Vision-Language Models
by: Han, Xu, et al.
Published: (2025)
by: Han, Xu, et al.
Published: (2025)
Behind Every Domain There is a Shift: Adapting Distortion-aware Vision Transformers for Panoramic Semantic Segmentation
by: Zhang, Jiaming, et al.
Published: (2022)
by: Zhang, Jiaming, et al.
Published: (2022)
See Silhouettes in Motion with Neuromorphic Vision
by: Zhang, Pei, et al.
Published: (2026)
by: Zhang, Pei, et al.
Published: (2026)
End-to-End Action Segmentation Transformer
by: Wang, Tieqiao, et al.
Published: (2025)
by: Wang, Tieqiao, et al.
Published: (2025)
SELECT: A Submodular Approach for Active LiDAR Semantic Segmentation
by: Mao, Ruiyu, et al.
Published: (2025)
by: Mao, Ruiyu, et al.
Published: (2025)
Vision-based automatic fruit counting with UAV
by: Szolc, Hubert, et al.
Published: (2025)
by: Szolc, Hubert, et al.
Published: (2025)
Deep Learning Aided Vision System for Planetary Rovers
by: Relia, Lomash, et al.
Published: (2026)
by: Relia, Lomash, et al.
Published: (2026)
Universal End-to-End Neural Network for Lossy Image Compression
by: Arezki, Bouzid, et al.
Published: (2024)
by: Arezki, Bouzid, et al.
Published: (2024)
VoxelPrompt: A Vision Agent for End-to-End Medical Image Analysis
by: Hoopes, Andrew, et al.
Published: (2024)
by: Hoopes, Andrew, et al.
Published: (2024)
MoRE: A Mixture-of-Experts-Based Task-Adaptive End-to-End Network for Multimodal MRI Reconstruction
by: Li, Yuyang, et al.
Published: (2026)
by: Li, Yuyang, et al.
Published: (2026)
TacShade A New 3D-printed Soft Optical Tactile Sensor Based on Light, Shadow and Greyscale for Shape Reconstruction
by: Lu, Zhenyu, et al.
Published: (2024)
by: Lu, Zhenyu, et al.
Published: (2024)
DHFP-PE: Dual-Precision Hybrid Floating Point Processing Element for AI Acceleration
by: Kumar, Shubham, et al.
Published: (2026)
by: Kumar, Shubham, et al.
Published: (2026)
Scene Representation using 360° Saliency Graph and its Application in Vision-based Indoor Navigation
by: Meena, Preeti, et al.
Published: (2026)
by: Meena, Preeti, et al.
Published: (2026)
End-to-End Differentiable Photon Counting CT
by: Wang, Sen, et al.
Published: (2026)
by: Wang, Sen, et al.
Published: (2026)
SDCM: Simulated Densifying and Compensatory Modeling Fusion for Radar-Vision 3-D Object Detection in Internet of Vehicles
by: Li, Shucong, et al.
Published: (2026)
by: Li, Shucong, et al.
Published: (2026)
Sensorless Remote Center of Motion Misalignment Estimation
by: Yang, Hao, et al.
Published: (2025)
by: Yang, Hao, et al.
Published: (2025)
Field Calibration of Hyperspectral Cameras for Terrain Inference
by: Hanson, Nathaniel, et al.
Published: (2025)
by: Hanson, Nathaniel, et al.
Published: (2025)
VISTA: A Benchmark for Real-Time Video Streaming under Network Impairments in Surgical Teleoperation
by: Deng, Zexin, et al.
Published: (2026)
by: Deng, Zexin, et al.
Published: (2026)
Invascal: Inverse-Vacuity Self-Calibration for Uncertainty-Aware LiDAR Range-View Semantic Segmentation
by: Turacan, Kerim, et al.
Published: (2026)
by: Turacan, Kerim, et al.
Published: (2026)
DriveNetBench: An Affordable and Configurable Single-Camera Benchmarking System for Autonomous Driving Networks
by: Al-Bustami, Ali, et al.
Published: (2025)
by: Al-Bustami, Ali, et al.
Published: (2025)
Noise Analysis and Modeling of the PMD Flexx2 Depth Camera for Robotic Applications
by: Cai, Yuke, et al.
Published: (2024)
by: Cai, Yuke, et al.
Published: (2024)
Web-based Augmented Reality with Auto-Scaling and Real-Time Head Tracking towards Markerless Neurointerventional Preoperative Planning and Training of Head-mounted Robotic Needle Insertion
by: Ho, Hon Lung, et al.
Published: (2024)
by: Ho, Hon Lung, et al.
Published: (2024)
SceneVGGT: VGGT-based online 3D semantic SLAM for indoor scene understanding and navigation
by: Gelencsér-Horváth, Anna, et al.
Published: (2026)
by: Gelencsér-Horváth, Anna, et al.
Published: (2026)
Robotic CBCT Meets Robotic Ultrasound
by: Li, Feng, et al.
Published: (2025)
by: Li, Feng, et al.
Published: (2025)
Active Optics for Hyperspectral Imaging of Reflective Agricultural Leaf Sensors
by: Burns, Dexter, et al.
Published: (2025)
by: Burns, Dexter, et al.
Published: (2025)
Hilti SLAM Challenge 2023: Benchmarking Single + Multi-session SLAM across Sensor Constellations in Construction
by: Nair, Ashish Devadas, et al.
Published: (2024)
by: Nair, Ashish Devadas, et al.
Published: (2024)
Training-Free Robot Pose Estimation using Off-the-Shelf Foundational Models
by: Liang, Laurence
Published: (2025)
by: Liang, Laurence
Published: (2025)
Two-Stage Camera Calibration Method for Multi-Camera Systems Using Scene Geometry
by: Abramov, Aleksandr
Published: (2025)
by: Abramov, Aleksandr
Published: (2025)
Iterative PnP and its application in 3D-2D vascular image registration for robot navigation
by: Song, Jingwei, et al.
Published: (2023)
by: Song, Jingwei, et al.
Published: (2023)
Similar Items
-
Monocular Event-Based Vision for Obstacle Avoidance with a Quadrotor
by: Bhattacharya, Anish, et al.
Published: (2024) -
Vision-Proprioception Fusion with Mamba2 in End-to-End Reinforcement Learning for Motion Control
by: Tao, Xiaowen, et al.
Published: (2025) -
LocoMamba: Vision-Driven Locomotion via End-to-End Deep Reinforcement Learning with Mamba
by: Wang, Yinuo, et al.
Published: (2025) -
Force-EvT: A Closer Look at Robotic Gripper Force Measurement with Event-based Vision Transformer
by: Guo, Qianyu, et al.
Published: (2024) -
The Era of End-to-End Autonomy: Transitioning from Rule-Based Driving to Large Driving Models
by: Nebot, Eduardo, et al.
Published: (2026)