Contrastive Learning for Enhancing Robust Scene Transfer in Vision-based Agile Flight
Fuente:
arXiv
Saved in:
| Main Authors: | Xing, Jiaxu, Bauersfeld, Leonard, Song, Yunlong, Xing, Chunwei, Scaramuzza, Davide |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bootstrapping Reinforcement Learning with Imitation for Vision-Based Agile Flight
by: Xing, Jiaxu, et al.
Published: (2024)
by: Xing, Jiaxu, et al.
Published: (2024)
Low-Latency Event-Based Velocimetry for Quadrotor Control in a Narrow Pipe
by: Bauersfeld, Leonard, et al.
Published: (2025)
by: Bauersfeld, Leonard, et al.
Published: (2025)
A Monocular Event-Camera Motion Capture System
by: Bauersfeld, Leonard, et al.
Published: (2025)
by: Bauersfeld, Leonard, et al.
Published: (2025)
Demonstrating Agile Flight from Pixels without State Estimation
by: Geles, Ismail, et al.
Published: (2024)
by: Geles, Ismail, et al.
Published: (2024)
ForesightNav: Learning Scene Imagination for Efficient Exploration
by: Shah, Hardik, et al.
Published: (2025)
by: Shah, Hardik, et al.
Published: (2025)
Generative Event Pretraining with Foundation Model Alignment
by: Cao, Jianwen, et al.
Published: (2026)
by: Cao, Jianwen, et al.
Published: (2026)
Sight Over Site: Perception-Aware Reinforcement Learning for Efficient Robotic Inspection
by: Kuhlmann, Richard, et al.
Published: (2025)
by: Kuhlmann, Richard, et al.
Published: (2025)
Approximate Imitation Learning for Event-based Quadrotor Flight in Cluttered Environments
by: Messikommer, Nico, et al.
Published: (2026)
by: Messikommer, Nico, et al.
Published: (2026)
Learning Agile Quadrotor Flight in the Real World
by: Ren, Yunfan, et al.
Published: (2026)
by: Ren, Yunfan, et al.
Published: (2026)
Image-Conditioned Adaptive Parameter Tuning for Visual Odometry Frontends
by: Nascivera, Simone, et al.
Published: (2026)
by: Nascivera, Simone, et al.
Published: (2026)
Vision-Based Agile Landing on Turbulent Waters
by: Angelis, Dimosthenis, et al.
Published: (2026)
by: Angelis, Dimosthenis, et al.
Published: (2026)
Reading in the Dark with Foveated Event Vision
by: Brander, Carl, et al.
Published: (2025)
by: Brander, Carl, et al.
Published: (2025)
Reinforcement Learning Meets Visual Odometry
by: Messikommer, Nico, et al.
Published: (2024)
by: Messikommer, Nico, et al.
Published: (2024)
Event-Aided Sharp Radiance Field Reconstruction for Fast-Flying Drones
by: Zou, Rong, et al.
Published: (2026)
by: Zou, Rong, et al.
Published: (2026)
Event Spectroscopy: Event-based Multispectral and Depth Sensing using Structured Light
by: Geckeler, Christian, et al.
Published: (2025)
by: Geckeler, Christian, et al.
Published: (2025)
Range, Endurance, and Optimal Speed Estimates for Multicopters
by: Bauersfeld, Leonard, et al.
Published: (2021)
by: Bauersfeld, Leonard, et al.
Published: (2021)
FaVoR: Features via Voxel Rendering for Camera Relocalization
by: Polizzi, Vincenzo, et al.
Published: (2024)
by: Polizzi, Vincenzo, et al.
Published: (2024)
Structure-Invariant Range-Visual-Inertial Odometry
by: Alberico, Ivan, et al.
Published: (2024)
by: Alberico, Ivan, et al.
Published: (2024)
Agile Robotics: Optimal Control, Reinforcement Learning, and Differentiable Simulation
by: Song, Yunlong, et al.
Published: (2024)
by: Song, Yunlong, et al.
Published: (2024)
All Eyes, no IMU: Learning Flight Attitude from Vision Alone
by: Hagenaars, Jesse J., et al.
Published: (2025)
by: Hagenaars, Jesse J., et al.
Published: (2025)
Motion-aware Event Suppression for Event Cameras
by: Pellerito, Roberto, et al.
Published: (2026)
by: Pellerito, Roberto, et al.
Published: (2026)
OpenESS: Event-based Semantic Scene Understanding with Open Vocabularies
by: Kong, Lingdong, et al.
Published: (2024)
by: Kong, Lingdong, et al.
Published: (2024)
Superhuman Safe and Agile Racing through Multi-Agent Reinforcement Learning
by: Geles, Ismail, et al.
Published: (2026)
by: Geles, Ismail, et al.
Published: (2026)
Weather-Robust Scene Semantics with Vision-Aligned 4D Radar
by: Hamilton, Kali, et al.
Published: (2026)
by: Hamilton, Kali, et al.
Published: (2026)
Estimating Scene Flow in Robot Surroundings with Distributed Miniaturized Time-of-Flight Sensors
by: Sander, Jack, et al.
Published: (2025)
by: Sander, Jack, et al.
Published: (2025)
RAPID: Robust and Agile Planner Using Inverse Reinforcement Learning for Vision-Based Drone Navigation
by: Kim, Minwoo, et al.
Published: (2025)
by: Kim, Minwoo, et al.
Published: (2025)
Recursive Visual Imagination and Adaptive Linguistic Grounding for Vision Language Navigation
by: Chen, Bolei, et al.
Published: (2025)
by: Chen, Bolei, et al.
Published: (2025)
CapVector: Learning Transferable Capability Vectors in Parametric Space for Vision-Language-Action Models
by: Song, Wenxuan, et al.
Published: (2026)
by: Song, Wenxuan, et al.
Published: (2026)
SeaSplat: Representing Underwater Scenes with 3D Gaussian Splatting and a Physically Grounded Image Formation Model
by: Yang, Daniel, et al.
Published: (2024)
by: Yang, Daniel, et al.
Published: (2024)
Efficient Multi-Task Scene Analysis with RGB-D Transformers
by: Fischedick, Söhnke Benedikt, et al.
Published: (2023)
by: Fischedick, Söhnke Benedikt, et al.
Published: (2023)
Chasing Day and Night: Towards Robust and Efficient All-Day Object Detection Guided by an Event Camera
by: Cao, Jiahang, et al.
Published: (2023)
by: Cao, Jiahang, et al.
Published: (2023)
DivScene: Towards Open-Vocabulary Object Navigation with Large Vision Language Models in Diverse Scenes
by: Wang, Zhaowei, et al.
Published: (2024)
by: Wang, Zhaowei, et al.
Published: (2024)
YoloTag: Vision-based Robust UAV Navigation with Fiducial Markers
by: Raxit, Sourav, et al.
Published: (2024)
by: Raxit, Sourav, et al.
Published: (2024)
Exosense: A Vision-Based Scene Understanding System For Exoskeletons
by: Wang, Jianeng, et al.
Published: (2024)
by: Wang, Jianeng, et al.
Published: (2024)
MISCGrasp: Leveraging Multiple Integrated Scales and Contrastive Learning for Enhanced Volumetric Grasping
by: Fan, Qingyu, et al.
Published: (2025)
by: Fan, Qingyu, et al.
Published: (2025)
Multi-Modal Graph Convolutional Network with Sinusoidal Encoding for Robust Human Action Segmentation
by: Xing, Hao, et al.
Published: (2025)
by: Xing, Hao, et al.
Published: (2025)
CLAP: Contrastive Latent Action Pretraining for Learning Vision-Language-Action Models from Human Videos
by: Zhang, Chubin, et al.
Published: (2026)
by: Zhang, Chubin, et al.
Published: (2026)
Talk2Event: Grounded Understanding of Dynamic Scenes from Event Cameras
by: Kong, Lingdong, et al.
Published: (2025)
by: Kong, Lingdong, et al.
Published: (2025)
LaMP: Learning Vision-Language-Action Policies with 3D Scene Flow as Latent Motion Prior
by: Wang, Xinkai, et al.
Published: (2026)
by: Wang, Xinkai, et al.
Published: (2026)
Pandora: Articulated 3D Scene Graphs from Egocentric Vision
by: Yu, Alan, et al.
Published: (2026)
by: Yu, Alan, et al.
Published: (2026)
Similar Items
-
Bootstrapping Reinforcement Learning with Imitation for Vision-Based Agile Flight
by: Xing, Jiaxu, et al.
Published: (2024) -
Low-Latency Event-Based Velocimetry for Quadrotor Control in a Narrow Pipe
by: Bauersfeld, Leonard, et al.
Published: (2025) -
A Monocular Event-Camera Motion Capture System
by: Bauersfeld, Leonard, et al.
Published: (2025) -
Demonstrating Agile Flight from Pixels without State Estimation
by: Geles, Ismail, et al.
Published: (2024) -
ForesightNav: Learning Scene Imagination for Efficient Exploration
by: Shah, Hardik, et al.
Published: (2025)