Saved in:
| Main Authors: | Cao, Deng, Zhang, Hongbo, Dhillon, Rajveer |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2405.21056 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AerialVLA: A Vision-Language-Action Model for UAV Navigation via Minimalist End-to-End Control
by: Xu, Peng, et al.
Published: (2026)
by: Xu, Peng, et al.
Published: (2026)
Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery
by: Tran, Chi-Nguyen, et al.
Published: (2026)
by: Tran, Chi-Nguyen, et al.
Published: (2026)
DiffSSC: Semantic LiDAR Scan Completion using Denoising Diffusion Probabilistic Models
by: Cao, Helin, et al.
Published: (2024)
by: Cao, Helin, et al.
Published: (2024)
DeepIPC: Deeply Integrated Perception and Control for an Autonomous Vehicle in Real Environments
by: Natan, Oskar, et al.
Published: (2022)
by: Natan, Oskar, et al.
Published: (2022)
SLCF-Net: Sequential LiDAR-Camera Fusion for Semantic Scene Completion using a 3D Recurrent U-Net
by: Cao, Helin, et al.
Published: (2024)
by: Cao, Helin, et al.
Published: (2024)
DeepKalPose: An Enhanced Deep-Learning Kalman Filter for Temporally Consistent Monocular Vehicle Pose Estimation
by: Di Bella, Leandro, et al.
Published: (2024)
by: Di Bella, Leandro, et al.
Published: (2024)
DeepIPCv2: LiDAR-powered Robust Environmental Perception and Navigational Control for Autonomous Vehicle
by: Natan, Oskar, et al.
Published: (2023)
by: Natan, Oskar, et al.
Published: (2023)
Continuous Vision-Language-Action Co-Learning with Semantic-Physical Alignment for Behavioral Cloning
by: Qi, Xiuxiu, et al.
Published: (2025)
by: Qi, Xiuxiu, et al.
Published: (2025)
Validation & Exploration of Multimodal Deep-Learning Camera-Lidar Calibration models
by: Karramreddy, Venkat, et al.
Published: (2024)
by: Karramreddy, Venkat, et al.
Published: (2024)
Autonomous Embodied Agents: When Robotics Meets Deep Learning Reasoning
by: Bigazzi, Roberto
Published: (2025)
by: Bigazzi, Roberto
Published: (2025)
Segment, Lift and Fit: Automatic 3D Shape Labeling from 2D Prompts
by: Li, Jianhao, et al.
Published: (2024)
by: Li, Jianhao, et al.
Published: (2024)
Improved LiDAR Odometry and Mapping using Deep Semantic Segmentation and Novel Outliers Detection
by: Afifi, Mohamed, et al.
Published: (2024)
by: Afifi, Mohamed, et al.
Published: (2024)
Measuring the Effect of Background on Classification and Feature Importance in Deep Learning for AV Perception
by: Sielemann, Anne, et al.
Published: (2025)
by: Sielemann, Anne, et al.
Published: (2025)
Environment-Driven Online LiDAR-Camera Extrinsic Calibration
by: Huang, Zhiwei, et al.
Published: (2025)
by: Huang, Zhiwei, et al.
Published: (2025)
The Safety Challenge of World Models for Embodied AI Agents: A Review
by: Baraldi, Lorenzo, et al.
Published: (2025)
by: Baraldi, Lorenzo, et al.
Published: (2025)
A Systematic Literature Review on Deep Learning-based Depth Estimation in Computer Vision
by: Rohan, Ali, et al.
Published: (2025)
by: Rohan, Ali, et al.
Published: (2025)
On Deep Learning for Geometric and Semantic Scene Understanding Using On-Vehicle 3D LiDAR
by: Li, Li
Published: (2024)
by: Li, Li
Published: (2024)
Quadrotor Navigation using Reinforcement Learning with Privileged Information
by: Lee, Jonathan, et al.
Published: (2025)
by: Lee, Jonathan, et al.
Published: (2025)
ContactGaussian-WM: Learning Physics-Grounded World Model from Videos
by: Wang, Meizhong, et al.
Published: (2026)
by: Wang, Meizhong, et al.
Published: (2026)
Securing the Skies: A Comprehensive Survey on Anti-UAV Methods, Benchmarking, and Future Directions
by: Dong, Yifei, et al.
Published: (2025)
by: Dong, Yifei, et al.
Published: (2025)
Real-Time Roadway Obstacle Detection for Electric Scooters Using Deep Learning and Multi-Sensor Fusion
by: Zheng, Zeyang, et al.
Published: (2025)
by: Zheng, Zeyang, et al.
Published: (2025)
Learning 3D Robotics Perception using Inductive Priors
by: Irshad, Muhammad Zubair
Published: (2024)
by: Irshad, Muhammad Zubair
Published: (2024)
A Deep Learning-based Pest Insect Monitoring System for Ultra-low Power Pocket-sized Drones
by: Crupi, Luca, et al.
Published: (2024)
by: Crupi, Luca, et al.
Published: (2024)
Keypoint Abstraction using Large Models for Object-Relative Imitation Learning
by: Fang, Xiaolin, et al.
Published: (2024)
by: Fang, Xiaolin, et al.
Published: (2024)
Seeing Through Uncertainty: A Free-Energy Approach for Real-Time Perceptual Adaptation in Robust Visual Navigation
by: Piriyajitakonkij, Maytus, et al.
Published: (2024)
by: Piriyajitakonkij, Maytus, et al.
Published: (2024)
OC-SOP: Enhancing Vision-Based 3D Semantic Occupancy Prediction by Object-Centric Awareness
by: Cao, Helin, et al.
Published: (2025)
by: Cao, Helin, et al.
Published: (2025)
HiRT: Enhancing Robotic Control with Hierarchical Robot Transformers
by: Zhang, Jianke, et al.
Published: (2024)
by: Zhang, Jianke, et al.
Published: (2024)
Is VLA Reasoning Faithful? Probing Safety of Chain-of-Causation in Autonomous Driving Models
by: Mayumu, Nicanor, et al.
Published: (2026)
by: Mayumu, Nicanor, et al.
Published: (2026)
Learning from Massive Human Videos for Universal Humanoid Pose Control
by: Mao, Jiageng, et al.
Published: (2024)
by: Mao, Jiageng, et al.
Published: (2024)
M4Diffuser: Multi-View Diffusion Policy with Manipulability-Aware Control for Robust Mobile Manipulation
by: Dong, Ju, et al.
Published: (2025)
by: Dong, Ju, et al.
Published: (2025)
MimicFunc: Imitating Tool Manipulation from a Single Human Video via Functional Correspondence
by: Tang, Chao, et al.
Published: (2025)
by: Tang, Chao, et al.
Published: (2025)
IMPACT-HOI: Supervisory Control for Onset-Anchored Partial HOI Event Construction
by: Zhang, Haoshen, et al.
Published: (2026)
by: Zhang, Haoshen, et al.
Published: (2026)
Privacy-Preserving Multi-Stage Fall Detection Framework with Semi-supervised Federated Learning and Robotic Vision Confirmation
by: Azghadi, Seyed Alireza Rahimi, et al.
Published: (2025)
by: Azghadi, Seyed Alireza Rahimi, et al.
Published: (2025)
DARTS: A Drone-Based AI-Powered Real-Time Traffic Incident Detection System
by: Li, Bai, et al.
Published: (2025)
by: Li, Bai, et al.
Published: (2025)
MCRL4OR: Multimodal Contrastive Representation Learning for Off-Road Environmental Perception
by: Yang, Yi, et al.
Published: (2025)
by: Yang, Yi, et al.
Published: (2025)
Learning to Manipulate Anywhere: A Visual Generalizable Framework For Reinforcement Learning
by: Yuan, Zhecheng, et al.
Published: (2024)
by: Yuan, Zhecheng, et al.
Published: (2024)
LLM-RG: Referential Grounding in Outdoor Scenarios using Large Language Models
by: Saxena, Pranav, et al.
Published: (2025)
by: Saxena, Pranav, et al.
Published: (2025)
Multiple Rotation Averaging with Constrained Reweighting Deep Matrix Factorization
by: Li, Shiqi, et al.
Published: (2024)
by: Li, Shiqi, et al.
Published: (2024)
Learning Robust Stereo Matching in the Wild with Selective Mixture-of-Experts
by: Wang, Yun, et al.
Published: (2025)
by: Wang, Yun, et al.
Published: (2025)
SWA-SOP: Spatially-aware Window Attention for Semantic Occupancy Prediction in Autonomous Driving
by: Cao, Helin, et al.
Published: (2025)
by: Cao, Helin, et al.
Published: (2025)
Similar Items
-
AerialVLA: A Vision-Language-Action Model for UAV Navigation via Minimalist End-to-End Control
by: Xu, Peng, et al.
Published: (2026) -
Weather-Robust Cross-View Geo-Localization via Prototype-Based Semantic Part Discovery
by: Tran, Chi-Nguyen, et al.
Published: (2026) -
DiffSSC: Semantic LiDAR Scan Completion using Denoising Diffusion Probabilistic Models
by: Cao, Helin, et al.
Published: (2024) -
DeepIPC: Deeply Integrated Perception and Control for an Autonomous Vehicle in Real Environments
by: Natan, Oskar, et al.
Published: (2022) -
SLCF-Net: Sequential LiDAR-Camera Fusion for Semantic Scene Completion using a 3D Recurrent U-Net
by: Cao, Helin, et al.
Published: (2024)