BEVTraj: Map-Free End-to-End Trajectory Prediction in Bird's-Eye View with Deformable Attention and Sparse Goal Proposals
Fuente:
arXiv
Saved in:
| Main Authors: | Kong, Minsang, Kim, Myeongjun, Kang, Sang Gu, Lu, Hejiu, Zhong, Yupeng, Lee, Sang Hun |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
by: Raoufi, Behnam, et al.
Published: (2025)
by: Raoufi, Behnam, et al.
Published: (2025)
TinyNav: End-to-End TinyML for Real-Time Autonomous Navigation on Microcontrollers
by: Roy, Pooria, et al.
Published: (2026)
by: Roy, Pooria, et al.
Published: (2026)
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
by: Chahine, Makram, et al.
Published: (2024)
by: Chahine, Makram, et al.
Published: (2024)
Botany Meets Robotics in Alpine Scree Monitoring
by: De Benedittis, Davide, et al.
Published: (2025)
by: De Benedittis, Davide, et al.
Published: (2025)
Have We Mastered Scale in Deep Monocular Visual SLAM? The ScaleMaster Dataset and Benchmark
by: Ju, Hyoseok, et al.
Published: (2026)
by: Ju, Hyoseok, et al.
Published: (2026)
Is Single-View Mesh Reconstruction Ready for Robotics?
by: Nolte, Frederik, et al.
Published: (2025)
by: Nolte, Frederik, et al.
Published: (2025)
FCBV-Net: Category-Level Robotic Garment Smoothing via Feature-Conditioned Bimanual Value Prediction
by: Daba, Mohammed, et al.
Published: (2025)
by: Daba, Mohammed, et al.
Published: (2025)
YOLO-APD: Enhancing YOLOv8 for Robust Pedestrian Detection on Complex Road Geometries
by: Joctum, Aquino, et al.
Published: (2025)
by: Joctum, Aquino, et al.
Published: (2025)
Single-Shot Metric Depth from Focused Plenoptic Cameras
by: Lasheras-Hernandez, Blanca, et al.
Published: (2024)
by: Lasheras-Hernandez, Blanca, et al.
Published: (2024)
Evaluating the Impact of Synthetic Data on Object Detection Tasks in Autonomous Driving
by: Özeren, Enes, et al.
Published: (2025)
by: Özeren, Enes, et al.
Published: (2025)
Dynamic Trajectory Adaptation for Efficient UAV Inspections of Wind Energy Units
by: Svystun, Serhii, et al.
Published: (2024)
by: Svystun, Serhii, et al.
Published: (2024)
Perception-to-Pursuit: Track-Centric Temporal Reasoning for Open-World Drone Detection and Autonomous Chasing
by: Oruganti, Venkatakrishna Reddy
Published: (2026)
by: Oruganti, Venkatakrishna Reddy
Published: (2026)
SGLoc: Semantic Localization System for Camera Pose Estimation from 3D Gaussian Splatting Representation
by: Xu, Beining, et al.
Published: (2025)
by: Xu, Beining, et al.
Published: (2025)
Adaptive Thresholding for Visual Place Recognition using Negative Gaussian Mixture Statistics
by: Trinh, Nick, et al.
Published: (2025)
by: Trinh, Nick, et al.
Published: (2025)
Rethinking Camera Choice: An Empirical Study on Fisheye Camera Properties in Robotic Manipulation
by: Xue, Han, et al.
Published: (2026)
by: Xue, Han, et al.
Published: (2026)
MAP: End-to-End Autonomous Driving with Map-Assisted Planning
by: Yin, Huilin, et al.
Published: (2025)
by: Yin, Huilin, et al.
Published: (2025)
Unveiling the Potential of iMarkers: Invisible Fiducial Markers for Advanced Robotics
by: Tourani, Ali, et al.
Published: (2025)
by: Tourani, Ali, et al.
Published: (2025)
SuperPoint-SLAM3: Augmenting ORB-SLAM3 with Deep Features, Adaptive NMS, and Learning-Based Loop Closure
by: Syed, Shahram Najam, et al.
Published: (2025)
by: Syed, Shahram Najam, et al.
Published: (2025)
Large Language Models and 3D Vision for Intelligent Robotic Perception and Autonomy
by: Mehta, Vinit, et al.
Published: (2025)
by: Mehta, Vinit, et al.
Published: (2025)
Autonomous Underwater Cognitive System for Adaptive Navigation: A SLAM-Integrated Cognitive Architecture
by: Jayarathne, K. A. I. N, et al.
Published: (2025)
by: Jayarathne, K. A. I. N, et al.
Published: (2025)
vS-Graphs: Tightly Coupling Visual SLAM and 3D Scene Graphs Exploiting Hierarchical Scene Understanding
by: Tourani, Ali, et al.
Published: (2025)
by: Tourani, Ali, et al.
Published: (2025)
Temporally Consistent Object 6D Pose Estimation for Robot Control
by: Zorina, Kateryna, et al.
Published: (2026)
by: Zorina, Kateryna, et al.
Published: (2026)
High-Frequency Semantics and Geometric Priors for End-to-End Detection Transformers in Challenging UAV Imagery
by: Peng, Hongxing, et al.
Published: (2025)
by: Peng, Hongxing, et al.
Published: (2025)
Systematic Comparison of Projection Methods for Monocular 3D Human Pose Estimation on Fisheye Images
by: Käs, Stephanie, et al.
Published: (2025)
by: Käs, Stephanie, et al.
Published: (2025)
MARS: Multi-Agent Robotic System with Multimodal Large Language Models for Assistive Intelligence
by: Gao, Renjun
Published: (2025)
by: Gao, Renjun
Published: (2025)
Distributed Intelligent System Architecture for UAV-Assisted Monitoring of Wind Energy Infrastructure
by: Svystun, Serhii, et al.
Published: (2024)
by: Svystun, Serhii, et al.
Published: (2024)
CARScenes: Semantic VLM Dataset for Safe Autonomous Driving
by: He, Yuankai, et al.
Published: (2025)
by: He, Yuankai, et al.
Published: (2025)
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models
by: Chen, Yiteng, et al.
Published: (2025)
by: Chen, Yiteng, et al.
Published: (2025)
AntiGrounding: Lifting Robotic Actions into VLM Representation Space for Decision Making
by: Li, Wenbo, et al.
Published: (2025)
by: Li, Wenbo, et al.
Published: (2025)
BEVLoc: Cross-View Localization and Matching via Birds-Eye-View Synthesis
by: Klammer, Christopher, et al.
Published: (2024)
by: Klammer, Christopher, et al.
Published: (2024)
FACT: Multinomial Misalignment Classification for Point Cloud Registration
by: Dillén, Ludvig, et al.
Published: (2025)
by: Dillén, Ludvig, et al.
Published: (2025)
StemVLA:An Open-Source Vision-Language-Action Model with Future 3D Spatial Geometry Knowledge and 4D Historical Representation
by: Xiao, Jiasong, et al.
Published: (2026)
by: Xiao, Jiasong, et al.
Published: (2026)
FlowDet: Overcoming Perspective and Scale Challenges in Real-Time End-to-End Traffic Detection
by: Wang, Zixing, et al.
Published: (2025)
by: Wang, Zixing, et al.
Published: (2025)
Learning from Watching: Scalable Extraction of Manipulation Trajectories from Human Videos
by: Hu, X., et al.
Published: (2025)
by: Hu, X., et al.
Published: (2025)
Can VLMs Unlock Semantic Anomaly Detection? A Framework for Structured Reasoning
by: Brusnicki, Roberto, et al.
Published: (2025)
by: Brusnicki, Roberto, et al.
Published: (2025)
Shaken, Not Stirred: A Novel Dataset for Visual Understanding of Glasses in Human-Robot Bartending Tasks
by: Gajdošech, Lukáš, et al.
Published: (2025)
by: Gajdošech, Lukáš, et al.
Published: (2025)
End-to-End Imitation Learning for Optimal Asteroid Proximity Operations
by: Quinn, Patrick, et al.
Published: (2025)
by: Quinn, Patrick, et al.
Published: (2025)
SmartDate: AI-Driven Precision Sorting and Quality Control in Date Fruits
by: Eskaf, Khaled
Published: (2025)
by: Eskaf, Khaled
Published: (2025)
Thermal and RGB Images Work Better Together in Wind Turbine Damage Detection
by: Svystun, Serhii, et al.
Published: (2024)
by: Svystun, Serhii, et al.
Published: (2024)
End-to-end example-based sim-to-real RL policy transfer based on neural stylisation with application to robotic cutting
by: Hathaway, Jamie, et al.
Published: (2026)
by: Hathaway, Jamie, et al.
Published: (2026)
Similar Items
-
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
by: Raoufi, Behnam, et al.
Published: (2025) -
TinyNav: End-to-End TinyML for Real-Time Autonomous Navigation on Microcontrollers
by: Roy, Pooria, et al.
Published: (2026) -
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
by: Chahine, Makram, et al.
Published: (2024) -
Botany Meets Robotics in Alpine Scree Monitoring
by: De Benedittis, Davide, et al.
Published: (2025) -
Have We Mastered Scale in Deep Monocular Visual SLAM? The ScaleMaster Dataset and Benchmark
by: Ju, Hyoseok, et al.
Published: (2026)