SwapTransformer: highway overtaking tactical planner model via imitation learning on OSHA dataset
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shamsoshoara, Alireza, Salih, Safin B, Aghazadeh, Pedram |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Goal-conditioned dual-action imitation learning for dexterous dual-arm robot manipulation
von: Kim, Heecheol, et al.
Veröffentlicht: (2022)
von: Kim, Heecheol, et al.
Veröffentlicht: (2022)
EmbodiSwap for Zero-Shot Robot Imitation Learning
von: Dessalene, Eadom, et al.
Veröffentlicht: (2025)
von: Dessalene, Eadom, et al.
Veröffentlicht: (2025)
TransForSeg: A Multitask Stereo ViT for Joint Stereo Segmentation and 3D Force Estimation in Catheterization
von: Fekri, Pedram, et al.
Veröffentlicht: (2025)
von: Fekri, Pedram, et al.
Veröffentlicht: (2025)
Accelerating Transformer-Based Monocular SLAM via Geometric Utility Scoring
von: Xiong, Xinmiao, et al.
Veröffentlicht: (2026)
von: Xiong, Xinmiao, et al.
Veröffentlicht: (2026)
Point-LN: A Lightweight Framework for Efficient Point Cloud Classification Using Non-Parametric Positional Encoding
von: Mohammadi, Marzieh, et al.
Veröffentlicht: (2025)
von: Mohammadi, Marzieh, et al.
Veröffentlicht: (2025)
Enhancing 3D Point Cloud Classification with ModelNet-R and Point-SkipNet
von: Saeid, Mohammad, et al.
Veröffentlicht: (2025)
von: Saeid, Mohammad, et al.
Veröffentlicht: (2025)
DINO-CVA: A Multimodal Goal-Conditioned Vision-to-Action Model for Autonomous Catheter Navigation
von: Fekri, Pedram, et al.
Veröffentlicht: (2025)
von: Fekri, Pedram, et al.
Veröffentlicht: (2025)
AYDIV: Adaptable Yielding 3D Object Detection via Integrated Contextual Vision Transformer
von: Dam, Tanmoy, et al.
Veröffentlicht: (2024)
von: Dam, Tanmoy, et al.
Veröffentlicht: (2024)
Aug3D: Augmenting large scale outdoor datasets for Generalizable Novel View Synthesis
von: Rauniyar, Aditya, et al.
Veröffentlicht: (2025)
von: Rauniyar, Aditya, et al.
Veröffentlicht: (2025)
Look, Focus, Act: Efficient and Robust Robot Learning via Human Gaze and Foveated Vision Transformers
von: Chuang, Ian, et al.
Veröffentlicht: (2025)
von: Chuang, Ian, et al.
Veröffentlicht: (2025)
MTA-RL: Robust Urban Driving via Multi-modal Transformer-based 3D Affordances and Reinforcement Learning
von: Chen, Guangli, et al.
Veröffentlicht: (2026)
von: Chen, Guangli, et al.
Veröffentlicht: (2026)
DVGT: Driving Visual Geometry Transformer
von: Zuo, Sicheng, et al.
Veröffentlicht: (2025)
von: Zuo, Sicheng, et al.
Veröffentlicht: (2025)
On-Device Diffusion Transformer Policy for Efficient Robot Manipulation
von: Wu, Yiming, et al.
Veröffentlicht: (2025)
von: Wu, Yiming, et al.
Veröffentlicht: (2025)
VIT-Ped: Visionary Intention Transformer for Pedestrian Behavior Analysis
von: Elkammar, Aly R., et al.
Veröffentlicht: (2026)
von: Elkammar, Aly R., et al.
Veröffentlicht: (2026)
HiRT: Enhancing Robotic Control with Hierarchical Robot Transformers
von: Zhang, Jianke, et al.
Veröffentlicht: (2024)
von: Zhang, Jianke, et al.
Veröffentlicht: (2024)
Weakly Supervised Point Clouds Transformer for 3D Object Detection
von: Tang, Zuojin, et al.
Veröffentlicht: (2023)
von: Tang, Zuojin, et al.
Veröffentlicht: (2023)
LIAM: Multimodal Transformer for Language Instructions, Images, Actions and Semantic Maps
von: Wang, Yihao, et al.
Veröffentlicht: (2025)
von: Wang, Yihao, et al.
Veröffentlicht: (2025)
Hierarchical place recognition with omnidirectional images and curriculum learning-based loss functions
von: Alfaro, Marcos, et al.
Veröffentlicht: (2024)
von: Alfaro, Marcos, et al.
Veröffentlicht: (2024)
Privacy-Preserving Multi-Stage Fall Detection Framework with Semi-supervised Federated Learning and Robotic Vision Confirmation
von: Azghadi, Seyed Alireza Rahimi, et al.
Veröffentlicht: (2025)
von: Azghadi, Seyed Alireza Rahimi, et al.
Veröffentlicht: (2025)
Understanding Video Transformers via Universal Concept Discovery
von: Kowal, Matthew, et al.
Veröffentlicht: (2024)
von: Kowal, Matthew, et al.
Veröffentlicht: (2024)
Region-Transformer: Self-Attention Region Based Class-Agnostic Point Cloud Segmentation
von: Gyawali, Dipesh, et al.
Veröffentlicht: (2024)
von: Gyawali, Dipesh, et al.
Veröffentlicht: (2024)
InterACT: Inter-dependency Aware Action Chunking with Hierarchical Attention Transformers for Bimanual Manipulation
von: Lee, Andrew, et al.
Veröffentlicht: (2024)
von: Lee, Andrew, et al.
Veröffentlicht: (2024)
CAPT: Category-level Articulation Estimation from a Single Point Cloud Using Transformer
von: Fu, Lian, et al.
Veröffentlicht: (2024)
von: Fu, Lian, et al.
Veröffentlicht: (2024)
M2DA: Multi-Modal Fusion Transformer Incorporating Driver Attention for Autonomous Driving
von: Xu, Dongyang, et al.
Veröffentlicht: (2024)
von: Xu, Dongyang, et al.
Veröffentlicht: (2024)
TransFusionOdom: Interpretable Transformer-based LiDAR-Inertial Fusion Odometry Estimation
von: Sun, Leyuan, et al.
Veröffentlicht: (2023)
von: Sun, Leyuan, et al.
Veröffentlicht: (2023)
X-VLA: Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model
von: Zheng, Jinliang, et al.
Veröffentlicht: (2025)
von: Zheng, Jinliang, et al.
Veröffentlicht: (2025)
Integrating Features for Recognizing Human Activities through Optimized Parameters in Graph Convolutional Networks and Transformer Architectures
von: Belal, Mohammad, et al.
Veröffentlicht: (2024)
von: Belal, Mohammad, et al.
Veröffentlicht: (2024)
Pair-VPR: Place-Aware Pre-training and Contrastive Pair Classification for Visual Place Recognition with Vision Transformers
von: Hausler, Stephen, et al.
Veröffentlicht: (2024)
von: Hausler, Stephen, et al.
Veröffentlicht: (2024)
DaFoEs: Mixing Datasets towards the generalization of vision-state deep-learning Force Estimation in Minimally Invasive Robotic Surgery
von: Reyzabal, Mikel De Iturrate, et al.
Veröffentlicht: (2024)
von: Reyzabal, Mikel De Iturrate, et al.
Veröffentlicht: (2024)
H-Net: A Multitask Architecture for Simultaneous 3D Force Estimation and Stereo Semantic Segmentation in Intracardiac Catheters
von: Fekri, Pedram, et al.
Veröffentlicht: (2024)
von: Fekri, Pedram, et al.
Veröffentlicht: (2024)
OmniNOCS: A unified NOCS dataset and model for 3D lifting of 2D objects
von: Krishnan, Akshay, et al.
Veröffentlicht: (2024)
von: Krishnan, Akshay, et al.
Veröffentlicht: (2024)
Validation & Exploration of Multimodal Deep-Learning Camera-Lidar Calibration models
von: Karramreddy, Venkat, et al.
Veröffentlicht: (2024)
von: Karramreddy, Venkat, et al.
Veröffentlicht: (2024)
Grounding Driving VLA via Inverse Kinematics
von: Park, Junsung, et al.
Veröffentlicht: (2026)
von: Park, Junsung, et al.
Veröffentlicht: (2026)
Planning with the Views via Scene Self-Exploration
von: Wang, Kangrui, et al.
Veröffentlicht: (2026)
von: Wang, Kangrui, et al.
Veröffentlicht: (2026)
Extending Deep Event Visual Odometry with Sparse Point-Cloud Export
von: Safdari, Alireza, et al.
Veröffentlicht: (2026)
von: Safdari, Alireza, et al.
Veröffentlicht: (2026)
Leveraging Foundation Models To learn the shape of semi-fluid deformable objects
von: Assal, Omar El, et al.
Veröffentlicht: (2024)
von: Assal, Omar El, et al.
Veröffentlicht: (2024)
PhotoBot: Reference-Guided Interactive Photography via Natural Language
von: Limoyo, Oliver, et al.
Veröffentlicht: (2024)
von: Limoyo, Oliver, et al.
Veröffentlicht: (2024)
sam-llm: interpretable lane change trajectoryprediction via parametric finetuning
von: Cao, Zhuo, et al.
Veröffentlicht: (2025)
von: Cao, Zhuo, et al.
Veröffentlicht: (2025)
Efficient Robotic Policy Learning via Latent Space Backward Planning
von: Liu, Dongxiu, et al.
Veröffentlicht: (2025)
von: Liu, Dongxiu, et al.
Veröffentlicht: (2025)
RoboSafe: Safeguarding Embodied Agents via Executable Safety Logic
von: Wang, Le, et al.
Veröffentlicht: (2025)
von: Wang, Le, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Goal-conditioned dual-action imitation learning for dexterous dual-arm robot manipulation
von: Kim, Heecheol, et al.
Veröffentlicht: (2022) -
EmbodiSwap for Zero-Shot Robot Imitation Learning
von: Dessalene, Eadom, et al.
Veröffentlicht: (2025) -
TransForSeg: A Multitask Stereo ViT for Joint Stereo Segmentation and 3D Force Estimation in Catheterization
von: Fekri, Pedram, et al.
Veröffentlicht: (2025) -
Accelerating Transformer-Based Monocular SLAM via Geometric Utility Scoring
von: Xiong, Xinmiao, et al.
Veröffentlicht: (2026) -
Point-LN: A Lightweight Framework for Efficient Point Cloud Classification Using Non-Parametric Positional Encoding
von: Mohammadi, Marzieh, et al.
Veröffentlicht: (2025)