Transformer-Based Model for Monocular Visual Odometry: A Video Understanding Approach
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Françani, André O., Maximo, Marcos R. O. A. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Motion Consistency Loss for Monocular Visual Odometry with Attention-Based Deep Learning
von: Françani, André O., et al.
Veröffentlicht: (2024)
von: Françani, André O., et al.
Veröffentlicht: (2024)
Building Brain Tumor Segmentation Networks with User-Assisted Filter Estimation and Selection
von: Cerqueira, Matheus A., et al.
Veröffentlicht: (2024)
von: Cerqueira, Matheus A., et al.
Veröffentlicht: (2024)
Hybrid SIFT-SNN for Efficient Anomaly Detection of Traffic Flow-Control Infrastructure
von: Rathee, Munish, et al.
Veröffentlicht: (2025)
von: Rathee, Munish, et al.
Veröffentlicht: (2025)
How Well Do Vision-Language Models Understand Sequential Driving Scenes? A Sensitivity Study
von: Brusnicki, Roberto, et al.
Veröffentlicht: (2026)
von: Brusnicki, Roberto, et al.
Veröffentlicht: (2026)
Explainable Image Similarity: Integrating Siamese Networks and Grad-CAM
von: Livieris, Ioannis E., et al.
Veröffentlicht: (2023)
von: Livieris, Ioannis E., et al.
Veröffentlicht: (2023)
Poisson Flow Consistency Training
von: Zhang, Anthony, et al.
Veröffentlicht: (2025)
von: Zhang, Anthony, et al.
Veröffentlicht: (2025)
Memory-augmented Online Video Anomaly Detection
von: Rossi, Leonardo, et al.
Veröffentlicht: (2023)
von: Rossi, Leonardo, et al.
Veröffentlicht: (2023)
Z-Order Transformer for Feed-Forward Gaussian Splatting
von: Wang, Can, et al.
Veröffentlicht: (2026)
von: Wang, Can, et al.
Veröffentlicht: (2026)
E Pluribus Unum Interpretable Convolutional Neural Networks
von: Dimas, George, et al.
Veröffentlicht: (2022)
von: Dimas, George, et al.
Veröffentlicht: (2022)
Interactive Image Selection and Training for Brain Tumor Segmentation Network
von: Cerqueira, Matheus A., et al.
Veröffentlicht: (2024)
von: Cerqueira, Matheus A., et al.
Veröffentlicht: (2024)
Performance Decay in Deepfake Detection: The Limitations of Training on Outdated Data
von: Richings, Jack, et al.
Veröffentlicht: (2025)
von: Richings, Jack, et al.
Veröffentlicht: (2025)
Unlocking UML Class Diagram Understanding in Vision Language Models
von: Naboichenko, Artem, et al.
Veröffentlicht: (2026)
von: Naboichenko, Artem, et al.
Veröffentlicht: (2026)
Point, Detect, Count: Multi-Task Medical Image Understanding with Instruction-Tuned Vision-Language Models
von: Gautam, Sushant, et al.
Veröffentlicht: (2025)
von: Gautam, Sushant, et al.
Veröffentlicht: (2025)
TWIG: Two-Step Image Generation using Segmentation Masks in Diffusion Models
von: Rakib, Mazharul Islam, et al.
Veröffentlicht: (2025)
von: Rakib, Mazharul Islam, et al.
Veröffentlicht: (2025)
RailSafeNet: Visual Scene Understanding for Tram Safety
von: Valach, Ondřej, et al.
Veröffentlicht: (2025)
von: Valach, Ondřej, et al.
Veröffentlicht: (2025)
Overcoming Catastrophic Forgetting in Federated Class-Incremental Learning via Federated Global Twin Generator
von: Nguyen, Thinh, et al.
Veröffentlicht: (2024)
von: Nguyen, Thinh, et al.
Veröffentlicht: (2024)
Depth Priors in Removal Neural Radiance Fields
von: Guo, Zhihao, et al.
Veröffentlicht: (2024)
von: Guo, Zhihao, et al.
Veröffentlicht: (2024)
EatGAN: An Edge-Attention Guided Generative Adversarial Network for Single Image Super-Resolution
von: Rao, Penghao, et al.
Veröffentlicht: (2025)
von: Rao, Penghao, et al.
Veröffentlicht: (2025)
Isolated Sign Language Recognition with Segmentation and Pose Estimation
von: Perkins, Daniel, et al.
Veröffentlicht: (2025)
von: Perkins, Daniel, et al.
Veröffentlicht: (2025)
Optimizing MoE Routers: Design, Implementation, and Evaluation in Transformer Models
von: Harvey, Daniel Fidel, et al.
Veröffentlicht: (2025)
von: Harvey, Daniel Fidel, et al.
Veröffentlicht: (2025)
MB-DSMIL-CL-PL: Scalable Weakly Supervised Ovarian Cancer Subtype Classification and Localisation Using Contrastive and Prototype Learning with Frozen Patch Features
von: Jenkins, Marcus, et al.
Veröffentlicht: (2026)
von: Jenkins, Marcus, et al.
Veröffentlicht: (2026)
On the Inherent Robustness of One-Stage Object Detection against Out-of-Distribution Data
von: Martinez-Seras, Aitor, et al.
Veröffentlicht: (2024)
von: Martinez-Seras, Aitor, et al.
Veröffentlicht: (2024)
Socratic Planner: Self-QA-Based Zero-Shot Planning for Embodied Instruction Following
von: Shin, Suyeon, et al.
Veröffentlicht: (2024)
von: Shin, Suyeon, et al.
Veröffentlicht: (2024)
OptiRoulette Optimizer: A New Stochastic Meta-Optimizer for up to 5.3x Faster Convergence
von: Mastromichalakis, Stamatis
Veröffentlicht: (2026)
von: Mastromichalakis, Stamatis
Veröffentlicht: (2026)
VideoHEDGE: Entropy-Based Hallucination Detection for Video-VLMs via Semantic Clustering and Spatiotemporal Perturbations
von: Gautam, Sushant, et al.
Veröffentlicht: (2026)
von: Gautam, Sushant, et al.
Veröffentlicht: (2026)
Attention Maps in 3D Shape Classification for Dental Stage Estimation with Class Node Graph Attention Networks
von: Buyukcakir, Barkin, et al.
Veröffentlicht: (2025)
von: Buyukcakir, Barkin, et al.
Veröffentlicht: (2025)
Detection and Measurement of Hailstones with Multimodal Large Language Models
von: Alker, Moritz, et al.
Veröffentlicht: (2025)
von: Alker, Moritz, et al.
Veröffentlicht: (2025)
Rethinking Visual Intelligence: Insights from Video Pretraining
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025)
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025)
Computational Imaging Priors for Wireless Capsule Endoscopy: Monte Carlo-Guided Hemoglobin Mapping for Rare-Anomaly Detection
von: Yang, Chengshuai, et al.
Veröffentlicht: (2026)
von: Yang, Chengshuai, et al.
Veröffentlicht: (2026)
LRVS-Fashion: Extending Visual Search with Referring Instructions
von: Lepage, Simon, et al.
Veröffentlicht: (2023)
von: Lepage, Simon, et al.
Veröffentlicht: (2023)
AUTHENTICATION: Identifying Rare Failure Modes in Autonomous Vehicle Perception Systems using Adversarially Guided Diffusion Models
von: Zarei, Mohammad, et al.
Veröffentlicht: (2025)
von: Zarei, Mohammad, et al.
Veröffentlicht: (2025)
Method of UAV Inspection of Photovoltaic Modules Using Thermal and RGB Data Fusion
von: Lysyi, Andrii, et al.
Veröffentlicht: (2025)
von: Lysyi, Andrii, et al.
Veröffentlicht: (2025)
Harmony: A Joint Self-Supervised and Weakly-Supervised Framework for Learning General Purpose Visual Representations
von: Baharoon, Mohammed, et al.
Veröffentlicht: (2024)
von: Baharoon, Mohammed, et al.
Veröffentlicht: (2024)
VLM@school -- Evaluation of AI image understanding on German middle school knowledge
von: Peinl, René, et al.
Veröffentlicht: (2025)
von: Peinl, René, et al.
Veröffentlicht: (2025)
TransDAE: Dual Attention Mechanism in a Hierarchical Transformer for Efficient Medical Image Segmentation
von: Azad, Bobby, et al.
Veröffentlicht: (2024)
von: Azad, Bobby, et al.
Veröffentlicht: (2024)
Spatial-ViLT: Enhancing Visual Spatial Reasoning through Multi-Task Learning
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2025)
von: Islam, Chashi Mahiul, et al.
Veröffentlicht: (2025)
YOLOv10 with Kolmogorov-Arnold networks and vision-language foundation models for interpretable object detection and trustworthy multimodal AI in computer vision perception
von: Impraimakis, Marios, et al.
Veröffentlicht: (2026)
von: Impraimakis, Marios, et al.
Veröffentlicht: (2026)
YOLO Ensemble for UAV-based Multispectral Defect Detection in Wind Turbine Components
von: Svystun, Serhii, et al.
Veröffentlicht: (2025)
von: Svystun, Serhii, et al.
Veröffentlicht: (2025)
DOD-SA: Infrared-Visible Decoupled Object Detection with Single-Modality Annotations
von: Jin, Hang, et al.
Veröffentlicht: (2025)
von: Jin, Hang, et al.
Veröffentlicht: (2025)
Multimodal Integration Challenges in Emotionally Expressive Child Avatars for Training Applications
von: Salehi, Pegah, et al.
Veröffentlicht: (2025)
von: Salehi, Pegah, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Motion Consistency Loss for Monocular Visual Odometry with Attention-Based Deep Learning
von: Françani, André O., et al.
Veröffentlicht: (2024) -
Building Brain Tumor Segmentation Networks with User-Assisted Filter Estimation and Selection
von: Cerqueira, Matheus A., et al.
Veröffentlicht: (2024) -
Hybrid SIFT-SNN for Efficient Anomaly Detection of Traffic Flow-Control Infrastructure
von: Rathee, Munish, et al.
Veröffentlicht: (2025) -
How Well Do Vision-Language Models Understand Sequential Driving Scenes? A Sensitivity Study
von: Brusnicki, Roberto, et al.
Veröffentlicht: (2026) -
Explainable Image Similarity: Integrating Siamese Networks and Grad-CAM
von: Livieris, Ioannis E., et al.
Veröffentlicht: (2023)