FeudalNav: A Simple Framework for Visual Navigation
Fuente:
arXiv
Salvato in:
| Autori principali: | Johnson, Faith, Cao, Bryan Bo, Jain, Shubham, Ashok, Ashwin, Dana, Kristin |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
A Landmark-Aware Visual Navigation Dataset
di: Johnson, Faith, et al.
Pubblicazione: (2024)
di: Johnson, Faith, et al.
Pubblicazione: (2024)
MerNav: A Highly Generalizable Memory-Execute-Review Framework for Zero-Shot Object Goal Navigation
di: Qi, Dekang, et al.
Pubblicazione: (2026)
di: Qi, Dekang, et al.
Pubblicazione: (2026)
FT-NCFM: An Influence-Aware Data Distillation Framework for Efficient VLA Models
di: Chen, Kewei, et al.
Pubblicazione: (2025)
di: Chen, Kewei, et al.
Pubblicazione: (2025)
ATAAT: Adaptive Threat-Aware Adversarial Tuning Framework against Backdoor Attacks on Vision-Language-Action Models
di: Chen, Kewei, et al.
Pubblicazione: (2026)
di: Chen, Kewei, et al.
Pubblicazione: (2026)
ABot-Claw: A Foundation for Persistent, Cooperative, and Self-Evolving Robotic Agents
di: Huo, Dongjie, et al.
Pubblicazione: (2026)
di: Huo, Dongjie, et al.
Pubblicazione: (2026)
CulinaryCut-VLAP: A Vision-Language-Action-Physics Framework for Food Cutting via a Force-Aware Material Point Method
di: Koh, Hyunseo, et al.
Pubblicazione: (2026)
di: Koh, Hyunseo, et al.
Pubblicazione: (2026)
Measuring What Matters: Scenario-Driven Evaluation for Trajectory Predictors in Autonomous Driving
di: Da, Longchao, et al.
Pubblicazione: (2025)
di: Da, Longchao, et al.
Pubblicazione: (2025)
Key-Scan-Based Mobile Robot Navigation: Integrated Mapping, Planning, and Control using Graphs of Scan Regions
di: Latha, Dharshan Bashkaran, et al.
Pubblicazione: (2024)
di: Latha, Dharshan Bashkaran, et al.
Pubblicazione: (2024)
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
di: Chahine, Makram, et al.
Pubblicazione: (2024)
di: Chahine, Makram, et al.
Pubblicazione: (2024)
PathFormer: A Transformer with 3D Grid Constraints for Digital Twin Robot-Arm Trajectory Generation
di: Alanazi, Ahmed, et al.
Pubblicazione: (2025)
di: Alanazi, Ahmed, et al.
Pubblicazione: (2025)
Evaluating Visual Mathematics in Multimodal LLMs: A Multilingual Benchmark Based on the Kangaroo Tests
di: Sáez, Arnau Igualde, et al.
Pubblicazione: (2025)
di: Sáez, Arnau Igualde, et al.
Pubblicazione: (2025)
Context-Dependent Affordance Computation in Vision-Language Models
di: Farzulla, Murad
Pubblicazione: (2026)
di: Farzulla, Murad
Pubblicazione: (2026)
Method of UAV Inspection of Photovoltaic Modules Using Thermal and RGB Data Fusion
di: Lysyi, Andrii, et al.
Pubblicazione: (2025)
di: Lysyi, Andrii, et al.
Pubblicazione: (2025)
A Survey on Vision-Language-Action Models for Embodied AI
di: Ma, Yueen, et al.
Pubblicazione: (2024)
di: Ma, Yueen, et al.
Pubblicazione: (2024)
Cooperative Perception: A Resource-Efficient Framework for Multi-Drone 3D Scene Reconstruction Using Federated Diffusion and NeRF
di: Pourmandi, Massoud
Pubblicazione: (2025)
di: Pourmandi, Massoud
Pubblicazione: (2025)
Agentic UAVs: LLM-Driven Autonomy with Integrated Tool-Calling and Cognitive Reasoning
di: Koubaa, Anis, et al.
Pubblicazione: (2025)
di: Koubaa, Anis, et al.
Pubblicazione: (2025)
Task and Motion Planning in Hierarchical 3D Scene Graphs
di: Ray, Aaron, et al.
Pubblicazione: (2024)
di: Ray, Aaron, et al.
Pubblicazione: (2024)
RACAS: Controlling Diverse Robots With a Single Agentic System
di: Ashley, Dylan R., et al.
Pubblicazione: (2026)
di: Ashley, Dylan R., et al.
Pubblicazione: (2026)
Experimental Evaluation of Road-Crossing Decisions by Autonomous Wheelchairs against Environmental Factors
di: Corradini, Franca, et al.
Pubblicazione: (2024)
di: Corradini, Franca, et al.
Pubblicazione: (2024)
SemanticFeels: Semantic Labeling during In-Hand Manipulation
di: Khalil, Anas Al Shikh, et al.
Pubblicazione: (2026)
di: Khalil, Anas Al Shikh, et al.
Pubblicazione: (2026)
Unpacking Hateful Memes: Presupposed Context and False Claims
di: Cai, Weibin, et al.
Pubblicazione: (2025)
di: Cai, Weibin, et al.
Pubblicazione: (2025)
Tricks and Plug-ins for Gradient Boosting in Image Classification
di: Fang, Biyi, et al.
Pubblicazione: (2025)
di: Fang, Biyi, et al.
Pubblicazione: (2025)
ExpReS-VLA: Specializing Vision-Language-Action Models Through Experience Replay and Retrieval
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
YOLO Ensemble for UAV-based Multispectral Defect Detection in Wind Turbine Components
di: Svystun, Serhii, et al.
Pubblicazione: (2025)
di: Svystun, Serhii, et al.
Pubblicazione: (2025)
Multimodal Generative AI for Story Point Estimation in Software Development
di: Islam, Mohammad Rubyet, et al.
Pubblicazione: (2025)
di: Islam, Mohammad Rubyet, et al.
Pubblicazione: (2025)
Balanced conic rectified flow
di: Kim, Shin Seong, et al.
Pubblicazione: (2025)
di: Kim, Shin Seong, et al.
Pubblicazione: (2025)
ReasonPlan: Unified Scene Prediction and Decision Reasoning for Closed-loop Autonomous Driving
di: Liu, Xueyi, et al.
Pubblicazione: (2025)
di: Liu, Xueyi, et al.
Pubblicazione: (2025)
Mitigating Covariate Shift in Imitation Learning for Autonomous Vehicles Using Latent Space Generative World Models
di: Popov, Alexander, et al.
Pubblicazione: (2024)
di: Popov, Alexander, et al.
Pubblicazione: (2024)
TowerVision: Understanding and Improving Multilinguality in Vision-Language Models
di: Viveiros, André G., et al.
Pubblicazione: (2025)
di: Viveiros, André G., et al.
Pubblicazione: (2025)
MORQA: Benchmarking Evaluation Metrics for Medical Open-Ended Question Answering
di: Yim, Wen-wai, et al.
Pubblicazione: (2025)
di: Yim, Wen-wai, et al.
Pubblicazione: (2025)
Think, Act, Learn: A Framework for Autonomous Robotic Agents using Closed-Loop Large Language Models
di: Menon, Anjali R., et al.
Pubblicazione: (2025)
di: Menon, Anjali R., et al.
Pubblicazione: (2025)
TableMoE: Neuro-Symbolic Routing for Structured Expert Reasoning in Multimodal Table Understanding
di: Zhang, Junwen, et al.
Pubblicazione: (2025)
di: Zhang, Junwen, et al.
Pubblicazione: (2025)
Towards Ubiquitous Mapping and Localization for Dynamic Indoor Environments
di: Djerroud, Halim, et al.
Pubblicazione: (2026)
di: Djerroud, Halim, et al.
Pubblicazione: (2026)
Perception-Consistency Multimodal Large Language Models Reasoning via Caption-Regularized Policy Optimization
di: Tu, Songjun, et al.
Pubblicazione: (2025)
di: Tu, Songjun, et al.
Pubblicazione: (2025)
AUTHENTICATION: Identifying Rare Failure Modes in Autonomous Vehicle Perception Systems using Adversarially Guided Diffusion Models
di: Zarei, Mohammad, et al.
Pubblicazione: (2025)
di: Zarei, Mohammad, et al.
Pubblicazione: (2025)
Incremental Bootstrapping and Classification of Structured Scenes in a Fuzzy Ontology
di: Buoncompagni, Luca, et al.
Pubblicazione: (2024)
di: Buoncompagni, Luca, et al.
Pubblicazione: (2024)
Semantically Guided Adversarial Testing of Vision Models Using Language Models
di: Filus, Katarzyna, et al.
Pubblicazione: (2025)
di: Filus, Katarzyna, et al.
Pubblicazione: (2025)
Rethinking Visual Intelligence: Insights from Video Pretraining
di: Acuaviva, Pablo, et al.
Pubblicazione: (2025)
di: Acuaviva, Pablo, et al.
Pubblicazione: (2025)
Inducing Causal World Models in LLMs for Zero-Shot Physical Reasoning
di: Sharma, Aditya, et al.
Pubblicazione: (2025)
di: Sharma, Aditya, et al.
Pubblicazione: (2025)
Surg$Σ$: A Spectrum of Large-Scale Multimodal Data and Foundation Models for Surgical Intelligence
di: Zeng, Zhitao, et al.
Pubblicazione: (2026)
di: Zeng, Zhitao, et al.
Pubblicazione: (2026)
Documenti analoghi
-
A Landmark-Aware Visual Navigation Dataset
di: Johnson, Faith, et al.
Pubblicazione: (2024) -
MerNav: A Highly Generalizable Memory-Execute-Review Framework for Zero-Shot Object Goal Navigation
di: Qi, Dekang, et al.
Pubblicazione: (2026) -
FT-NCFM: An Influence-Aware Data Distillation Framework for Efficient VLA Models
di: Chen, Kewei, et al.
Pubblicazione: (2025) -
ATAAT: Adaptive Threat-Aware Adversarial Tuning Framework against Backdoor Attacks on Vision-Language-Action Models
di: Chen, Kewei, et al.
Pubblicazione: (2026) -
ABot-Claw: A Foundation for Persistent, Cooperative, and Self-Evolving Robotic Agents
di: Huo, Dongjie, et al.
Pubblicazione: (2026)