FlySearch: Exploring how vision-language models explore
Fuente:
arXiv
Saved in:
| Main Authors: | Pardyl, Adam, Matuszek, Dominik, Przebieracz, Mateusz, Cygan, Marek, Zieliński, Bartosz, Wołczyk, Maciej |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AdaGlimpse: Active Visual Exploration with Arbitrary Glimpse Position and Scale
by: Pardyl, Adam, et al.
Published: (2024)
by: Pardyl, Adam, et al.
Published: (2024)
Beyond Grids: Exploring Elastic Input Sampling for Vision Transformers
by: Pardyl, Adam, et al.
Published: (2023)
by: Pardyl, Adam, et al.
Published: (2023)
Enhancing people localisation in drone imagery for better crowd management by utilising every pixel in high-resolution images
by: Ptak, Bartosz, et al.
Published: (2025)
by: Ptak, Bartosz, et al.
Published: (2025)
Improving trajectory continuity in drone-based crowd monitoring using a set of minimal-cost techniques and deep discriminative correlation filters
by: Ptak, Bartosz, et al.
Published: (2025)
by: Ptak, Bartosz, et al.
Published: (2025)
A Comparative Evaluation of Geometric Accuracy in NeRF and Gaussian Splatting
by: Zielinski, Mikolaj, et al.
Published: (2026)
by: Zielinski, Mikolaj, et al.
Published: (2026)
Keyframe-based Dense Mapping with the Graph of View-Dependent Local Maps
by: Zielinski, Krzysztof, et al.
Published: (2026)
by: Zielinski, Krzysztof, et al.
Published: (2026)
SpaRRTa: A Synthetic Benchmark for Evaluating Spatial Intelligence in Visual Foundation Models
by: Kargin, Turhan Can, et al.
Published: (2026)
by: Kargin, Turhan Can, et al.
Published: (2026)
A vision-language model and platform for temporally mapping surgery from video
by: Kiyasseh, Dani
Published: (2026)
by: Kiyasseh, Dani
Published: (2026)
Openfly: A comprehensive platform for aerial vision-language navigation
by: Gao, Yunpeng, et al.
Published: (2025)
by: Gao, Yunpeng, et al.
Published: (2025)
3D Reconstruction of non-visible surfaces of objects from a Single Depth View -- Comparative Study
by: Staszak, Rafał, et al.
Published: (2025)
by: Staszak, Rafał, et al.
Published: (2025)
Digital and Robotic Twinning for Validation of Proximity Operations and Formation Flying
by: Ahmed, Z., et al.
Published: (2025)
by: Ahmed, Z., et al.
Published: (2025)
EventFly: Event Camera Perception from Ground to the Sky
by: Kong, Lingdong, et al.
Published: (2025)
by: Kong, Lingdong, et al.
Published: (2025)
Learning on the Fly: Replay-Based Continual Object Perception for Indoor Drones
by: Nae, Sebastian-Ion, et al.
Published: (2026)
by: Nae, Sebastian-Ion, et al.
Published: (2026)
On-the-Fly SfM: What you capture is What you get
by: Zhan, Zongqian, et al.
Published: (2023)
by: Zhan, Zongqian, et al.
Published: (2023)
Event-Aided Sharp Radiance Field Reconstruction for Fast-Flying Drones
by: Zou, Rong, et al.
Published: (2026)
by: Zou, Rong, et al.
Published: (2026)
FlyPose: Towards Robust Human Pose Estimation From Aerial Views
by: Farooq, Hassaan, et al.
Published: (2026)
by: Farooq, Hassaan, et al.
Published: (2026)
Adaptive Articulated Object Manipulation On The Fly with Foundation Model Reasoning and Part Grounding
by: Zhang, Xiaojie, et al.
Published: (2025)
by: Zhang, Xiaojie, et al.
Published: (2025)
CarLLaVA: Vision language models for camera-only closed-loop driving
by: Renz, Katrin, et al.
Published: (2024)
by: Renz, Katrin, et al.
Published: (2024)
High-fidelity 3D reconstruction for planetary exploration
by: Martínez-Petersen, Alfonso, et al.
Published: (2026)
by: Martínez-Petersen, Alfonso, et al.
Published: (2026)
Exploring Conditions for Diffusion models in Robotic Control
by: Shin, Heeseong, et al.
Published: (2025)
by: Shin, Heeseong, et al.
Published: (2025)
Multi-Agent 3D Map Reconstruction and Change Detection in Microgravity with Free-Flying Robots
by: Dinkel, Holly, et al.
Published: (2023)
by: Dinkel, Holly, et al.
Published: (2023)
UAV-Flow Colosseo: A Real-World Benchmark for Flying-on-a-Word UAV Imitation Learning
by: Wang, Xiangyu, et al.
Published: (2025)
by: Wang, Xiangyu, et al.
Published: (2025)
iFlyBot-VLA Technical Report
by: Zhang, Yuan, et al.
Published: (2025)
by: Zhang, Yuan, et al.
Published: (2025)
Self-Supervised Learning to Fly using Efficient Semantic Segmentation and Metric Depth Estimation for Low-Cost Autonomous UAVs
by: Mocanu, Sebastian, et al.
Published: (2025)
by: Mocanu, Sebastian, et al.
Published: (2025)
Beyond [cls]: Exploring the true potential of Masked Image Modeling representations
by: Przewięźlikowski, Marcin, et al.
Published: (2024)
by: Przewięźlikowski, Marcin, et al.
Published: (2024)
Augmentation-aware Self-supervised Learning with Conditioned Projector
by: Przewięźlikowski, Marcin, et al.
Published: (2023)
by: Przewięźlikowski, Marcin, et al.
Published: (2023)
Graph-based Robot Localization Using a Graph Neural Network with a Floor Camera and a Feature Rich Industrial Floor
by: Brämer, Dominik, et al.
Published: (2025)
by: Brämer, Dominik, et al.
Published: (2025)
Object Depth and Size Estimation using Stereo-vision and Integration with SLAM
by: Hamad, Layth, et al.
Published: (2024)
by: Hamad, Layth, et al.
Published: (2024)
MoManipVLA: Transferring Vision-language-action Models for General Mobile Manipulation
by: Wu, Zhenyu, et al.
Published: (2025)
by: Wu, Zhenyu, et al.
Published: (2025)
Robot Localization Using a Learned Keypoint Detector and Descriptor with a Floor Camera and a Feature Rich Industrial Floor
by: Brömmel, Piet, et al.
Published: (2025)
by: Brömmel, Piet, et al.
Published: (2025)
CaLiV: LiDAR-to-Vehicle Calibration of Arbitrary Sensor Setups
by: Tahiraj, Ilir, et al.
Published: (2025)
by: Tahiraj, Ilir, et al.
Published: (2025)
Touch begins where vision ends: Generalizable policies for contact-rich manipulation
by: Zhao, Zifan, et al.
Published: (2025)
by: Zhao, Zifan, et al.
Published: (2025)
Multi-vision-based Picking Point Localisation of Target Fruit for Harvesting Robots
by: Beldek, C., et al.
Published: (2025)
by: Beldek, C., et al.
Published: (2025)
An indoor DSO-based ceiling-vision odometry system for indoor industrial environments
by: Bougouffa, Abdelhak, et al.
Published: (2024)
by: Bougouffa, Abdelhak, et al.
Published: (2024)
Quantitative evaluation of brain-inspired vision sensors in high-speed robotic perception
by: Wang, Taoyi, et al.
Published: (2025)
by: Wang, Taoyi, et al.
Published: (2025)
SurgeMOD: Translating image-space tissue motions into vision-based surgical forces
by: Reyzabal, Mikel De Iturrate, et al.
Published: (2024)
by: Reyzabal, Mikel De Iturrate, et al.
Published: (2024)
WildOS: Open-Vocabulary Object Search in the Wild
by: Shah, Hardik, et al.
Published: (2026)
by: Shah, Hardik, et al.
Published: (2026)
NeuralLabeling: A versatile toolset for labeling vision datasets using Neural Radiance Fields
by: Erich, Floris, et al.
Published: (2023)
by: Erich, Floris, et al.
Published: (2023)
Device-Conditioned Neural Architecture Search for Efficient Robotic Manipulation
by: Wu, Yiming, et al.
Published: (2026)
by: Wu, Yiming, et al.
Published: (2026)
STORM: Search-Guided Generative World Models for Robotic Manipulation
by: Lin, Wenjun, et al.
Published: (2025)
by: Lin, Wenjun, et al.
Published: (2025)
Similar Items
-
AdaGlimpse: Active Visual Exploration with Arbitrary Glimpse Position and Scale
by: Pardyl, Adam, et al.
Published: (2024) -
Beyond Grids: Exploring Elastic Input Sampling for Vision Transformers
by: Pardyl, Adam, et al.
Published: (2023) -
Enhancing people localisation in drone imagery for better crowd management by utilising every pixel in high-resolution images
by: Ptak, Bartosz, et al.
Published: (2025) -
Improving trajectory continuity in drone-based crowd monitoring using a set of minimal-cost techniques and deep discriminative correlation filters
by: Ptak, Bartosz, et al.
Published: (2025) -
A Comparative Evaluation of Geometric Accuracy in NeRF and Gaussian Splatting
by: Zielinski, Mikolaj, et al.
Published: (2026)