Shaken, Not Stirred: A Novel Dataset for Visual Understanding of Glasses in Human-Robot Bartending Tasks
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gajdošech, Lukáš, Ali, Hassan, Habekost, Jan-Gerrit, Madaras, Martin, Kerzel, Matthias, Wermter, Stefan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Detecting 3D Line Segments for 6DoF Pose Estimation with Limited Data
von: Mok, Matej, et al.
Veröffentlicht: (2026)
von: Mok, Matej, et al.
Veröffentlicht: (2026)
Novel Synthetic Data Tool for Data-Driven Cardboard Box Localization
von: Gajdošech, Lukáš, et al.
Veröffentlicht: (2023)
von: Gajdošech, Lukáš, et al.
Veröffentlicht: (2023)
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
von: Chahine, Makram, et al.
Veröffentlicht: (2024)
von: Chahine, Makram, et al.
Veröffentlicht: (2024)
Pointing-Guided Target Estimation via Transformer-Based Attention
von: Müller, Luca, et al.
Veröffentlicht: (2025)
von: Müller, Luca, et al.
Veröffentlicht: (2025)
Have We Mastered Scale in Deep Monocular Visual SLAM? The ScaleMaster Dataset and Benchmark
von: Ju, Hyoseok, et al.
Veröffentlicht: (2026)
von: Ju, Hyoseok, et al.
Veröffentlicht: (2026)
Safe Road-Crossing by Autonomous Wheelchairs: a Novel Dataset and its Experimental Evaluation
von: Grigioni, Carlo, et al.
Veröffentlicht: (2024)
von: Grigioni, Carlo, et al.
Veröffentlicht: (2024)
Visual Categorization Across Minds and Models: Cognitive Analysis of Human Labeling and Neuro-Symbolic Integration
von: Kabgere, Chethana Prasad
Veröffentlicht: (2025)
von: Kabgere, Chethana Prasad
Veröffentlicht: (2025)
Botany Meets Robotics in Alpine Scree Monitoring
von: De Benedittis, Davide, et al.
Veröffentlicht: (2025)
von: De Benedittis, Davide, et al.
Veröffentlicht: (2025)
Cooperative Perception: A Resource-Efficient Framework for Multi-Drone 3D Scene Reconstruction Using Federated Diffusion and NeRF
von: Pourmandi, Massoud
Veröffentlicht: (2025)
von: Pourmandi, Massoud
Veröffentlicht: (2025)
Experimental Evaluation of Road-Crossing Decisions by Autonomous Wheelchairs against Environmental Factors
von: Corradini, Franca, et al.
Veröffentlicht: (2024)
von: Corradini, Franca, et al.
Veröffentlicht: (2024)
Method of UAV Inspection of Photovoltaic Modules Using Thermal and RGB Data Fusion
von: Lysyi, Andrii, et al.
Veröffentlicht: (2025)
von: Lysyi, Andrii, et al.
Veröffentlicht: (2025)
Comparing Apples to Oranges: LLM-powered Multimodal Intention Prediction in an Object Categorization Task
von: Ali, Hassan, et al.
Veröffentlicht: (2024)
von: Ali, Hassan, et al.
Veröffentlicht: (2024)
SemanticFeels: Semantic Labeling during In-Hand Manipulation
von: Khalil, Anas Al Shikh, et al.
Veröffentlicht: (2026)
von: Khalil, Anas Al Shikh, et al.
Veröffentlicht: (2026)
Evaluating Model-Agnostic Meta-Learning on MetaWorld ML10 Benchmark: Fast Adaptation in Robotic Manipulation Tasks
von: Atamuradov, Sanjar
Veröffentlicht: (2025)
von: Atamuradov, Sanjar
Veröffentlicht: (2025)
Decentralized Privacy-Preserving Federal Learning of Computer Vision Models on Edge Devices
von: Harenčák, Damian, et al.
Veröffentlicht: (2026)
von: Harenčák, Damian, et al.
Veröffentlicht: (2026)
YOLO Ensemble for UAV-based Multispectral Defect Detection in Wind Turbine Components
von: Svystun, Serhii, et al.
Veröffentlicht: (2025)
von: Svystun, Serhii, et al.
Veröffentlicht: (2025)
vS-Graphs: Tightly Coupling Visual SLAM and 3D Scene Graphs Exploiting Hierarchical Scene Understanding
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models
von: Chen, Yiteng, et al.
Veröffentlicht: (2025)
von: Chen, Yiteng, et al.
Veröffentlicht: (2025)
Towards Ubiquitous Mapping and Localization for Dynamic Indoor Environments
von: Djerroud, Halim, et al.
Veröffentlicht: (2026)
von: Djerroud, Halim, et al.
Veröffentlicht: (2026)
FCBV-Net: Category-Level Robotic Garment Smoothing via Feature-Conditioned Bimanual Value Prediction
von: Daba, Mohammed, et al.
Veröffentlicht: (2025)
von: Daba, Mohammed, et al.
Veröffentlicht: (2025)
Task and Motion Planning in Hierarchical 3D Scene Graphs
von: Ray, Aaron, et al.
Veröffentlicht: (2024)
von: Ray, Aaron, et al.
Veröffentlicht: (2024)
Evaluating the Impact of Synthetic Data on Object Detection Tasks in Autonomous Driving
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
von: Özeren, Enes, et al.
Veröffentlicht: (2025)
A Landmark-Aware Visual Navigation Dataset
von: Johnson, Faith, et al.
Veröffentlicht: (2024)
von: Johnson, Faith, et al.
Veröffentlicht: (2024)
Rethinking Visual Intelligence: Insights from Video Pretraining
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025)
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025)
What do we learn from a large-scale study of pre-trained visual representations in sim and real environments?
von: Silwal, Sneha, et al.
Veröffentlicht: (2023)
von: Silwal, Sneha, et al.
Veröffentlicht: (2023)
PathFormer: A Transformer with 3D Grid Constraints for Digital Twin Robot-Arm Trajectory Generation
von: Alanazi, Ahmed, et al.
Veröffentlicht: (2025)
von: Alanazi, Ahmed, et al.
Veröffentlicht: (2025)
Rethinking Camera Choice: An Empirical Study on Fisheye Camera Properties in Robotic Manipulation
von: Xue, Han, et al.
Veröffentlicht: (2026)
von: Xue, Han, et al.
Veröffentlicht: (2026)
Unveiling the Potential of iMarkers: Invisible Fiducial Markers for Advanced Robotics
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
von: Tourani, Ali, et al.
Veröffentlicht: (2025)
Temporally Consistent Object 6D Pose Estimation for Robot Control
von: Zorina, Kateryna, et al.
Veröffentlicht: (2026)
von: Zorina, Kateryna, et al.
Veröffentlicht: (2026)
Adaptive Thresholding for Visual Place Recognition using Negative Gaussian Mixture Statistics
von: Trinh, Nick, et al.
Veröffentlicht: (2025)
von: Trinh, Nick, et al.
Veröffentlicht: (2025)
Large Language Models and 3D Vision for Intelligent Robotic Perception and Autonomy
von: Mehta, Vinit, et al.
Veröffentlicht: (2025)
von: Mehta, Vinit, et al.
Veröffentlicht: (2025)
LLM-Guided Task- and Affordance-Level Exploration in Reinforcement Learning
von: Luijkx, Jelle, et al.
Veröffentlicht: (2025)
von: Luijkx, Jelle, et al.
Veröffentlicht: (2025)
On Representation of 3D Rotation in the Context of Deep Learning
von: Pravdová, Viktória, et al.
Veröffentlicht: (2024)
von: Pravdová, Viktória, et al.
Veröffentlicht: (2024)
CARScenes: Semantic VLM Dataset for Safe Autonomous Driving
von: He, Yuankai, et al.
Veröffentlicht: (2025)
von: He, Yuankai, et al.
Veröffentlicht: (2025)
Is Single-View Mesh Reconstruction Ready for Robotics?
von: Nolte, Frederik, et al.
Veröffentlicht: (2025)
von: Nolte, Frederik, et al.
Veröffentlicht: (2025)
Multi-User Personalisation in Human-Robot Interaction: Resolving Preference Conflicts Using Gradual Argumentation
von: Civit, Aniol, et al.
Veröffentlicht: (2025)
von: Civit, Aniol, et al.
Veröffentlicht: (2025)
Lifting Vision: Ground to Aerial Localization with Reasoning Guided Planning
von: Pahari, Soham, et al.
Veröffentlicht: (2025)
von: Pahari, Soham, et al.
Veröffentlicht: (2025)
Go Big or Go Home: Simulating Mobbing Behavior with Braitenbergian Robots
von: Sanoubari, Elaheh
Veröffentlicht: (2026)
von: Sanoubari, Elaheh
Veröffentlicht: (2026)
Emotion estimation from video footage with LSTM
von: Attrah, Samer
Veröffentlicht: (2025)
von: Attrah, Samer
Veröffentlicht: (2025)
RoboGrind: Intuitive and Interactive Surface Treatment with Industrial Robots
von: Alt, Benjamin, et al.
Veröffentlicht: (2024)
von: Alt, Benjamin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Detecting 3D Line Segments for 6DoF Pose Estimation with Limited Data
von: Mok, Matej, et al.
Veröffentlicht: (2026) -
Novel Synthetic Data Tool for Data-Driven Cardboard Box Localization
von: Gajdošech, Lukáš, et al.
Veröffentlicht: (2023) -
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
von: Chahine, Makram, et al.
Veröffentlicht: (2024) -
Pointing-Guided Target Estimation via Transformer-Based Attention
von: Müller, Luca, et al.
Veröffentlicht: (2025) -
Have We Mastered Scale in Deep Monocular Visual SLAM? The ScaleMaster Dataset and Benchmark
von: Ju, Hyoseok, et al.
Veröffentlicht: (2026)