SemanticFeels: Semantic Labeling during In-Hand Manipulation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Khalil, Anas Al Shikh, Qi, Haozhi, Calandra, Roberto |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Cooperative Perception: A Resource-Efficient Framework for Multi-Drone 3D Scene Reconstruction Using Federated Diffusion and NeRF
von: Pourmandi, Massoud
Veröffentlicht: (2025)
von: Pourmandi, Massoud
Veröffentlicht: (2025)
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
von: Chahine, Makram, et al.
Veröffentlicht: (2024)
von: Chahine, Makram, et al.
Veröffentlicht: (2024)
Method of UAV Inspection of Photovoltaic Modules Using Thermal and RGB Data Fusion
von: Lysyi, Andrii, et al.
Veröffentlicht: (2025)
von: Lysyi, Andrii, et al.
Veröffentlicht: (2025)
CulinaryCut-VLAP: A Vision-Language-Action-Physics Framework for Food Cutting via a Force-Aware Material Point Method
von: Koh, Hyunseo, et al.
Veröffentlicht: (2026)
von: Koh, Hyunseo, et al.
Veröffentlicht: (2026)
FT-NCFM: An Influence-Aware Data Distillation Framework for Efficient VLA Models
von: Chen, Kewei, et al.
Veröffentlicht: (2025)
von: Chen, Kewei, et al.
Veröffentlicht: (2025)
ATAAT: Adaptive Threat-Aware Adversarial Tuning Framework against Backdoor Attacks on Vision-Language-Action Models
von: Chen, Kewei, et al.
Veröffentlicht: (2026)
von: Chen, Kewei, et al.
Veröffentlicht: (2026)
Experimental Evaluation of Road-Crossing Decisions by Autonomous Wheelchairs against Environmental Factors
von: Corradini, Franca, et al.
Veröffentlicht: (2024)
von: Corradini, Franca, et al.
Veröffentlicht: (2024)
Safe Road-Crossing by Autonomous Wheelchairs: a Novel Dataset and its Experimental Evaluation
von: Grigioni, Carlo, et al.
Veröffentlicht: (2024)
von: Grigioni, Carlo, et al.
Veröffentlicht: (2024)
ExpReS-VLA: Specializing Vision-Language-Action Models Through Experience Replay and Retrieval
von: Syed, Shahram Najam, et al.
Veröffentlicht: (2025)
von: Syed, Shahram Najam, et al.
Veröffentlicht: (2025)
Measuring What Matters: Scenario-Driven Evaluation for Trajectory Predictors in Autonomous Driving
von: Da, Longchao, et al.
Veröffentlicht: (2025)
von: Da, Longchao, et al.
Veröffentlicht: (2025)
Towards Ubiquitous Mapping and Localization for Dynamic Indoor Environments
von: Djerroud, Halim, et al.
Veröffentlicht: (2026)
von: Djerroud, Halim, et al.
Veröffentlicht: (2026)
Key-Scan-Based Mobile Robot Navigation: Integrated Mapping, Planning, and Control using Graphs of Scan Regions
von: Latha, Dharshan Bashkaran, et al.
Veröffentlicht: (2024)
von: Latha, Dharshan Bashkaran, et al.
Veröffentlicht: (2024)
RoboMoRe: LLM-based Robot Co-design via Joint Optimization of Morphology and Reward
von: Fang, Jiawei, et al.
Veröffentlicht: (2025)
von: Fang, Jiawei, et al.
Veröffentlicht: (2025)
YOLO Ensemble for UAV-based Multispectral Defect Detection in Wind Turbine Components
von: Svystun, Serhii, et al.
Veröffentlicht: (2025)
von: Svystun, Serhii, et al.
Veröffentlicht: (2025)
Visual Categorization Across Minds and Models: Cognitive Analysis of Human Labeling and Neuro-Symbolic Integration
von: Kabgere, Chethana Prasad
Veröffentlicht: (2025)
von: Kabgere, Chethana Prasad
Veröffentlicht: (2025)
Agentic UAVs: LLM-Driven Autonomy with Integrated Tool-Calling and Cognitive Reasoning
von: Koubaa, Anis, et al.
Veröffentlicht: (2025)
von: Koubaa, Anis, et al.
Veröffentlicht: (2025)
Short-Window Sliding Learning for Real-Time Violence Detection via LLM-based Auto-Labeling
von: Jung, Seoik, et al.
Veröffentlicht: (2025)
von: Jung, Seoik, et al.
Veröffentlicht: (2025)
PathFormer: A Transformer with 3D Grid Constraints for Digital Twin Robot-Arm Trajectory Generation
von: Alanazi, Ahmed, et al.
Veröffentlicht: (2025)
von: Alanazi, Ahmed, et al.
Veröffentlicht: (2025)
Multimodal Generative AI for Story Point Estimation in Software Development
von: Islam, Mohammad Rubyet, et al.
Veröffentlicht: (2025)
von: Islam, Mohammad Rubyet, et al.
Veröffentlicht: (2025)
Predictive Modeling of Maritime Radar Data Using Transformer Architecture
von: Qesaraku, Bjorna, et al.
Veröffentlicht: (2025)
von: Qesaraku, Bjorna, et al.
Veröffentlicht: (2025)
COBRA-PPM: A Causal Bayesian Reasoning Architecture Using Probabilistic Programming for Robot Manipulation Under Uncertainty
von: Cannizzaro, Ricardo, et al.
Veröffentlicht: (2024)
von: Cannizzaro, Ricardo, et al.
Veröffentlicht: (2024)
Is Single-View Mesh Reconstruction Ready for Robotics?
von: Nolte, Frederik, et al.
Veröffentlicht: (2025)
von: Nolte, Frederik, et al.
Veröffentlicht: (2025)
Evaluating Model-Agnostic Meta-Learning on MetaWorld ML10 Benchmark: Fast Adaptation in Robotic Manipulation Tasks
von: Atamuradov, Sanjar
Veröffentlicht: (2025)
von: Atamuradov, Sanjar
Veröffentlicht: (2025)
A Survey on Vision-Language-Action Models for Embodied AI
von: Ma, Yueen, et al.
Veröffentlicht: (2024)
von: Ma, Yueen, et al.
Veröffentlicht: (2024)
Do Generative Metrics Predict YOLO Performance? An Evaluation Across Models, Augmentation Ratios, and Dataset Complexity
von: Marian, Vasile, et al.
Veröffentlicht: (2026)
von: Marian, Vasile, et al.
Veröffentlicht: (2026)
AUTHENTICATION: Identifying Rare Failure Modes in Autonomous Vehicle Perception Systems using Adversarially Guided Diffusion Models
von: Zarei, Mohammad, et al.
Veröffentlicht: (2025)
von: Zarei, Mohammad, et al.
Veröffentlicht: (2025)
Connectivity-Aware Representations for Constrained Motion Planning via Multi-Scale Contrastive Learning
von: Jeon, Suhyun, et al.
Veröffentlicht: (2026)
von: Jeon, Suhyun, et al.
Veröffentlicht: (2026)
ThinkJEPA: Empowering Latent World Models with Large Vision-Language Reasoning Model
von: Zhang, Haichao, et al.
Veröffentlicht: (2026)
von: Zhang, Haichao, et al.
Veröffentlicht: (2026)
CARScenes: Semantic VLM Dataset for Safe Autonomous Driving
von: He, Yuankai, et al.
Veröffentlicht: (2025)
von: He, Yuankai, et al.
Veröffentlicht: (2025)
T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models
von: Chen, Yiteng, et al.
Veröffentlicht: (2025)
von: Chen, Yiteng, et al.
Veröffentlicht: (2025)
MerNav: A Highly Generalizable Memory-Execute-Review Framework for Zero-Shot Object Goal Navigation
von: Qi, Dekang, et al.
Veröffentlicht: (2026)
von: Qi, Dekang, et al.
Veröffentlicht: (2026)
Think, Act, Learn: A Framework for Autonomous Robotic Agents using Closed-Loop Large Language Models
von: Menon, Anjali R., et al.
Veröffentlicht: (2025)
von: Menon, Anjali R., et al.
Veröffentlicht: (2025)
What do we learn from a large-scale study of pre-trained visual representations in sim and real environments?
von: Silwal, Sneha, et al.
Veröffentlicht: (2023)
von: Silwal, Sneha, et al.
Veröffentlicht: (2023)
Mitigating Covariate Shift in Imitation Learning for Autonomous Vehicles Using Latent Space Generative World Models
von: Popov, Alexander, et al.
Veröffentlicht: (2024)
von: Popov, Alexander, et al.
Veröffentlicht: (2024)
From Demonstrations to Safe Deployment: Path-Consistent Safety Filtering for Diffusion Policies
von: Römer, Ralf, et al.
Veröffentlicht: (2025)
von: Römer, Ralf, et al.
Veröffentlicht: (2025)
ElasticFlow: One-Step Physics-Consistent Policy with Elastic Time Horizons for Language-Guided Manipulation
von: Chen, Kewei, et al.
Veröffentlicht: (2026)
von: Chen, Kewei, et al.
Veröffentlicht: (2026)
To Whom are You Talking? A Deep Learning Model to Endow Social Robots with Addressee Estimation Skills
von: Mazzola, Carlo, et al.
Veröffentlicht: (2023)
von: Mazzola, Carlo, et al.
Veröffentlicht: (2023)
Rethinking Visual Intelligence: Insights from Video Pretraining
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025)
von: Acuaviva, Pablo, et al.
Veröffentlicht: (2025)
Dream to Fly: Model-Based Reinforcement Learning for Vision-Based Drone Flight
von: Romero, Angel, et al.
Veröffentlicht: (2025)
von: Romero, Angel, et al.
Veröffentlicht: (2025)
WayFASTER: a Self-Supervised Traversability Prediction for Increased Navigation Awareness
von: Gasparino, Mateus Valverde, et al.
Veröffentlicht: (2024)
von: Gasparino, Mateus Valverde, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Cooperative Perception: A Resource-Efficient Framework for Multi-Drone 3D Scene Reconstruction Using Federated Diffusion and NeRF
von: Pourmandi, Massoud
Veröffentlicht: (2025) -
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
von: Chahine, Makram, et al.
Veröffentlicht: (2024) -
Method of UAV Inspection of Photovoltaic Modules Using Thermal and RGB Data Fusion
von: Lysyi, Andrii, et al.
Veröffentlicht: (2025) -
CulinaryCut-VLAP: A Vision-Language-Action-Physics Framework for Food Cutting via a Force-Aware Material Point Method
von: Koh, Hyunseo, et al.
Veröffentlicht: (2026) -
FT-NCFM: An Influence-Aware Data Distillation Framework for Efficient VLA Models
von: Chen, Kewei, et al.
Veröffentlicht: (2025)