CulinaryCut-VLAP: A Vision-Language-Action-Physics Framework for Food Cutting via a Force-Aware Material Point Method
Fuente:
arXiv
Salvato in:
| Autori principali: | Koh, Hyunseo, Song, Chang-Yong, Choi, Youngjae, Viveiros, Misa, Hyde, David, Kim, Heewon |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
ATAAT: Adaptive Threat-Aware Adversarial Tuning Framework against Backdoor Attacks on Vision-Language-Action Models
di: Chen, Kewei, et al.
Pubblicazione: (2026)
di: Chen, Kewei, et al.
Pubblicazione: (2026)
ExpReS-VLA: Specializing Vision-Language-Action Models Through Experience Replay and Retrieval
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025)
FT-NCFM: An Influence-Aware Data Distillation Framework for Efficient VLA Models
di: Chen, Kewei, et al.
Pubblicazione: (2025)
di: Chen, Kewei, et al.
Pubblicazione: (2025)
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
di: Chahine, Makram, et al.
Pubblicazione: (2024)
di: Chahine, Makram, et al.
Pubblicazione: (2024)
A Survey on Vision-Language-Action Models for Embodied AI
di: Ma, Yueen, et al.
Pubblicazione: (2024)
di: Ma, Yueen, et al.
Pubblicazione: (2024)
A Landmark-Aware Visual Navigation Dataset
di: Johnson, Faith, et al.
Pubblicazione: (2024)
di: Johnson, Faith, et al.
Pubblicazione: (2024)
Think, Act, Learn: A Framework for Autonomous Robotic Agents using Closed-Loop Large Language Models
di: Menon, Anjali R., et al.
Pubblicazione: (2025)
di: Menon, Anjali R., et al.
Pubblicazione: (2025)
RoboMoRe: LLM-based Robot Co-design via Joint Optimization of Morphology and Reward
di: Fang, Jiawei, et al.
Pubblicazione: (2025)
di: Fang, Jiawei, et al.
Pubblicazione: (2025)
Inducing Causal World Models in LLMs for Zero-Shot Physical Reasoning
di: Sharma, Aditya, et al.
Pubblicazione: (2025)
di: Sharma, Aditya, et al.
Pubblicazione: (2025)
Method of UAV Inspection of Photovoltaic Modules Using Thermal and RGB Data Fusion
di: Lysyi, Andrii, et al.
Pubblicazione: (2025)
di: Lysyi, Andrii, et al.
Pubblicazione: (2025)
PathFormer: A Transformer with 3D Grid Constraints for Digital Twin Robot-Arm Trajectory Generation
di: Alanazi, Ahmed, et al.
Pubblicazione: (2025)
di: Alanazi, Ahmed, et al.
Pubblicazione: (2025)
Mitigating Covariate Shift in Imitation Learning for Autonomous Vehicles Using Latent Space Generative World Models
di: Popov, Alexander, et al.
Pubblicazione: (2024)
di: Popov, Alexander, et al.
Pubblicazione: (2024)
Cooperative Perception: A Resource-Efficient Framework for Multi-Drone 3D Scene Reconstruction Using Federated Diffusion and NeRF
di: Pourmandi, Massoud
Pubblicazione: (2025)
di: Pourmandi, Massoud
Pubblicazione: (2025)
To Whom are You Talking? A Deep Learning Model to Endow Social Robots with Addressee Estimation Skills
di: Mazzola, Carlo, et al.
Pubblicazione: (2023)
di: Mazzola, Carlo, et al.
Pubblicazione: (2023)
Evaluating Model-Agnostic Meta-Learning on MetaWorld ML10 Benchmark: Fast Adaptation in Robotic Manipulation Tasks
di: Atamuradov, Sanjar
Pubblicazione: (2025)
di: Atamuradov, Sanjar
Pubblicazione: (2025)
Tricks and Plug-ins for Gradient Boosting in Image Classification
di: Fang, Biyi, et al.
Pubblicazione: (2025)
di: Fang, Biyi, et al.
Pubblicazione: (2025)
Emergence of Goal-Directed Behaviors via Active Inference with Self-Prior
di: Kim, Dongmin, et al.
Pubblicazione: (2025)
di: Kim, Dongmin, et al.
Pubblicazione: (2025)
SemanticFeels: Semantic Labeling during In-Hand Manipulation
di: Khalil, Anas Al Shikh, et al.
Pubblicazione: (2026)
di: Khalil, Anas Al Shikh, et al.
Pubblicazione: (2026)
Multimodal Generative AI for Story Point Estimation in Software Development
di: Islam, Mohammad Rubyet, et al.
Pubblicazione: (2025)
di: Islam, Mohammad Rubyet, et al.
Pubblicazione: (2025)
Unpacking Hateful Memes: Presupposed Context and False Claims
di: Cai, Weibin, et al.
Pubblicazione: (2025)
di: Cai, Weibin, et al.
Pubblicazione: (2025)
Training for X-Ray Vision: Amodal Segmentation, Amodal Content Completion, and View-Invariant Object Representation from Multi-Camera Video
di: Moore, Alexander, et al.
Pubblicazione: (2025)
di: Moore, Alexander, et al.
Pubblicazione: (2025)
Agentic UAVs: LLM-Driven Autonomy with Integrated Tool-Calling and Cognitive Reasoning
di: Koubaa, Anis, et al.
Pubblicazione: (2025)
di: Koubaa, Anis, et al.
Pubblicazione: (2025)
RACAS: Controlling Diverse Robots With a Single Agentic System
di: Ashley, Dylan R., et al.
Pubblicazione: (2026)
di: Ashley, Dylan R., et al.
Pubblicazione: (2026)
Connectivity-Aware Representations for Constrained Motion Planning via Multi-Scale Contrastive Learning
di: Jeon, Suhyun, et al.
Pubblicazione: (2026)
di: Jeon, Suhyun, et al.
Pubblicazione: (2026)
ABot-Claw: A Foundation for Persistent, Cooperative, and Self-Evolving Robotic Agents
di: Huo, Dongjie, et al.
Pubblicazione: (2026)
di: Huo, Dongjie, et al.
Pubblicazione: (2026)
SoMA: A Real-to-Sim Neural Simulator for Robotic Soft-body Manipulation
di: Huang, Mu, et al.
Pubblicazione: (2026)
di: Huang, Mu, et al.
Pubblicazione: (2026)
LLM-Guided Task- and Affordance-Level Exploration in Reinforcement Learning
di: Luijkx, Jelle, et al.
Pubblicazione: (2025)
di: Luijkx, Jelle, et al.
Pubblicazione: (2025)
Active Inference with a Self-Prior in the Mirror-Mark Task
di: Kim, Dongmin, et al.
Pubblicazione: (2026)
di: Kim, Dongmin, et al.
Pubblicazione: (2026)
Autonomous Multi-Robot Infrastructure for AI-Enabled Healthcare Delivery and Diagnostics
di: Kalaivanan, Nakhul, et al.
Pubblicazione: (2025)
di: Kalaivanan, Nakhul, et al.
Pubblicazione: (2025)
Human-Aligned Skill Discovery: Balancing Behaviour Exploration and Alignment
di: Hussonnois, Maxence, et al.
Pubblicazione: (2025)
di: Hussonnois, Maxence, et al.
Pubblicazione: (2025)
Perception-Consistency Multimodal Large Language Models Reasoning via Caption-Regularized Policy Optimization
di: Tu, Songjun, et al.
Pubblicazione: (2025)
di: Tu, Songjun, et al.
Pubblicazione: (2025)
COBRA-PPM: A Causal Bayesian Reasoning Architecture Using Probabilistic Programming for Robot Manipulation Under Uncertainty
di: Cannizzaro, Ricardo, et al.
Pubblicazione: (2024)
di: Cannizzaro, Ricardo, et al.
Pubblicazione: (2024)
Training a Student Expert via Semi-Supervised Foundation Model Distillation
di: Taghavi, Pardis, et al.
Pubblicazione: (2026)
di: Taghavi, Pardis, et al.
Pubblicazione: (2026)
Semantically Guided Adversarial Testing of Vision Models Using Language Models
di: Filus, Katarzyna, et al.
Pubblicazione: (2025)
di: Filus, Katarzyna, et al.
Pubblicazione: (2025)
VA-$π$: Variational Policy Alignment for Pixel-Aware Autoregressive Generation
di: Liao, Xinyao, et al.
Pubblicazione: (2025)
di: Liao, Xinyao, et al.
Pubblicazione: (2025)
TensLoRA: Tensor Alternatives for Low-Rank Adaptation
di: Marmoret, Axel, et al.
Pubblicazione: (2025)
di: Marmoret, Axel, et al.
Pubblicazione: (2025)
Measuring What Matters: Scenario-Driven Evaluation for Trajectory Predictors in Autonomous Driving
di: Da, Longchao, et al.
Pubblicazione: (2025)
di: Da, Longchao, et al.
Pubblicazione: (2025)
Go Big or Go Home: Simulating Mobbing Behavior with Braitenbergian Robots
di: Sanoubari, Elaheh
Pubblicazione: (2026)
di: Sanoubari, Elaheh
Pubblicazione: (2026)
Spectral Integrated Gradients for Coarse-to-Fine Feature Attribution
di: Kim, Soyeon, et al.
Pubblicazione: (2026)
di: Kim, Soyeon, et al.
Pubblicazione: (2026)
Manifold-Aligned Guided Integrated Gradients for Reliable Feature Attribution
di: Kim, Soyeon, et al.
Pubblicazione: (2026)
di: Kim, Soyeon, et al.
Pubblicazione: (2026)
Documenti analoghi
-
ATAAT: Adaptive Threat-Aware Adversarial Tuning Framework against Backdoor Attacks on Vision-Language-Action Models
di: Chen, Kewei, et al.
Pubblicazione: (2026) -
ExpReS-VLA: Specializing Vision-Language-Action Models Through Experience Replay and Retrieval
di: Syed, Shahram Najam, et al.
Pubblicazione: (2025) -
FT-NCFM: An Influence-Aware Data Distillation Framework for Efficient VLA Models
di: Chen, Kewei, et al.
Pubblicazione: (2025) -
Flex: End-to-End Text-Instructed Visual Navigation from Foundation Model Features
di: Chahine, Makram, et al.
Pubblicazione: (2024) -
A Survey on Vision-Language-Action Models for Embodied AI
di: Ma, Yueen, et al.
Pubblicazione: (2024)