MLLM-Fabric: Multimodal Large Language Model-Driven Robotic Framework for Fabric Sorting and Selection
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Liman, Zhong, Hanyang, Wang, Tianyuan, Luo, Shan, Zhu, Jihong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LLM-SAP: Large Language Models Situational Awareness Based Planning
by: Wang, Liman, et al.
Published: (2023)
by: Wang, Liman, et al.
Published: (2023)
Sensorimotor Self-Recognition in Multimodal Large Language Model-Driven Robots
by: Varela, Iñaki Dellibarda, et al.
Published: (2025)
by: Varela, Iñaki Dellibarda, et al.
Published: (2025)
MLLM-Search: A Zero-Shot Approach to Finding People using Multimodal Large Language Models
by: Fung, Angus, et al.
Published: (2024)
by: Fung, Angus, et al.
Published: (2024)
Balancing Rigor and Utility: Mitigating Cognitive Biases in Large Language Models for Multiple-Choice Questions
by: Zhong, Hanyang, et al.
Published: (2024)
by: Zhong, Hanyang, et al.
Published: (2024)
FENet: Focusing Enhanced Network for Lane Detection
by: Wang, Liman, et al.
Published: (2023)
by: Wang, Liman, et al.
Published: (2023)
GSON: A Group-based Social Navigation Framework with Large Multimodal Model
by: Luo, Shangyi, et al.
Published: (2024)
by: Luo, Shangyi, et al.
Published: (2024)
Multimodal Large Language Model Driven Scenario Testing for Autonomous Vehicles
by: Lu, Qiujing, et al.
Published: (2024)
by: Lu, Qiujing, et al.
Published: (2024)
Large Language Models for Robotics: A Survey
by: Zeng, Fanlong, et al.
Published: (2023)
by: Zeng, Fanlong, et al.
Published: (2023)
AIC MLLM: Autonomous Interactive Correction MLLM for Robust Robotic Manipulation
by: Xiong, Chuyan, et al.
Published: (2024)
by: Xiong, Chuyan, et al.
Published: (2024)
COHERENT: Collaboration of Heterogeneous Multi-Robot System with Large Language Models
by: Liu, Kehui, et al.
Published: (2024)
by: Liu, Kehui, et al.
Published: (2024)
DeeR-VLA: Dynamic Inference of Multimodal Large Language Models for Efficient Robot Execution
by: Yue, Yang, et al.
Published: (2024)
by: Yue, Yang, et al.
Published: (2024)
OpenNav: Open-World Navigation with Multimodal Large Language Models
by: Yuan, Mingfeng, et al.
Published: (2025)
by: Yuan, Mingfeng, et al.
Published: (2025)
Hybrid Framework for Robotic Manipulation: Integrating Reinforcement Learning and Large Language Models
by: Saad, Md, et al.
Published: (2026)
by: Saad, Md, et al.
Published: (2026)
Kaiwu: A Multimodal Manipulation Dataset and Framework for Robot Learning and Human-Robot Interaction
by: Jiang, Shuo, et al.
Published: (2025)
by: Jiang, Shuo, et al.
Published: (2025)
Large Language Models for Robotics: Opportunities, Challenges, and Perspectives
by: Wang, Jiaqi, et al.
Published: (2024)
by: Wang, Jiaqi, et al.
Published: (2024)
LADEV: A Language-Driven Testing and Evaluation Platform for Vision-Language-Action Models in Robotic Manipulation
by: Wang, Zhijie, et al.
Published: (2024)
by: Wang, Zhijie, et al.
Published: (2024)
SayCoNav: Utilizing Large Language Models for Adaptive Collaboration in Decentralized Multi-Robot Navigation
by: Rajvanshi, Abhinav, et al.
Published: (2025)
by: Rajvanshi, Abhinav, et al.
Published: (2025)
GameVLM: A Decision-making Framework for Robotic Task Planning Based on Visual Language Models and Zero-sum Games
by: Mei, Aoran, et al.
Published: (2024)
by: Mei, Aoran, et al.
Published: (2024)
Realistic Corner Case Generation for Autonomous Vehicles with Multimodal Large Language Model
by: Lu, Qiujing, et al.
Published: (2024)
by: Lu, Qiujing, et al.
Published: (2024)
Large Reward Models: Generalizable Online Robot Reward Generation with Vision-Language Models
by: Wu, Yanru, et al.
Published: (2026)
by: Wu, Yanru, et al.
Published: (2026)
Large Language Models for Orchestrating Bimanual Robots
by: Chu, Kun, et al.
Published: (2024)
by: Chu, Kun, et al.
Published: (2024)
Exploring the Adversarial Vulnerabilities of Vision-Language-Action Models in Robotics
by: Wang, Taowen, et al.
Published: (2024)
by: Wang, Taowen, et al.
Published: (2024)
CoPAL: Corrective Planning of Robot Actions with Large Language Models
by: Joublin, Frank, et al.
Published: (2023)
by: Joublin, Frank, et al.
Published: (2023)
Proposition of Affordance-Driven Environment Recognition Framework Using Symbol Networks in Large Language Models
by: Arii, Kazuma, et al.
Published: (2025)
by: Arii, Kazuma, et al.
Published: (2025)
Enhancing Robotic Manipulation with AI Feedback from Multimodal Large Language Models
by: Liu, Jinyi, et al.
Published: (2024)
by: Liu, Jinyi, et al.
Published: (2024)
Natural Selection via Foundation Models for Soft Robot Evolution
by: Chen, Changhe, et al.
Published: (2025)
by: Chen, Changhe, et al.
Published: (2025)
DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation
by: Su, Taiyi, et al.
Published: (2026)
by: Su, Taiyi, et al.
Published: (2026)
FCRF: Flexible Constructivism Reflection for Long-Horizon Robotic Task Planning with Large Language Models
by: Song, Yufan, et al.
Published: (2025)
by: Song, Yufan, et al.
Published: (2025)
AnyBipe: An End-to-End Framework for Training and Deploying Bipedal Robots Guided by Large Language Models
by: Yao, Yifei, et al.
Published: (2024)
by: Yao, Yifei, et al.
Published: (2024)
Towards Interactive and Learnable Cooperative Driving Automation: a Large Language Model-Driven Decision-Making Framework
by: Fang, Shiyu, et al.
Published: (2024)
by: Fang, Shiyu, et al.
Published: (2024)
Large Language Models for Multi-Robot Systems: A Survey
by: Li, Peihan, et al.
Published: (2025)
by: Li, Peihan, et al.
Published: (2025)
Multi-Scenario Reasoning: Unlocking Cognitive Autonomy in Humanoid Robots for Multimodal Understanding
by: Wang, Libo
Published: (2024)
by: Wang, Libo
Published: (2024)
RoboAlign-R1: Distilled Multimodal Reward Alignment for Robot Video World Models
by: Wu, Hao, et al.
Published: (2026)
by: Wu, Hao, et al.
Published: (2026)
Destination-to-Chutes Task Mapping Optimization for Multi-Robot Coordination in Robotic Sorting Systems
by: Zhang, Yulun, et al.
Published: (2025)
by: Zhang, Yulun, et al.
Published: (2025)
CubeRobot: Grounding Language in Rubik's Cube Manipulation via Vision-Language Model
by: Wang, Feiyang, et al.
Published: (2025)
by: Wang, Feiyang, et al.
Published: (2025)
Safety Aware Task Planning via Large Language Models in Robotics
by: Khan, Azal Ahmad, et al.
Published: (2025)
by: Khan, Azal Ahmad, et al.
Published: (2025)
INGRID: Intelligent Generative Robotic Design Using Large Language Models
by: Jia, Guanglu, et al.
Published: (2025)
by: Jia, Guanglu, et al.
Published: (2025)
Manual2Skill++: Connector-Aware General Robotic Assembly from Instruction Manuals via Vision-Language Models
by: Tie, Chenrui, et al.
Published: (2025)
by: Tie, Chenrui, et al.
Published: (2025)
Topology-Driven Anti-Entanglement Control for Soft Robots
by: Le, Haoyang, et al.
Published: (2026)
by: Le, Haoyang, et al.
Published: (2026)
STaR: Scalable Task-Conditioned Retrieval for Long-Horizon Multimodal Robot Memory
by: Yuan, Mingfeng, et al.
Published: (2026)
by: Yuan, Mingfeng, et al.
Published: (2026)
Similar Items
-
LLM-SAP: Large Language Models Situational Awareness Based Planning
by: Wang, Liman, et al.
Published: (2023) -
Sensorimotor Self-Recognition in Multimodal Large Language Model-Driven Robots
by: Varela, Iñaki Dellibarda, et al.
Published: (2025) -
MLLM-Search: A Zero-Shot Approach to Finding People using Multimodal Large Language Models
by: Fung, Angus, et al.
Published: (2024) -
Balancing Rigor and Utility: Mitigating Cognitive Biases in Large Language Models for Multiple-Choice Questions
by: Zhong, Hanyang, et al.
Published: (2024) -
FENet: Focusing Enhanced Network for Lane Detection
by: Wang, Liman, et al.
Published: (2023)