Visual Reasoning and Multi-Agent Approach in Multimodal Large Language Models (MLLMs): Solving TSP and mTSP Combinatorial Challenges
Fuente:
arXiv
Salvato in:
| Autori principali: | Elhenawy, Mohammed, Abutahoun, Ahmad, Alhadidi, Taqwa I., Jaber, Ahmed, Ashqar, Huthaifa I., Jaradat, Shadi, Abdelhay, Ahmed, Glaser, Sebastien, Rakotonirainy, Andry |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Eyeballing Combinatorial Problems: A Case Study of Using Multimodal Large Language Models to Solve Traveling Salesman Problems
di: Elhenawy, Mohammed, et al.
Pubblicazione: (2024)
di: Elhenawy, Mohammed, et al.
Pubblicazione: (2024)
Object Detection using Oriented Window Learning Vi-sion Transformer: Roadway Assets Recognition
di: Alhadidi, Taqwa, et al.
Pubblicazione: (2024)
di: Alhadidi, Taqwa, et al.
Pubblicazione: (2024)
Zero-Shot Scene Understanding with Multimodal Large Language Models for Automated Vehicles
di: Elhenawy, Mohammed, et al.
Pubblicazione: (2025)
di: Elhenawy, Mohammed, et al.
Pubblicazione: (2025)
Advancing Object Detection in Transportation with Multimodal Large Language Models (MLLMs): A Comprehensive Review and Empirical Testing
di: Ashqar, Huthaifa I., et al.
Pubblicazione: (2024)
di: Ashqar, Huthaifa I., et al.
Pubblicazione: (2024)
Enhancing Pavement Crack Classification with Bidirectional Cascaded Neural Networks
di: Alhadidi, Taqwa I., et al.
Pubblicazione: (2025)
di: Alhadidi, Taqwa I., et al.
Pubblicazione: (2025)
Vision-Language Models for Autonomous Driving: CLIP-Based Dynamic Scene Understanding
di: Elhenawy, Mohammed, et al.
Pubblicazione: (2025)
di: Elhenawy, Mohammed, et al.
Pubblicazione: (2025)
Exploring Traffic Crash Narratives in Jordan Using Text Mining Analytics
di: Jaradat, Shadi, et al.
Pubblicazione: (2024)
di: Jaradat, Shadi, et al.
Pubblicazione: (2024)
The Use of Multimodal Large Language Models to Detect Objects from Thermal Images: Transportation Applications
di: Ashqar, Huthaifa I., et al.
Pubblicazione: (2024)
di: Ashqar, Huthaifa I., et al.
Pubblicazione: (2024)
Transformer Models in Education: Summarizing Science Textbooks with AraBART, MT5, AraT5, and mBART
di: Masri, Sari, et al.
Pubblicazione: (2024)
di: Masri, Sari, et al.
Pubblicazione: (2024)
Capability-Priced Micro-Markets: A Micro-Economic Framework for the Agentic Web over HTTP 402
di: Huang, Ken, et al.
Pubblicazione: (2026)
di: Huang, Ken, et al.
Pubblicazione: (2026)
Trust-MARL: Trust-Based Multi-Agent Reinforcement Learning Framework for Cooperative On-Ramp Merging Control in Heterogeneous Traffic Flow
di: Pan, Jie, et al.
Pubblicazione: (2025)
di: Pan, Jie, et al.
Pubblicazione: (2025)
A Gossip-Enhanced Communication Substrate for Agentic AI: Toward Decentralized Coordination in Large-Scale Multi-Agent Systems
di: Khan, Nafiul I., et al.
Pubblicazione: (2025)
di: Khan, Nafiul I., et al.
Pubblicazione: (2025)
Mathematical Modelling of Ethical AI Use in Higher Education: A Coordination Game Framework for Future-Facing Learning
di: Ogbo, Ndidi Bianca, et al.
Pubblicazione: (2026)
di: Ogbo, Ndidi Bianca, et al.
Pubblicazione: (2026)
Decentralized Optimal Equilibrium Learning in Stochastic Games via Single-bit Feedback
di: Kiremitci, Seref Taha, et al.
Pubblicazione: (2026)
di: Kiremitci, Seref Taha, et al.
Pubblicazione: (2026)
Exploring Combinatorial Problem Solving with Large Language Models: A Case Study on the Travelling Salesman Problem Using GPT-3.5 Turbo
di: Masoud, Mahmoud, et al.
Pubblicazione: (2024)
di: Masoud, Mahmoud, et al.
Pubblicazione: (2024)
Urban Air Mobility as a System of Systems: An LLM-Enhanced Holonic Approach
di: Sadik, Ahmed R., et al.
Pubblicazione: (2025)
di: Sadik, Ahmed R., et al.
Pubblicazione: (2025)
What Do Agents Think One Another Want? Level-2 Inverse Games for Inferring Agents' Estimates of Others' Objectives
di: Khan, Hamzah I., et al.
Pubblicazione: (2025)
di: Khan, Hamzah I., et al.
Pubblicazione: (2025)
A Hybrid Classical-Quantum Annealing Algorithm for the TSP
di: Hu, Siwei, et al.
Pubblicazione: (2026)
di: Hu, Siwei, et al.
Pubblicazione: (2026)
Modular Autonomous Vehicle in Heterogeneous Traffic Flow: Modeling, Simulation, and Implication
di: Ye, Lanhang, et al.
Pubblicazione: (2024)
di: Ye, Lanhang, et al.
Pubblicazione: (2024)
AGENTSAFE: A Unified Framework for Ethical Assurance and Governance in Agentic AI
di: Khan, Rafflesia, et al.
Pubblicazione: (2025)
di: Khan, Rafflesia, et al.
Pubblicazione: (2025)
Q-RESTORE: Quantum-Driven Framework for Resilient and Equitable Transportation Network Restoration
di: Udekwe, Daniel, et al.
Pubblicazione: (2025)
di: Udekwe, Daniel, et al.
Pubblicazione: (2025)
Content Caching-Assisted Vehicular Edge Computing Using Multi-Agent Graph Attention Reinforcement Learning
di: Shen, Jinjin, et al.
Pubblicazione: (2024)
di: Shen, Jinjin, et al.
Pubblicazione: (2024)
Understanding Iterative Combinatorial Auction Designs via Multi-Agent Reinforcement Learning
di: d'Eon, Greg, et al.
Pubblicazione: (2024)
di: d'Eon, Greg, et al.
Pubblicazione: (2024)
LLM-Ehnanced Holonic Architecture for Ad-Hoc Scalable SoS
di: Ashfaq, Muhammad, et al.
Pubblicazione: (2025)
di: Ashfaq, Muhammad, et al.
Pubblicazione: (2025)
Synergistic Simulations: Multi-Agent Problem Solving with Large Language Models
di: Sprigler, Asher, et al.
Pubblicazione: (2024)
di: Sprigler, Asher, et al.
Pubblicazione: (2024)
Fusion Intelligence: Confluence of Natural and Artificial Intelligence for Enhanced Problem-Solving Efficiency
di: Kalavakonda, Rohan Reddy, et al.
Pubblicazione: (2024)
di: Kalavakonda, Rohan Reddy, et al.
Pubblicazione: (2024)
On the Power of Perturbation under Sampling in Solving Extensive-Form Games
di: Masaka, Wataru, et al.
Pubblicazione: (2025)
di: Masaka, Wataru, et al.
Pubblicazione: (2025)
SciDataCopilot: An Agentic Data Preparation Framework for AGI-driven Scientific Discovery
di: Rao, Jiyong, et al.
Pubblicazione: (2026)
di: Rao, Jiyong, et al.
Pubblicazione: (2026)
Large Language Models (LLMs) as Traffic Control Systems at Urban Intersections: A New Paradigm
di: Masri, Sari, et al.
Pubblicazione: (2024)
di: Masri, Sari, et al.
Pubblicazione: (2024)
Automated Question Generation for Science Tests in Arabic Language Using NLP Techniques
di: Tami, Mohammad, et al.
Pubblicazione: (2024)
di: Tami, Mohammad, et al.
Pubblicazione: (2024)
Leveraging Large Language Models (LLMs) for Traffic Management at Urban Intersections: The Case of Mixed Traffic Scenarios
di: Masri, Sari, et al.
Pubblicazione: (2024)
di: Masri, Sari, et al.
Pubblicazione: (2024)
Visual Reasoning at Urban Intersections: FineTuning GPT-4o for Traffic Conflict Detection
di: Masri, Sari, et al.
Pubblicazione: (2025)
di: Masri, Sari, et al.
Pubblicazione: (2025)
Words are not Wind -- How Public Joint Commitment and Reputation Solve the Prisoner's Dilemma
di: Krellner, Marcus, et al.
Pubblicazione: (2023)
di: Krellner, Marcus, et al.
Pubblicazione: (2023)
Solving Zero-Sum Convex Markov Games
di: Kalogiannis, Fivos, et al.
Pubblicazione: (2025)
di: Kalogiannis, Fivos, et al.
Pubblicazione: (2025)
SMEVCA: Stable Matching-based EV Charging Assignment in Subscription-Based Models
di: Khanda, Arindam, et al.
Pubblicazione: (2024)
di: Khanda, Arindam, et al.
Pubblicazione: (2024)
Optimal Energy-Aware Service Management in Future Networks with a Gamified Incentives Mechanism
di: Varsos, Konstantinos, et al.
Pubblicazione: (2026)
di: Varsos, Konstantinos, et al.
Pubblicazione: (2026)
MI9: An Integrated Runtime Governance Framework for Agentic AI
di: Wang, Charles L., et al.
Pubblicazione: (2025)
di: Wang, Charles L., et al.
Pubblicazione: (2025)
Autonomous Traffic Signal Optimization Using Digital Twin and Agentic AI for Real-Time Decision-Making
di: Jan, Salman, et al.
Pubblicazione: (2026)
di: Jan, Salman, et al.
Pubblicazione: (2026)
A Replica for our Democracies? On Using Digital Twins to Enhance Deliberative Democracy
di: Novelli, Claudio, et al.
Pubblicazione: (2025)
di: Novelli, Claudio, et al.
Pubblicazione: (2025)
One Step is Enough: Multi-Agent Reinforcement Learning based on One-Step Policy Optimization for Order Dispatch on Ride-Sharing Platforms
di: Zhao, Zijian, et al.
Pubblicazione: (2025)
di: Zhao, Zijian, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Eyeballing Combinatorial Problems: A Case Study of Using Multimodal Large Language Models to Solve Traveling Salesman Problems
di: Elhenawy, Mohammed, et al.
Pubblicazione: (2024) -
Object Detection using Oriented Window Learning Vi-sion Transformer: Roadway Assets Recognition
di: Alhadidi, Taqwa, et al.
Pubblicazione: (2024) -
Zero-Shot Scene Understanding with Multimodal Large Language Models for Automated Vehicles
di: Elhenawy, Mohammed, et al.
Pubblicazione: (2025) -
Advancing Object Detection in Transportation with Multimodal Large Language Models (MLLMs): A Comprehensive Review and Empirical Testing
di: Ashqar, Huthaifa I., et al.
Pubblicazione: (2024) -
Enhancing Pavement Crack Classification with Bidirectional Cascaded Neural Networks
di: Alhadidi, Taqwa I., et al.
Pubblicazione: (2025)