Leveraging Multimodal-LLMs Assisted by Instance Segmentation for Intelligent Traffic Monitoring
Fuente:
arXiv
Salvato in:
| Autori principali: | Onsu, Murat Arda, Lohan, Poonam, Kantarci, Burak, Syed, Aisha, Andrews, Matthew, Kennedy, Sean |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Leveraging Edge Intelligence and LLMs to Advance 6G-Enabled Internet of Automated Defense Vehicles
di: Onsu, Murat Arda, et al.
Pubblicazione: (2024)
di: Onsu, Murat Arda, et al.
Pubblicazione: (2024)
A Lightweight Digital-Twin-Based Framework for Edge-Assisted Vehicle Tracking and Collision Prediction
di: Onsu, Murat Arda, et al.
Pubblicazione: (2026)
di: Onsu, Murat Arda, et al.
Pubblicazione: (2026)
Spatiotemporal Semantic V2X Framework for Cooperative Collision Prediction
di: Onsu, Murat Arda, et al.
Pubblicazione: (2026)
di: Onsu, Murat Arda, et al.
Pubblicazione: (2026)
Semantic Edge-Cloud Communication for Real-Time Urban Traffic Surveillance with ViT and LLMs over Mobile Networks
di: Onsu, Murat Arda, et al.
Pubblicazione: (2025)
di: Onsu, Murat Arda, et al.
Pubblicazione: (2025)
Integrating Language Models for Enhanced Network State Monitoring in DRL-Based SFC Provisioning
di: Moshiri, Parisa Fard, et al.
Pubblicazione: (2025)
di: Moshiri, Parisa Fard, et al.
Pubblicazione: (2025)
GenAI Assistance for Deep Reinforcement Learning-based VNF Placement and SFC Provisioning in 5G Cores
di: Onsu, Murat Arda, et al.
Pubblicazione: (2024)
di: Onsu, Murat Arda, et al.
Pubblicazione: (2024)
Scalability Assurance in SFC provisioning via Distributed Design for Deep Reinforcement Learning
di: Onsu, Murat Arda, et al.
Pubblicazione: (2024)
di: Onsu, Murat Arda, et al.
Pubblicazione: (2024)
A New Realistic Platform for Benchmarking and Performance Evaluation of DRL-Driven and Reconfigurable SFC Provisioning Solutions
di: Onsu, Murat Arda, et al.
Pubblicazione: (2024)
di: Onsu, Murat Arda, et al.
Pubblicazione: (2024)
Structure-Aware NL-to-SQL for SFC Provisioning via AST-Masking Empowered Language Models
di: Zhu, Xinyu, et al.
Pubblicazione: (2026)
di: Zhu, Xinyu, et al.
Pubblicazione: (2026)
LiLM-RDB-SFC: Lightweight Language Model with Relational Database-Guided DRL for Optimized SFC Provisioning
di: Moshiri, Parisa Fard, et al.
Pubblicazione: (2025)
di: Moshiri, Parisa Fard, et al.
Pubblicazione: (2025)
Intent2QoS: Language Model-Driven Automation of Traffic Shaping Configurations
di: Acharya, Sudipta, et al.
Pubblicazione: (2026)
di: Acharya, Sudipta, et al.
Pubblicazione: (2026)
TrafficLens: Multi-Camera Traffic Video Analysis Using LLMs
di: Arefeen, Md Adnan, et al.
Pubblicazione: (2025)
di: Arefeen, Md Adnan, et al.
Pubblicazione: (2025)
Rethinking the Mixture of Vision Encoders Paradigm for Enhanced Visual Understanding in Multimodal LLMs
di: Azadani, Mozhgan Nasr, et al.
Pubblicazione: (2025)
di: Azadani, Mozhgan Nasr, et al.
Pubblicazione: (2025)
Phrase-Instance Alignment for Generalized Referring Segmentation
di: Nguyen, E-Ro, et al.
Pubblicazione: (2024)
di: Nguyen, E-Ro, et al.
Pubblicazione: (2024)
Holistic Evaluation of Multimodal LLMs on Spatial Intelligence
di: Cai, Zhongang, et al.
Pubblicazione: (2025)
di: Cai, Zhongang, et al.
Pubblicazione: (2025)
Multi-Agent Deep Reinforcement Learning for Optimized Multi-UAV Coverage and Power-Efficient UE Connectivity
di: Cai, Xuli, et al.
Pubblicazione: (2025)
di: Cai, Xuli, et al.
Pubblicazione: (2025)
A Novel Joint DRL-Based Utility Optimization for UAV Data Services
di: Cai, Xuli, et al.
Pubblicazione: (2024)
di: Cai, Xuli, et al.
Pubblicazione: (2024)
FLARE: Flying Learning Agents for Resource Efficiency in Next-Gen UAV Networks
di: Cai, Xuli, et al.
Pubblicazione: (2025)
di: Cai, Xuli, et al.
Pubblicazione: (2025)
Rethinking Detection Based Table Structure Recognition for Visually Rich Document Images
di: Xiao, Bin, et al.
Pubblicazione: (2023)
di: Xiao, Bin, et al.
Pubblicazione: (2023)
Pre-Training Multimodal Hallucination Detectors with Corrupted Grounding Data
di: Whitehead, Spencer, et al.
Pubblicazione: (2024)
di: Whitehead, Spencer, et al.
Pubblicazione: (2024)
Multimodal Large Language Models for Enhanced Traffic Safety: A Comprehensive Review and Future Trends
di: Tami, Mohammad Abu, et al.
Pubblicazione: (2025)
di: Tami, Mohammad Abu, et al.
Pubblicazione: (2025)
PersonaVLM: Long-Term Personalized Multimodal LLMs
di: Nie, Chang, et al.
Pubblicazione: (2026)
di: Nie, Chang, et al.
Pubblicazione: (2026)
Chitranuvad: Adapting Multi-Lingual LLMs for Multimodal Translation
di: Khan, Shaharukh, et al.
Pubblicazione: (2025)
di: Khan, Shaharukh, et al.
Pubblicazione: (2025)
EmbodiedEval: Evaluate Multimodal LLMs as Embodied Agents
di: Cheng, Zhili, et al.
Pubblicazione: (2025)
di: Cheng, Zhili, et al.
Pubblicazione: (2025)
Understanding Alignment in Multimodal LLMs: A Comprehensive Study
di: Amirloo, Elmira, et al.
Pubblicazione: (2024)
di: Amirloo, Elmira, et al.
Pubblicazione: (2024)
Clinical Cognition Alignment for Gastrointestinal Diagnosis with Multimodal LLMs
di: Zheng, Huan, et al.
Pubblicazione: (2026)
di: Zheng, Huan, et al.
Pubblicazione: (2026)
Leveraging Entity Information for Cross-Modality Correlation Learning: The Entity-Guided Multimodal Summarization
di: Zhang, Yanghai, et al.
Pubblicazione: (2024)
di: Zhang, Yanghai, et al.
Pubblicazione: (2024)
Leveraging LLMs for On-the-Fly Instruction Guided Image Editing
di: Santos, Rodrigo, et al.
Pubblicazione: (2024)
di: Santos, Rodrigo, et al.
Pubblicazione: (2024)
Catch Me If You Can Describe Me: Open-Vocabulary Camouflaged Instance Segmentation with Diffusion
di: Vu, Tuan-Anh, et al.
Pubblicazione: (2023)
di: Vu, Tuan-Anh, et al.
Pubblicazione: (2023)
LLMs as Bridges: Reformulating Grounded Multimodal Named Entity Recognition
di: Li, Jinyuan, et al.
Pubblicazione: (2024)
di: Li, Jinyuan, et al.
Pubblicazione: (2024)
Leveraging Large Language Models (LLMs) for Traffic Management at Urban Intersections: The Case of Mixed Traffic Scenarios
di: Masri, Sari, et al.
Pubblicazione: (2024)
di: Masri, Sari, et al.
Pubblicazione: (2024)
Advancing Autonomous Vehicle Intelligence: Deep Learning and Multimodal LLM for Traffic Sign Recognition and Robust Lane Detection
di: Sah, Chandan Kumar, et al.
Pubblicazione: (2025)
di: Sah, Chandan Kumar, et al.
Pubblicazione: (2025)
Implicit Multimodal Alignment: On the Generalization of Frozen LLMs to Multimodal Inputs
di: Shukor, Mustafa, et al.
Pubblicazione: (2024)
di: Shukor, Mustafa, et al.
Pubblicazione: (2024)
Insight Over Sight: Exploring the Vision-Knowledge Conflicts in Multimodal LLMs
di: Liu, Xiaoyuan, et al.
Pubblicazione: (2024)
di: Liu, Xiaoyuan, et al.
Pubblicazione: (2024)
Can Multimodal LLMs do Visual Temporal Understanding and Reasoning? The answer is No!
di: Imam, Mohamed Fazli, et al.
Pubblicazione: (2025)
di: Imam, Mohamed Fazli, et al.
Pubblicazione: (2025)
Evaluating Linguistic Capabilities of Multimodal LLMs in the Lens of Few-Shot Learning
di: Dogan, Mustafa, et al.
Pubblicazione: (2024)
di: Dogan, Mustafa, et al.
Pubblicazione: (2024)
MIA-Bench: Towards Better Instruction Following Evaluation of Multimodal LLMs
di: Qian, Yusu, et al.
Pubblicazione: (2024)
di: Qian, Yusu, et al.
Pubblicazione: (2024)
ViCA: Efficient Multimodal LLMs with Vision-Only Cross-Attention
di: Liu, Wenjie, et al.
Pubblicazione: (2026)
di: Liu, Wenjie, et al.
Pubblicazione: (2026)
CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs
di: Wang, Zirui, et al.
Pubblicazione: (2024)
di: Wang, Zirui, et al.
Pubblicazione: (2024)
Vision Enhancing LLMs: Empowering Multimodal Knowledge Storage and Sharing in LLMs
di: Li, Yunxin, et al.
Pubblicazione: (2023)
di: Li, Yunxin, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Leveraging Edge Intelligence and LLMs to Advance 6G-Enabled Internet of Automated Defense Vehicles
di: Onsu, Murat Arda, et al.
Pubblicazione: (2024) -
A Lightweight Digital-Twin-Based Framework for Edge-Assisted Vehicle Tracking and Collision Prediction
di: Onsu, Murat Arda, et al.
Pubblicazione: (2026) -
Spatiotemporal Semantic V2X Framework for Cooperative Collision Prediction
di: Onsu, Murat Arda, et al.
Pubblicazione: (2026) -
Semantic Edge-Cloud Communication for Real-Time Urban Traffic Surveillance with ViT and LLMs over Mobile Networks
di: Onsu, Murat Arda, et al.
Pubblicazione: (2025) -
Integrating Language Models for Enhanced Network State Monitoring in DRL-Based SFC Provisioning
di: Moshiri, Parisa Fard, et al.
Pubblicazione: (2025)