All You Need for Object Detection: From Pixels, Points, and Prompts to Next-Gen Fusion and Multimodal LLMs/VLMs in Autonomous Vehicles
Fuente:
arXiv
Saved in:
| Main Authors: | Boroujeni, Sayed Pedram Haeri, Mehrabi, Niloufar, Alzorgan, Hazim, Fazeli, Mahlagha, Razi, Abolfazl |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Meta-Adaptive Beam Search Planning for Transformer-Based Reinforcement Learning Control of UAVs with Overhead Manipulators under Flight Disturbances
by: Alzorgan, Hazim, et al.
Published: (2026)
by: Alzorgan, Hazim, et al.
Published: (2026)
Enhanced Cooperative Perception for Autonomous Vehicles Using Imperfect Communication
by: Sarlak, Ahmad, et al.
Published: (2024)
by: Sarlak, Ahmad, et al.
Published: (2024)
From Talking Words to Sharing Thoughts: Scalable Multi-LLM Aggregation via Structured Message Passing
by: Mehrabi, Niloufar, et al.
Published: (2026)
by: Mehrabi, Niloufar, et al.
Published: (2026)
Don't Waste Bits! Adaptive KV-Cache Quantization for Lightweight On-Device LLMs
by: Boroujeni, Sayed Pedram Haeri, et al.
Published: (2026)
by: Boroujeni, Sayed Pedram Haeri, et al.
Published: (2026)
Monte Carlo Beam Search for Actor-Critic Reinforcement Learning in Continuous Control
by: Alzorgan, Hazim, et al.
Published: (2025)
by: Alzorgan, Hazim, et al.
Published: (2025)
Integrating Random Regret Minimization-Based Discrete Choice Models with Mixed Integer Linear Programming for Revenue Optimization
by: Talebi, Amirreza, et al.
Published: (2024)
by: Talebi, Amirreza, et al.
Published: (2024)
Opinion Dynamics in Social Multiplex Networks with Mono and Bi-directional Interactions in the Presence of Leaders
by: Talebi, Amirreza, et al.
Published: (2024)
by: Talebi, Amirreza, et al.
Published: (2024)
Driving Towards Inclusion: A Systematic Review of AI-powered Accessibility Enhancements for People with Disability in Autonomous Vehicles
by: Bastola, Ashish, et al.
Published: (2024)
by: Bastola, Ashish, et al.
Published: (2024)
Adaptive Data Transport Mechanism for UAV Surveillance Missions in Lossy Environments
by: Mehrabi, Niloufar, et al.
Published: (2024)
by: Mehrabi, Niloufar, et al.
Published: (2024)
Eyes on the Environment: AI-Driven Analysis for Fire and Smoke Classification, Segmentation, and Detection
by: Boroujeni, Sayed Pedram Haeri, et al.
Published: (2025)
by: Boroujeni, Sayed Pedram Haeri, et al.
Published: (2025)
Enhancing Graph Neural Networks in Large-scale Traffic Incident Analysis with Concurrency Hypothesis
by: Chen, Xiwen, et al.
Published: (2024)
by: Chen, Xiwen, et al.
Published: (2024)
From Shadow to Light: Toward Safe and Efficient Policy Learning Across MPC, DeePC, RL, and LLM Agents
by: Vahidi-Moghaddam, Amin, et al.
Published: (2025)
by: Vahidi-Moghaddam, Amin, et al.
Published: (2025)
FLAME Diffuser: Wildfire Image Synthesis using Mask Guided Diffusion
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
Spatial Optimization of Autonomous Vehicle Assignment Based on Distance-Driven Demand and Customer Patience
by: Boroujeni, Niloufar Mirzavand, et al.
Published: (2025)
by: Boroujeni, Niloufar Mirzavand, et al.
Published: (2025)
Extended Visibility of Autonomous Vehicles via Optimized Cooperative Perception under Imperfect Communication
by: Sarlak, Ahmad, et al.
Published: (2025)
by: Sarlak, Ahmad, et al.
Published: (2025)
A comprehensive survey of research towards AI-enabled unmanned aerial systems in pre-, active-, and post-wildfire management
by: Boroujeni, Sayed Pedram Haeri, et al.
Published: (2024)
by: Boroujeni, Sayed Pedram Haeri, et al.
Published: (2024)
Graph Based Deep Reinforcement Learning Aided by Transformers for Multi-Agent Cooperation
by: Elrod, Michael, et al.
Published: (2025)
by: Elrod, Michael, et al.
Published: (2025)
Reliability-Aware Determinantal Point Processes for Robust Informative Data Selection in Large Language Models
by: Sarlak, Ahmad, et al.
Published: (2026)
by: Sarlak, Ahmad, et al.
Published: (2026)
An Improved Metaheuristic Algorithm for On-site Workshop Availability Cost Problem
by: Boroujeni, Niloufar Mirzavand, et al.
Published: (2025)
by: Boroujeni, Niloufar Mirzavand, et al.
Published: (2025)
Two-echelon Electric Vehicle Routing Problem in Parcel Delivery: A Literature Review
by: Moradi, Nima, et al.
Published: (2024)
by: Moradi, Nima, et al.
Published: (2024)
Anomaly Detection in Cooperative Vehicle Perception Systems under Imperfect Communication
by: Bastola, Ashish, et al.
Published: (2025)
by: Bastola, Ashish, et al.
Published: (2025)
Large Language Models for Lossless Image Compression: Next-Pixel Prediction in Language Space is All You Need
by: Chen, Kecheng, et al.
Published: (2024)
by: Chen, Kecheng, et al.
Published: (2024)
Comparative Analysis of Patch Attack on VLM-Based Autonomous Driving Architectures
by: Fernandez, David, et al.
Published: (2026)
by: Fernandez, David, et al.
Published: (2026)
Emu3: Next-Token Prediction is All You Need
by: Wang, Xinlong, et al.
Published: (2024)
by: Wang, Xinlong, et al.
Published: (2024)
Is Discretization Fusion All You Need for Collaborative Perception?
by: Yang, Kang, et al.
Published: (2025)
by: Yang, Kang, et al.
Published: (2025)
Fusion or Confusion? Multimodal Complexity Is Not All You Need
by: Rheude, Tillmann, et al.
Published: (2025)
by: Rheude, Tillmann, et al.
Published: (2025)
Decentralized Signaling Mechanisms
by: Boroujeni, Niloufar Mirzavand, et al.
Published: (2025)
by: Boroujeni, Niloufar Mirzavand, et al.
Published: (2025)
Moving Object Segmentation: All You Need Is SAM (and Flow)
by: Xie, Junyu, et al.
Published: (2024)
by: Xie, Junyu, et al.
Published: (2024)
One Pixel is All I Need
by: Siqin, Deng, et al.
Published: (2024)
by: Siqin, Deng, et al.
Published: (2024)
Is Intermediate Fusion All You Need for UAV-based Collaborative Perception?
by: Hao, Jiuwu, et al.
Published: (2025)
by: Hao, Jiuwu, et al.
Published: (2025)
Indirect Prompt Injections: Are Firewalls All You Need, or Stronger Benchmarks?
by: Bhagwatkar, Rishika, et al.
Published: (2025)
by: Bhagwatkar, Rishika, et al.
Published: (2025)
All You Need is Group Actions: Advancing Robust Autonomous Planning
by: Basco, Vincenzo
Published: (2024)
by: Basco, Vincenzo
Published: (2024)
Cognition is All You Need -- The Next Layer of AI Above Large Language Models
by: Spivack, Nova, et al.
Published: (2024)
by: Spivack, Nova, et al.
Published: (2024)
SVAC: Scaling Is All You Need For Referring Video Object Segmentation
by: Zhang, Li, et al.
Published: (2025)
by: Zhang, Li, et al.
Published: (2025)
Attention is All You Need Until You Need Retention
by: Yaslioglu, M. Murat
Published: (2025)
by: Yaslioglu, M. Murat
Published: (2025)
GenFormer -- Generated Images are All You Need to Improve Robustness of Transformers on Small Datasets
by: Oehri, Sven, et al.
Published: (2024)
by: Oehri, Sven, et al.
Published: (2024)
Challenges in designing ethical rules for Infrastructures in Internet of Vehicles
by: Iqbal, Razi
Published: (2025)
by: Iqbal, Razi
Published: (2025)
xGen-MM-Vid (BLIP-3-Video): You Only Need 32 Tokens to Represent a Video Even in VLMs
by: Ryoo, Michael S., et al.
Published: (2024)
by: Ryoo, Michael S., et al.
Published: (2024)
Choice of PEFT Technique in Continual Learning: Prompt Tuning is Not All You Need
by: Wistuba, Martin, et al.
Published: (2024)
by: Wistuba, Martin, et al.
Published: (2024)
Text Is Not All You Need: Multimodal Prompting Helps LLMs Understand Humor
by: Baluja, Ashwin
Published: (2024)
by: Baluja, Ashwin
Published: (2024)
Similar Items
-
Meta-Adaptive Beam Search Planning for Transformer-Based Reinforcement Learning Control of UAVs with Overhead Manipulators under Flight Disturbances
by: Alzorgan, Hazim, et al.
Published: (2026) -
Enhanced Cooperative Perception for Autonomous Vehicles Using Imperfect Communication
by: Sarlak, Ahmad, et al.
Published: (2024) -
From Talking Words to Sharing Thoughts: Scalable Multi-LLM Aggregation via Structured Message Passing
by: Mehrabi, Niloufar, et al.
Published: (2026) -
Don't Waste Bits! Adaptive KV-Cache Quantization for Lightweight On-Device LLMs
by: Boroujeni, Sayed Pedram Haeri, et al.
Published: (2026) -
Monte Carlo Beam Search for Actor-Critic Reinforcement Learning in Continuous Control
by: Alzorgan, Hazim, et al.
Published: (2025)