Advancing Vehicle Plate Recognition: Multitasking Visual Language Models with VehiclePaliGemma
Fuente:
arXiv
Saved in:
| Main Authors: | AlDahoul, Nouar, Tan, Myles Joshua Toledo, Tera, Raghava Reddy, Karim, Hezerul Abdul, Lim, Chee How, Mishra, Manish Kumar, Zaki, Yasir |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Benchmarking the Medical Understanding and Reasoning of Large Language Models in Arabic Healthcare Tasks
by: AlDahoul, Nouar, et al.
Published: (2025)
by: AlDahoul, Nouar, et al.
Published: (2025)
Benchmarking the Legal Reasoning of LLMs in Arabic Islamic Inheritance Cases
by: AlDahoul, Nouar, et al.
Published: (2025)
by: AlDahoul, Nouar, et al.
Published: (2025)
Detecting Hope, Hate, and Emotion in Arabic Textual Speech and Multi-modal Memes Using Large Language Models
by: AlDahoul, Nouar, et al.
Published: (2025)
by: AlDahoul, Nouar, et al.
Published: (2025)
A Conceptual Exploration of Generative AI-Induced Cognitive Dissonance and its Emergence in University-Level Academic Writing
by: Seran, Carl Errol, et al.
Published: (2025)
by: Seran, Carl Errol, et al.
Published: (2025)
Advancing Content Moderation: Evaluating Large Language Models for Detecting Sensitive Content Across Text, Images, and Videos
by: AlDahoul, Nouar, et al.
Published: (2024)
by: AlDahoul, Nouar, et al.
Published: (2024)
Exploring Vision Language Models for Facial Attribute Recognition: Emotion, Race, Gender, and Age
by: AlDahoul, Nouar, et al.
Published: (2024)
by: AlDahoul, Nouar, et al.
Published: (2024)
Fine-tuned Vision Language Model for Localization of Parasitic Eggs in Microscopic Images
by: Sien, Chan Hao, et al.
Published: (2026)
by: Sien, Chan Hao, et al.
Published: (2026)
A Novel BERT-based Classifier to Detect Political Leaning of YouTube Videos based on their Titles
by: AlDahoul, Nouar, et al.
Published: (2024)
by: AlDahoul, Nouar, et al.
Published: (2024)
AI-generated faces influence gender stereotypes and racial homogenization
by: AlDahoul, Nouar, et al.
Published: (2024)
by: AlDahoul, Nouar, et al.
Published: (2024)
Empowering the Grid: Collaborative Edge Artificial Intelligence for Decentralized Energy Systems
by: Paula Jr, Eddie de, et al.
Published: (2025)
by: Paula Jr, Eddie de, et al.
Published: (2025)
Multitask Mayhem: Unveiling and Mitigating Safety Gaps in LLMs Fine-tuning
by: Jan, Essa, et al.
Published: (2024)
by: Jan, Essa, et al.
Published: (2024)
Self-Reflection Makes Large Language Models Safer, Less Biased, and Ideologically Neutral
by: Liu, Fengyuan, et al.
Published: (2024)
by: Liu, Fengyuan, et al.
Published: (2024)
Inclusive content reduces racial and gender biases, yet non-inclusive content dominates popular culture
by: AlDahoul, Nouar, et al.
Published: (2024)
by: AlDahoul, Nouar, et al.
Published: (2024)
Can Personalized Medicine Coexist with Health Equity? Examining the Cost Barrier and Ethical Implications
by: Francisco, Kishi Kobe Yee, et al.
Published: (2024)
by: Francisco, Kishi Kobe Yee, et al.
Published: (2024)
Neutralizing the Narrative: AI-Powered Debiasing of Online News Articles
by: Kuo, Chen Wei, et al.
Published: (2025)
by: Kuo, Chen Wei, et al.
Published: (2025)
Semantic-Aware Advanced Persistent Threat Detection Using Autoencoders on LLM-Encoded System Logs
by: Mohammed, Waleed Khan, et al.
Published: (2026)
by: Mohammed, Waleed Khan, et al.
Published: (2026)
A Longitudinal Analysis of Racial and Gender Bias in New York Times and Fox News Images and Articles
by: Ibrahim, Hazem, et al.
Published: (2024)
by: Ibrahim, Hazem, et al.
Published: (2024)
Enhancing Password Security Through a High-Accuracy Scoring Framework Using Random Forests
by: Mazelan, Muhammed El Mustaqeem, et al.
Published: (2025)
by: Mazelan, Muhammed El Mustaqeem, et al.
Published: (2025)
Toward a Safer Web: Multilingual Multi-Agent LLMs for Mitigating Adversarial Misinformation Attacks
by: Aldahoul, Nouar, et al.
Published: (2025)
by: Aldahoul, Nouar, et al.
Published: (2025)
Real-Time Human Detection for Aerial Captured Video Sequences via Deep Models
by: AlDahoul, Nouar, et al.
Published: (2026)
by: AlDahoul, Nouar, et al.
Published: (2026)
An explainable Recursive Feature Elimination to detect Advanced Persistent Threats using Random Forest classifier
by: Mutalib, Noor Hazlina Abdul, et al.
Published: (2025)
by: Mutalib, Noor Hazlina Abdul, et al.
Published: (2025)
PaliGemma 2: A Family of Versatile VLMs for Transfer
by: Steiner, Andreas, et al.
Published: (2024)
by: Steiner, Andreas, et al.
Published: (2024)
PaliGemma: A versatile 3B VLM for transfer
by: Beyer, Lucas, et al.
Published: (2024)
by: Beyer, Lucas, et al.
Published: (2024)
Who Gets Seen in the Age of AI? Adoption Patterns of Large Language Models in Scholarly Writing and Citation Outcomes
by: Khan, Farhan Kamrul, et al.
Published: (2025)
by: Khan, Farhan Kamrul, et al.
Published: (2025)
Physics-Grounded Monocular Vehicle Distance Estimation Using Standardized License Plate Typography
by: Reddy, Manognya Lokesh, et al.
Published: (2026)
by: Reddy, Manognya Lokesh, et al.
Published: (2026)
Large Language Models are often politically extreme, usually ideologically inconsistent, and persuasive even in informational contexts
by: Aldahoul, Nouar, et al.
Published: (2025)
by: Aldahoul, Nouar, et al.
Published: (2025)
PaliGemma-CXR: A Multi-task Multimodal Model for TB Chest X-ray Interpretation
by: Musinguzi, Denis, et al.
Published: (2025)
by: Musinguzi, Denis, et al.
Published: (2025)
PixLift: Accelerating Web Browsing via AI Upscaling
by: Atinafu, Yonas, et al.
Published: (2025)
by: Atinafu, Yonas, et al.
Published: (2025)
Multitask Vehicle Signal Recognition With Dual‐Speed Adaptive Weighting
by: Dianjing Cheng, et al.
Published: (2025)
by: Dianjing Cheng, et al.
Published: (2025)
TikTok's recommendations skewed towards Republican content during the 2024 U.S. presidential race
by: Ibrahim, Hazem, et al.
Published: (2025)
by: Ibrahim, Hazem, et al.
Published: (2025)
Schadenfreude in the Digital Public Sphere: A cross-national and decade-long analysis of Facebook news engagement
by: Aldahoul, Nouar, et al.
Published: (2026)
by: Aldahoul, Nouar, et al.
Published: (2026)
Evaluating the Efficacy of Next.js: A Comparative Analysis with React.js on Performance, SEO, and Global Network Equity
by: Pati, Swostik, et al.
Published: (2025)
by: Pati, Swostik, et al.
Published: (2025)
Formal Safety Guarantees for Autonomous Vehicles using Barrier Certificates
by: Barhoumi, Oumaima, et al.
Published: (2026)
by: Barhoumi, Oumaima, et al.
Published: (2026)
Clustering Study of Vehicle Behaviors Using License Plate Recognition
by: Bolaños-Martinez, Daniel, et al.
Published: (2022)
by: Bolaños-Martinez, Daniel, et al.
Published: (2022)
YOLO and Mask R-CNN for Vehicle Number Plate Identification
by: Ganjoo, Siddharth
Published: (2022)
by: Ganjoo, Siddharth
Published: (2022)
Addressing Intersectionality, Explainability, and Ethics in AI-Driven Diagnostics: A Rebuttal and Call for Transdiciplinary Action
by: Tan, Myles Joshua Toledo, et al.
Published: (2025)
by: Tan, Myles Joshua Toledo, et al.
Published: (2025)
Planning Autonomous Vehicle Maneuvering in Work Zones Through Game-Theoretic Trajectory Generation
by: Nour, Mayar, et al.
Published: (2026)
by: Nour, Mayar, et al.
Published: (2026)
Semantic Segmentation for Real-World and Synthetic Vehicle's Forward-Facing Camera Images
by: Nguyen, Tuan T., et al.
Published: (2024)
by: Nguyen, Tuan T., et al.
Published: (2024)
Toward Unified Fine-Grained Vehicle Classification and Automatic License Plate Recognition
by: Lima, Gabriel E., et al.
Published: (2026)
by: Lima, Gabriel E., et al.
Published: (2026)
CodeGemma: Open Code Models Based on Gemma
by: CodeGemma Team, et al.
Published: (2024)
by: CodeGemma Team, et al.
Published: (2024)
Similar Items
-
Benchmarking the Medical Understanding and Reasoning of Large Language Models in Arabic Healthcare Tasks
by: AlDahoul, Nouar, et al.
Published: (2025) -
Benchmarking the Legal Reasoning of LLMs in Arabic Islamic Inheritance Cases
by: AlDahoul, Nouar, et al.
Published: (2025) -
Detecting Hope, Hate, and Emotion in Arabic Textual Speech and Multi-modal Memes Using Large Language Models
by: AlDahoul, Nouar, et al.
Published: (2025) -
A Conceptual Exploration of Generative AI-Induced Cognitive Dissonance and its Emergence in University-Level Academic Writing
by: Seran, Carl Errol, et al.
Published: (2025) -
Advancing Content Moderation: Evaluating Large Language Models for Detecting Sensitive Content Across Text, Images, and Videos
by: AlDahoul, Nouar, et al.
Published: (2024)