LP-LLM: End-to-End Real-World Degraded License Plate Text Recognition via Large Multimodal Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gong, Haoyan, Liu, Hongbin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CharDiff-LP: A Diffusion Model with Character-Level Guidance for License Plate Image Restoration
von: Na, Kihyun, et al.
Veröffentlicht: (2025)
von: Na, Kihyun, et al.
Veröffentlicht: (2025)
Forensic License Plate Recognition with Compression-Informed Transformers
von: Moussa, Denise, et al.
Veröffentlicht: (2022)
von: Moussa, Denise, et al.
Veröffentlicht: (2022)
PatrolVision: Automated License Plate Recognition in the wild
von: Singhal, Anmol Singhal Navya
Veröffentlicht: (2025)
von: Singhal, Anmol Singhal Navya
Veröffentlicht: (2025)
X-World: Controllable Ego-Centric Multi-Camera World Models for Scalable End-to-End Driving
von: Zheng, Chaoda, et al.
Veröffentlicht: (2026)
von: Zheng, Chaoda, et al.
Veröffentlicht: (2026)
USAD: End-to-End Human Activity Recognition via Diffusion Model with Spatiotemporal Attention
von: Xiao, Hang, et al.
Veröffentlicht: (2025)
von: Xiao, Hang, et al.
Veröffentlicht: (2025)
A Dataset and Model for Realistic License Plate Deblurring
von: Gong, Haoyan, et al.
Veröffentlicht: (2024)
von: Gong, Haoyan, et al.
Veröffentlicht: (2024)
Efficient License Plate Recognition via Pseudo-Labeled Supervision with Grounding DINO and YOLOv8
von: Vargoorani, Zahra Ebrahimi, et al.
Veröffentlicht: (2025)
von: Vargoorani, Zahra Ebrahimi, et al.
Veröffentlicht: (2025)
LMGenDrive: Bridging Multimodal Understanding and Generative World Modeling for End-to-End Driving
von: Shao, Hao, et al.
Veröffentlicht: (2026)
von: Shao, Hao, et al.
Veröffentlicht: (2026)
DOTA: Deformable Optimized Transformer Architecture for End-to-End Text Recognition with Retrieval-Augmented Generation
von: Nithisopa, Naphat, et al.
Veröffentlicht: (2025)
von: Nithisopa, Naphat, et al.
Veröffentlicht: (2025)
An End-to-End, Segmentation-Free, Arabic Handwritten Recognition Model on KHATT
von: Aabed, Sondos, et al.
Veröffentlicht: (2024)
von: Aabed, Sondos, et al.
Veröffentlicht: (2024)
SimpleLLM4AD: An End-to-End Vision-Language Model with Graph Visual Question Answering for Autonomous Driving
von: Zheng, Peiru, et al.
Veröffentlicht: (2024)
von: Zheng, Peiru, et al.
Veröffentlicht: (2024)
Video-based Vehicle Surveillance in the Wild: License Plate, Make, and Model Recognition with Self Reflective Vision-Language Models
von: Parsa, Pouya, et al.
Veröffentlicht: (2025)
von: Parsa, Pouya, et al.
Veröffentlicht: (2025)
TDATR: Improving End-to-End Table Recognition via Table Detail-Aware Learning and Cell-Level Visual Alignment
von: Qin, Chunxia, et al.
Veröffentlicht: (2026)
von: Qin, Chunxia, et al.
Veröffentlicht: (2026)
PDF-WuKong: A Large Multimodal Model for Efficient Long PDF Reading with End-to-End Sparse Sampling
von: Xie, Xudong, et al.
Veröffentlicht: (2024)
von: Xie, Xudong, et al.
Veröffentlicht: (2024)
Enabling Intelligent Traffic Systems: A Deep Learning Method for Accurate Arabic License Plate Recognition
von: Sayedelahl, M. A.
Veröffentlicht: (2024)
von: Sayedelahl, M. A.
Veröffentlicht: (2024)
MF-LPR$^2$: Multi-Frame License Plate Image Restoration and Recognition using Optical Flow
von: Na, Kihyun, et al.
Veröffentlicht: (2025)
von: Na, Kihyun, et al.
Veröffentlicht: (2025)
License Plate Images Generation with Diffusion Models
von: Shpir, Mariia, et al.
Veröffentlicht: (2025)
von: Shpir, Mariia, et al.
Veröffentlicht: (2025)
VLM-E2E: Enhancing End-to-End Autonomous Driving with Multimodal Driver Attention Fusion
von: Liu, Pei, et al.
Veröffentlicht: (2025)
von: Liu, Pei, et al.
Veröffentlicht: (2025)
Risk-Aware World Model Predictive Control for Generalizable End-to-End Autonomous Driving
von: Sun, Jiangxin, et al.
Veröffentlicht: (2026)
von: Sun, Jiangxin, et al.
Veröffentlicht: (2026)
DIAL: Decoupling Intent and Action via Latent World Modeling for End-to-End VLA
von: Chen, Yi, et al.
Veröffentlicht: (2026)
von: Chen, Yi, et al.
Veröffentlicht: (2026)
An End-to-End Two-Stream Network Based on RGB Flow and Representation Flow for Human Action Recognition
von: Lai, Song-Jiang, et al.
Veröffentlicht: (2024)
von: Lai, Song-Jiang, et al.
Veröffentlicht: (2024)
End to End AI System for Surgical Gesture Sequence Recognition and Clinical Outcome Prediction
von: Li, Xi, et al.
Veröffentlicht: (2025)
von: Li, Xi, et al.
Veröffentlicht: (2025)
End-to-End Human Instance Matting
von: Liu, Qinglin, et al.
Veröffentlicht: (2024)
von: Liu, Qinglin, et al.
Veröffentlicht: (2024)
Guiding Attention in End-to-End Driving Models
von: Porres, Diego, et al.
Veröffentlicht: (2024)
von: Porres, Diego, et al.
Veröffentlicht: (2024)
ContourFormer: Real-Time Contour-Based End-to-End Instance Segmentation Transformer
von: Yao, Weiwei, et al.
Veröffentlicht: (2025)
von: Yao, Weiwei, et al.
Veröffentlicht: (2025)
End-To-End Underwater Video Enhancement: Dataset and Model
von: Du, Dazhao, et al.
Veröffentlicht: (2024)
von: Du, Dazhao, et al.
Veröffentlicht: (2024)
Next-Generation License Plate Detection and Recognition System using YOLOv8
von: Amin, Arslan, et al.
Veröffentlicht: (2025)
von: Amin, Arslan, et al.
Veröffentlicht: (2025)
YOLO26: An Analysis of NMS-Free End to End Framework for Real-Time Object Detection
von: Chakrabarty, Sudip
Veröffentlicht: (2026)
von: Chakrabarty, Sudip
Veröffentlicht: (2026)
EMMA: End-to-End Multimodal Model for Autonomous Driving
von: Hwang, Jyh-Jing, et al.
Veröffentlicht: (2024)
von: Hwang, Jyh-Jing, et al.
Veröffentlicht: (2024)
A Paradigm Shift: Fully End-to-End Training for Temporal Sentence Grounding in Videos
von: He, Allen, et al.
Veröffentlicht: (2026)
von: He, Allen, et al.
Veröffentlicht: (2026)
Automatically Generating Visual Hallucination Test Cases for Multimodal Large Language Models
von: Liu, Zhongye, et al.
Veröffentlicht: (2024)
von: Liu, Zhongye, et al.
Veröffentlicht: (2024)
Training Multi-Image Vision Agents via End2End Reinforcement Learning
von: Dong, Chengqi, et al.
Veröffentlicht: (2025)
von: Dong, Chengqi, et al.
Veröffentlicht: (2025)
iPad: Iterative Proposal-centric End-to-End Autonomous Driving
von: Guo, Ke, et al.
Veröffentlicht: (2025)
von: Guo, Ke, et al.
Veröffentlicht: (2025)
Fose: Fusion of One-Step Diffusion and End-to-End Network for Pansharpening
von: Liu, Kai, et al.
Veröffentlicht: (2025)
von: Liu, Kai, et al.
Veröffentlicht: (2025)
Multi-Dimensional Insights: Benchmarking Real-World Personalization in Large Multimodal Models
von: Zhang, YiFan, et al.
Veröffentlicht: (2024)
von: Zhang, YiFan, et al.
Veröffentlicht: (2024)
SamLP: A Customized Segment Anything Model for License Plate Detection
von: Ding, Haoxuan, et al.
Veröffentlicht: (2024)
von: Ding, Haoxuan, et al.
Veröffentlicht: (2024)
Raw2Drive: Reinforcement Learning with Aligned World Models for End-to-End Autonomous Driving (in CARLA v2)
von: Yang, Zhenjie, et al.
Veröffentlicht: (2025)
von: Yang, Zhenjie, et al.
Veröffentlicht: (2025)
Mapping License Plate Recoverability Under Extreme Viewing Angles for Oppor-tunistic Urban Sensing
von: Adamenko, Igor, et al.
Veröffentlicht: (2026)
von: Adamenko, Igor, et al.
Veröffentlicht: (2026)
Inter-Class Relational Loss for Small Object Detection: A Case Study on License Plates
von: Ning, Dian, et al.
Veröffentlicht: (2025)
von: Ning, Dian, et al.
Veröffentlicht: (2025)
SyncVSR: Data-Efficient Visual Speech Recognition with End-to-End Crossmodal Audio Token Synchronization
von: Ahn, Young Jin, et al.
Veröffentlicht: (2024)
von: Ahn, Young Jin, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
CharDiff-LP: A Diffusion Model with Character-Level Guidance for License Plate Image Restoration
von: Na, Kihyun, et al.
Veröffentlicht: (2025) -
Forensic License Plate Recognition with Compression-Informed Transformers
von: Moussa, Denise, et al.
Veröffentlicht: (2022) -
PatrolVision: Automated License Plate Recognition in the wild
von: Singhal, Anmol Singhal Navya
Veröffentlicht: (2025) -
X-World: Controllable Ego-Centric Multi-Camera World Models for Scalable End-to-End Driving
von: Zheng, Chaoda, et al.
Veröffentlicht: (2026) -
USAD: End-to-End Human Activity Recognition via Diffusion Model with Spatiotemporal Attention
von: Xiao, Hang, et al.
Veröffentlicht: (2025)