ERPA: Efficient RPA Model Integrating OCR and LLMs for Intelligent Document Processing
Fuente:
arXiv
Saved in:
| Main Authors: | Abdellaif, Osama, Nader, Abdelrahman, Hamdi, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LMRPA: Large Language Model-Driven Efficient Robotic Process Automation for OCR
by: Abdellaif, Osama Hosam, et al.
Published: (2024)
by: Abdellaif, Osama Hosam, et al.
Published: (2024)
Spacewalker: Traversing Representation Spaces for Fast Interactive Exploration and Annotation of Unstructured Data
by: Heine, Lukas, et al.
Published: (2024)
by: Heine, Lukas, et al.
Published: (2024)
CLAS: A Machine Learning Enhanced Framework for Exploring Large 3D Design Datasets
by: Zhang, XiuYu, et al.
Published: (2024)
by: Zhang, XiuYu, et al.
Published: (2024)
Positive-First Most Ambiguous: A Simple Active Learning Criterion for Interactive Retrieval of Rare Categories
by: Zaher, Kawtar, et al.
Published: (2026)
by: Zaher, Kawtar, et al.
Published: (2026)
Revisiting Human-in-the-Loop Object Retrieval with Pre-Trained Vision Transformers
by: Zaher, Kawtar, et al.
Published: (2026)
by: Zaher, Kawtar, et al.
Published: (2026)
A Versatile Dataset of Mouse and Eye Movements on Search Engine Results Pages
by: Latifzadeh, Kayhan, et al.
Published: (2025)
by: Latifzadeh, Kayhan, et al.
Published: (2025)
Civiverse: A Dataset for Analyzing User Engagement with Open-Source Text-to-Image Models
by: Palmini, Maria-Teresa De Rosa, et al.
Published: (2024)
by: Palmini, Maria-Teresa De Rosa, et al.
Published: (2024)
ModalChorus: Visual Probing and Alignment of Multi-modal Embeddings via Modal Fusion Map
by: Ye, Yilin, et al.
Published: (2024)
by: Ye, Yilin, et al.
Published: (2024)
ArtCognition: A Multimodal AI Framework for Affective State Sensing from Visual and Kinematic Drawing Cues
by: Binaei-Haghighi, Behrad, et al.
Published: (2026)
by: Binaei-Haghighi, Behrad, et al.
Published: (2026)
Evaluating VisualRAG: Quantifying Cross-Modal Performance in Enterprise Document Understanding
by: Mannam, Varun, et al.
Published: (2025)
by: Mannam, Varun, et al.
Published: (2025)
Generative AI for Video Trailer Synthesis: From Extractive Heuristics to Autoregressive Creativity
by: Dharmaratnakar, Abhishek, et al.
Published: (2026)
by: Dharmaratnakar, Abhishek, et al.
Published: (2026)
Seeing Faces in Things: A Model and Dataset for Pareidolia
by: Hamilton, Mark, et al.
Published: (2024)
by: Hamilton, Mark, et al.
Published: (2024)
SymbioticRAG: Enhancing Document Intelligence Through Human-LLM Symbiotic Collaboration
by: Sun, Qiang, et al.
Published: (2025)
by: Sun, Qiang, et al.
Published: (2025)
Chaining text-to-image and large language model: A novel approach for generating personalized e-commerce banners
by: Vashishtha, Shanu, et al.
Published: (2024)
by: Vashishtha, Shanu, et al.
Published: (2024)
Digitization of Document and Information Extraction using OCR
by: Sinha, Rasha, et al.
Published: (2025)
by: Sinha, Rasha, et al.
Published: (2025)
The Klarna Product Page Dataset: Web Element Nomination with Graph Neural Networks and Large Language Models
by: Hotti, Alexandra, et al.
Published: (2021)
by: Hotti, Alexandra, et al.
Published: (2021)
How Do Experts Make Sense of Integrated Process Models?
by: Chen, Tianwa, et al.
Published: (2025)
by: Chen, Tianwa, et al.
Published: (2025)
Efficient and Responsible Adaptation of Large Language Models for Robust Top-k Recommendations
by: Kaur, Kirandeep, et al.
Published: (2024)
by: Kaur, Kirandeep, et al.
Published: (2024)
BibSonomy Meets ChatLLMs for Publication Management: From Chat to Publication Management: Organizing your related work using BibSonomy & LLMs
by: Völker, Tom, et al.
Published: (2024)
by: Völker, Tom, et al.
Published: (2024)
Fetch-A-Set: A Large-Scale OCR-Free Benchmark for Historical Document Retrieval
by: Molina, Adrià, et al.
Published: (2024)
by: Molina, Adrià, et al.
Published: (2024)
Doc To The Future: Infomorphs for Interactive, Multimodal Document Transformation and Generation
by: Kumaravel, Balasaravanan Thoravi
Published: (2025)
by: Kumaravel, Balasaravanan Thoravi
Published: (2025)
Using LLMs to Investigate Correlations of Conversational Follow-up Queries with User Satisfaction
by: Kim, Hyunwoo, et al.
Published: (2024)
by: Kim, Hyunwoo, et al.
Published: (2024)
Cross-Format Retrieval-Augmented Generation in XR with LLMs for Context-Aware Maintenance Assistance
by: Nagy, Akos, et al.
Published: (2025)
by: Nagy, Akos, et al.
Published: (2025)
LLMs in HCI Data Work: Bridging the Gap Between Information Retrieval and Responsible Research Practices
by: Serajeh, Neda Taghizadeh, et al.
Published: (2024)
by: Serajeh, Neda Taghizadeh, et al.
Published: (2024)
Agentic AutoSurvey: Let LLMs Survey LLMs
by: Liu, Yixin, et al.
Published: (2025)
by: Liu, Yixin, et al.
Published: (2025)
CourtNav: Voice-Guided, Anchor-Accurate Navigation of Long Legal Documents in Courtrooms
by: Khadloya, Sai, et al.
Published: (2025)
by: Khadloya, Sai, et al.
Published: (2025)
PoultryTalk: A Multi-modal Retrieval-Augmented Generation (RAG) System for Intelligent Poultry Management and Decision Support
by: Khanal, Kapalik, et al.
Published: (2025)
by: Khanal, Kapalik, et al.
Published: (2025)
Filtering Discomforting Recommendations with Large Language Models
by: Liu, Jiahao, et al.
Published: (2024)
by: Liu, Jiahao, et al.
Published: (2024)
On Explaining Recommendations with Large Language Models: A Review
by: Said, Alan
Published: (2024)
by: Said, Alan
Published: (2024)
Can Instructed Retrieval Models Really Support Exploration?
by: Maheshwari, Piyush, et al.
Published: (2026)
by: Maheshwari, Piyush, et al.
Published: (2026)
Understanding Mental Models of Generative Conversational Search and The Effect of Interface Transparency
by: Degachi, Chadha, et al.
Published: (2025)
by: Degachi, Chadha, et al.
Published: (2025)
From Latent to Observable Position-Based Click Models in Carousel Interfaces
by: de Leon-Martinez, Santiago, et al.
Published: (2026)
by: de Leon-Martinez, Santiago, et al.
Published: (2026)
General-Purpose User Modeling with Behavioral Logs: A Snapchat Case Study
by: Fang, Qixiang, et al.
Published: (2023)
by: Fang, Qixiang, et al.
Published: (2023)
Task Supportive and Personalized Human-Large Language Model Interaction: A User Study
by: Wang, Ben, et al.
Published: (2024)
by: Wang, Ben, et al.
Published: (2024)
Multi-TAP: Multi-criteria Target Adaptive Persona Modeling for Cross-Domain Recommendation
by: Kang, Daehee, et al.
Published: (2026)
by: Kang, Daehee, et al.
Published: (2026)
Gender Biased Legal Case Retrieval System on Users' Decision Process
by: Zhang, Ruizhe, et al.
Published: (2024)
by: Zhang, Ruizhe, et al.
Published: (2024)
Balancing Information Perception with Yin-Yang: Agent-Based Information Neutrality Model for Recommendation Systems
by: Wang, Mengyan, et al.
Published: (2024)
by: Wang, Mengyan, et al.
Published: (2024)
Moving Beyond LDA: A Comparison of Unsupervised Topic Modelling Techniques for Qualitative Data Analysis of Online Communities
by: Kaur, Amandeep, et al.
Published: (2024)
by: Kaur, Amandeep, et al.
Published: (2024)
Resource-Efficient Gesture Recognition using Low-Resolution Thermal Camera via Spiking Neural Networks and Sparse Segmentation
by: Safa, Ali, et al.
Published: (2024)
by: Safa, Ali, et al.
Published: (2024)
VIBE: Annotation-Free Video-to-Text Information Bottleneck Evaluation for TL;DR
by: Chen, Shenghui, et al.
Published: (2025)
by: Chen, Shenghui, et al.
Published: (2025)
Similar Items
-
LMRPA: Large Language Model-Driven Efficient Robotic Process Automation for OCR
by: Abdellaif, Osama Hosam, et al.
Published: (2024) -
Spacewalker: Traversing Representation Spaces for Fast Interactive Exploration and Annotation of Unstructured Data
by: Heine, Lukas, et al.
Published: (2024) -
CLAS: A Machine Learning Enhanced Framework for Exploring Large 3D Design Datasets
by: Zhang, XiuYu, et al.
Published: (2024) -
Positive-First Most Ambiguous: A Simple Active Learning Criterion for Interactive Retrieval of Rare Categories
by: Zaher, Kawtar, et al.
Published: (2026) -
Revisiting Human-in-the-Loop Object Retrieval with Pre-Trained Vision Transformers
by: Zaher, Kawtar, et al.
Published: (2026)