AMuRD: Annotated Arabic-English Receipt Dataset for Key Information Extraction and Classification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Abdallah, Abdelrahman, Abdalla, Mahmoud, Elkasaby, Mohamed, Elbendary, Yasser, Jatowt, Adam |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ArabicaQA: A Comprehensive Dataset for Arabic Question Answering
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
ReceiptSense: Beyond Traditional OCR -- A Dataset for Receipt Understanding
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
Generator-Retriever-Generator Approach for Open-Domain Question Answering
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023)
ComplexTempQA:A 100m Dataset for Complex Temporal Question Answering
von: Gruber, Raphael, et al.
Veröffentlicht: (2024)
von: Gruber, Raphael, et al.
Veröffentlicht: (2024)
Wrong Answers Can Also Be Useful: PlausibleQA -- A Large-Scale QA Dataset with Answer Plausibility Scores
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
ASRank: Zero-Shot Re-Ranking with Answer Scent for Document Retrieval
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
Detecting Temporal Ambiguity in Questions
von: Piryani, Bhawna, et al.
Veröffentlicht: (2024)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2024)
DynRank: Improving Passage Retrieval with Dynamic Zero-Shot Prompting Based on Question Classification
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
Transformers and Language Models in Form Understanding: A Comprehensive Review of Scanned Document Analysis
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
DeAR: Dual-Stage Document Reranking with Reasoning Agents via LLM Distillation
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
HintEval: A Comprehensive Framework for Hint Generation and Evaluation for Questions
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
Exploring Hint Generation Approaches in Open-Domain Question Answering
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
From Retrieval to Generation: Comparing Different Approaches
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
Evaluating Robustness of LLMs in Question Answering on Multilingual Noisy OCR Data
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
A Study into Investigating Temporal Robustness of LLMs
von: Wallat, Jonas, et al.
Veröffentlicht: (2025)
von: Wallat, Jonas, et al.
Veröffentlicht: (2025)
WikiHint: A Human-Annotated Dataset for Hint Ranking and Generation
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
It's High Time: A Survey of Temporal Question Answering
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
TempRetriever: Fusion-based Temporal Dense Passage Retrieval for Time-Sensitive Questions
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
How Good are LLM-based Rerankers? An Empirical Analysis of State-of-the-Art Reranking Models
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
Rankify: A Comprehensive Python Toolkit for Retrieval, Re-Ranking, and Retrieval-Augmented Generation
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
REGREACT: Self-Correcting Multi-Agent Pipelines for Structured Regulatory Information Extraction
von: Ali, Mohammed, et al.
Veröffentlicht: (2026)
von: Ali, Mohammed, et al.
Veröffentlicht: (2026)
Automated Analysis of Sustainability Reports: Using Large Language Models for the Extraction and Prediction of EU Taxonomy-Compliant KPIs
von: Schmoll, Jonathan, et al.
Veröffentlicht: (2025)
von: Schmoll, Jonathan, et al.
Veröffentlicht: (2025)
EmoHopeSpeech: An Annotated Dataset of Emotions and Hope Speech in English and Arabic
von: Zaghouani, Wajdi, et al.
Veröffentlicht: (2025)
von: Zaghouani, Wajdi, et al.
Veröffentlicht: (2025)
Arabic Dataset for LLM Safeguard Evaluation
von: Ashraf, Yasser, et al.
Veröffentlicht: (2024)
von: Ashraf, Yasser, et al.
Veröffentlicht: (2024)
!MSA at BAREC Shared Task 2025: Ensembling Arabic Transformers for Readability Assessment
von: Basem, Mohamed, et al.
Veröffentlicht: (2025)
von: Basem, Mohamed, et al.
Veröffentlicht: (2025)
TriviaHG: A Dataset for Automatic Hint Generation from Factoid Questions
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
Guidelines for Fine-grained Sentence-level Arabic Readability Annotation
von: Habash, Nizar, et al.
Veröffentlicht: (2024)
von: Habash, Nizar, et al.
Veröffentlicht: (2024)
SustainableQA: A Comprehensive Question Answering Dataset for Corporate Sustainability and EU Taxonomy Reporting
von: Ali, Mohammed, et al.
Veröffentlicht: (2025)
von: Ali, Mohammed, et al.
Veröffentlicht: (2025)
ChroniclingAmericaQA: A Large-scale Question Answering Dataset based on Historical American Newspaper Pages
von: Piryani, Bhawna, et al.
Veröffentlicht: (2024)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2024)
Comparison of Multilingual and Bilingual Models for Satirical News Detection of Arabic and English
von: Abdalla, Omar W., et al.
Veröffentlicht: (2024)
von: Abdalla, Omar W., et al.
Veröffentlicht: (2024)
Exploring Retrieval Augmented Generation in Arabic
von: El-Beltagy, Samhaa R., et al.
Veröffentlicht: (2024)
von: El-Beltagy, Samhaa R., et al.
Veröffentlicht: (2024)
Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction
von: Rashad, Mohamed
Veröffentlicht: (2024)
von: Rashad, Mohamed
Veröffentlicht: (2024)
Negative Sampling Techniques in Information Retrieval: A Survey
von: Wischounig, Laurin, et al.
Veröffentlicht: (2026)
von: Wischounig, Laurin, et al.
Veröffentlicht: (2026)
Exploring NLP Benchmarks in an Extremely Low-Resource Setting
von: Nuha, Ulin, et al.
Veröffentlicht: (2025)
von: Nuha, Ulin, et al.
Veröffentlicht: (2025)
RankArena: A Unified Platform for Evaluating Retrieval, Reranking and RAG with Human and LLM Feedback
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
ArPoMeme: An Annotated Arabic Multimodal Dataset for Political Ideology and Polarization
von: Zaghouani, Wajdi, et al.
Veröffentlicht: (2026)
von: Zaghouani, Wajdi, et al.
Veröffentlicht: (2026)
Analyzing the Role of Context in Forecasting with Large Language Models
von: Mutschlechner, Gerrit, et al.
Veröffentlicht: (2025)
von: Mutschlechner, Gerrit, et al.
Veröffentlicht: (2025)
Navigating Tomorrow: Reliably Assessing Large Language Models Performance on Future Event Prediction
von: Nako, Petraq, et al.
Veröffentlicht: (2025)
von: Nako, Petraq, et al.
Veröffentlicht: (2025)
WOLI at SemEval-2020 Task 12: Arabic Offensive Language Identification on Different Twitter Datasets
von: Otiefy, Yasser, et al.
Veröffentlicht: (2020)
von: Otiefy, Yasser, et al.
Veröffentlicht: (2020)
EgMM-Corpus: A Multimodal Vision-Language Dataset for Egyptian Culture
von: Gamil, Mohamed, et al.
Veröffentlicht: (2025)
von: Gamil, Mohamed, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
ArabicaQA: A Comprehensive Dataset for Arabic Question Answering
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024) -
ReceiptSense: Beyond Traditional OCR -- A Dataset for Receipt Understanding
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024) -
Generator-Retriever-Generator Approach for Open-Domain Question Answering
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023) -
ComplexTempQA:A 100m Dataset for Complex Temporal Question Answering
von: Gruber, Raphael, et al.
Veröffentlicht: (2024) -
Wrong Answers Can Also Be Useful: PlausibleQA -- A Large-Scale QA Dataset with Answer Plausibility Scores
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)