ReceiptSense: Beyond Traditional OCR -- A Dataset for Receipt Understanding
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Abdallah, Abdelrahman, Mounis, Mohamed, Abdalla, Mahmoud, Kasem, Mahmoud SalahEldin, Mahmoud, Mohamed, Abdelhalim, Ibrahim, Elkasaby, Mohamed, ElBendary, Yasser, Jatowt, Adam |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AMuRD: Annotated Arabic-English Receipt Dataset for Key Information Extraction and Classification
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023)
ArabicaQA: A Comprehensive Dataset for Arabic Question Answering
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
BRIDGE: Multimodal-to-Text Retrieval via Reinforcement-Learned Query Alignment
von: Mounis, Mohamed Darwish, et al.
Veröffentlicht: (2026)
von: Mounis, Mohamed Darwish, et al.
Veröffentlicht: (2026)
A Comprehensive Survey of Masked Faces: Recognition, Detection, and Unmasking
von: Mahmoud, Mohamed, et al.
Veröffentlicht: (2024)
von: Mahmoud, Mohamed, et al.
Veröffentlicht: (2024)
MM-BRIGHT: A Multi-Task Multimodal Benchmark for Reasoning-Intensive Retrieval
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026)
MARVEL: Multimodal Adaptive Reasoning-intensiVe Expand-rerank and retrievaL
von: Kasem, Mahmoud SalahEldin, et al.
Veröffentlicht: (2026)
von: Kasem, Mahmoud SalahEldin, et al.
Veröffentlicht: (2026)
HIVE: Query, Hypothesize, Verify An LLM Framework for Multimodal Reasoning-Intensive Retrieval
von: Abdalla, Mahmoud, et al.
Veröffentlicht: (2026)
von: Abdalla, Mahmoud, et al.
Veröffentlicht: (2026)
Performance enhancement of concrete using Dokhan volcanics powder as cement replacement and nano‐banded iron formation as additive: Individual and combined effects
von: A. Serag Faried, et al.
Veröffentlicht: (2026)
von: A. Serag Faried, et al.
Veröffentlicht: (2026)
CoT3DRef: Chain-of-Thoughts Data-Efficient 3D Visual Grounding
von: Abdelrahman, Eslam, et al.
Veröffentlicht: (2023)
von: Abdelrahman, Eslam, et al.
Veröffentlicht: (2023)
Generator-Retriever-Generator Approach for Open-Domain Question Answering
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023)
Evaluating Robustness of LLMs in Question Answering on Multilingual Noisy OCR Data
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2025)
Datacenters in the Desert: Feasibility and Sustainability of LLM Inference in the Middle East
von: Hassan, Lara, et al.
Veröffentlicht: (2025)
von: Hassan, Lara, et al.
Veröffentlicht: (2025)
Foundation Models as Class-Incremental Learners for Dermatological Image Classification
von: Elkhayat, Mohamed, et al.
Veröffentlicht: (2025)
von: Elkhayat, Mohamed, et al.
Veröffentlicht: (2025)
A Quantum-Secure Voting Framework Using QKD, Dual-Key Symmetric Encryption, and Verifiable Receipts
von: Mahmoud, Taha M., et al.
Veröffentlicht: (2025)
von: Mahmoud, Taha M., et al.
Veröffentlicht: (2025)
Beyond Traditional Single Object Tracking: A Survey
von: Abdelaziz, Omar, et al.
Veröffentlicht: (2024)
von: Abdelaziz, Omar, et al.
Veröffentlicht: (2024)
RankArena: A Unified Platform for Evaluating Retrieval, Reranking and RAG with Human and LLM Feedback
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
Defense That Attacks: How Robust Models Become Better Attackers
von: Awad, Mohamed, et al.
Veröffentlicht: (2025)
von: Awad, Mohamed, et al.
Veröffentlicht: (2025)
Kestrel: 3D Multimodal LLM for Part-Aware Grounded Description
von: Ahmed, Mahmoud, et al.
Veröffentlicht: (2024)
von: Ahmed, Mahmoud, et al.
Veröffentlicht: (2024)
3DCoMPaT200: Language-Grounded Compositional Understanding of Parts and Materials of 3D Shapes
von: Ahmed, Mahmoud, et al.
Veröffentlicht: (2025)
von: Ahmed, Mahmoud, et al.
Veröffentlicht: (2025)
MMIS: Multimodal Dataset for Interior Scene Visual Generation and Recognition
von: Kassab, Hozaifa, et al.
Veröffentlicht: (2024)
von: Kassab, Hozaifa, et al.
Veröffentlicht: (2024)
$100,000 or the Robot Gets it! Tech Workers' Resistance Guide: Tech Worker Actions, History, Risks, Impacts, and the Case for a Radical Flank
von: Abdalla, Mohamed
Veröffentlicht: (2025)
von: Abdalla, Mohamed
Veröffentlicht: (2025)
Efficient Hate Speech Detection: A Three-Layer LoRA-Tuned BERTweet Framework
von: El-Bahnasawi, Mahmoud
Veröffentlicht: (2025)
von: El-Bahnasawi, Mahmoud
Veröffentlicht: (2025)
Doppler-Enhanced Deep Learning: Improving Thyroid Nodule Segmentation with YOLOv5 Instance Segmentation
von: Hussieni, Mahmoud El
Veröffentlicht: (2025)
von: Hussieni, Mahmoud El
Veröffentlicht: (2025)
Transformers and Language Models in Form Understanding: A Comprehensive Review of Scanned Document Analysis
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024)
FPGA-based Lane Detection System incorporating Temperature and Light Control Units
von: Qamar, Ibrahim, et al.
Veröffentlicht: (2025)
von: Qamar, Ibrahim, et al.
Veröffentlicht: (2025)
MedDChest: A Content-Aware Multimodal Foundational Vision Model for Thoracic Imaging
von: Soliman, Mahmoud, et al.
Veröffentlicht: (2025)
von: Soliman, Mahmoud, et al.
Veröffentlicht: (2025)
What Do You See? Enhancing Zero-Shot Image Classification with Multimodal Large Language Models
von: Abdelhamed, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Abdelhamed, Abdelrahman, et al.
Veröffentlicht: (2024)
ASRank: Zero-Shot Re-Ranking with Answer Scent for Document Retrieval
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
Detecting Temporal Ambiguity in Questions
von: Piryani, Bhawna, et al.
Veröffentlicht: (2024)
von: Piryani, Bhawna, et al.
Veröffentlicht: (2024)
ComplexTempQA:A 100m Dataset for Complex Temporal Question Answering
von: Gruber, Raphael, et al.
Veröffentlicht: (2024)
von: Gruber, Raphael, et al.
Veröffentlicht: (2024)
Causes of High-school dropouts in Somalia
von: Anisa Mahmoud Mohamed
Veröffentlicht: (2024)
von: Anisa Mahmoud Mohamed
Veröffentlicht: (2024)
InfiniBench: A Benchmark for Large Multi-Modal Models in Long-Form Movies and TV Shows
von: Ataallah, Kirolos, et al.
Veröffentlicht: (2024)
von: Ataallah, Kirolos, et al.
Veröffentlicht: (2024)
Comparative Efficacy of Platelet‐Rich Plasma, Platelet‐Rich Fibrin, and Conventional Therapy in Chronic Non‐Healing Skin Ulcers: A Randomised Clinical Trial
von: Eman Mohamed Salah, et al.
Veröffentlicht: (2026)
von: Eman Mohamed Salah, et al.
Veröffentlicht: (2026)
Arabic-Nougat: Fine-Tuning Vision Transformers for Arabic OCR and Markdown Extraction
von: Rashad, Mohamed
Veröffentlicht: (2024)
von: Rashad, Mohamed
Veröffentlicht: (2024)
GNN-MoE: Context-Aware Patch Routing using GNNs for Parameter-Efficient Domain Generalization
von: Soliman, Mahmoud, et al.
Veröffentlicht: (2025)
von: Soliman, Mahmoud, et al.
Veröffentlicht: (2025)
Wrong Answers Can Also Be Useful: PlausibleQA -- A Large-Scale QA Dataset with Answer Plausibility Scores
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
DeAR: Dual-Stage Document Reranking with Reasoning Agents via LLM Distillation
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2025)
HintEval: A Comprehensive Framework for Hint Generation and Evaluation for Questions
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2025)
Exploring Hint Generation Approaches in Open-Domain Question Answering
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
von: Mozafari, Jamshid, et al.
Veröffentlicht: (2024)
Brain Tumor Radiogenomic Classification
von: Mohamed, Amr, et al.
Veröffentlicht: (2024)
von: Mohamed, Amr, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
AMuRD: Annotated Arabic-English Receipt Dataset for Key Information Extraction and Classification
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2023) -
ArabicaQA: A Comprehensive Dataset for Arabic Question Answering
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2024) -
BRIDGE: Multimodal-to-Text Retrieval via Reinforcement-Learned Query Alignment
von: Mounis, Mohamed Darwish, et al.
Veröffentlicht: (2026) -
A Comprehensive Survey of Masked Faces: Recognition, Detection, and Unmasking
von: Mahmoud, Mohamed, et al.
Veröffentlicht: (2024) -
MM-BRIGHT: A Multi-Task Multimodal Benchmark for Reasoning-Intensive Retrieval
von: Abdallah, Abdelrahman, et al.
Veröffentlicht: (2026)