TabSniper: Towards Accurate Table Detection & Structure Recognition for Bank Statements
Fuente:
arXiv
Guardado en:
| Autores principales: | Trivedi, Abhishek, Mukherjee, Sourajit, Singh, Rajat Kumar, Agarwal, Vani, Ramakrishnan, Sriranjani, Bhatt, Himanshu S. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Improving Shift Invariance in Convolutional Neural Networks with Translation Invariant Polyphase Sampling
por: Saha, Sourajit, et al.
Publicado: (2024)
por: Saha, Sourajit, et al.
Publicado: (2024)
SPRINT: Script-agnostic Structure Recognition in Tables
por: Kudale, Dhruv, et al.
Publicado: (2025)
por: Kudale, Dhruv, et al.
Publicado: (2025)
Advancing Toward Robust and Scalable Fingerprint Orientation Estimation: From Gradients to Deep Learning
por: Trivedi, Amit Kumar, et al.
Publicado: (2020)
por: Trivedi, Amit Kumar, et al.
Publicado: (2020)
UniTabNet: Bridging Vision and Language Models for Enhanced Table Structure Recognition
por: Zhang, Zhenrong, et al.
Publicado: (2024)
por: Zhang, Zhenrong, et al.
Publicado: (2024)
ZK-WAGON: Imperceptible Watermark for Image Generation Models using ZK-SNARKs
por: Ramakrishnan, Aadarsh Anantha, et al.
Publicado: (2025)
por: Ramakrishnan, Aadarsh Anantha, et al.
Publicado: (2025)
NanoVLMs: How small can we go and still make coherent Vision Language Models?
por: Agarwalla, Mukund, et al.
Publicado: (2025)
por: Agarwalla, Mukund, et al.
Publicado: (2025)
Side Effects of Erasing Concepts from Diffusion Models
por: Saha, Shaswati, et al.
Publicado: (2025)
por: Saha, Shaswati, et al.
Publicado: (2025)
GAViD: A Large-Scale Multimodal Dataset for Context-Aware Group Affect Recognition from Videos
por: Kumar, Deepak, et al.
Publicado: (2026)
por: Kumar, Deepak, et al.
Publicado: (2026)
Snowy Scenes,Clear Detections: A Robust Model for Traffic Light Detection in Adverse Weather Conditions
por: Garg, Shivank, et al.
Publicado: (2024)
por: Garg, Shivank, et al.
Publicado: (2024)
TabPedia: Towards Comprehensive Visual Table Understanding with Concept Synergy
por: Zhao, Weichao, et al.
Publicado: (2024)
por: Zhao, Weichao, et al.
Publicado: (2024)
Hierarchical Modeling Approach to Fast and Accurate Table Recognition
por: Kawakatsu, Takaya
Publicado: (2025)
por: Kawakatsu, Takaya
Publicado: (2025)
TrueSkin: Towards Fair and Accurate Skin Tone Recognition and Generation
por: Lu, Haoming
Publicado: (2025)
por: Lu, Haoming
Publicado: (2025)
DenTab: A Dataset for Table Recognition and Visual QA on Real-World Dental Estimates
por: Hamdi, Laziz, et al.
Publicado: (2026)
por: Hamdi, Laziz, et al.
Publicado: (2026)
On Occlusions in Video Action Detection: Benchmark Datasets And Training Recipes
por: Modi, Rajat, et al.
Publicado: (2024)
por: Modi, Rajat, et al.
Publicado: (2024)
TC-OCR: TableCraft OCR for Efficient Detection & Recognition of Table Structure & Content
por: Anand, Avinash, et al.
Publicado: (2024)
por: Anand, Avinash, et al.
Publicado: (2024)
Detection of Virus and Small Cell Patches in Foci Images Using Switchable Convolution and Feature Pyramid Networks
por: Singh, Amrita, et al.
Publicado: (2026)
por: Singh, Amrita, et al.
Publicado: (2026)
AQD: Towards Accurate Fully-Quantized Object Detection
por: Chen, Peng, et al.
Publicado: (2020)
por: Chen, Peng, et al.
Publicado: (2020)
Sketch and Refine: Towards Fast and Accurate Lane Detection
por: Chen, Chao, et al.
Publicado: (2024)
por: Chen, Chao, et al.
Publicado: (2024)
Efficient Human Pose Estimation: Leveraging Advanced Techniques with MediaPipe
por: Sengar, Sandeep Singh, et al.
Publicado: (2024)
por: Sengar, Sandeep Singh, et al.
Publicado: (2024)
Scale-Invariant Object Detection by Adaptive Convolution with Unified Global-Local Context
por: Singh, Amrita, et al.
Publicado: (2024)
por: Singh, Amrita, et al.
Publicado: (2024)
Advanced Gesture Recognition for Autism Spectrum Disorder Detection: Integrating YOLOv7, Video Augmentation, and VideoMAE for Naturalistic Video Analysis
por: Singh, Amit Kumar, et al.
Publicado: (2024)
por: Singh, Amit Kumar, et al.
Publicado: (2024)
FashionSD-X: Multimodal Fashion Garment Synthesis using Latent Diffusion
por: Singh, Abhishek Kumar, et al.
Publicado: (2024)
por: Singh, Abhishek Kumar, et al.
Publicado: (2024)
Towards Accurate One-Stage Object Detection with AP-Loss
por: Chen, Kean, et al.
Publicado: (2019)
por: Chen, Kean, et al.
Publicado: (2019)
On the Effect of Image Resolution on Semantic Segmentation
por: Singh, Ritambhara, et al.
Publicado: (2024)
por: Singh, Ritambhara, et al.
Publicado: (2024)
Towards More Accurate Fake Detection on Images Generated from Advanced Generative and Neural Rendering Models
por: Dong, Chengdong, et al.
Publicado: (2024)
por: Dong, Chengdong, et al.
Publicado: (2024)
Asynchronous Perception Machine For Efficient Test-Time-Training
por: Modi, Rajat, et al.
Publicado: (2024)
por: Modi, Rajat, et al.
Publicado: (2024)
The Black Ninjas and the Sniper: On Robustness of Population Protocols
por: Lossin, Benno, et al.
Publicado: (2024)
por: Lossin, Benno, et al.
Publicado: (2024)
Uncertainty Quantification in Table Structure Recognition
por: Ajayi, Kehinde, et al.
Publicado: (2024)
por: Ajayi, Kehinde, et al.
Publicado: (2024)
Face Detection: Present State and Research Directions
por: Prabhat, Purnendu, et al.
Publicado: (2024)
por: Prabhat, Purnendu, et al.
Publicado: (2024)
Towards Explainable LiDAR Point Cloud Semantic Segmentation via Gradient Based Target Localization
por: Kuriyal, Abhishek, et al.
Publicado: (2024)
por: Kuriyal, Abhishek, et al.
Publicado: (2024)
Google-MedGemma Based Abnormality Detection in Musculoskeletal radiographs
por: Maity, Soumyajit, et al.
Publicado: (2025)
por: Maity, Soumyajit, et al.
Publicado: (2025)
Images as Tables: In-Context Learning with TabPFN for Low-Data Detection of AI-Generated Images
por: Walter, Jan Philip, et al.
Publicado: (2026)
por: Walter, Jan Philip, et al.
Publicado: (2026)
Rethinking Detection Based Table Structure Recognition for Visually Rich Document Images
por: Xiao, Bin, et al.
Publicado: (2023)
por: Xiao, Bin, et al.
Publicado: (2023)
Towards Accurate Camouflaged Object Detection with Mixture Convolution and Interactive Fusion
por: Chen, Geng, et al.
Publicado: (2021)
por: Chen, Geng, et al.
Publicado: (2021)
DIFEM: Key-points Interaction based Feature Extraction Module for Violence Recognition in Videos
por: Mittal, Himanshu, et al.
Publicado: (2024)
por: Mittal, Himanshu, et al.
Publicado: (2024)
Improving Video Question Answering through query-based frame selection
por: Patil, Himanshu, et al.
Publicado: (2026)
por: Patil, Himanshu, et al.
Publicado: (2026)
Recall to Predict: Grounding Motion Forecasting in Interpretable Motion Bank
por: Vivekanandan, Abhishek, et al.
Publicado: (2026)
por: Vivekanandan, Abhishek, et al.
Publicado: (2026)
MCDDPM: Multichannel Conditional Denoising Diffusion Model for Unsupervised Anomaly Detection in Brain MRI
por: Trivedi, Vivek Kumar, et al.
Publicado: (2024)
por: Trivedi, Vivek Kumar, et al.
Publicado: (2024)
InstructTable: Improving Table Structure Recognition Through Instructions
por: Chen, Boming, et al.
Publicado: (2026)
por: Chen, Boming, et al.
Publicado: (2026)
Benchmarking Vision-Language Models on Optical Character Recognition in Dynamic Video Environments
por: Nagaonkar, Sankalp, et al.
Publicado: (2025)
por: Nagaonkar, Sankalp, et al.
Publicado: (2025)
Ejemplares similares
-
Improving Shift Invariance in Convolutional Neural Networks with Translation Invariant Polyphase Sampling
por: Saha, Sourajit, et al.
Publicado: (2024) -
SPRINT: Script-agnostic Structure Recognition in Tables
por: Kudale, Dhruv, et al.
Publicado: (2025) -
Advancing Toward Robust and Scalable Fingerprint Orientation Estimation: From Gradients to Deep Learning
por: Trivedi, Amit Kumar, et al.
Publicado: (2020) -
UniTabNet: Bridging Vision and Language Models for Enhanced Table Structure Recognition
por: Zhang, Zhenrong, et al.
Publicado: (2024) -
ZK-WAGON: Imperceptible Watermark for Image Generation Models using ZK-SNARKs
por: Ramakrishnan, Aadarsh Anantha, et al.
Publicado: (2025)