GraphKD: Exploring Knowledge Distillation Towards Document Object Detection with Structured Graph Creation
Fuente:
arXiv
Salvato in:
| Autori principali: | Banerjee, Ayan, Biswas, Sanket, Lladós, Josep, Pal, Umapada |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CraftSVG: Multi-Object Text-to-SVG Synthesis via Layout Guided Diffusion
di: Banerjee, Ayan, et al.
Pubblicazione: (2024)
di: Banerjee, Ayan, et al.
Pubblicazione: (2024)
DistilDoc: Knowledge Distillation for Visually-Rich Document Applications
di: Van Landeghem, Jordy, et al.
Pubblicazione: (2024)
di: Van Landeghem, Jordy, et al.
Pubblicazione: (2024)
DocRevive: A Unified Pipeline for Document Text Restoration
di: Purkayastha, Kunal, et al.
Pubblicazione: (2026)
di: Purkayastha, Kunal, et al.
Pubblicazione: (2026)
Diving into the Depths of Spotting Text in Multi-Domain Noisy Scenes
di: Das, Alloy, et al.
Pubblicazione: (2023)
di: Das, Alloy, et al.
Pubblicazione: (2023)
TaleDiffusion: Multi-Character Story Generation with Dialogue Rendering
di: Banerjee, Ayan, et al.
Pubblicazione: (2025)
di: Banerjee, Ayan, et al.
Pubblicazione: (2025)
FastTextSpotter: A High-Efficiency Transformer for Multilingual Scene Text Spotting
di: Das, Alloy, et al.
Pubblicazione: (2024)
di: Das, Alloy, et al.
Pubblicazione: (2024)
Towards Robust Cross-Dataset Object Detection Generalization under Domain Specificity
di: Chakraborty, Ritabrata, et al.
Pubblicazione: (2026)
di: Chakraborty, Ritabrata, et al.
Pubblicazione: (2026)
LayeredDoc: Domain Adaptive Document Restoration with a Layer Separation Approach
di: Pilligua, Maria, et al.
Pubblicazione: (2024)
di: Pilligua, Maria, et al.
Pubblicazione: (2024)
CraftGraffiti: Exploring Human Identity with Custom Graffiti Art via Facial-Preserving Diffusion Models
di: Banerjee, Ayan, et al.
Pubblicazione: (2025)
di: Banerjee, Ayan, et al.
Pubblicazione: (2025)
SketchGPT: Autoregressive Modeling for Sketch Generation and Recognition
di: Tiwari, Adarsh, et al.
Pubblicazione: (2024)
di: Tiwari, Adarsh, et al.
Pubblicazione: (2024)
GeoContrastNet: Contrastive Key-Value Edge Learning for Language-Agnostic Document Understanding
di: Biescas, Nil, et al.
Pubblicazione: (2024)
di: Biescas, Nil, et al.
Pubblicazione: (2024)
NoTeS-Bank: Benchmarking Neural Transcription and Search for Scientific Notes Understanding
di: Pal, Aniket, et al.
Pubblicazione: (2025)
di: Pal, Aniket, et al.
Pubblicazione: (2025)
DocSynthv2: A Practical Autoregressive Modeling for Document Generation
di: Biswas, Sanket, et al.
Pubblicazione: (2024)
di: Biswas, Sanket, et al.
Pubblicazione: (2024)
Towards Generative Class Prompt Learning for Fine-grained Visual Recognition
di: Chattopadhyay, Soumitri, et al.
Pubblicazione: (2024)
di: Chattopadhyay, Soumitri, et al.
Pubblicazione: (2024)
FASTER: A Font-Agnostic Scene Text Editing and Rendering Framework
di: Das, Alloy, et al.
Pubblicazione: (2023)
di: Das, Alloy, et al.
Pubblicazione: (2023)
MIO : Mutual Information Optimization using Self-Supervised Binary Contrastive Learning
di: Manna, Siladittya, et al.
Pubblicazione: (2021)
di: Manna, Siladittya, et al.
Pubblicazione: (2021)
AI-KD: Adversarial learning and Implicit regularization for self-Knowledge Distillation
di: Kim, Hyungmin, et al.
Pubblicazione: (2022)
di: Kim, Hyungmin, et al.
Pubblicazione: (2022)
m2mKD: Module-to-Module Knowledge Distillation for Modular Transformers
di: Lo, Ka Man, et al.
Pubblicazione: (2024)
di: Lo, Ka Man, et al.
Pubblicazione: (2024)
The OCR Quest for Generalization: Learning to recognize low-resource alphabets with model editing
di: Rodríguez, Adrià Molina, et al.
Pubblicazione: (2025)
di: Rodríguez, Adrià Molina, et al.
Pubblicazione: (2025)
AuG-KD: Anchor-Based Mixup Generation for Out-of-Domain Knowledge Distillation
di: Tang, Zihao, et al.
Pubblicazione: (2024)
di: Tang, Zihao, et al.
Pubblicazione: (2024)
A Transformer Based Handwriting Recognition System Jointly Using Online and Offline Features
di: Lodh, Ayush, et al.
Pubblicazione: (2025)
di: Lodh, Ayush, et al.
Pubblicazione: (2025)
Conformal uncertainty quantification to evaluate predictive fairness of foundation AI model for skin lesion classes across patient demographics
di: Bhattacharyya, Swarnava, et al.
Pubblicazione: (2025)
di: Bhattacharyya, Swarnava, et al.
Pubblicazione: (2025)
DeepKD: A Deeply Decoupled and Denoised Knowledge Distillation Trainer
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
di: Huang, Haiduo, et al.
Pubblicazione: (2025)
Adv-KD: Adversarial Knowledge Distillation for Faster Diffusion Sampling
di: Mekonnen, Kidist Amde, et al.
Pubblicazione: (2024)
di: Mekonnen, Kidist Amde, et al.
Pubblicazione: (2024)
Optimizing YOLOv5s Object Detection through Knowledge Distillation algorithm
di: Huang, Guanming, et al.
Pubblicazione: (2024)
di: Huang, Guanming, et al.
Pubblicazione: (2024)
KD-OCT: Efficient Knowledge Distillation for Clinical-Grade Retinal OCT Classification
di: Nourbakhsh, Erfan, et al.
Pubblicazione: (2025)
di: Nourbakhsh, Erfan, et al.
Pubblicazione: (2025)
Dynamically Scaled Temperature in Self-Supervised Contrastive Learning
di: Manna, Siladittya, et al.
Pubblicazione: (2023)
di: Manna, Siladittya, et al.
Pubblicazione: (2023)
CrossKD: Cross-Head Knowledge Distillation for Object Detection
di: Wang, Jiabao, et al.
Pubblicazione: (2023)
di: Wang, Jiabao, et al.
Pubblicazione: (2023)
GlobalDoc: A Cross-Modal Vision-Language Framework for Real-World Document Image Retrieval and Classification
di: Bakkali, Souhail, et al.
Pubblicazione: (2023)
di: Bakkali, Souhail, et al.
Pubblicazione: (2023)
Graph Query Networks for Object Detection with Automotive Radar
di: Saini, Loveneet, et al.
Pubblicazione: (2025)
di: Saini, Loveneet, et al.
Pubblicazione: (2025)
LIB-KD: Teaching Inductive Bias for Efficient Vision Transformer Distillation and Compression
di: Habib, Gousia, et al.
Pubblicazione: (2023)
di: Habib, Gousia, et al.
Pubblicazione: (2023)
HyperKD: Distilling Cross-Spectral Knowledge in Masked Autoencoders via Inverse Domain Shift with Spatial-Aware Masking and Specialized Loss
di: Matin, Abdul, et al.
Pubblicazione: (2025)
di: Matin, Abdul, et al.
Pubblicazione: (2025)
DisCoM-KD: Cross-Modal Knowledge Distillation via Disentanglement Representation and Adversarial Learning
di: Ienco, Dino, et al.
Pubblicazione: (2024)
di: Ienco, Dino, et al.
Pubblicazione: (2024)
Lecture Video Visual Objects (LVVO) Dataset: A Benchmark for Visual Object Detection in Educational Videos
di: Biswas, Dipayan, et al.
Pubblicazione: (2025)
di: Biswas, Dipayan, et al.
Pubblicazione: (2025)
SpectralKD: A Unified Framework for Interpreting and Distilling Vision Transformers via Spectral Analysis
di: Tian, Huiyuan, et al.
Pubblicazione: (2024)
di: Tian, Huiyuan, et al.
Pubblicazione: (2024)
CountLoop: Training-Free High-Instance Image Generation via Iterative Agent Guidance
di: Mondal, Anindya, et al.
Pubblicazione: (2025)
di: Mondal, Anindya, et al.
Pubblicazione: (2025)
Interpretable Dynamic Graph Neural Networks for Small Occluded Object Detection and Tracking
di: Soudeep, Shahriar, et al.
Pubblicazione: (2024)
di: Soudeep, Shahriar, et al.
Pubblicazione: (2024)
SAFE-KD: Risk-Controlled Early-Exit Distillation for Vision Backbones
di: Khazem, Salim
Pubblicazione: (2026)
di: Khazem, Salim
Pubblicazione: (2026)
Detection-Fusion for Knowledge Graph Extraction from Videos
di: Das, Taniya, et al.
Pubblicazione: (2024)
di: Das, Taniya, et al.
Pubblicazione: (2024)
Measuring the Impact of Scene Level Objects on Object Detection: Towards Quantitative Explanations of Detection Decisions
di: Haar, Lynn Vonder, et al.
Pubblicazione: (2024)
di: Haar, Lynn Vonder, et al.
Pubblicazione: (2024)
Documenti analoghi
-
CraftSVG: Multi-Object Text-to-SVG Synthesis via Layout Guided Diffusion
di: Banerjee, Ayan, et al.
Pubblicazione: (2024) -
DistilDoc: Knowledge Distillation for Visually-Rich Document Applications
di: Van Landeghem, Jordy, et al.
Pubblicazione: (2024) -
DocRevive: A Unified Pipeline for Document Text Restoration
di: Purkayastha, Kunal, et al.
Pubblicazione: (2026) -
Diving into the Depths of Spotting Text in Multi-Domain Noisy Scenes
di: Das, Alloy, et al.
Pubblicazione: (2023) -
TaleDiffusion: Multi-Character Story Generation with Dialogue Rendering
di: Banerjee, Ayan, et al.
Pubblicazione: (2025)