AI-Generated Lecture Slides for Improving Slide Element Detection and Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Maniyar, Suyash, Trivedi, Vishvesh, Mondal, Ajoy, Mishra, Anand, Jawahar, C. V. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Generating Narrated Lecture Videos from Slides with Synchronized Highlights
by: Holmberg, Alexander
Published: (2025)
by: Holmberg, Alexander
Published: (2025)
Multiple Instance Learning for Glioma Diagnosis using Hematoxylin and Eosin Whole Slide Images: An Indian Cohort Study
by: Chauhan, Ekansh, et al.
Published: (2024)
by: Chauhan, Ekansh, et al.
Published: (2024)
SlideCoder: Layout-aware RAG-enhanced Hierarchical Slide Generation from Design
by: Tang, Wenxin, et al.
Published: (2025)
by: Tang, Wenxin, et al.
Published: (2025)
PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajectory Prediction
by: Mishra, Naman, et al.
Published: (2026)
by: Mishra, Naman, et al.
Published: (2026)
Script-to-Slide Grounding: Grounding Script Sentences to Slide Objects for Automatic Instructional Video Generation
by: Suzuki, Rena, et al.
Published: (2026)
by: Suzuki, Rena, et al.
Published: (2026)
IndicSTR12: A Dataset for Indic Scene Text Recognition
by: Lunia, Harsh, et al.
Published: (2024)
by: Lunia, Harsh, et al.
Published: (2024)
Advancing Question Answering on Handwritten Documents: A State-of-the-Art Recognition-Based Model for HW-SQuAD
by: Pal, Aniket, et al.
Published: (2024)
by: Pal, Aniket, et al.
Published: (2024)
ChatBCG: Can AI Read Your Slide Deck?
by: Singh, Nikita, et al.
Published: (2024)
by: Singh, Nikita, et al.
Published: (2024)
DesignLab: Designing Slides Through Iterative Detection and Correction
by: Yun, Jooyeol, et al.
Published: (2025)
by: Yun, Jooyeol, et al.
Published: (2025)
SlideChat: A Large Vision-Language Assistant for Whole-Slide Pathology Image Understanding
by: Chen, Ying, et al.
Published: (2024)
by: Chen, Ying, et al.
Published: (2024)
Towards Deployable OCR models for Indic languages
by: Mathew, Minesh, et al.
Published: (2022)
by: Mathew, Minesh, et al.
Published: (2022)
HW-MLVQA: Elucidating Multilingual Handwritten Document Understanding with a Comprehensive VQA Benchmark
by: Pal, Aniket, et al.
Published: (2025)
by: Pal, Aniket, et al.
Published: (2025)
CRAFT: Critic-Refined Adaptive Key-Frame Targeting for Multimodal Video Question Answering
by: Bhosale, Mahesh, et al.
Published: (2026)
by: Bhosale, Mahesh, et al.
Published: (2026)
WSI-VQA: Interpreting Whole Slide Images by Generative Visual Question Answering
by: Chen, Pingyi, et al.
Published: (2024)
by: Chen, Pingyi, et al.
Published: (2024)
PathNavigate: A Training-Free Pathology Agent with Surprise-Guided Scan and Shared Slide Memory for Whole-Slide Image VQA
by: Yang, Chunze, et al.
Published: (2026)
by: Yang, Chunze, et al.
Published: (2026)
Temporal Object-Aware Vision Transformer for Few-Shot Video Object Detection
by: Kumar, Yogesh, et al.
Published: (2025)
by: Kumar, Yogesh, et al.
Published: (2025)
Spatial Blindness in Whole-Slide Multiple Instance Learning
by: Li, Xiangyu, et al.
Published: (2026)
by: Li, Xiangyu, et al.
Published: (2026)
Transcriptomics-guided Slide Representation Learning in Computational Pathology
by: Jaume, Guillaume, et al.
Published: (2024)
by: Jaume, Guillaume, et al.
Published: (2024)
Hypergraph Mamba for Efficient Whole Slide Image Understanding
by: Lu, Jiaxuan, et al.
Published: (2025)
by: Lu, Jiaxuan, et al.
Published: (2025)
SWAT: Sliding Window Adversarial Training for Gradual Domain Adaptation
by: Wang, Zixi, et al.
Published: (2025)
by: Wang, Zixi, et al.
Published: (2025)
DriveSafe: A Framework for Risk Detection and Safety Suggestions in Driving Scenarios
by: Artham, Sainithin, et al.
Published: (2026)
by: Artham, Sainithin, et al.
Published: (2026)
Efficient AI-Driven Multi-Section Whole Slide Image Analysis for Biochemical Recurrence Prediction in Prostate Cancer
by: Cho, Yesung, et al.
Published: (2026)
by: Cho, Yesung, et al.
Published: (2026)
Domain-Specific Pre-training Improves Confidence in Whole Slide Image Classification
by: Chitnis, Soham Rohit, et al.
Published: (2023)
by: Chitnis, Soham Rohit, et al.
Published: (2023)
PATHS: A Hierarchical Transformer for Efficient Whole Slide Image Analysis
by: Buzzard, Zak, et al.
Published: (2024)
by: Buzzard, Zak, et al.
Published: (2024)
TAKT: Target-Aware Knowledge Transfer for Whole Slide Image Classification
by: Xiong, Conghao, et al.
Published: (2023)
by: Xiong, Conghao, et al.
Published: (2023)
Cross-Patient Pseudo Bags Generation and Curriculum Contrastive Learning for Imbalanced Multiclassification of Whole Slide Image
by: Wu, Yonghuang, et al.
Published: (2024)
by: Wu, Yonghuang, et al.
Published: (2024)
Improving Quality Control of Whole Slide Images by Explicit Artifact Augmentation
by: Jurgas, Artur, et al.
Published: (2024)
by: Jurgas, Artur, et al.
Published: (2024)
WsiCaption: Multiple Instance Generation of Pathology Reports for Gigapixel Whole-Slide Images
by: Chen, Pingyi, et al.
Published: (2023)
by: Chen, Pingyi, et al.
Published: (2023)
DECKBench: Benchmarking Multi-Agent Frameworks for Academic Slide Generation and Editing
by: Jang, Daesik, et al.
Published: (2026)
by: Jang, Daesik, et al.
Published: (2026)
VLM-SlideEval: Evaluating VLMs on Structured Comprehension and Perturbation Sensitivity in PPT
by: Kang, Hyeonsu, et al.
Published: (2025)
by: Kang, Hyeonsu, et al.
Published: (2025)
Enhancing Whole Slide Image Classification through Supervised Contrastive Domain Adaptation
by: Carretero, Ilán, et al.
Published: (2024)
by: Carretero, Ilán, et al.
Published: (2024)
Finding Regions of Interest in Whole Slide Images Using Multiple Instance Learning
by: Afonso, Martim, et al.
Published: (2024)
by: Afonso, Martim, et al.
Published: (2024)
KPIs 2024 Challenge: Advancing Glomerular Segmentation from Patch- to Slide-Level
by: Deng, Ruining, et al.
Published: (2025)
by: Deng, Ruining, et al.
Published: (2025)
Agent Aggregator with Mask Denoise Mechanism for Histopathology Whole Slide Image Analysis
by: Ling, Xitong, et al.
Published: (2024)
by: Ling, Xitong, et al.
Published: (2024)
Self-Contrastive Weakly Supervised Learning Framework for Prognostic Prediction Using Whole Slide Images
by: Fuster, Saul, et al.
Published: (2024)
by: Fuster, Saul, et al.
Published: (2024)
Zero-Shot Whole Slide Image Retrieval in Histopathology Using Embeddings of Foundation Models
by: Alfasly, Saghir, et al.
Published: (2024)
by: Alfasly, Saghir, et al.
Published: (2024)
GCE-MIL: Faithful and Recoverable Evidence for Multiple Instance Learning in Whole-Slide Imaging
by: Li, Xiangyu, et al.
Published: (2026)
by: Li, Xiangyu, et al.
Published: (2026)
MambaBack: Bridging Local Features and Global Contexts in Whole Slide Image Analysis
by: Chen, Sicheng, et al.
Published: (2026)
by: Chen, Sicheng, et al.
Published: (2026)
CARMIL: Context-Aware Regularization on Multiple Instance Learning models for Whole Slide Images
by: Saada, Thiziri Nait, et al.
Published: (2024)
by: Saada, Thiziri Nait, et al.
Published: (2024)
Exploiting Low-Dimensional Manifold of Features for Few-Shot Whole Slide Image Classification
by: Xiong, Conghao, et al.
Published: (2025)
by: Xiong, Conghao, et al.
Published: (2025)
Similar Items
-
Generating Narrated Lecture Videos from Slides with Synchronized Highlights
by: Holmberg, Alexander
Published: (2025) -
Multiple Instance Learning for Glioma Diagnosis using Hematoxylin and Eosin Whole Slide Images: An Indian Cohort Study
by: Chauhan, Ekansh, et al.
Published: (2024) -
SlideCoder: Layout-aware RAG-enhanced Hierarchical Slide Generation from Design
by: Tang, Wenxin, et al.
Published: (2025) -
PEDESTRIANQA: A Benchmark for Vision-Language Models on Pedestrian Intention and Trajectory Prediction
by: Mishra, Naman, et al.
Published: (2026) -
Script-to-Slide Grounding: Grounding Script Sentences to Slide Objects for Automatic Instructional Video Generation
by: Suzuki, Rena, et al.
Published: (2026)