Gespeichert in:
| Hauptverfasser: | Singh, Nikita, Balian, Rob, Martinelli, Lukas |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2407.12875 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SlideChat: A Large Vision-Language Assistant for Whole-Slide Pathology Image Understanding
von: Chen, Ying, et al.
Veröffentlicht: (2024)
von: Chen, Ying, et al.
Veröffentlicht: (2024)
AI-Generated Lecture Slides for Improving Slide Element Detection and Retrieval
von: Maniyar, Suyash, et al.
Veröffentlicht: (2025)
von: Maniyar, Suyash, et al.
Veröffentlicht: (2025)
TRINS: Towards Multimodal Language Models that Can Read
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
Transcriptomics-guided Slide Representation Learning in Computational Pathology
von: Jaume, Guillaume, et al.
Veröffentlicht: (2024)
von: Jaume, Guillaume, et al.
Veröffentlicht: (2024)
Your AI-Generated Image Detector Can Secretly Achieve SOTA Accuracy, If Calibrated
von: Yang, Muli, et al.
Veröffentlicht: (2026)
von: Yang, Muli, et al.
Veröffentlicht: (2026)
DetailMaster: Can Your Text-to-Image Model Handle Long Prompts?
von: Jiao, Qirui, et al.
Veröffentlicht: (2025)
von: Jiao, Qirui, et al.
Veröffentlicht: (2025)
A Hybrid Machine Learning Model for Cerebral Palsy Detection
von: Singh, Karan Kumar, et al.
Veröffentlicht: (2026)
von: Singh, Karan Kumar, et al.
Veröffentlicht: (2026)
Your One-Stop Solution for AI-Generated Video Detection
von: Ma, Long, et al.
Veröffentlicht: (2026)
von: Ma, Long, et al.
Veröffentlicht: (2026)
SlideCoder: Layout-aware RAG-enhanced Hierarchical Slide Generation from Design
von: Tang, Wenxin, et al.
Veröffentlicht: (2025)
von: Tang, Wenxin, et al.
Veröffentlicht: (2025)
Your Vision-Language Model Can't Even Count to 20: Exposing the Failures of VLMs in Compositional Counting
von: Guo, Xuyang, et al.
Veröffentlicht: (2025)
von: Guo, Xuyang, et al.
Veröffentlicht: (2025)
Script-to-Slide Grounding: Grounding Script Sentences to Slide Objects for Automatic Instructional Video Generation
von: Suzuki, Rena, et al.
Veröffentlicht: (2026)
von: Suzuki, Rena, et al.
Veröffentlicht: (2026)
Can I Trust Your Answer? Visually Grounded Video Question Answering
von: Xiao, Junbin, et al.
Veröffentlicht: (2023)
von: Xiao, Junbin, et al.
Veröffentlicht: (2023)
Efficient AI-Driven Multi-Section Whole Slide Image Analysis for Biochemical Recurrence Prediction in Prostate Cancer
von: Cho, Yesung, et al.
Veröffentlicht: (2026)
von: Cho, Yesung, et al.
Veröffentlicht: (2026)
ChatGen: Automatic Text-to-Image Generation From FreeStyle Chatting
von: Jia, Chengyou, et al.
Veröffentlicht: (2024)
von: Jia, Chengyou, et al.
Veröffentlicht: (2024)
Is My Data in Your AI? Membership Inference Test (MINT) applied to Face Biometrics
von: DeAlcala, Daniel, et al.
Veröffentlicht: (2024)
von: DeAlcala, Daniel, et al.
Veröffentlicht: (2024)
Can ChatGPT Perform Image Splicing Detection? A Preliminary Study
von: Nath, Souradip
Veröffentlicht: (2025)
von: Nath, Souradip
Veröffentlicht: (2025)
TextCraftor: Your Text Encoder Can be Image Quality Controller
von: Li, Yanyu, et al.
Veröffentlicht: (2024)
von: Li, Yanyu, et al.
Veröffentlicht: (2024)
Can Your Generative Model Detect Out-of-Distribution Covariate Shift?
von: Viviers, Christiaan, et al.
Veröffentlicht: (2024)
von: Viviers, Christiaan, et al.
Veröffentlicht: (2024)
PathNavigate: A Training-Free Pathology Agent with Surprise-Guided Scan and Shared Slide Memory for Whole-Slide Image VQA
von: Yang, Chunze, et al.
Veröffentlicht: (2026)
von: Yang, Chunze, et al.
Veröffentlicht: (2026)
Can AI Assistance Aid in the Grading of Handwritten Answer Sheets?
von: Sil, Pritam, et al.
Veröffentlicht: (2024)
von: Sil, Pritam, et al.
Veröffentlicht: (2024)
Spatial Blindness in Whole-Slide Multiple Instance Learning
von: Li, Xiangyu, et al.
Veröffentlicht: (2026)
von: Li, Xiangyu, et al.
Veröffentlicht: (2026)
Hypergraph Mamba for Efficient Whole Slide Image Understanding
von: Lu, Jiaxuan, et al.
Veröffentlicht: (2025)
von: Lu, Jiaxuan, et al.
Veröffentlicht: (2025)
A Multicenter Benchmark of Multiple Instance Learning Models for Lymphoma Subtyping from HE-stained Whole Slide Images
von: Umer, Rao Muhammad, et al.
Veröffentlicht: (2025)
von: Umer, Rao Muhammad, et al.
Veröffentlicht: (2025)
AI-Generated Content Enhanced Computer-Aided Diagnosis Model for Thyroid Nodules: A ChatGPT-Style Assistant
von: Yao, Jincao, et al.
Veröffentlicht: (2024)
von: Yao, Jincao, et al.
Veröffentlicht: (2024)
SWAT: Sliding Window Adversarial Training for Gradual Domain Adaptation
von: Wang, Zixi, et al.
Veröffentlicht: (2025)
von: Wang, Zixi, et al.
Veröffentlicht: (2025)
DesignLab: Designing Slides Through Iterative Detection and Correction
von: Yun, Jooyeol, et al.
Veröffentlicht: (2025)
von: Yun, Jooyeol, et al.
Veröffentlicht: (2025)
Generating Narrated Lecture Videos from Slides with Synchronized Highlights
von: Holmberg, Alexander
Veröffentlicht: (2025)
von: Holmberg, Alexander
Veröffentlicht: (2025)
Can ChatGPT Learn My Life From a Week of First-Person Video?
von: Harris, Keegan
Veröffentlicht: (2025)
von: Harris, Keegan
Veröffentlicht: (2025)
LLaVA-Read: Enhancing Reading Ability of Multimodal Language Models
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024)
Beyond the First Read: AI-Assisted Perceptual Error Detection in Chest Radiography Accounting for Interobserver Variability
von: Vutukuri, Adhrith, et al.
Veröffentlicht: (2025)
von: Vutukuri, Adhrith, et al.
Veröffentlicht: (2025)
Is ChatGPT-5 Ready for Mammogram VQA?
von: Li, Qiang, et al.
Veröffentlicht: (2025)
von: Li, Qiang, et al.
Veröffentlicht: (2025)
PATHS: A Hierarchical Transformer for Efficient Whole Slide Image Analysis
von: Buzzard, Zak, et al.
Veröffentlicht: (2024)
von: Buzzard, Zak, et al.
Veröffentlicht: (2024)
TAKT: Target-Aware Knowledge Transfer for Whole Slide Image Classification
von: Xiong, Conghao, et al.
Veröffentlicht: (2023)
von: Xiong, Conghao, et al.
Veröffentlicht: (2023)
Assessing Greenspace Attractiveness with ChatGPT, Claude, and Gemini: Do AI Models Reflect Human Perceptions?
von: Malekzadeh, Milad, et al.
Veröffentlicht: (2025)
von: Malekzadeh, Milad, et al.
Veröffentlicht: (2025)
GPTDrawer: Enhancing Visual Synthesis through ChatGPT
von: Li, Kun, et al.
Veröffentlicht: (2024)
von: Li, Kun, et al.
Veröffentlicht: (2024)
Enhancing Whole Slide Image Classification through Supervised Contrastive Domain Adaptation
von: Carretero, Ilán, et al.
Veröffentlicht: (2024)
von: Carretero, Ilán, et al.
Veröffentlicht: (2024)
Finding Regions of Interest in Whole Slide Images Using Multiple Instance Learning
von: Afonso, Martim, et al.
Veröffentlicht: (2024)
von: Afonso, Martim, et al.
Veröffentlicht: (2024)
Agent Aggregator with Mask Denoise Mechanism for Histopathology Whole Slide Image Analysis
von: Ling, Xitong, et al.
Veröffentlicht: (2024)
von: Ling, Xitong, et al.
Veröffentlicht: (2024)
WSI-VQA: Interpreting Whole Slide Images by Generative Visual Question Answering
von: Chen, Pingyi, et al.
Veröffentlicht: (2024)
von: Chen, Pingyi, et al.
Veröffentlicht: (2024)
PVChat: Personalized Video Chat with One-Shot Learning
von: Shi, Yufei, et al.
Veröffentlicht: (2025)
von: Shi, Yufei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SlideChat: A Large Vision-Language Assistant for Whole-Slide Pathology Image Understanding
von: Chen, Ying, et al.
Veröffentlicht: (2024) -
AI-Generated Lecture Slides for Improving Slide Element Detection and Retrieval
von: Maniyar, Suyash, et al.
Veröffentlicht: (2025) -
TRINS: Towards Multimodal Language Models that Can Read
von: Zhang, Ruiyi, et al.
Veröffentlicht: (2024) -
Transcriptomics-guided Slide Representation Learning in Computational Pathology
von: Jaume, Guillaume, et al.
Veröffentlicht: (2024) -
Your AI-Generated Image Detector Can Secretly Achieve SOTA Accuracy, If Calibrated
von: Yang, Muli, et al.
Veröffentlicht: (2026)