Stitching Gaps: Fusing Situated Perceptual Knowledge with Vision Transformers for High-Level Image Classification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pandiani, Delfina Sol Martinez, Lazzari, Nicolas, Presutti, Valentina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Seeing the Intangible: Survey of Image Classification into High-Level and Abstract Categories
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2023)
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2023)
Situated Ground Truths: Enhancing Bias-Aware AI by Situating Data Labels with SituAnnotate
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2024)
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2024)
Automatic Modeling of Social Concepts Evoked by Art Images as Multimodal Frames
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2021)
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2021)
Toxic Memes: A Survey of Computational Perspectives on the Detection and Explanation of Meme Toxicities
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2024)
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2024)
HybridStitch: Pixel and Timestep Level Model Stitching for Diffusion Acceleration
von: Sun, Desen, et al.
Veröffentlicht: (2026)
von: Sun, Desen, et al.
Veröffentlicht: (2026)
Sensitive Image Classification by Vision Transformers
von: He, Hanxian, et al.
Veröffentlicht: (2024)
von: He, Hanxian, et al.
Veröffentlicht: (2024)
Deep Neural Networks Fused with Textures for Image Classification
von: Bera, Asish, et al.
Veröffentlicht: (2023)
von: Bera, Asish, et al.
Veröffentlicht: (2023)
Adaptive Knowledge Distillation for Classification of Hand Images using Explainable Vision Transformers
von: Nguyen, Thanh Thi, et al.
Veröffentlicht: (2024)
von: Nguyen, Thanh Thi, et al.
Veröffentlicht: (2024)
To Neuro-Symbolic Classification and Beyond by Compiling Description Logic Ontologies to Probabilistic Circuits
von: Lazzari, Nicolas, et al.
Veröffentlicht: (2026)
von: Lazzari, Nicolas, et al.
Veröffentlicht: (2026)
Hierarchical Vision Transformer Enhanced by Graph Convolutional Network for Image Classification
von: Jiao, Haibin
Veröffentlicht: (2026)
von: Jiao, Haibin
Veröffentlicht: (2026)
StabStitch++: Unsupervised Online Video Stitching with Spatiotemporal Bidirectional Warps
von: Nie, Lang, et al.
Veröffentlicht: (2025)
von: Nie, Lang, et al.
Veröffentlicht: (2025)
Human and AI Perceptual Differences in Image Classification Errors
von: Liu, Minghao, et al.
Veröffentlicht: (2023)
von: Liu, Minghao, et al.
Veröffentlicht: (2023)
LMLT: Low-to-high Multi-Level Vision Transformer for Image Super-Resolution
von: Kim, Jeongsoo, et al.
Veröffentlicht: (2024)
von: Kim, Jeongsoo, et al.
Veröffentlicht: (2024)
Sandra -- A Neuro-Symbolic Reasoner Based On Descriptions And Situations
von: Lazzari, Nicolas, et al.
Veröffentlicht: (2024)
von: Lazzari, Nicolas, et al.
Veröffentlicht: (2024)
Effective Damage Data Generation by Fusing Imagery with Human Knowledge Using Vision-Language Models
von: Wei, Jie, et al.
Veröffentlicht: (2025)
von: Wei, Jie, et al.
Veröffentlicht: (2025)
Probing Perceptual Constancy in Large Vision-Language Models
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
von: Sun, Haoran, et al.
Veröffentlicht: (2025)
Probing the Efficacy of Federated Parameter-Efficient Fine-Tuning of Vision Transformers for Medical Image Classification
von: Alkhunaizi, Naif, et al.
Veröffentlicht: (2024)
von: Alkhunaizi, Naif, et al.
Veröffentlicht: (2024)
Evaluating Deep Learning Models for African Wildlife Image Classification: From DenseNet to Vision Transformers
von: Aliyu, Lukman Jibril, et al.
Veröffentlicht: (2025)
von: Aliyu, Lukman Jibril, et al.
Veröffentlicht: (2025)
Stitch: Training-Free Position Control in Multimodal Diffusion Transformers
von: Bader, Jessica, et al.
Veröffentlicht: (2025)
von: Bader, Jessica, et al.
Veröffentlicht: (2025)
Fuse Before Transfer: Knowledge Fusion for Heterogeneous Distillation
von: Li, Guopeng, et al.
Veröffentlicht: (2024)
von: Li, Guopeng, et al.
Veröffentlicht: (2024)
HSVLT: Hierarchical Scale-Aware Vision-Language Transformer for Multi-Label Image Classification
von: Ouyang, Shuyi, et al.
Veröffentlicht: (2024)
von: Ouyang, Shuyi, et al.
Veröffentlicht: (2024)
Salient Mask-Guided Vision Transformer for Fine-Grained Classification
von: Demidov, Dmitry, et al.
Veröffentlicht: (2023)
von: Demidov, Dmitry, et al.
Veröffentlicht: (2023)
Uncertainty Quantification in Detection Transformers: Object-Level Calibration and Image-Level Reliability
von: Park, Young-Jin, et al.
Veröffentlicht: (2024)
von: Park, Young-Jin, et al.
Veröffentlicht: (2024)
Fusing Echocardiography Images and Medical Records for Continuous Patient Stratification
von: Painchaud, Nathan, et al.
Veröffentlicht: (2024)
von: Painchaud, Nathan, et al.
Veröffentlicht: (2024)
Siamese Transformer Networks for Few-shot Image Classification
von: Jiang, Weihao, et al.
Veröffentlicht: (2024)
von: Jiang, Weihao, et al.
Veröffentlicht: (2024)
Rethinking Plant Disease Diagnosis: Bridging the Academic-Practical Gap with Vision Transformers and Zero-Shot Learning
von: Benabbas, Wassim, et al.
Veröffentlicht: (2025)
von: Benabbas, Wassim, et al.
Veröffentlicht: (2025)
IoT Botnet Detection: Application of Vision Transformer to Classification of Network Flow Traffic
von: Wasswa, Hassan, et al.
Veröffentlicht: (2025)
von: Wasswa, Hassan, et al.
Veröffentlicht: (2025)
A Data-Centric Vision Transformer Baseline for SAR Sea Ice Classification
von: Mike-Ewewie, David, et al.
Veröffentlicht: (2026)
von: Mike-Ewewie, David, et al.
Veröffentlicht: (2026)
Dynamic Weight Adjustment for Knowledge Distillation: Leveraging Vision Transformer for High-Accuracy Lung Cancer Detection and Real-Time Deployment
von: Khan, Saif Ur Rehman, et al.
Veröffentlicht: (2025)
von: Khan, Saif Ur Rehman, et al.
Veröffentlicht: (2025)
Disentangling Visual Transformers: Patch-level Interpretability for Image Classification
von: Jeanneret, Guillaume, et al.
Veröffentlicht: (2025)
von: Jeanneret, Guillaume, et al.
Veröffentlicht: (2025)
Towards Optimal Trade-offs in Knowledge Distillation for CNNs and Vision Transformers at the Edge
von: Violos, John, et al.
Veröffentlicht: (2024)
von: Violos, John, et al.
Veröffentlicht: (2024)
TAKT: Target-Aware Knowledge Transfer for Whole Slide Image Classification
von: Xiong, Conghao, et al.
Veröffentlicht: (2023)
von: Xiong, Conghao, et al.
Veröffentlicht: (2023)
Refine-IQA: Multi-Stage Reinforcement Finetuning for Perceptual Image Quality Assessment
von: Jia, Ziheng, et al.
Veröffentlicht: (2025)
von: Jia, Ziheng, et al.
Veröffentlicht: (2025)
HiRes-FusedMIM: A High-Resolution RGB-DSM Pre-trained Model for Building-Level Remote Sensing Applications
von: Mutreja, Guneet, et al.
Veröffentlicht: (2025)
von: Mutreja, Guneet, et al.
Veröffentlicht: (2025)
Intersectional Fairness in Vision-Language Models for Medical Image Disease Classification
von: Zhang, Yupeng, et al.
Veröffentlicht: (2025)
von: Zhang, Yupeng, et al.
Veröffentlicht: (2025)
Surformer v1: Transformer-Based Surface Classification Using Tactile and Vision Features
von: Kansana, Manish, et al.
Veröffentlicht: (2025)
von: Kansana, Manish, et al.
Veröffentlicht: (2025)
Cross-Task Multi-Branch Vision Transformer for Facial Expression and Mask Wearing Classification
von: Zhu, Armando, et al.
Veröffentlicht: (2024)
von: Zhu, Armando, et al.
Veröffentlicht: (2024)
A Simple Interpretable Transformer for Fine-Grained Image Classification and Analysis
von: Paul, Dipanjyoti, et al.
Veröffentlicht: (2023)
von: Paul, Dipanjyoti, et al.
Veröffentlicht: (2023)
Effective Fine-Tuning of Vision Transformers with Low-Rank Adaptation for Privacy-Preserving Image Classification
von: Lin, Haiwei, et al.
Veröffentlicht: (2025)
von: Lin, Haiwei, et al.
Veröffentlicht: (2025)
Tiny-ViT: A Compact Vision Transformer for Efficient and Explainable Potato Leaf Disease Classification
von: Mia, Shakil, et al.
Veröffentlicht: (2026)
von: Mia, Shakil, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Seeing the Intangible: Survey of Image Classification into High-Level and Abstract Categories
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2023) -
Situated Ground Truths: Enhancing Bias-Aware AI by Situating Data Labels with SituAnnotate
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2024) -
Automatic Modeling of Social Concepts Evoked by Art Images as Multimodal Frames
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2021) -
Toxic Memes: A Survey of Computational Perspectives on the Detection and Explanation of Meme Toxicities
von: Pandiani, Delfina Sol Martinez, et al.
Veröffentlicht: (2024) -
HybridStitch: Pixel and Timestep Level Model Stitching for Diffusion Acceleration
von: Sun, Desen, et al.
Veröffentlicht: (2026)