Guardado en:
| Autores principales: | Soroka, Emi, Arzyn, Artem |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2511.03046 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
DatUS^2: Data-driven Unsupervised Semantic Segmentation with Pre-trained Self-supervised Vision Transformer
por: Kumar, Sonal, et al.
Publicado: (2024)
por: Kumar, Sonal, et al.
Publicado: (2024)
Domain-Specific Self-Supervised Pre-training for Agricultural Disease Classification: A Hierarchical Vision Transformer Study
por: Sonavane, Arnav S.
Publicado: (2026)
por: Sonavane, Arnav S.
Publicado: (2026)
SynGen-Vision: Synthetic Data Generation for training industrial vision models
por: Dubey, Alpana, et al.
Publicado: (2025)
por: Dubey, Alpana, et al.
Publicado: (2025)
DIET-CP: Lightweight and Data Efficient Self Supervised Continued Pretraining
por: Rodas, Bryan, et al.
Publicado: (2025)
por: Rodas, Bryan, et al.
Publicado: (2025)
Do All Vision Transformers Need Registers? A Cross-Architectural Reassessment
por: Baxevanakis, Spiros, et al.
Publicado: (2026)
por: Baxevanakis, Spiros, et al.
Publicado: (2026)
Massively Multi-Person 3D Human Motion Forecasting with Scene Context
por: Mueller, Felix B, et al.
Publicado: (2024)
por: Mueller, Felix B, et al.
Publicado: (2024)
PyCAT4: A Hierarchical Vision Transformer-based Framework for 3D Human Pose Estimation
por: Yang, Zongyou, et al.
Publicado: (2025)
por: Yang, Zongyou, et al.
Publicado: (2025)
On the Domain Robustness of Contrastive Vision-Language Models
por: Koddenbrock, Mario, et al.
Publicado: (2025)
por: Koddenbrock, Mario, et al.
Publicado: (2025)
Unified Local and Global Attention Interaction Modeling for Vision Transformers
por: Nguyen, Tan, et al.
Publicado: (2024)
por: Nguyen, Tan, et al.
Publicado: (2024)
Streetscape Analysis with Generative AI (SAGAI): Vision-Language Assessment and Mapping of Urban Scenes
por: Perez, Joan, et al.
Publicado: (2025)
por: Perez, Joan, et al.
Publicado: (2025)
SynthEnsemble: A Fusion of CNN, Vision Transformer, and Hybrid Models for Multi-Label Chest X-Ray Classification
por: Ashraf, S. M. Nabil, et al.
Publicado: (2023)
por: Ashraf, S. M. Nabil, et al.
Publicado: (2023)
A Review of Pseudo-Labeling for Computer Vision
por: Kage, Patrick, et al.
Publicado: (2024)
por: Kage, Patrick, et al.
Publicado: (2024)
Exploring Visual Embedding Spaces Induced by Vision Transformers for Online Auto Parts Marketplaces
por: Armijo, Cameron, et al.
Publicado: (2025)
por: Armijo, Cameron, et al.
Publicado: (2025)
VT-FSL: Bridging Vision and Text with LLMs for Few-Shot Learning
por: Li, Wenhao, et al.
Publicado: (2025)
por: Li, Wenhao, et al.
Publicado: (2025)
In Context Learning with Vision Transformers: Case Study
por: Zhao, Antony, et al.
Publicado: (2025)
por: Zhao, Antony, et al.
Publicado: (2025)
Attention-Aware Transformer-Based Aggregation Network for Video Periocular Recognition
por: Carreira, Luiz G F, et al.
Publicado: (2026)
por: Carreira, Luiz G F, et al.
Publicado: (2026)
Application of Generative Adversarial Network (GAN) for Synthetic Training Data Creation to improve performance of ANN Classifier for extracting Built-Up pixels from Landsat Satellite Imagery
por: Mukherjee, Amritendu, et al.
Publicado: (2025)
por: Mukherjee, Amritendu, et al.
Publicado: (2025)
Residual Vision Transformer (ResViT) Based Self-Supervised Learning Model for Brain Tumor Classification
por: Karagoz, Meryem Altin, et al.
Publicado: (2024)
por: Karagoz, Meryem Altin, et al.
Publicado: (2024)
Vision Transformer-based Model for Severity Quantification of Lung Pneumonia Using Chest X-ray Images
por: Slika, Bouthaina, et al.
Publicado: (2023)
por: Slika, Bouthaina, et al.
Publicado: (2023)
High-Entropy Tokens as Multimodal Failure Points in Vision-Language Models
por: He, Mengqi, et al.
Publicado: (2025)
por: He, Mengqi, et al.
Publicado: (2025)
Disentangling Generation and Regression in Stochastic Interpolants for Controllable Image Restoration
por: Liu, Yi, et al.
Publicado: (2026)
por: Liu, Yi, et al.
Publicado: (2026)
Fusion and Grouping Strategies in Deep Learning for Local Climate Zone Classification of Multimodal Remote Sensing Data
por: Thomas, Ancymol, et al.
Publicado: (2026)
por: Thomas, Ancymol, et al.
Publicado: (2026)
Comparative Analysis of Vision Transformers and Convolutional Neural Networks for Medical Image Classification
por: Kawadkar, Kunal
Publicado: (2025)
por: Kawadkar, Kunal
Publicado: (2025)
Data-driven Super-Resolution of Flood Inundation Maps using Synthetic Simulations
por: Aravamudan, Akshay, et al.
Publicado: (2025)
por: Aravamudan, Akshay, et al.
Publicado: (2025)
Kolmogorov-Arnold Attention: Is Learnable Attention Better For Vision Transformers?
por: Maity, Subhajit, et al.
Publicado: (2025)
por: Maity, Subhajit, et al.
Publicado: (2025)
DesertFormer: Transformer-Based Semantic Segmentation for Off-Road Desert Terrain Classification in Autonomous Navigation Systems
por: Chebolu, Yasaswini
Publicado: (2026)
por: Chebolu, Yasaswini
Publicado: (2026)
Serpent: Scalable and Efficient Image Restoration via Multi-scale Structured State Space Models
por: Sepehri, Mohammad Shahab, et al.
Publicado: (2024)
por: Sepehri, Mohammad Shahab, et al.
Publicado: (2024)
Diffusing DeBias: Synthetic Bias Amplification for Model Debiasing
por: Ciranni, Massimiliano, et al.
Publicado: (2025)
por: Ciranni, Massimiliano, et al.
Publicado: (2025)
LeDiFlow: Learned Distribution-guided Flow Matching to Accelerate Image Generation
por: Zwick, Pascal, et al.
Publicado: (2025)
por: Zwick, Pascal, et al.
Publicado: (2025)
Looking at Model Debiasing through the Lens of Anomaly Detection
por: Pastore, Vito Paolo, et al.
Publicado: (2024)
por: Pastore, Vito Paolo, et al.
Publicado: (2024)
Global-Local Similarity for Efficient Fine-Grained Image Recognition with Vision Transformers
por: Rios, Edwin Arkel, et al.
Publicado: (2024)
por: Rios, Edwin Arkel, et al.
Publicado: (2024)
FutureHuman3D: Forecasting Complex Long-Term 3D Human Behavior from Video Observations
por: Diller, Christian, et al.
Publicado: (2022)
por: Diller, Christian, et al.
Publicado: (2022)
A Spitting Image: Modular Superpixel Tokenization in Vision Transformers
por: Aasan, Marius, et al.
Publicado: (2024)
por: Aasan, Marius, et al.
Publicado: (2024)
Boosting Model Resilience via Implicit Adversarial Data Augmentation
por: Zhou, Xiaoling, et al.
Publicado: (2024)
por: Zhou, Xiaoling, et al.
Publicado: (2024)
A Genealogy of Foundation Models in Remote Sensing
por: Lane, Kevin, et al.
Publicado: (2025)
por: Lane, Kevin, et al.
Publicado: (2025)
Comparison Study: Glacier Calving Front Delineation in Synthetic Aperture Radar Images With Deep Learning
por: Gourmelon, Nora, et al.
Publicado: (2025)
por: Gourmelon, Nora, et al.
Publicado: (2025)
Label Delay in Online Continual Learning
por: Csaba, Botos, et al.
Publicado: (2023)
por: Csaba, Botos, et al.
Publicado: (2023)
An Immersive Multi-Elevation Multi-Seasonal Dataset for 3D Reconstruction and Visualization
por: Liu, Xijun, et al.
Publicado: (2024)
por: Liu, Xijun, et al.
Publicado: (2024)
Simplifying Source-Free Domain Adaptation for Object Detection: Effective Self-Training Strategies and Performance Insights
por: Hao, Yan, et al.
Publicado: (2024)
por: Hao, Yan, et al.
Publicado: (2024)
From Misclassifications to Outliers: Joint Reliability Assessment in Classification
por: Li, Yang, et al.
Publicado: (2026)
por: Li, Yang, et al.
Publicado: (2026)
Ejemplares similares
-
DatUS^2: Data-driven Unsupervised Semantic Segmentation with Pre-trained Self-supervised Vision Transformer
por: Kumar, Sonal, et al.
Publicado: (2024) -
Domain-Specific Self-Supervised Pre-training for Agricultural Disease Classification: A Hierarchical Vision Transformer Study
por: Sonavane, Arnav S.
Publicado: (2026) -
SynGen-Vision: Synthetic Data Generation for training industrial vision models
por: Dubey, Alpana, et al.
Publicado: (2025) -
DIET-CP: Lightweight and Data Efficient Self Supervised Continued Pretraining
por: Rodas, Bryan, et al.
Publicado: (2025) -
Do All Vision Transformers Need Registers? A Cross-Architectural Reassessment
por: Baxevanakis, Spiros, et al.
Publicado: (2026)