Unsupervised Training of Vision Transformers with Synthetic Negatives
Fuente:
arXiv
Saved in:
| Main Authors: | Giakoumoglou, Nikolaos, Floros, Andreas, Papadopoulos, Kleanthis Marios, Stathaki, Tania |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fake & Square: Training Self-Supervised Vision Transformers with Synthetic Data and Synthetic Hard Negatives
by: Giakoumoglou, Nikolaos, et al.
Published: (2025)
by: Giakoumoglou, Nikolaos, et al.
Published: (2025)
Cluster Contrast for Unsupervised Visual Representation Learning
by: Giakoumoglou, Nikolaos, et al.
Published: (2025)
by: Giakoumoglou, Nikolaos, et al.
Published: (2025)
SynCo: Synthetic Hard Negatives for Contrastive Visual Representation Learning
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)
A Review on Discriminative Self-supervised Learning Methods in Computer Vision
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)
Relational Representation Distillation
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)
Discriminative and Consistent Representation Distillation
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)
Distilling Invariant Representations with Dual Augmentation
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)
Comparing ImageNet Pre-training with Digital Pathology Foundation Models for Whole Slide Image-Based Survival Analysis
by: Papadopoulos, Kleanthis Marios, et al.
Published: (2024)
by: Papadopoulos, Kleanthis Marios, et al.
Published: (2024)
Caption-Matching: A Multimodal Approach for Cross-Domain Image Retrieval
by: Iijima, Lucas, et al.
Published: (2024)
by: Iijima, Lucas, et al.
Published: (2024)
Towards Optimal Trade-offs in Knowledge Distillation for CNNs and Vision Transformers at the Edge
by: Violos, John, et al.
Published: (2024)
by: Violos, John, et al.
Published: (2024)
DiffusionPrint: Learning Generative Fingerprints for Diffusion-Based Inpainting Localization
by: Giakoumoglou, Paschalis, et al.
Published: (2026)
by: Giakoumoglou, Paschalis, et al.
Published: (2026)
Training-Free Unsupervised Prompt for Vision-Language Models
by: Long, Sifan, et al.
Published: (2024)
by: Long, Sifan, et al.
Published: (2024)
SIDBench: A Python Framework for Reliably Assessing Synthetic Image Detection Methods
by: Schinas, Manos, et al.
Published: (2024)
by: Schinas, Manos, et al.
Published: (2024)
TGIF2: Extended Text-Guided Inpainting Forgery Dataset & Benchmark
by: Mareen, Hannes, et al.
Published: (2026)
by: Mareen, Hannes, et al.
Published: (2026)
Unsupervised Synthetic Image Attribution: Alignment and Disentanglement
by: Liu, Zongfang, et al.
Published: (2026)
by: Liu, Zongfang, et al.
Published: (2026)
Vision Transformers Don't Need Trained Registers
by: Jiang, Nick, et al.
Published: (2025)
by: Jiang, Nick, et al.
Published: (2025)
FALCON: False-Negative Aware Learning of Contrastive Negatives in Vision-Language Alignment
by: Kim, Myunsoo, et al.
Published: (2025)
by: Kim, Myunsoo, et al.
Published: (2025)
TNG-CLIP:Training-Time Negation Data Generation for Negation Awareness of CLIP
by: Cai, Yuliang, et al.
Published: (2025)
by: Cai, Yuliang, et al.
Published: (2025)
LLaVA-CKD: Bottom-Up Cascaded Knowledge Distillation for Vision-Language Models
by: Gkalelis, Nikolaos, et al.
Published: (2026)
by: Gkalelis, Nikolaos, et al.
Published: (2026)
G-FARS: Gradient-Field-based Auto-Regressive Sampling for 3D Part Grouping
by: Cheng, Junfeng, et al.
Published: (2024)
by: Cheng, Junfeng, et al.
Published: (2024)
Size Aware Cross-shape Scribble Supervision for Medical Image Segmentation
by: Yuan, Jing, et al.
Published: (2024)
by: Yuan, Jing, et al.
Published: (2024)
EFTViT: Efficient Federated Training of Vision Transformers with Masked Images on Resource-Constrained Clients
by: Wu, Meihan, et al.
Published: (2024)
by: Wu, Meihan, et al.
Published: (2024)
JetViT: Efficient High-Resolution Vision Transformer with Post-Training Attention Search
by: Zou, Dongyun, et al.
Published: (2026)
by: Zou, Dongyun, et al.
Published: (2026)
Trio-ViT: Post-Training Quantization and Acceleration for Softmax-Free Efficient Vision Transformer
by: Shi, Huihong, et al.
Published: (2024)
by: Shi, Huihong, et al.
Published: (2024)
MAFA: Managing False Negatives for Vision-Language Pre-training
by: Byun, Jaeseok, et al.
Published: (2023)
by: Byun, Jaeseok, et al.
Published: (2023)
SynDroneVision: A Synthetic Dataset for Image-Based Drone Detection
by: Lenhard, Tamara R., et al.
Published: (2024)
by: Lenhard, Tamara R., et al.
Published: (2024)
EUDA: An Efficient Unsupervised Domain Adaptation via Self-Supervised Vision Transformer
by: Abedi, Ali, et al.
Published: (2024)
by: Abedi, Ali, et al.
Published: (2024)
IPTQ-ViT: Post-Training Quantization of Non-linear Functions for Integer-only Vision Transformers
by: Kim, Gihwan, et al.
Published: (2025)
by: Kim, Gihwan, et al.
Published: (2025)
Progressive Fine-to-Coarse Reconstruction for Accurate Low-Bit Post-Training Quantization in Vision Transformers
by: Ding, Rui, et al.
Published: (2024)
by: Ding, Rui, et al.
Published: (2024)
Feature Fusion Transferability Aware Transformer for Unsupervised Domain Adaptation
by: Yu, Xiaowei, et al.
Published: (2024)
by: Yu, Xiaowei, et al.
Published: (2024)
Vision-Based Neurosurgical Guidance: Unsupervised Localization and Camera-Pose Prediction
by: Sarwin, Gary, et al.
Published: (2024)
by: Sarwin, Gary, et al.
Published: (2024)
Hide and Seek: Investigating Redundancy in Earth Observation Imagery
by: Papazafeiropoulos, Tasos, et al.
Published: (2026)
by: Papazafeiropoulos, Tasos, et al.
Published: (2026)
Repurposing Stable Diffusion Attention for Training-Free Unsupervised Interactive Segmentation
by: Karmann, Markus, et al.
Published: (2024)
by: Karmann, Markus, et al.
Published: (2024)
Learning to Mask and Permute Visual Tokens for Vision Transformer Pre-Training
by: Baraldi, Lorenzo, et al.
Published: (2023)
by: Baraldi, Lorenzo, et al.
Published: (2023)
Gaze-Informed Vision Transformers: Predicting Driving Decisions Under Uncertainty
by: Koorathota, Sharath, et al.
Published: (2023)
by: Koorathota, Sharath, et al.
Published: (2023)
Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations
by: Xue, Zhiyu, et al.
Published: (2025)
by: Xue, Zhiyu, et al.
Published: (2025)
Adversarial Robustness of Vision in Open Foundation Models
by: Fox, Jonathon, et al.
Published: (2025)
by: Fox, Jonathon, et al.
Published: (2025)
T-TAME: Trainable Attention Mechanism for Explaining Convolutional Networks and Vision Transformers
by: Ntrougkas, Mariano V., et al.
Published: (2024)
by: Ntrougkas, Mariano V., et al.
Published: (2024)
Investigation of Accuracy and Bias in Face Recognition Trained with Synthetic Data
by: Korshunov, Pavel, et al.
Published: (2025)
by: Korshunov, Pavel, et al.
Published: (2025)
Position: Quo Vadis, Unsupervised Time Series Anomaly Detection?
by: Sarfraz, M. Saquib, et al.
Published: (2024)
by: Sarfraz, M. Saquib, et al.
Published: (2024)
Similar Items
-
Fake & Square: Training Self-Supervised Vision Transformers with Synthetic Data and Synthetic Hard Negatives
by: Giakoumoglou, Nikolaos, et al.
Published: (2025) -
Cluster Contrast for Unsupervised Visual Representation Learning
by: Giakoumoglou, Nikolaos, et al.
Published: (2025) -
SynCo: Synthetic Hard Negatives for Contrastive Visual Representation Learning
by: Giakoumoglou, Nikolaos, et al.
Published: (2024) -
A Review on Discriminative Self-supervised Learning Methods in Computer Vision
by: Giakoumoglou, Nikolaos, et al.
Published: (2024) -
Relational Representation Distillation
by: Giakoumoglou, Nikolaos, et al.
Published: (2024)