Implicit to Explicit Entropy Regularization: Benchmarking ViT Fine-tuning under Noisy Labels
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Marrium, Maria, Mahmood, Arif, Bennamoun, Mohammed |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DFQ-ViT: Data-Free Quantization for Vision Transformers without Fine-tuning
von: Tong, Yujia, et al.
Veröffentlicht: (2025)
von: Tong, Yujia, et al.
Veröffentlicht: (2025)
AquaticCLIP: A Vision-Language Foundation Model for Underwater Scene Analysis
von: Alawode, Basit, et al.
Veröffentlicht: (2025)
von: Alawode, Basit, et al.
Veröffentlicht: (2025)
Purrturbed but Stable: Human-Cat Invariant Representations Across CNNs, ViTs and Self-Supervised ViTs
von: Shah, Arya, et al.
Veröffentlicht: (2025)
von: Shah, Arya, et al.
Veröffentlicht: (2025)
ViT-Lens: Towards Omni-modal Representations
von: Lei, Weixian, et al.
Veröffentlicht: (2023)
von: Lei, Weixian, et al.
Veröffentlicht: (2023)
SFMViT: SlowFast Meet ViT in Chaotic World
von: Lin, Jiaying, et al.
Veröffentlicht: (2024)
von: Lin, Jiaying, et al.
Veröffentlicht: (2024)
MMeViT: Multi-Modal ensemble ViT for Post-Stroke Rehabilitation Action Recognition
von: Kim, Ye-eun, et al.
Veröffentlicht: (2025)
von: Kim, Ye-eun, et al.
Veröffentlicht: (2025)
HydraViT: Stacking Heads for a Scalable ViT
von: Haberer, Janek, et al.
Veröffentlicht: (2024)
von: Haberer, Janek, et al.
Veröffentlicht: (2024)
Sub-token ViT Embedding via Stochastic Resonance Transformers
von: Lao, Dong, et al.
Veröffentlicht: (2023)
von: Lao, Dong, et al.
Veröffentlicht: (2023)
NT-VOT211: A Large-Scale Benchmark for Night-time Visual Object Tracking
von: Liu, Yu, et al.
Veröffentlicht: (2024)
von: Liu, Yu, et al.
Veröffentlicht: (2024)
CLAMP-ViT: Contrastive Data-Free Learning for Adaptive Post-Training Quantization of ViTs
von: Ramachandran, Akshat, et al.
Veröffentlicht: (2024)
von: Ramachandran, Akshat, et al.
Veröffentlicht: (2024)
Knowledge Distillation in YOLOX-ViT for Side-Scan Sonar Object Detection
von: Aubard, Martin, et al.
Veröffentlicht: (2024)
von: Aubard, Martin, et al.
Veröffentlicht: (2024)
SAC-ViT: Semantic-Aware Clustering Vision Transformer with Early Exit
von: Hu, Youbing, et al.
Veröffentlicht: (2025)
von: Hu, Youbing, et al.
Veröffentlicht: (2025)
Hybrid CNN-ViT Framework for Motion-Blurred Scene Text Restoration
von: Rashid, Umar, et al.
Veröffentlicht: (2025)
von: Rashid, Umar, et al.
Veröffentlicht: (2025)
ViT-Linearizer: Distilling Quadratic Knowledge into Linear-Time Vision Models
von: Wei, Guoyizhe, et al.
Veröffentlicht: (2025)
von: Wei, Guoyizhe, et al.
Veröffentlicht: (2025)
ViT-ProtoNet for Few-Shot Image Classification: A Multi-Benchmark Evaluation
von: Mutlu, Abdulvahap, et al.
Veröffentlicht: (2025)
von: Mutlu, Abdulvahap, et al.
Veröffentlicht: (2025)
FNBench: Benchmarking Robust Federated Learning against Noisy Labels
von: Jiang, Xuefeng, et al.
Veröffentlicht: (2025)
von: Jiang, Xuefeng, et al.
Veröffentlicht: (2025)
LF-ViT: Reducing Spatial Redundancy in Vision Transformer for Efficient Image Recognition
von: Hu, Youbing, et al.
Veröffentlicht: (2024)
von: Hu, Youbing, et al.
Veröffentlicht: (2024)
An Intermediate Fusion ViT Enables Efficient Text-Image Alignment in Diffusion Models
von: Hu, Zizhao, et al.
Veröffentlicht: (2024)
von: Hu, Zizhao, et al.
Veröffentlicht: (2024)
Filtered-ViT: A Robust Defense Against Multiple Adversarial Patch Attacks
von: Khanal, Aja, et al.
Veröffentlicht: (2025)
von: Khanal, Aja, et al.
Veröffentlicht: (2025)
VIVID-Med: LLM-Supervised Structured Pretraining for Deployable Medical ViTs
von: Wang, Xiyao, et al.
Veröffentlicht: (2026)
von: Wang, Xiyao, et al.
Veröffentlicht: (2026)
GTP-ViT: Efficient Vision Transformers via Graph-based Token Propagation
von: Xu, Xuwei, et al.
Veröffentlicht: (2023)
von: Xu, Xuwei, et al.
Veröffentlicht: (2023)
Concept-Guided Fine-Tuning: Steering ViTs away from Spurious Correlations to Improve Robustness
von: Elisha, Yehonatan, et al.
Veröffentlicht: (2026)
von: Elisha, Yehonatan, et al.
Veröffentlicht: (2026)
Trio-ViT: Post-Training Quantization and Acceleration for Softmax-Free Efficient Vision Transformer
von: Shi, Huihong, et al.
Veröffentlicht: (2024)
von: Shi, Huihong, et al.
Veröffentlicht: (2024)
ViTs are Everywhere: A Comprehensive Study Showcasing Vision Transformers in Different Domain
von: Mia, Md Sohag, et al.
Veröffentlicht: (2023)
von: Mia, Md Sohag, et al.
Veröffentlicht: (2023)
IMEX-Reg: Implicit-Explicit Regularization in the Function Space for Continual Learning
von: Bhat, Prashant, et al.
Veröffentlicht: (2024)
von: Bhat, Prashant, et al.
Veröffentlicht: (2024)
DiffPoint: Single and Multi-view Point Cloud Reconstruction with ViT Based Diffusion Model
von: Feng, Yu, et al.
Veröffentlicht: (2024)
von: Feng, Yu, et al.
Veröffentlicht: (2024)
How Can Multimodal Remote Sensing Datasets Transform Classification via SpatialNet-ViT?
von: Kashyap, Gautam Siddharth, et al.
Veröffentlicht: (2025)
von: Kashyap, Gautam Siddharth, et al.
Veröffentlicht: (2025)
Tiny-ViT: A Compact Vision Transformer for Efficient and Explainable Potato Leaf Disease Classification
von: Mia, Shakil, et al.
Veröffentlicht: (2026)
von: Mia, Shakil, et al.
Veröffentlicht: (2026)
IPTQ-ViT: Post-Training Quantization of Non-linear Functions for Integer-only Vision Transformers
von: Kim, Gihwan, et al.
Veröffentlicht: (2025)
von: Kim, Gihwan, et al.
Veröffentlicht: (2025)
Training-Free Acceleration of ViTs with Delayed Spatial Merging
von: Heo, Jung Hwan, et al.
Veröffentlicht: (2023)
von: Heo, Jung Hwan, et al.
Veröffentlicht: (2023)
Octic Vision Transformers: Quicker ViTs Through Equivariance
von: Nordström, David, et al.
Veröffentlicht: (2025)
von: Nordström, David, et al.
Veröffentlicht: (2025)
LaCViT: A Label-aware Contrastive Fine-tuning Framework for Vision Transformers
von: Long, Zijun, et al.
Veröffentlicht: (2023)
von: Long, Zijun, et al.
Veröffentlicht: (2023)
Case-Enhanced Vision Transformer: Improving Explanations of Image Similarity with a ViT-based Similarity Metric
von: Zhao, Ziwei, et al.
Veröffentlicht: (2024)
von: Zhao, Ziwei, et al.
Veröffentlicht: (2024)
MobilePlantViT: A Mobile-friendly Hybrid ViT for Generalized Plant Disease Image Classification
von: Tonmoy, Moshiur Rahman, et al.
Veröffentlicht: (2025)
von: Tonmoy, Moshiur Rahman, et al.
Veröffentlicht: (2025)
ConcatPlexer: Additional Dim1 Batching for Faster ViTs
von: Han, Donghoon, et al.
Veröffentlicht: (2023)
von: Han, Donghoon, et al.
Veröffentlicht: (2023)
H-CNN-ViT: A Hierarchical Gated Attention Multi-Branch Model for Bladder Cancer Recurrence Prediction
von: Li, Xueyang, et al.
Veröffentlicht: (2025)
von: Li, Xueyang, et al.
Veröffentlicht: (2025)
Noisy Label Processing for Classification: A Survey
von: Li, Mengting, et al.
Veröffentlicht: (2024)
von: Li, Mengting, et al.
Veröffentlicht: (2024)
Quasar-ViT: Hardware-Oriented Quantization-Aware Architecture Search for Vision Transformers
von: Li, Zhengang, et al.
Veröffentlicht: (2024)
von: Li, Zhengang, et al.
Veröffentlicht: (2024)
TAP-ViTs: Task-Adaptive Pruning for On-Device Deployment of Vision Transformers
von: Wang, Zhibo, et al.
Veröffentlicht: (2026)
von: Wang, Zhibo, et al.
Veröffentlicht: (2026)
Communication Efficient Split Learning of ViTs with Attention-based Double Compression
von: Alvetreti, Federico, et al.
Veröffentlicht: (2025)
von: Alvetreti, Federico, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
DFQ-ViT: Data-Free Quantization for Vision Transformers without Fine-tuning
von: Tong, Yujia, et al.
Veröffentlicht: (2025) -
AquaticCLIP: A Vision-Language Foundation Model for Underwater Scene Analysis
von: Alawode, Basit, et al.
Veröffentlicht: (2025) -
Purrturbed but Stable: Human-Cat Invariant Representations Across CNNs, ViTs and Self-Supervised ViTs
von: Shah, Arya, et al.
Veröffentlicht: (2025) -
ViT-Lens: Towards Omni-modal Representations
von: Lei, Weixian, et al.
Veröffentlicht: (2023) -
SFMViT: SlowFast Meet ViT in Chaotic World
von: Lin, Jiaying, et al.
Veröffentlicht: (2024)