Margin and Consistency Supervision for Calibrated and Robust Vision Models
Fuente:
arXiv
Saved in:
| Main Author: | Khazem, Salim |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SAFE-KD: Risk-Controlled Early-Exit Distillation for Vision Backbones
by: Khazem, Salim
Published: (2026)
by: Khazem, Salim
Published: (2026)
AdapterTune: Zero-Initialized Low-Rank Adapters for Frozen Vision Transformers
by: Khazem, Salim
Published: (2026)
by: Khazem, Salim
Published: (2026)
TopoLoRA-SAM: Topology-Aware Parameter-Efficient Adaptation of Foundation Segmenters for Thin-Structure and Cross-Domain Binary Semantic Segmentation
by: Khazem, Salim
Published: (2026)
by: Khazem, Salim
Published: (2026)
PolygoNet: Leveraging Simplified Polygonal Representation for Effective Image Classification
by: Khazem, Salim, et al.
Published: (2025)
by: Khazem, Salim, et al.
Published: (2025)
MC-RFM: Geometry-Aware Few-Shot Adaptation via Mixed-Curvature Riemannian Flow Matching
by: Khazem, Salim, et al.
Published: (2026)
by: Khazem, Salim, et al.
Published: (2026)
Cyclical Temporal Encoding and Hybrid Deep Ensembles for Multistep Energy Forecasting
by: Khazem, Salim, et al.
Published: (2025)
by: Khazem, Salim, et al.
Published: (2025)
Detecting Brittle Decisions for Free: Leveraging Margin Consistency in Deep Robust Classifiers
by: Ngnawé, Jonas, et al.
Published: (2024)
by: Ngnawé, Jonas, et al.
Published: (2024)
Consistency Calibration: Improving Uncertainty Calibration via Consistency among Perturbed Neighbors
by: Tao, Linwei, et al.
Published: (2024)
by: Tao, Linwei, et al.
Published: (2024)
BYO-Eval: Build Your Own Dataset for Fine-Grained Visual Assessment of Multimodal Language Models
by: Arnould, Ludovic, et al.
Published: (2025)
by: Arnould, Ludovic, et al.
Published: (2025)
From Theory to Decision Rule: Calibrating the Noisy-Label Crossover for Vision-Language Model Weak Supervision Across Three Medical-Imaging Benchmarks
by: Xu, Bruce Changlong, et al.
Published: (2026)
by: Xu, Bruce Changlong, et al.
Published: (2026)
Robust Representation Consistency Model via Contrastive Denoising
by: Lei, Jiachen, et al.
Published: (2025)
by: Lei, Jiachen, et al.
Published: (2025)
Self-Calibrated Tuning of Vision-Language Models for Out-of-Distribution Detection
by: Yu, Geng, et al.
Published: (2024)
by: Yu, Geng, et al.
Published: (2024)
Dataset Distillation for Pre-Trained Self-Supervised Vision Models
by: Cazenavette, George, et al.
Published: (2025)
by: Cazenavette, George, et al.
Published: (2025)
Robust CLIP: Unsupervised Adversarial Fine-Tuning of Vision Embeddings for Robust Large Vision-Language Models
by: Schlarmann, Christian, et al.
Published: (2024)
by: Schlarmann, Christian, et al.
Published: (2024)
Unified Supervision For Vision-Language Modeling in 3D Computed Tomography
by: Lee, Hao-Chih, et al.
Published: (2025)
by: Lee, Hao-Chih, et al.
Published: (2025)
CP-MoE: Consistency-Preserving Mixture-of-Experts for Continual Learning
by: Liu, Yang, et al.
Published: (2026)
by: Liu, Yang, et al.
Published: (2026)
Mitigating Hallucinations via Inter-Layer Consistency Aggregation in Large Vision-Language Models
by: Tang, Kai, et al.
Published: (2025)
by: Tang, Kai, et al.
Published: (2025)
Hierarchically Robust Zero-shot Vision-language Models
by: Dong, Junhao, et al.
Published: (2026)
by: Dong, Junhao, et al.
Published: (2026)
Robust Alzheimer's Progression Modeling using Cross-Domain Self-Supervised Deep Learning
by: Dadsetan, Saba, et al.
Published: (2022)
by: Dadsetan, Saba, et al.
Published: (2022)
Distilling Out-of-Distribution Robustness from Vision-Language Foundation Models
by: Zhou, Andy, et al.
Published: (2023)
by: Zhou, Andy, et al.
Published: (2023)
Truncated Consistency Models
by: Lee, Sangyun, et al.
Published: (2024)
by: Lee, Sangyun, et al.
Published: (2024)
A Survey of the Self Supervised Learning Mechanisms for Vision Transformers
by: Khan, Asifullah, et al.
Published: (2024)
by: Khan, Asifullah, et al.
Published: (2024)
Pre-training Vision Transformers with Formula-driven Supervised Learning
by: Kataoka, Hirokatsu, et al.
Published: (2022)
by: Kataoka, Hirokatsu, et al.
Published: (2022)
Logit Calibration and Feature Contrast for Robust Federated Learning on Non-IID Data
by: Qiao, Yu, et al.
Published: (2024)
by: Qiao, Yu, et al.
Published: (2024)
Scaling Laws for Robust Comparison of Open Foundation Language-Vision Models and Datasets
by: Nezhurina, Marianna, et al.
Published: (2025)
by: Nezhurina, Marianna, et al.
Published: (2025)
Balancing Accuracy, Calibration, and Efficiency in Active Learning with Vision Transformers Under Label Noise
by: Mots'oehli, Moseli, et al.
Published: (2025)
by: Mots'oehli, Moseli, et al.
Published: (2025)
When Does Supervised Training Pay Off? The Hidden Economics of Object Detection in the Era of Vision-Language Models
by: Al-Hamadani, Samer
Published: (2025)
by: Al-Hamadani, Samer
Published: (2025)
Imperfect Vision Encoders: Efficient and Robust Tuning for Vision-Language Models
by: Panos, Aristeidis, et al.
Published: (2024)
by: Panos, Aristeidis, et al.
Published: (2024)
Multi-Scale Visual Prompting for Lightweight Small-Image Classification
by: Khazem, Salim
Published: (2025)
by: Khazem, Salim
Published: (2025)
Semi-Supervised Masked Autoencoders: Unlocking Vision Transformer Potential with Limited Data
by: Faysal, Atik, et al.
Published: (2026)
by: Faysal, Atik, et al.
Published: (2026)
EUDA: An Efficient Unsupervised Domain Adaptation via Self-Supervised Vision Transformer
by: Abedi, Ali, et al.
Published: (2024)
by: Abedi, Ali, et al.
Published: (2024)
DINORANKCLIP: DINOv3 Distillation and Injection for Vision-Language Pretraining with High-Order Ranking Consistency
by: Jiang, Shuyang, et al.
Published: (2026)
by: Jiang, Shuyang, et al.
Published: (2026)
Self-Captioning Multimodal Interaction Tuning: Amplifying Exploitable Redundancies for Robust Vision Language Models
by: Ryan, Yuriel, et al.
Published: (2026)
by: Ryan, Yuriel, et al.
Published: (2026)
ClipGrader: Leveraging Vision-Language Models for Robust Label Quality Assessment in Object Detection
by: Lu, Hong, et al.
Published: (2025)
by: Lu, Hong, et al.
Published: (2025)
One Prompt Word is Enough to Boost Adversarial Robustness for Pre-trained Vision-Language Models
by: Li, Lin, et al.
Published: (2024)
by: Li, Lin, et al.
Published: (2024)
When Multi-Task Learning Meets Partial Supervision: A Computer Vision Review
by: Fontana, Maxime, et al.
Published: (2023)
by: Fontana, Maxime, et al.
Published: (2023)
Coordinated Robustness Evaluation Framework for Vision-Language Models
by: Babu, Ashwin Ramesh, et al.
Published: (2025)
by: Babu, Ashwin Ramesh, et al.
Published: (2025)
Generalized Consistency Trajectory Models for Image Manipulation
by: Kim, Beomsu, et al.
Published: (2024)
by: Kim, Beomsu, et al.
Published: (2024)
Improving Consistency Models with Generator-Augmented Flows
by: Issenhuth, Thibaut, et al.
Published: (2024)
by: Issenhuth, Thibaut, et al.
Published: (2024)
Towards Adversarially Robust Vision-Language Models: Insights from Design Choices and Prompt Formatting Techniques
by: Bhagwatkar, Rishika, et al.
Published: (2024)
by: Bhagwatkar, Rishika, et al.
Published: (2024)
Similar Items
-
SAFE-KD: Risk-Controlled Early-Exit Distillation for Vision Backbones
by: Khazem, Salim
Published: (2026) -
AdapterTune: Zero-Initialized Low-Rank Adapters for Frozen Vision Transformers
by: Khazem, Salim
Published: (2026) -
TopoLoRA-SAM: Topology-Aware Parameter-Efficient Adaptation of Foundation Segmenters for Thin-Structure and Cross-Domain Binary Semantic Segmentation
by: Khazem, Salim
Published: (2026) -
PolygoNet: Leveraging Simplified Polygonal Representation for Effective Image Classification
by: Khazem, Salim, et al.
Published: (2025) -
MC-RFM: Geometry-Aware Few-Shot Adaptation via Mixed-Curvature Riemannian Flow Matching
by: Khazem, Salim, et al.
Published: (2026)