Improving Model Safety by Targeted Error Correction
Fuente:
arXiv
Saved in:
| Main Authors: | Mohammadi-Seif, Abolfazl, Baeza-Yates, Ricardo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Face Density as a Proxy for Data Complexity: Quantifying the Hardness of Instance Count
by: Mohammadi-Seif, Abolfazl, et al.
Published: (2026)
by: Mohammadi-Seif, Abolfazl, et al.
Published: (2026)
Risk-Calibrated Learning: Minimizing Fatal Errors in Medical AI
by: Mohammadi-Seif, Abolfazl, et al.
Published: (2026)
by: Mohammadi-Seif, Abolfazl, et al.
Published: (2026)
Adversarial Error Correction for Visual Autoregressive Generation
by: Bi, Ligong, et al.
Published: (2026)
by: Bi, Ligong, et al.
Published: (2026)
Region-Aware Deformable Convolutions
by: Maleki, Abolfazl Saheban, et al.
Published: (2025)
by: Maleki, Abolfazl Saheban, et al.
Published: (2025)
Targeted Attack Improves Protection against Unauthorized Diffusion Customization
by: Zheng, Boyang, et al.
Published: (2023)
by: Zheng, Boyang, et al.
Published: (2023)
MedErr-CT: A Visual Question Answering Benchmark for Identifying and Correcting Errors in CT Reports
by: Kyung, Sunggu, et al.
Published: (2025)
by: Kyung, Sunggu, et al.
Published: (2025)
Continual Error Correction on Low-Resource Devices
by: Paramonov, Kirill, et al.
Published: (2025)
by: Paramonov, Kirill, et al.
Published: (2025)
Ordinal Adaptive Correction: A Data-Centric Approach to Ordinal Image Classification with Noisy Labels
by: Moghaddam, Alireza Sedighi, et al.
Published: (2025)
by: Moghaddam, Alireza Sedighi, et al.
Published: (2025)
FloodVision: Urban Flood Depth Estimation Using Foundation Vision-Language Models and Domain Knowledge Graph
by: Liu, Zhangding, et al.
Published: (2025)
by: Liu, Zhangding, et al.
Published: (2025)
Panoptic Segmentation of Mammograms with Text-To-Image Diffusion Model
by: Zhao, Kun, et al.
Published: (2024)
by: Zhao, Kun, et al.
Published: (2024)
Lost in UNet: Improving Infrared Small Target Detection by Underappreciated Local Features
by: Quan, Wuzhou, et al.
Published: (2024)
by: Quan, Wuzhou, et al.
Published: (2024)
AI-Driven Virtual Teacher for Enhanced Educational Efficiency: Leveraging Large Pretrain Models for Autonomous Error Analysis and Correction
by: Xu, Tianlong, et al.
Published: (2024)
by: Xu, Tianlong, et al.
Published: (2024)
Correcting Class Imbalances with Self-Training for Improved Universal Lesion Detection and Tagging
by: Shieh, Alexander, et al.
Published: (2025)
by: Shieh, Alexander, et al.
Published: (2025)
A Comparison of Human and Machine Learning Errors in Face Recognition
by: Estévez-Almenzar, Marina, et al.
Published: (2025)
by: Estévez-Almenzar, Marina, et al.
Published: (2025)
Safety Alignment for Vision Language Models
by: Liu, Zhendong, et al.
Published: (2024)
by: Liu, Zhendong, et al.
Published: (2024)
Demonstrating and Reducing Shortcuts in Vision-Language Representation Learning
by: Bleeker, Maurits, et al.
Published: (2024)
by: Bleeker, Maurits, et al.
Published: (2024)
Two Heads Are Better Than One: Averaging along Fine-Tuning to Improve Targeted Transferability
by: Zeng, Hui, et al.
Published: (2024)
by: Zeng, Hui, et al.
Published: (2024)
DiReCT: Disentangled Regularization of Contrastive Trajectories for Physics-Refined Video Generation
by: Meyarian, Abolfazl, et al.
Published: (2026)
by: Meyarian, Abolfazl, et al.
Published: (2026)
Can Hallucination Correction Improve Video-Language Alignment?
by: Zhao, Lingjun, et al.
Published: (2025)
by: Zhao, Lingjun, et al.
Published: (2025)
Learning Through Retrospection: Improving Trajectory Prediction for Automated Driving with Error Feedback
by: Hagedorn, Steffen, et al.
Published: (2025)
by: Hagedorn, Steffen, et al.
Published: (2025)
Mirror Target YOLO: An Improved YOLOv8 Method with Indirect Vision for Heritage Buildings Fire Detection
by: Liang, Jian, et al.
Published: (2024)
by: Liang, Jian, et al.
Published: (2024)
Seeing Through the Noise: Improving Infrared Small Target Detection and Segmentation from Noise Suppression Perspective
by: Yuan, Maoxun, et al.
Published: (2025)
by: Yuan, Maoxun, et al.
Published: (2025)
HoliSafe: Holistic Safety Benchmarking and Modeling for Vision-Language Model
by: Lee, Youngwan, et al.
Published: (2025)
by: Lee, Youngwan, et al.
Published: (2025)
Personalized Safety Alignment for Text-to-Image Diffusion Models
by: Lei, Yu, et al.
Published: (2025)
by: Lei, Yu, et al.
Published: (2025)
Beyond the Mean: Distribution-Aware Loss Functions for Bimodal Regression
by: Mohammadi-Seif, Abolfazl, et al.
Published: (2026)
by: Mohammadi-Seif, Abolfazl, et al.
Published: (2026)
QNCD: Quantization Noise Correction for Diffusion Models
by: Chu, Huanpeng, et al.
Published: (2024)
by: Chu, Huanpeng, et al.
Published: (2024)
DiffGuard: Text-Based Safety Checker for Diffusion Models
by: Khader, Massine El, et al.
Published: (2024)
by: Khader, Massine El, et al.
Published: (2024)
Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations
by: Xue, Zhiyu, et al.
Published: (2025)
by: Xue, Zhiyu, et al.
Published: (2025)
Target Prompting for Information Extraction with Vision Language Model
by: Medhi, Dipankar
Published: (2024)
by: Medhi, Dipankar
Published: (2024)
Multi-Label Classification Framework for Hurricane Damage Assessment
by: Liu, Zhangding, et al.
Published: (2025)
by: Liu, Zhangding, et al.
Published: (2025)
MCANet: A Multi-Scale Class-Specific Attention Network for Multi-Label Post-Hurricane Damage Assessment using UAV Imagery
by: Liu, Zhangding, et al.
Published: (2025)
by: Liu, Zhangding, et al.
Published: (2025)
Enhancing Few-Shot Image Classification through Learnable Multi-Scale Embedding and Attention Mechanisms
by: Askari, Fatemeh, et al.
Published: (2024)
by: Askari, Fatemeh, et al.
Published: (2024)
HiBug2: Efficient and Interpretable Error Slice Discovery for Comprehensive Model Debugging
by: Chen, Muxi, et al.
Published: (2025)
by: Chen, Muxi, et al.
Published: (2025)
Evaluation of Safety Cognition Capability in Vision-Language Models for Autonomous Driving
by: Zhang, Enming, et al.
Published: (2025)
by: Zhang, Enming, et al.
Published: (2025)
Multimodal Prompt Decoupling Attack on the Safety Filters in Text-to-Image Models
by: Peng, Xingkai, et al.
Published: (2025)
by: Peng, Xingkai, et al.
Published: (2025)
SIA: Enhancing Safety via Intent Awareness for Vision-Language Models
by: Na, Youngjin, et al.
Published: (2025)
by: Na, Youngjin, et al.
Published: (2025)
From Evaluation to Defense: Advancing Safety in Video Large Language Models
by: Sun, Yiwei, et al.
Published: (2025)
by: Sun, Yiwei, et al.
Published: (2025)
SafeVid: Toward Safety Aligned Video Large Multimodal Models
by: Wang, Yixu, et al.
Published: (2025)
by: Wang, Yixu, et al.
Published: (2025)
TrojFlow: Flow Models are Natural Targets for Trojan Attacks
by: Qi, Zhengyang, et al.
Published: (2024)
by: Qi, Zhengyang, et al.
Published: (2024)
Prior2Posterior: Model Prior Correction for Long-Tailed Learning
by: Bhat, S Divakar, et al.
Published: (2024)
by: Bhat, S Divakar, et al.
Published: (2024)
Similar Items
-
Face Density as a Proxy for Data Complexity: Quantifying the Hardness of Instance Count
by: Mohammadi-Seif, Abolfazl, et al.
Published: (2026) -
Risk-Calibrated Learning: Minimizing Fatal Errors in Medical AI
by: Mohammadi-Seif, Abolfazl, et al.
Published: (2026) -
Adversarial Error Correction for Visual Autoregressive Generation
by: Bi, Ligong, et al.
Published: (2026) -
Region-Aware Deformable Convolutions
by: Maleki, Abolfazl Saheban, et al.
Published: (2025) -
Targeted Attack Improves Protection against Unauthorized Diffusion Customization
by: Zheng, Boyang, et al.
Published: (2023)