The Path to Reconciling Quality and Safety in Text-to-Image Generation: Dataset, Method, and Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Ruan, Shouwei, Wu, Zhenyu, Huang, Yao, Zhang, Ruochen, Sun, Yitong, Kang, Caixin, Zhao, Shiji, Wei, Xingxing |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
NDM: A Noise-driven Detection and Mitigation Framework against Implicit Sexual Intentions in Text-to-Image Generation
by: Sun, Yitong, et al.
Published: (2025)
by: Sun, Yitong, et al.
Published: (2025)
OODFace: Benchmarking Robustness of Face Recognition under Common Corruptions and Appearance Variations
by: Kang, Caixin, et al.
Published: (2024)
by: Kang, Caixin, et al.
Published: (2024)
When Lighting Deceives: Exposing Vision-Language Models' Illumination Vulnerability Through Illumination Transformation Attack
by: Liu, Hanqing, et al.
Published: (2025)
by: Liu, Hanqing, et al.
Published: (2025)
Breaking the Ceiling: Exploring the Potential of Jailbreak Attacks through Expanding Strategy Space
by: Huang, Yao, et al.
Published: (2025)
by: Huang, Yao, et al.
Published: (2025)
AdvDreamer Unveils: Are Vision-Language Models Truly Ready for Real-World 3D Variations?
by: Ruan, Shouwei, et al.
Published: (2024)
by: Ruan, Shouwei, et al.
Published: (2024)
From reactive to cognitive: brain-inspired spatial intelligence for embodied agents
by: Ruan, Shouwei, et al.
Published: (2025)
by: Ruan, Shouwei, et al.
Published: (2025)
DIFFender: Diffusion-Based Adversarial Defense against Patch Attacks
by: Kang, Caixin, et al.
Published: (2023)
by: Kang, Caixin, et al.
Published: (2023)
Real-world Adversarial Defense against Patch Attacks based on Diffusion Model
by: Wei, Xingxing, et al.
Published: (2024)
by: Wei, Xingxing, et al.
Published: (2024)
Jailbreaking Multimodal Large Language Models via Shuffle Inconsistency
by: Zhao, Shiji, et al.
Published: (2025)
by: Zhao, Shiji, et al.
Published: (2025)
Mind over Space: Can Multimodal Large Language Models Mentally Navigate?
by: Zhu, Qihui, et al.
Published: (2026)
by: Zhu, Qihui, et al.
Published: (2026)
Embodied Laser Attack:Leveraging Scene Priors to Achieve Agent-based Robust Non-contact Attacks
by: Sun, Yitong, et al.
Published: (2023)
by: Sun, Yitong, et al.
Published: (2023)
DeceptionBench: A Comprehensive Benchmark for AI Deception Behaviors in Real-world Scenarios
by: Huang, Yao, et al.
Published: (2025)
by: Huang, Yao, et al.
Published: (2025)
Towards Transferable Targeted 3D Adversarial Attack in the Physical World
by: Huang, Yao, et al.
Published: (2023)
by: Huang, Yao, et al.
Published: (2023)
Mitigating Overthinking in Large Reasoning Models via Manifold Steering
by: Huang, Yao, et al.
Published: (2025)
by: Huang, Yao, et al.
Published: (2025)
Omniview-Tuning: Boosting Viewpoint Invariance of Vision-Language Pre-training Models
by: Ruan, Shouwei, et al.
Published: (2024)
by: Ruan, Shouwei, et al.
Published: (2024)
Mitigating Accuracy-Robustness Trade-off via Balanced Multi-Teacher Adversarial Distillation
by: Zhao, Shiji, et al.
Published: (2023)
by: Zhao, Shiji, et al.
Published: (2023)
Revisiting the Trade-off between Accuracy and Robustness via Weight Distribution of Filters
by: Wei, Xingxing, et al.
Published: (2023)
by: Wei, Xingxing, et al.
Published: (2023)
Improving Safety Alignment via Balanced Direct Preference Optimization
by: Zhao, Shiji, et al.
Published: (2026)
by: Zhao, Shiji, et al.
Published: (2026)
An Amplitude-Encoding-Based Classical-Quantum Transfer Learning framework: Outperforming Classical Methods in Image Recognition
by: Hu, Shouwei, et al.
Published: (2025)
by: Hu, Shouwei, et al.
Published: (2025)
MoAPT: Mixture of Adversarial Prompt Tuning for Vision-Language Models
by: Zhao, Shiji, et al.
Published: (2025)
by: Zhao, Shiji, et al.
Published: (2025)
MESA: Improving MoE Safety Alignment via Decentralized Expertise
by: Sun, Yitong, et al.
Published: (2026)
by: Sun, Yitong, et al.
Published: (2026)
Improving Adversarial Robust Fairness via Anti-Bias Soft Label Distillation
by: Zhao, Shiji, et al.
Published: (2023)
by: Zhao, Shiji, et al.
Published: (2023)
A High-Quality Dataset and Reliable Evaluation for Interleaved Image-Text Generation
by: Feng, Yukang, et al.
Published: (2025)
by: Feng, Yukang, et al.
Published: (2025)
Visual Adversarial Attacks and Defenses in the Physical World: A Survey
by: Wei, Xingxing, et al.
Published: (2022)
by: Wei, Xingxing, et al.
Published: (2022)
VRSA: Jailbreaking Multimodal Large Language Models through Visual Reasoning Sequential Attack
by: Zhao, Shiji, et al.
Published: (2025)
by: Zhao, Shiji, et al.
Published: (2025)
Towards Class-wise Fair Adversarial Training via Anti-Bias Soft Label Distillation
by: Zhao, Shiji, et al.
Published: (2025)
by: Zhao, Shiji, et al.
Published: (2025)
Knowledge Visualization: A Benchmark and Method for Knowledge-Intensive Text-to-Image Generation
by: Zhao, Ran, et al.
Published: (2026)
by: Zhao, Ran, et al.
Published: (2026)
World2Mind: Cognition Toolkit for Allocentric Spatial Reasoning in Foundation Models
by: Ruan, Shouwei, et al.
Published: (2026)
by: Ruan, Shouwei, et al.
Published: (2026)
SocialDirector: Training-Free Social Interaction Control for Multi-Person Video Generation
by: Ouyang, Liangyang, et al.
Published: (2026)
by: Ouyang, Liangyang, et al.
Published: (2026)
Strata-Sword: A Hierarchical Safety Evaluation towards LLMs based on Reasoning Complexity of Jailbreak Instructions
by: Zhao, Shiji, et al.
Published: (2025)
by: Zhao, Shiji, et al.
Published: (2025)
An Effective Data Augmentation Method by Asking Questions about Scene Text Images
by: Yao, Xu, et al.
Published: (2026)
by: Yao, Xu, et al.
Published: (2026)
Compositional Text-to-Image Generation Via Region-aware Bimodal Direct Preference Optimization
by: Liu, Zhuohan, et al.
Published: (2026)
by: Liu, Zhuohan, et al.
Published: (2026)
Reconciling Safety Measurement and Dynamic Assurance
by: Denney, Ewen, et al.
Published: (2024)
by: Denney, Ewen, et al.
Published: (2024)
A Rapid Low‐Temperature Internal Heating Method for Lithium‐Ion Batteries Using Pulse Charging–Discharging Circuit
by: Xingxing Zhao, et al.
Published: (2025)
by: Xingxing Zhao, et al.
Published: (2025)
Knowledge-Guided Adversarial Training for Infrared Object Detection via Thermal Radiation Modeling
by: Zhao, Shiji, et al.
Published: (2026)
by: Zhao, Shiji, et al.
Published: (2026)
MAGMA-Edu: Multi-Agent Generative Multimodal Framework for Text-Diagram Educational Question Generation
by: Wu, Zhenyu, et al.
Published: (2025)
by: Wu, Zhenyu, et al.
Published: (2025)
Seeing Through the Noise: Improving Infrared Small Target Detection and Segmentation from Noise Suppression Perspective
by: Yuan, Maoxun, et al.
Published: (2025)
by: Yuan, Maoxun, et al.
Published: (2025)
Eva-VLA: Evaluating Vision-Language-Action Models' Robustness Under Real-World Physical Variations
by: Liu, Hanqing, et al.
Published: (2025)
by: Liu, Hanqing, et al.
Published: (2025)
Text-Visual Semantic Constrained AI-Generated Image Quality Assessment
by: Li, Qiang, et al.
Published: (2025)
by: Li, Qiang, et al.
Published: (2025)
Advancing Prompt-Based Methods for Replay-Independent General Continual Learning
by: Kang, Zhiqi, et al.
Published: (2025)
by: Kang, Zhiqi, et al.
Published: (2025)
Similar Items
-
NDM: A Noise-driven Detection and Mitigation Framework against Implicit Sexual Intentions in Text-to-Image Generation
by: Sun, Yitong, et al.
Published: (2025) -
OODFace: Benchmarking Robustness of Face Recognition under Common Corruptions and Appearance Variations
by: Kang, Caixin, et al.
Published: (2024) -
When Lighting Deceives: Exposing Vision-Language Models' Illumination Vulnerability Through Illumination Transformation Attack
by: Liu, Hanqing, et al.
Published: (2025) -
Breaking the Ceiling: Exploring the Potential of Jailbreak Attacks through Expanding Strategy Space
by: Huang, Yao, et al.
Published: (2025) -
AdvDreamer Unveils: Are Vision-Language Models Truly Ready for Real-World 3D Variations?
by: Ruan, Shouwei, et al.
Published: (2024)