CluCERT: Certifying LLM Robustness via Clustering-Guided Denoising Smoothing
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Zixia, Jin, Gaojie, Hu, Jia, Mu, Ronghui |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reconcile Certified Robustness and Accuracy for DNN-based Smoothed Majority Vote Classifier
by: Jin, Gaojie, et al.
Published: (2025)
by: Jin, Gaojie, et al.
Published: (2025)
Principal Eigenvalue Regularization for Improved Worst-Class Certified Robustness of Smoothed Classifiers
by: Jin, Gaojie, et al.
Published: (2025)
by: Jin, Gaojie, et al.
Published: (2025)
POT: Inducing Overthinking in LLMs via Black-Box Iterative Optimization
by: Li, Xinyu, et al.
Published: (2025)
by: Li, Xinyu, et al.
Published: (2025)
Invariant Correlation of Representation with Label: Enhancing Domain Generalization in Noisy Environments
by: Jin, Gaojie, et al.
Published: (2024)
by: Jin, Gaojie, et al.
Published: (2024)
Safety of Embodied Navigation: A Survey
by: Wang, Zixia, et al.
Published: (2025)
by: Wang, Zixia, et al.
Published: (2025)
Certified Adversarial Robustness via Partition-based Randomized Smoothing
by: Goli, Hossein, et al.
Published: (2024)
by: Goli, Hossein, et al.
Published: (2024)
Certified Robustness for Deep Equilibrium Models via Serialized Random Smoothing
by: Gao, Weizhi, et al.
Published: (2024)
by: Gao, Weizhi, et al.
Published: (2024)
scSiameseClu: A Siamese Clustering Framework for Interpreting single-cell RNA Sequencing Data
by: Xu, Ping, et al.
Published: (2025)
by: Xu, Ping, et al.
Published: (2025)
DifCluE: Generating Counterfactual Explanations with Diffusion Autoencoders and modal clustering
by: Jain, Suparshva, et al.
Published: (2025)
by: Jain, Suparshva, et al.
Published: (2025)
Certifying Language Model Robustness with Fuzzed Randomized Smoothing: An Efficient Defense Against Backdoor Attacks
by: He, Bowei, et al.
Published: (2025)
by: He, Bowei, et al.
Published: (2025)
Margin-Adaptive Confidence Ranking for Reliable LLM Judgement
by: Jin, Gaojie, et al.
Published: (2026)
by: Jin, Gaojie, et al.
Published: (2026)
Enhancing Robust Fairness via Confusional Spectral Regularization
by: Jin, Gaojie, et al.
Published: (2025)
by: Jin, Gaojie, et al.
Published: (2025)
Pixel-level Certified Explanations via Randomized Smoothing
by: Anani, Alaa, et al.
Published: (2025)
by: Anani, Alaa, et al.
Published: (2025)
Certified Robustness against Sparse Adversarial Perturbations via Data Localization
by: Pal, Ambar, et al.
Published: (2024)
by: Pal, Ambar, et al.
Published: (2024)
Certified Robustness via Dynamic Margin Maximization and Improved Lipschitz Regularization
by: Fazlyab, Mahyar, et al.
Published: (2023)
by: Fazlyab, Mahyar, et al.
Published: (2023)
Robust Privacy: Inference-Time Privacy through Certified Robustness
by: Jin, Jiankai, et al.
Published: (2026)
by: Jin, Jiankai, et al.
Published: (2026)
How Catastrophic is Your LLM? Certifying Risk in Conversation
by: Wang, Chengxiao, et al.
Published: (2025)
by: Wang, Chengxiao, et al.
Published: (2025)
COLEP: Certifiably Robust Learning-Reasoning Conformal Prediction via Probabilistic Circuits
by: Kang, Mintong, et al.
Published: (2024)
by: Kang, Mintong, et al.
Published: (2024)
Lipschitz-aware Linearity Grafting for Certified Robustness
by: Han, Yongjin, et al.
Published: (2025)
by: Han, Yongjin, et al.
Published: (2025)
CEAR: Certified Ensemble Adversarial Robustness in DNNs
by: Sadig, Daniel, et al.
Published: (2026)
by: Sadig, Daniel, et al.
Published: (2026)
Certifying Global Robustness for Deep Neural Networks
by: Li, You, et al.
Published: (2024)
by: Li, You, et al.
Published: (2024)
Certifiably Byzantine-Robust Federated Conformal Prediction
by: Kang, Mintong, et al.
Published: (2024)
by: Kang, Mintong, et al.
Published: (2024)
CERT-ED: Certifiably Robust Text Classification for Edit Distance
by: Huang, Zhuoqun, et al.
Published: (2024)
by: Huang, Zhuoqun, et al.
Published: (2024)
Beyond Sharp Minima: Robust LLM Unlearning via Feedback-Guided Multi-Point Optimization
by: Wu, Wenhan, et al.
Published: (2025)
by: Wu, Wenhan, et al.
Published: (2025)
TF-TransUNet1D: Time-Frequency Guided Transformer U-Net for Robust ECG Denoising in Digital Twin
by: Wang, Shijie, et al.
Published: (2025)
by: Wang, Shijie, et al.
Published: (2025)
Adaptive Diffusion Denoised Smoothing : Certified Robustness via Randomized Smoothing with Differentially Private Guided Denoising Diffusion
by: Shpilevskiy, Frederick, et al.
Published: (2025)
by: Shpilevskiy, Frederick, et al.
Published: (2025)
OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents
by: Li, Xinyu, et al.
Published: (2026)
by: Li, Xinyu, et al.
Published: (2026)
Similarity and Dissimilarity Guided Co-association Matrix Construction for Ensemble Clustering
by: Zhang, Xu, et al.
Published: (2024)
by: Zhang, Xu, et al.
Published: (2024)
SPAM: Spike-Aware Adam with Momentum Reset for Stable LLM Training
by: Huang, Tianjin, et al.
Published: (2025)
by: Huang, Tianjin, et al.
Published: (2025)
Toward a Unified Lyapunov-Certified ODE Convergence Analysis of Smooth Q-Learning with p-Norms
by: Lee, Donghwan, et al.
Published: (2024)
by: Lee, Donghwan, et al.
Published: (2024)
Confidence-aware Denoised Fine-tuning of Off-the-shelf Models for Certified Robustness
by: Jang, Suhyeok, et al.
Published: (2024)
by: Jang, Suhyeok, et al.
Published: (2024)
Certified Signed Graph Unlearning
by: Zhao, Junpeng, et al.
Published: (2025)
by: Zhao, Junpeng, et al.
Published: (2025)
Certifying Counterfactual Bias in LLMs
by: Chaudhary, Isha, et al.
Published: (2024)
by: Chaudhary, Isha, et al.
Published: (2024)
ECPO: Evidence-Coupled Policy Optimization for Evidence-Certified Candidate Ranking
by: Hu, Miaobo, et al.
Published: (2026)
by: Hu, Miaobo, et al.
Published: (2026)
GF-Score: Certified Class-Conditional Robustness Evaluation with Fairness Guarantees
by: Shah, Arya, et al.
Published: (2026)
by: Shah, Arya, et al.
Published: (2026)
Learning to Denoise Biomedical Knowledge Graph for Robust Molecular Interaction Prediction
by: Ma, Tengfei, et al.
Published: (2023)
by: Ma, Tengfei, et al.
Published: (2023)
Kurtosis-Guided Denoising Score Matching for Tabular Anomaly Detection
by: Livernoche, Victor, et al.
Published: (2026)
by: Livernoche, Victor, et al.
Published: (2026)
It's Not You, It's Clipping: A Soft Trust-Region via Probability Smoothing for LLM RL
by: Dwyer, Madeleine, et al.
Published: (2025)
by: Dwyer, Madeleine, et al.
Published: (2025)
Certified Robustness from Approximate Gaussian Mixture Structures in Pretrained Latent Spaces
by: Emmanouilidis, Konstantinos, et al.
Published: (2026)
by: Emmanouilidis, Konstantinos, et al.
Published: (2026)
Certified Robustness Under Bounded Levenshtein Distance
by: Rocamora, Elias Abad, et al.
Published: (2025)
by: Rocamora, Elias Abad, et al.
Published: (2025)
Similar Items
-
Reconcile Certified Robustness and Accuracy for DNN-based Smoothed Majority Vote Classifier
by: Jin, Gaojie, et al.
Published: (2025) -
Principal Eigenvalue Regularization for Improved Worst-Class Certified Robustness of Smoothed Classifiers
by: Jin, Gaojie, et al.
Published: (2025) -
POT: Inducing Overthinking in LLMs via Black-Box Iterative Optimization
by: Li, Xinyu, et al.
Published: (2025) -
Invariant Correlation of Representation with Label: Enhancing Domain Generalization in Noisy Environments
by: Jin, Gaojie, et al.
Published: (2024) -
Safety of Embodied Navigation: A Survey
by: Wang, Zixia, et al.
Published: (2025)