Saved in:
| Main Authors: | Xu, Changming, Singh, Gagandeep |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2506.11992 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Cross-Input Certified Training for Universal Perturbations
by: Xu, Changming, et al.
Published: (2024)
by: Xu, Changming, et al.
Published: (2024)
Support is All You Need for Certified VAE Training
by: Xu, Changming, et al.
Published: (2025)
by: Xu, Changming, et al.
Published: (2025)
Bypassing the Safety Training of Open-Source LLMs with Priming Attacks
by: Vega, Jason, et al.
Published: (2023)
by: Vega, Jason, et al.
Published: (2023)
Towards Generalized Certified Robustness with Multi-Norm Training
by: Jiang, Enyi, et al.
Published: (2024)
by: Jiang, Enyi, et al.
Published: (2024)
Certifying Knowledge Comprehension in LLMs
by: Chaudhary, Isha, et al.
Published: (2024)
by: Chaudhary, Isha, et al.
Published: (2024)
Formal Synthesis of Certifiably Robust Neural Lyapunov-Barrier Certificates
by: Wang, Chengxiao, et al.
Published: (2026)
by: Wang, Chengxiao, et al.
Published: (2026)
Certifying Counterfactual Bias in LLMs
by: Chaudhary, Isha, et al.
Published: (2024)
by: Chaudhary, Isha, et al.
Published: (2024)
How Catastrophic is Your LLM? Certifying Risk in Conversation
by: Wang, Chengxiao, et al.
Published: (2025)
by: Wang, Chengxiao, et al.
Published: (2025)
ARQ: A Mixed-Precision Quantization Framework for Accurate and Certifiably Robust DNNs
by: Yang, Yuchen, et al.
Published: (2024)
by: Yang, Yuchen, et al.
Published: (2024)
SEVerA: Verified Synthesis of Self-Evolving Agents
by: Banerjee, Debangshu, et al.
Published: (2026)
by: Banerjee, Debangshu, et al.
Published: (2026)
Robust Thompson Sampling Algorithms Against Reward Poisoning Attacks
by: Xu, Yinglun, et al.
Published: (2024)
by: Xu, Yinglun, et al.
Published: (2024)
Relational DNN Verification With Cross Executional Bound Refinement
by: Banerjee, Debangshu, et al.
Published: (2024)
by: Banerjee, Debangshu, et al.
Published: (2024)
RAMP: Boosting Adversarial Robustness Against Multiple $l_p$ Perturbations for Universal Robustness
by: Jiang, Enyi, et al.
Published: (2024)
by: Jiang, Enyi, et al.
Published: (2024)
Universal Black-Box Reward Poisoning Attack against Offline Reinforcement Learning
by: Xu, Yinglun, et al.
Published: (2024)
by: Xu, Yinglun, et al.
Published: (2024)
Binary Reward Labeling: Bridging Offline Preference and Reward-Based Reinforcement Learning
by: Xu, Yinglun, et al.
Published: (2024)
by: Xu, Yinglun, et al.
Published: (2024)
Cost-Driven Synthesis of Sound Abstract Interpreters
by: Gu, Qiuhan, et al.
Published: (2025)
by: Gu, Qiuhan, et al.
Published: (2025)
Sample Compression for Self Certified Continual Learning
by: Comeau, Jacob, et al.
Published: (2025)
by: Comeau, Jacob, et al.
Published: (2025)
Data Shifts Hurt CoT: A Theoretical Study
by: Yin, Lang, et al.
Published: (2025)
by: Yin, Lang, et al.
Published: (2025)
Probabilistic Trust Intervals for Out of Distribution Detection
by: Singh, Gagandeep, et al.
Published: (2021)
by: Singh, Gagandeep, et al.
Published: (2021)
Learning a Pessimistic Reward Model in RLHF
by: Xu, Yinglun, et al.
Published: (2025)
by: Xu, Yinglun, et al.
Published: (2025)
Is The Watermarking Of LLM-Generated Code Robust?
by: Suresh, Tarun, et al.
Published: (2024)
by: Suresh, Tarun, et al.
Published: (2024)
Efficient Distributed Training through Gradient Compression with Sparsification and Quantization Techniques
by: Singh, Shruti, et al.
Published: (2024)
by: Singh, Shruti, et al.
Published: (2024)
On Using Certified Training towards Empirical Robustness
by: De Palma, Alessandro, et al.
Published: (2024)
by: De Palma, Alessandro, et al.
Published: (2024)
CTBENCH: A Library and Benchmark for Certified Training
by: Mao, Yuhao, et al.
Published: (2024)
by: Mao, Yuhao, et al.
Published: (2024)
Understanding Certified Training with Interval Bound Propagation
by: Mao, Yuhao, et al.
Published: (2023)
by: Mao, Yuhao, et al.
Published: (2023)
Correct-By-Construction: Certified Individual Fairness through Neural Network Training
by: Zhang, Ruihan, et al.
Published: (2025)
by: Zhang, Ruihan, et al.
Published: (2025)
Gaussian Loss Smoothing Enables Certified Training with Tight Convex Relaxations
by: Balauca, Stefan, et al.
Published: (2024)
by: Balauca, Stefan, et al.
Published: (2024)
CRANE: Reasoning with constrained LLM generation
by: Banerjee, Debangshu, et al.
Published: (2025)
by: Banerjee, Debangshu, et al.
Published: (2025)
DAQ: Delta-Aware Quantization for Post-Training LLM Weight Compression
by: Yu, Xiaoming, et al.
Published: (2026)
by: Yu, Xiaoming, et al.
Published: (2026)
Enhancing Sign Language Detection through Mediapipe and Convolutional Neural Networks (CNN)
by: Verma, Aditya Raj, et al.
Published: (2024)
by: Verma, Aditya Raj, et al.
Published: (2024)
Certified Self-Consistency: Statistical Guarantees and Test-Time Training for Reliable Reasoning in LLMs
by: Cordero-Encinar, Paula, et al.
Published: (2025)
by: Cordero-Encinar, Paula, et al.
Published: (2025)
Quantum Interval Bound Propagation for Certified Training of Quantum Neural Networks
by: Andrews, Emma, et al.
Published: (2026)
by: Andrews, Emma, et al.
Published: (2026)
DINGO: Constrained Inference for Diffusion LLMs
by: Suresh, Tarun, et al.
Published: (2025)
by: Suresh, Tarun, et al.
Published: (2025)
IterGen: Iterative Semantic-aware Structured LLM Generation with Backtracking
by: Ugare, Shubham, et al.
Published: (2024)
by: Ugare, Shubham, et al.
Published: (2024)
Incremental Randomized Smoothing Certification
by: Ugare, Shubham, et al.
Published: (2023)
by: Ugare, Shubham, et al.
Published: (2023)
Revealing Interpretable Failure Modes of VLMs
by: Chaudhary, Isha, et al.
Published: (2026)
by: Chaudhary, Isha, et al.
Published: (2026)
Certified Robust Invariant Polytope Training in Neural Controlled ODEs
by: Harapanahalli, Akash, et al.
Published: (2024)
by: Harapanahalli, Akash, et al.
Published: (2024)
Stochastic Monkeys at Play: Random Augmentations Cheaply Break LLM Safety Alignment
by: Vega, Jason, et al.
Published: (2024)
by: Vega, Jason, et al.
Published: (2024)
Training Transformers for KV Cache Compressibility
by: Gelberg, Yoav, et al.
Published: (2026)
by: Gelberg, Yoav, et al.
Published: (2026)
Is Adversarial Training with Compressed Datasets Effective?
by: Chen, Tong, et al.
Published: (2024)
by: Chen, Tong, et al.
Published: (2024)
Similar Items
-
Cross-Input Certified Training for Universal Perturbations
by: Xu, Changming, et al.
Published: (2024) -
Support is All You Need for Certified VAE Training
by: Xu, Changming, et al.
Published: (2025) -
Bypassing the Safety Training of Open-Source LLMs with Priming Attacks
by: Vega, Jason, et al.
Published: (2023) -
Towards Generalized Certified Robustness with Multi-Norm Training
by: Jiang, Enyi, et al.
Published: (2024) -
Certifying Knowledge Comprehension in LLMs
by: Chaudhary, Isha, et al.
Published: (2024)