Statistical Runtime Verification for LLMs via Robustness Estimation
Fuente:
arXiv
Saved in:
| Main Authors: | Levy, Natan, Ashrov, Adiel, Katz, Guy |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DEM: A Method for Certifying Deep Neural Network Classifier Outputs in Aerospace
by: Katz, Guy, et al.
Published: (2024)
by: Katz, Guy, et al.
Published: (2024)
Exploring and Evaluating Interplays of BPpy with Deep Reinforcement Learning and Formal Methods
by: Yaacov, Tom, et al.
Published: (2025)
by: Yaacov, Tom, et al.
Published: (2025)
ML Study of MaliciousTransactions in Ethereum
by: Katz, Natan
Published: (2024)
by: Katz, Natan
Published: (2024)
On Improving Deep Active Learning with Formal Verification
by: Spiegelman, Jonathan, et al.
Published: (2025)
by: Spiegelman, Jonathan, et al.
Published: (2025)
Input Validation for Neural Networks via Runtime Local Robustness Verification
by: Liu, Jiangchao, et al.
Published: (2020)
by: Liu, Jiangchao, et al.
Published: (2020)
Verification-Guided Shielding for Deep Reinforcement Learning
by: Corsi, Davide, et al.
Published: (2024)
by: Corsi, Davide, et al.
Published: (2024)
Optimizing Computational-Statistical Runtime for Wasserstein Distance Estimation
by: Jacobs, Peter Matthew, et al.
Published: (2026)
by: Jacobs, Peter Matthew, et al.
Published: (2026)
Proof Minimization in Neural Network Verification
by: Isac, Omri, et al.
Published: (2025)
by: Isac, Omri, et al.
Published: (2025)
PICID: Proof-Driven Clause Learning in Neural Network Verification
by: Isac, Omri, et al.
Published: (2025)
by: Isac, Omri, et al.
Published: (2025)
Nearly Solved? Robust Deepfake Detection Requires More than Visual Forensics
by: Levy, Guy, et al.
Published: (2024)
by: Levy, Guy, et al.
Published: (2024)
Beyond the Surface: Uncovering Implicit Locations with LLMs for Personalized Local News
by: Katz, Gali, et al.
Published: (2025)
by: Katz, Gali, et al.
Published: (2025)
Distributionally Robust Statistical Verification with Imprecise Neural Networks
by: Dutta, Souradeep, et al.
Published: (2023)
by: Dutta, Souradeep, et al.
Published: (2023)
GameGen-Verifier: Parallel Keypoint-Based Verification for LLM-Generated Games via Runtime State Injection
by: Jia, Chaobo, et al.
Published: (2026)
by: Jia, Chaobo, et al.
Published: (2026)
Not All Invariants Are Equal: Curating Training Data to Accelerate Program Verification with SLMs
by: Pinto, Ido, et al.
Published: (2026)
by: Pinto, Ido, et al.
Published: (2026)
Talking with Verifiers: Automatic Specification Generation for Neural Network Verification
by: Elboher, Yizhak Y., et al.
Published: (2026)
by: Elboher, Yizhak Y., et al.
Published: (2026)
Statistical Inference for Responsiveness Verification
by: Cheon, Seung Hyun, et al.
Published: (2025)
by: Cheon, Seung Hyun, et al.
Published: (2025)
Analyzing Adversarial Inputs in Deep Reinforcement Learning
by: Corsi, Davide, et al.
Published: (2024)
by: Corsi, Davide, et al.
Published: (2024)
Bridging Efficiency and Safety: Formal Verification of Neural Networks with Early Exits
by: Elboher, Yizhak Yisrael, et al.
Published: (2025)
by: Elboher, Yizhak Yisrael, et al.
Published: (2025)
Adjusted Wasserstein Distributionally Robust Estimator in Statistical Learning
by: Xie, Yiling, et al.
Published: (2023)
by: Xie, Yiling, et al.
Published: (2023)
NLP Verification: Towards a General Methodology for Certifying Robustness
by: Casadio, Marco, et al.
Published: (2024)
by: Casadio, Marco, et al.
Published: (2024)
Bring Your Own (Non-Robust) Algorithm to Solve Robust MDPs by Estimating The Worst Kernel
by: Wang, Kaixin, et al.
Published: (2023)
by: Wang, Kaixin, et al.
Published: (2023)
Probabilistic Runtime Verification, Evaluation and Risk Assessment of Visual Deep Learning Systems
by: Torpmann-Hagen, Birk, et al.
Published: (2025)
by: Torpmann-Hagen, Birk, et al.
Published: (2025)
Towards a Certified Proof Checker for Deep Neural Network Verification
by: Desmartin, Remi, et al.
Published: (2023)
by: Desmartin, Remi, et al.
Published: (2023)
ThrowBench: Benchmarking LLMs by Predicting Runtime Exceptions
by: Prenner, Julian Aron, et al.
Published: (2025)
by: Prenner, Julian Aron, et al.
Published: (2025)
Treatment of Statistical Estimation Problems in Randomized Smoothing for Adversarial Robustness
by: Voracek, Vaclav
Published: (2024)
by: Voracek, Vaclav
Published: (2024)
In-context Learning and Gradient Descent Revisited
by: Deutch, Gilad, et al.
Published: (2023)
by: Deutch, Gilad, et al.
Published: (2023)
Local vs. Global Interpretability: A Computational Complexity Perspective
by: Bassan, Shahaf, et al.
Published: (2024)
by: Bassan, Shahaf, et al.
Published: (2024)
Hard to Explain: On the Computational Hardness of In-Distribution Model Interpretation
by: Amir, Guy, et al.
Published: (2024)
by: Amir, Guy, et al.
Published: (2024)
Do What You Say: Steering Vision-Language-Action Models via Runtime Reasoning-Action Alignment Verification
by: Wu, Yilin, et al.
Published: (2025)
by: Wu, Yilin, et al.
Published: (2025)
Additive Models Explained: A Computational Complexity Approach
by: Bassan, Shahaf, et al.
Published: (2025)
by: Bassan, Shahaf, et al.
Published: (2025)
Transpose Attack: Stealing Datasets with Bidirectional Training
by: Amit, Guy, et al.
Published: (2023)
by: Amit, Guy, et al.
Published: (2023)
Formal Mechanistic Interpretability: Automated Circuit Discovery with Provable Guarantees
by: Hadad, Itamar, et al.
Published: (2026)
by: Hadad, Itamar, et al.
Published: (2026)
Boosting Few-Pixel Robustness Verification via Covering Verification Designs
by: Shapira, Yuval, et al.
Published: (2024)
by: Shapira, Yuval, et al.
Published: (2024)
Cost-sensitive retraining via posterior learning debt
by: Katz, Harrison
Published: (2026)
by: Katz, Harrison
Published: (2026)
Factual Confidence of LLMs: on Reliability and Robustness of Current Estimators
by: Mahaut, Matéo, et al.
Published: (2024)
by: Mahaut, Matéo, et al.
Published: (2024)
Multitask Learning and Bandits via Robust Statistics
by: Xu, Kan, et al.
Published: (2021)
by: Xu, Kan, et al.
Published: (2021)
What makes an Ensemble (Un) Interpretable?
by: Bassan, Shahaf, et al.
Published: (2025)
by: Bassan, Shahaf, et al.
Published: (2025)
Robustness Implies Privacy in Statistical Estimation
by: Hopkins, Samuel B., et al.
Published: (2022)
by: Hopkins, Samuel B., et al.
Published: (2022)
Verifying the Generalization of Deep Learning to Out-of-Distribution Domains
by: Amir, Guy, et al.
Published: (2024)
by: Amir, Guy, et al.
Published: (2024)
LASER: Language Model Regression for Semi-Structured Workflow Resource and Runtime Estimation
by: Yin, Yuxuan, et al.
Published: (2025)
by: Yin, Yuxuan, et al.
Published: (2025)
Similar Items
-
DEM: A Method for Certifying Deep Neural Network Classifier Outputs in Aerospace
by: Katz, Guy, et al.
Published: (2024) -
Exploring and Evaluating Interplays of BPpy with Deep Reinforcement Learning and Formal Methods
by: Yaacov, Tom, et al.
Published: (2025) -
ML Study of MaliciousTransactions in Ethereum
by: Katz, Natan
Published: (2024) -
On Improving Deep Active Learning with Formal Verification
by: Spiegelman, Jonathan, et al.
Published: (2025) -
Input Validation for Neural Networks via Runtime Local Robustness Verification
by: Liu, Jiangchao, et al.
Published: (2020)