Performative Risk Control: Calibrating Models for Reliable Deployment under Performativity
Fuente:
arXiv
Salvato in:
| Autori principali: | Li, Victor, Chen, Baiting, Mao, Yuzhen, Lei, Qi, Deng, Zhun |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Statistical Inference under Performativity
di: Li, Xiang, et al.
Pubblicazione: (2025)
di: Li, Xiang, et al.
Pubblicazione: (2025)
Conformal Tail Risk Control for Large Language Model Alignment
di: Chen, Catherine Yu-Chi, et al.
Pubblicazione: (2025)
di: Chen, Catherine Yu-Chi, et al.
Pubblicazione: (2025)
Adversarially Robust Control of Conditional Value-at-Risk via Rockafellar-Uryasev Conformal Inference
di: Chen, Catherine, et al.
Pubblicazione: (2026)
di: Chen, Catherine, et al.
Pubblicazione: (2026)
Prompt Risk Control: A Rigorous Framework for Responsible Deployment of Large Language Models
di: Zollo, Thomas P., et al.
Pubblicazione: (2023)
di: Zollo, Thomas P., et al.
Pubblicazione: (2023)
ALIGN: Aligned Delegation with Performance Guarantees for Multi-Agent LLM Reasoning
di: Zhu, Tong, et al.
Pubblicazione: (2026)
di: Zhu, Tong, et al.
Pubblicazione: (2026)
Strategically Deceptive Model Deployment in Performative Prediction
di: Bautiste, Javier Sanguino, et al.
Pubblicazione: (2025)
di: Bautiste, Javier Sanguino, et al.
Pubblicazione: (2025)
Modeling and Controlling Deployment Reliability under Temporal Distribution Shift
di: Rahman, Naimur, et al.
Pubblicazione: (2026)
di: Rahman, Naimur, et al.
Pubblicazione: (2026)
Improving Predictor Reliability with Selective Recalibration
di: Zollo, Thomas P., et al.
Pubblicazione: (2024)
di: Zollo, Thomas P., et al.
Pubblicazione: (2024)
Bayesian Deployment Approval for Learned Landing Controllers under Finite Rollout Validation
di: Jiang, Fei, et al.
Pubblicazione: (2026)
di: Jiang, Fei, et al.
Pubblicazione: (2026)
Performance Control in Early Exiting to Deploy Large Models at the Same Cost of Smaller Ones
di: Mofakhami, Mehrnaz, et al.
Pubblicazione: (2024)
di: Mofakhami, Mehrnaz, et al.
Pubblicazione: (2024)
IceFormer: Accelerated Inference with Long-Sequence Transformers on CPUs
di: Mao, Yuzhen, et al.
Pubblicazione: (2024)
di: Mao, Yuzhen, et al.
Pubblicazione: (2024)
Reliably Detecting Model Failures in Deployment Without Labels
di: Nguyen, Viet, et al.
Pubblicazione: (2025)
di: Nguyen, Viet, et al.
Pubblicazione: (2025)
Balancing Classification and Calibration Performance in Decision-Making LLMs via Calibration Aware Reinforcement Learning
di: Yaldiz, Duygu Nur, et al.
Pubblicazione: (2026)
di: Yaldiz, Duygu Nur, et al.
Pubblicazione: (2026)
Efficiently Deploying LLMs with Controlled Risk
di: Zellinger, Michael J., et al.
Pubblicazione: (2024)
di: Zellinger, Michael J., et al.
Pubblicazione: (2024)
ADiff4TPP: Asynchronous Diffusion Models for Temporal Point Processes
di: Mukherjee, Amartya, et al.
Pubblicazione: (2025)
di: Mukherjee, Amartya, et al.
Pubblicazione: (2025)
Breaking SafetyCore: Exploring the Risks of On-Device AI Deployment
di: Guyomard, Victor, et al.
Pubblicazione: (2025)
di: Guyomard, Victor, et al.
Pubblicazione: (2025)
A Deployment Audit of Release-Side Risk in Conformal Triage under Prevalence Shift
di: Li, Chengze, et al.
Pubblicazione: (2026)
di: Li, Chengze, et al.
Pubblicazione: (2026)
Enhancing Performance and Calibration in Quantile Hyperparameter Optimization
di: Doyle, Riccardo
Pubblicazione: (2025)
di: Doyle, Riccardo
Pubblicazione: (2025)
High Performance, Low Reliability: Uncertainty Benchmarking for Tabular Foundation Models
di: Costa, José Lucas De Melo, et al.
Pubblicazione: (2026)
di: Costa, José Lucas De Melo, et al.
Pubblicazione: (2026)
Forget, Then Recall: Learnable Compression and Selective Unfolding via Gist Sparse Attention
di: Mao, Yuzhen, et al.
Pubblicazione: (2026)
di: Mao, Yuzhen, et al.
Pubblicazione: (2026)
Distribution-Free Statistical Dispersion Control for Societal Applications
di: Deng, Zhun, et al.
Pubblicazione: (2023)
di: Deng, Zhun, et al.
Pubblicazione: (2023)
Loss-Controlling Calibration for Predictive Models
di: Wang, Di, et al.
Pubblicazione: (2023)
di: Wang, Di, et al.
Pubblicazione: (2023)
Beyond Scaling Laws: Understanding Transformer Performance with Associative Memory
di: Niu, Xueyan, et al.
Pubblicazione: (2024)
di: Niu, Xueyan, et al.
Pubblicazione: (2024)
Calibrating Decision Robustness via Inverse Conformal Risk Control
di: Zhou, Wenbin, et al.
Pubblicazione: (2025)
di: Zhou, Wenbin, et al.
Pubblicazione: (2025)
Unveiling Uncertainty: A Deep Dive into Calibration and Performance of Multimodal Large Language Models
di: Chen, Zijun, et al.
Pubblicazione: (2024)
di: Chen, Zijun, et al.
Pubblicazione: (2024)
Architecture Selection via the Trade-off Between Accuracy and Robustness
di: Deng, Zhun, et al.
Pubblicazione: (2019)
di: Deng, Zhun, et al.
Pubblicazione: (2019)
IceCache: Memory-efficient KV-cache Management for Long-Sequence LLMs
di: Mao, Yuzhen, et al.
Pubblicazione: (2026)
di: Mao, Yuzhen, et al.
Pubblicazione: (2026)
On the Impact of Black-box Deployment Strategies for Edge AI on Latency and Model Performance
di: Singh, Jaskirat, et al.
Pubblicazione: (2024)
di: Singh, Jaskirat, et al.
Pubblicazione: (2024)
From Uncertainty to Precision: Enhancing Binary Classifier Performance through Calibration
di: Machado, Agathe Fernandes, et al.
Pubblicazione: (2024)
di: Machado, Agathe Fernandes, et al.
Pubblicazione: (2024)
Calibrated Similarity for Reliable Geometric Analysis of Embedding Spaces
di: Tacheny, Nicolas
Pubblicazione: (2026)
di: Tacheny, Nicolas
Pubblicazione: (2026)
Ranking-Aware Calibration for Reliable Multimodal Reinforcement Learning
di: Cui, Peng, et al.
Pubblicazione: (2026)
di: Cui, Peng, et al.
Pubblicazione: (2026)
Learning and Forgetting Unsafe Examples in Large Language Models
di: Zhao, Jiachen, et al.
Pubblicazione: (2023)
di: Zhao, Jiachen, et al.
Pubblicazione: (2023)
Conformal Prediction: A Data Perspective
di: Zhou, Xiaofan, et al.
Pubblicazione: (2024)
di: Zhou, Xiaofan, et al.
Pubblicazione: (2024)
SAFER: A Calibrated Risk-Aware Multimodal Recommendation Model for Dynamic Treatment Regimes
di: Shen, Yishan, et al.
Pubblicazione: (2025)
di: Shen, Yishan, et al.
Pubblicazione: (2025)
Learning to Trust Experience: A Monitor-Trust-Regulator Framework for Learning under Unobservable Feedback Reliability
di: Zhang, Zhipeng, et al.
Pubblicazione: (2026)
di: Zhang, Zhipeng, et al.
Pubblicazione: (2026)
On the Performance of Empirical Risk Minimization with Smoothed Data
di: Block, Adam, et al.
Pubblicazione: (2024)
di: Block, Adam, et al.
Pubblicazione: (2024)
Incentivizing Truthful Language Models via Peer Elicitation Games
di: Chen, Baiting, et al.
Pubblicazione: (2025)
di: Chen, Baiting, et al.
Pubblicazione: (2025)
Performance Analysis of Decentralized Federated Learning Deployments
di: Jiang, Chengyan, et al.
Pubblicazione: (2025)
di: Jiang, Chengyan, et al.
Pubblicazione: (2025)
ODD: Overlap-aware Estimation of Model Performance under Distribution Shift
di: Mishra, Aayush, et al.
Pubblicazione: (2025)
di: Mishra, Aayush, et al.
Pubblicazione: (2025)
Boulder2Vec: Modeling Climber Performances in Professional Bouldering Competitions
di: Baron, Ethan, et al.
Pubblicazione: (2024)
di: Baron, Ethan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Statistical Inference under Performativity
di: Li, Xiang, et al.
Pubblicazione: (2025) -
Conformal Tail Risk Control for Large Language Model Alignment
di: Chen, Catherine Yu-Chi, et al.
Pubblicazione: (2025) -
Adversarially Robust Control of Conditional Value-at-Risk via Rockafellar-Uryasev Conformal Inference
di: Chen, Catherine, et al.
Pubblicazione: (2026) -
Prompt Risk Control: A Rigorous Framework for Responsible Deployment of Large Language Models
di: Zollo, Thomas P., et al.
Pubblicazione: (2023) -
ALIGN: Aligned Delegation with Performance Guarantees for Multi-Agent LLM Reasoning
di: Zhu, Tong, et al.
Pubblicazione: (2026)