Scalable Delphi: Large Language Models for Structured Risk Estimation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lorenz, Tobias, Fritz, Mario |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MIBP-Cert: Certified Training against Data Perturbations with Mixed-Integer Bilinear Programs
von: Lorenz, Tobias, et al.
Veröffentlicht: (2024)
von: Lorenz, Tobias, et al.
Veröffentlicht: (2024)
FullCert: Deterministic End-to-End Certification for Training and Inference of Neural Networks
von: Lorenz, Tobias, et al.
Veröffentlicht: (2024)
von: Lorenz, Tobias, et al.
Veröffentlicht: (2024)
Scalable Task Planning via Large Language Models and Structured World Representations
von: Pérez-Dattari, Rodrigo, et al.
Veröffentlicht: (2024)
von: Pérez-Dattari, Rodrigo, et al.
Veröffentlicht: (2024)
Statistical Estimation of Adversarial Risk in Large Language Models under Best-of-N Sampling
von: Feng, Mingqian, et al.
Veröffentlicht: (2026)
von: Feng, Mingqian, et al.
Veröffentlicht: (2026)
Pixel-level Certified Explanations via Randomized Smoothing
von: Anani, Alaa, et al.
Veröffentlicht: (2025)
von: Anani, Alaa, et al.
Veröffentlicht: (2025)
Fundamental Risks in the Current Deployment of General-Purpose AI Models: What Have We (Not) Learnt From Cybersecurity?
von: Fritz, Mario
Veröffentlicht: (2024)
von: Fritz, Mario
Veröffentlicht: (2024)
A Scalable Pipeline for Estimating Verb Frame Frequencies Using Large Language Models
von: Morgan, Adam M., et al.
Veröffentlicht: (2025)
von: Morgan, Adam M., et al.
Veröffentlicht: (2025)
Transforming Expert Knowledge into Scalable Ontology via Large Language Models
von: Itoku, Ikkei, et al.
Veröffentlicht: (2025)
von: Itoku, Ikkei, et al.
Veröffentlicht: (2025)
Towards Scalable Schema Mapping using Large Language Models
von: Buss, Christopher, et al.
Veröffentlicht: (2025)
von: Buss, Christopher, et al.
Veröffentlicht: (2025)
An Interpretable and Scalable Framework for Evaluating Large Language Models
von: Qu, Xinhao, et al.
Veröffentlicht: (2026)
von: Qu, Xinhao, et al.
Veröffentlicht: (2026)
Certified Circuits: Stability Guarantees for Mechanistic Circuits
von: Anani, Alaa, et al.
Veröffentlicht: (2026)
von: Anani, Alaa, et al.
Veröffentlicht: (2026)
Post-training Large Language Models for Diverse High-Quality Responses
von: Chen, Yilei, et al.
Veröffentlicht: (2025)
von: Chen, Yilei, et al.
Veröffentlicht: (2025)
Estimating Tail Risks in Language Model Output Distributions
von: Angell, Rico, et al.
Veröffentlicht: (2026)
von: Angell, Rico, et al.
Veröffentlicht: (2026)
SAFER: Risk-Constrained Sample-then-Filter in Large Language Models
von: Wang, Qingni, et al.
Veröffentlicht: (2025)
von: Wang, Qingni, et al.
Veröffentlicht: (2025)
Optimization and Scalability of Collaborative Filtering Algorithms in Large Language Models
von: Yang, Haowei, et al.
Veröffentlicht: (2024)
von: Yang, Haowei, et al.
Veröffentlicht: (2024)
Risk Structures: Towards Engineering Risk-aware Autonomous Systems
von: Gleirscher, Mario
Veröffentlicht: (2019)
von: Gleirscher, Mario
Veröffentlicht: (2019)
Scalable and Explainable Learner-Video Interaction Prediction using Multimodal Large Language Models
von: Glandorf, Dominik, et al.
Veröffentlicht: (2026)
von: Glandorf, Dominik, et al.
Veröffentlicht: (2026)
Risk-Averse Finetuning of Large Language Models
von: Chaudhary, Sapana, et al.
Veröffentlicht: (2025)
von: Chaudhary, Sapana, et al.
Veröffentlicht: (2025)
Risks of Cultural Erasure in Large Language Models
von: Qadri, Rida, et al.
Veröffentlicht: (2025)
von: Qadri, Rida, et al.
Veröffentlicht: (2025)
CodePMP: Scalable Preference Model Pretraining for Large Language Model Reasoning
von: Yu, Huimu, et al.
Veröffentlicht: (2024)
von: Yu, Huimu, et al.
Veröffentlicht: (2024)
JudgeLM: Fine-tuned Large Language Models are Scalable Judges
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
GWT: Scalable Optimizer State Compression for Large Language Model Training
von: Wen, Ziqing, et al.
Veröffentlicht: (2025)
von: Wen, Ziqing, et al.
Veröffentlicht: (2025)
InverseScope: Scalable Activation Inversion for Interpreting Large Language Models
von: Luo, Yifan, et al.
Veröffentlicht: (2025)
von: Luo, Yifan, et al.
Veröffentlicht: (2025)
Selecting and Combining Large Language Models for Scalable Code Clone Detection
von: Chochlov, Muslim, et al.
Veröffentlicht: (2025)
von: Chochlov, Muslim, et al.
Veröffentlicht: (2025)
Language-Agnostic Suicidal Risk Detection Using Large Language Models
von: Kim, June-Woo, et al.
Veröffentlicht: (2025)
von: Kim, June-Woo, et al.
Veröffentlicht: (2025)
OptiMUS: Scalable Optimization Modeling with (MI)LP Solvers and Large Language Models
von: AhmadiTeshnizi, Ali, et al.
Veröffentlicht: (2024)
von: AhmadiTeshnizi, Ali, et al.
Veröffentlicht: (2024)
Error Detection and Correction for Interpretable Mathematics in Large Language Models
von: Yang, Yijin, et al.
Veröffentlicht: (2025)
von: Yang, Yijin, et al.
Veröffentlicht: (2025)
Language Models as Zero-shot Lossless Gradient Compressors: Towards General Neural Parameter Prior Models
von: Wang, Hui-Po, et al.
Veröffentlicht: (2024)
von: Wang, Hui-Po, et al.
Veröffentlicht: (2024)
LANTERN: Scalable Distillation of Large Language Models for Job-Person Fit and Explanation
von: Fu, Zhoutong, et al.
Veröffentlicht: (2025)
von: Fu, Zhoutong, et al.
Veröffentlicht: (2025)
Structural Reward Model: Enhancing Interpretability, Efficiency, and Scalability in Reward Modeling
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2025)
Scalable Token-Level Hallucination Detection in Large Language Models
von: Min, Rui, et al.
Veröffentlicht: (2026)
von: Min, Rui, et al.
Veröffentlicht: (2026)
VocalParse: Towards Unified and Scalable Singing Voice Transcription with Large Audio Language Models
von: Chen, Yukun, et al.
Veröffentlicht: (2026)
von: Chen, Yukun, et al.
Veröffentlicht: (2026)
Semantic Structure in Large Language Model Embeddings
von: Kozlowski, Austin C., et al.
Veröffentlicht: (2025)
von: Kozlowski, Austin C., et al.
Veröffentlicht: (2025)
Structured Chemistry Reasoning with Large Language Models
von: Ouyang, Siru, et al.
Veröffentlicht: (2023)
von: Ouyang, Siru, et al.
Veröffentlicht: (2023)
Credit Risk Meets Large Language Models: Building a Risk Indicator from Loan Descriptions in P2P Lending
von: Sanz-Guerrero, Mario, et al.
Veröffentlicht: (2024)
von: Sanz-Guerrero, Mario, et al.
Veröffentlicht: (2024)
The Cost of Thinking: Increased Jailbreak Risk in Large Language Models
von: Yang, Fan
Veröffentlicht: (2025)
von: Yang, Fan
Veröffentlicht: (2025)
Large Language Model Capabilities in Perioperative Risk Prediction and Prognostication
von: Chung, Philip, et al.
Veröffentlicht: (2024)
von: Chung, Philip, et al.
Veröffentlicht: (2024)
Understanding Privacy Risks of Embeddings Induced by Large Language Models
von: Zhu, Zhihao, et al.
Veröffentlicht: (2024)
von: Zhu, Zhihao, et al.
Veröffentlicht: (2024)
GRASPrune: Global Gating for Budgeted Structured Pruning of Large Language Models
von: Wang, Ziyang, et al.
Veröffentlicht: (2026)
von: Wang, Ziyang, et al.
Veröffentlicht: (2026)
LsrIF: Enhancing Logic-Structured Instruction Following of Large Language Models
von: Ren, Qingyu, et al.
Veröffentlicht: (2026)
von: Ren, Qingyu, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MIBP-Cert: Certified Training against Data Perturbations with Mixed-Integer Bilinear Programs
von: Lorenz, Tobias, et al.
Veröffentlicht: (2024) -
FullCert: Deterministic End-to-End Certification for Training and Inference of Neural Networks
von: Lorenz, Tobias, et al.
Veröffentlicht: (2024) -
Scalable Task Planning via Large Language Models and Structured World Representations
von: Pérez-Dattari, Rodrigo, et al.
Veröffentlicht: (2024) -
Statistical Estimation of Adversarial Risk in Large Language Models under Best-of-N Sampling
von: Feng, Mingqian, et al.
Veröffentlicht: (2026) -
Pixel-level Certified Explanations via Randomized Smoothing
von: Anani, Alaa, et al.
Veröffentlicht: (2025)