Statistical Hypothesis Testing for Auditing Robustness in Language Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Rauba, Paulius, Wei, Qiyao, van der Schaar, Mihaela |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Visualizing token importance for black-box language models
por: Rauba, Paulius, et al.
Publicado: (2025)
por: Rauba, Paulius, et al.
Publicado: (2025)
Quantifying perturbation impacts for large language models
por: Rauba, Paulius, et al.
Publicado: (2024)
por: Rauba, Paulius, et al.
Publicado: (2024)
Redefining Digital Health Interfaces with Large Language Models
por: Imrie, Fergus, et al.
Publicado: (2023)
por: Imrie, Fergus, et al.
Publicado: (2023)
Deep Hierarchical Learning with Nested Subspace Networks for Large Language Models
por: Rauba, Paulius, et al.
Publicado: (2025)
por: Rauba, Paulius, et al.
Publicado: (2025)
Context-Aware Testing: A New Paradigm for Model Testing with Large Language Models
por: Rauba, Paulius, et al.
Publicado: (2024)
por: Rauba, Paulius, et al.
Publicado: (2024)
Tiny Autoregressive Recursive Models
por: Rauba, Paulius, et al.
Publicado: (2026)
por: Rauba, Paulius, et al.
Publicado: (2026)
Multi-Agent Systems Should be Treated as Principal-Agent Problems
por: Rauba, Paulius, et al.
Publicado: (2026)
por: Rauba, Paulius, et al.
Publicado: (2026)
No More, No Less: Least-Privilege Language Models
por: Rauba, Paulius, et al.
Publicado: (2026)
por: Rauba, Paulius, et al.
Publicado: (2026)
Cascaded Language Models for Cost-effective Human-AI Decision-Making
por: Fanconi, Claudio, et al.
Publicado: (2025)
por: Fanconi, Claudio, et al.
Publicado: (2025)
Language Bottleneck Models for Qualitative Knowledge State Modeling
por: Berthon, Antonin, et al.
Publicado: (2025)
por: Berthon, Antonin, et al.
Publicado: (2025)
Inverse Reinforcement Learning Meets Large Language Model Post-Training: Basics, Advances, and Opportunities
por: Sun, Hao, et al.
Publicado: (2025)
por: Sun, Hao, et al.
Publicado: (2025)
Self-Healing Machine Learning: A Framework for Autonomous Adaptation in Real-World Environments
por: Rauba, Paulius, et al.
Publicado: (2024)
por: Rauba, Paulius, et al.
Publicado: (2024)
Continuously Updating Digital Twins using Large Language Models
por: Amad, Harry, et al.
Publicado: (2025)
por: Amad, Harry, et al.
Publicado: (2025)
Reusing Embeddings: Reproducible Reward Model Research in Large Language Model Alignment without GPUs
por: Sun, Hao, et al.
Publicado: (2025)
por: Sun, Hao, et al.
Publicado: (2025)
The Synergy of LLMs & RL Unlocks Offline Learning of Generalizable Language-Conditioned Policies with Low-fidelity Data
por: Pouplin, Thomas, et al.
Publicado: (2024)
por: Pouplin, Thomas, et al.
Publicado: (2024)
Query-Dependent Prompt Evaluation and Optimization with Offline Inverse RL
por: Sun, Hao, et al.
Publicado: (2023)
por: Sun, Hao, et al.
Publicado: (2023)
Active Task Disambiguation with LLMs
por: Kobalczyk, Katarzyna, et al.
Publicado: (2025)
por: Kobalczyk, Katarzyna, et al.
Publicado: (2025)
GameTalk: Training LLMs for Strategic Conversation
por: Vendrell, Victor Conchello, et al.
Publicado: (2026)
por: Vendrell, Victor Conchello, et al.
Publicado: (2026)
Delving into Multilingual Ethical Bias: The MSQAD with Statistical Hypothesis Tests for Large Language Models
por: Yu, Seunguk, et al.
Publicado: (2025)
por: Yu, Seunguk, et al.
Publicado: (2025)
On Error Propagation of Diffusion Models
por: Li, Yangming, et al.
Publicado: (2023)
por: Li, Yangming, et al.
Publicado: (2023)
OpenReview Should be Protected and Leveraged as a Community Asset for Research in the Era of Large Language Models
por: Sun, Hao, et al.
Publicado: (2025)
por: Sun, Hao, et al.
Publicado: (2025)
The AI Imperative: Scaling High-Quality Peer Review in Machine Learning
por: Wei, Qiyao, et al.
Publicado: (2025)
por: Wei, Qiyao, et al.
Publicado: (2025)
L2MAC: Large Language Model Automatic Computer for Extensive Code Generation
por: Holt, Samuel, et al.
Publicado: (2023)
por: Holt, Samuel, et al.
Publicado: (2023)
Semantic-KG: Using Knowledge Graphs to Construct Benchmarks for Measuring Semantic Similarity
por: Wei, Qiyao, et al.
Publicado: (2025)
por: Wei, Qiyao, et al.
Publicado: (2025)
Defining Expertise: Applications to Treatment Effect Estimation
por: Hüyük, Alihan, et al.
Publicado: (2024)
por: Hüyük, Alihan, et al.
Publicado: (2024)
Retrieval Augmented Thought Process for Private Data Handling in Healthcare
por: Pouplin, Thomas, et al.
Publicado: (2024)
por: Pouplin, Thomas, et al.
Publicado: (2024)
Matchmaker: Self-Improving Large Language Model Programs for Schema Matching
por: Seedat, Nabeel, et al.
Publicado: (2024)
por: Seedat, Nabeel, et al.
Publicado: (2024)
Inverse-RLignment: Large Language Model Alignment from Demonstrations through Inverse Reinforcement Learning
por: Sun, Hao, et al.
Publicado: (2024)
por: Sun, Hao, et al.
Publicado: (2024)
A Statistical Hypothesis Testing Framework for Data Misappropriation Detection in Large Language Models
por: Cai, Yinpeng, et al.
Publicado: (2025)
por: Cai, Yinpeng, et al.
Publicado: (2025)
Soft Mixture Denoising: Beyond the Expressive Bottleneck of Diffusion Models
por: Li, Yangming, et al.
Publicado: (2023)
por: Li, Yangming, et al.
Publicado: (2023)
Guideline-Grounded Evidence Accumulation for High-Stakes Agent Verification
por: Zhang, Yichi, et al.
Publicado: (2026)
por: Zhang, Yichi, et al.
Publicado: (2026)
The Cylindrical Representation Hypothesis for Language Model Steering
por: Gao, Lang, et al.
Publicado: (2026)
por: Gao, Lang, et al.
Publicado: (2026)
Distributionally Robust Reinforcement Learning with Human Feedback
por: Mandal, Debmalya, et al.
Publicado: (2025)
por: Mandal, Debmalya, et al.
Publicado: (2025)
Hypothesis Testing Prompting Improves Deductive Reasoning in Large Language Models
por: Li, Yitian, et al.
Publicado: (2024)
por: Li, Yitian, et al.
Publicado: (2024)
Improving LLM Agent Planning with In-Context Learning via Atomic Fact Augmentation and Lookahead Search
por: Holt, Samuel, et al.
Publicado: (2025)
por: Holt, Samuel, et al.
Publicado: (2025)
Assessing Agentic Large Language Models in Multilingual National Bias
por: Liu, Qianying, et al.
Publicado: (2025)
por: Liu, Qianying, et al.
Publicado: (2025)
Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis Refinement
por: Qiu, Linlu, et al.
Publicado: (2023)
por: Qiu, Linlu, et al.
Publicado: (2023)
AuditWen:An Open-Source Large Language Model for Audit
por: Huang, Jiajia, et al.
Publicado: (2024)
por: Huang, Jiajia, et al.
Publicado: (2024)
XRec: Large Language Models for Explainable Recommendation
por: Ma, Qiyao, et al.
Publicado: (2024)
por: Ma, Qiyao, et al.
Publicado: (2024)
DC-Check: A Data-Centric AI checklist to guide the development of reliable machine learning systems
por: Seedat, Nabeel, et al.
Publicado: (2022)
por: Seedat, Nabeel, et al.
Publicado: (2022)
Ejemplares similares
-
Visualizing token importance for black-box language models
por: Rauba, Paulius, et al.
Publicado: (2025) -
Quantifying perturbation impacts for large language models
por: Rauba, Paulius, et al.
Publicado: (2024) -
Redefining Digital Health Interfaces with Large Language Models
por: Imrie, Fergus, et al.
Publicado: (2023) -
Deep Hierarchical Learning with Nested Subspace Networks for Large Language Models
por: Rauba, Paulius, et al.
Publicado: (2025) -
Context-Aware Testing: A New Paradigm for Model Testing with Large Language Models
por: Rauba, Paulius, et al.
Publicado: (2024)