Scaling Up Active Testing to Large Language Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Berrada, Gabrielle, Kossen, Jannik, Smith, Freddie Bickford, Razzak, Muhammed, Gal, Yarin, Rainforth, Tom |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
In-Context Learning Learns Label Relationships but Is Not Conventional Learning
von: Kossen, Jannik, et al.
Veröffentlicht: (2023)
von: Kossen, Jannik, et al.
Veröffentlicht: (2023)
Fine-Tuning Large Language Models to Appropriately Abstain with Semantic Entropy
von: Tjandra, Benedict Aaron, et al.
Veröffentlicht: (2024)
von: Tjandra, Benedict Aaron, et al.
Veröffentlicht: (2024)
Loss-Driven Bayesian Active Learning
von: Huang, Zhuoyue, et al.
Veröffentlicht: (2026)
von: Huang, Zhuoyue, et al.
Veröffentlicht: (2026)
Making Better Use of Unlabelled Data in Bayesian Active Learning
von: Smith, Freddie Bickford, et al.
Veröffentlicht: (2024)
von: Smith, Freddie Bickford, et al.
Veröffentlicht: (2024)
Rethinking Aleatoric and Epistemic Uncertainty
von: Smith, Freddie Bickford, et al.
Veröffentlicht: (2024)
von: Smith, Freddie Bickford, et al.
Veröffentlicht: (2024)
Semantic Entropy Probes: Robust and Cheap Hallucination Detection in LLMs
von: Kossen, Jannik, et al.
Veröffentlicht: (2024)
von: Kossen, Jannik, et al.
Veröffentlicht: (2024)
Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities
von: Nikitin, Alexander, et al.
Veröffentlicht: (2024)
von: Nikitin, Alexander, et al.
Veröffentlicht: (2024)
Prediction-Oriented Subsampling from Data Streams
von: Mussati, Benedetta Lavinia, et al.
Veröffentlicht: (2025)
von: Mussati, Benedetta Lavinia, et al.
Veröffentlicht: (2025)
The Benefits and Risks of Transductive Approaches for AI Fairness
von: Razzak, Muhammed, et al.
Veröffentlicht: (2024)
von: Razzak, Muhammed, et al.
Veröffentlicht: (2024)
Active Learning with Task-Driven Representations for Messy Pools
von: Ashouritaklimi, Kianoosh, et al.
Veröffentlicht: (2025)
von: Ashouritaklimi, Kianoosh, et al.
Veröffentlicht: (2025)
BED-LLM: Intelligent Information Gathering with LLMs and Bayesian Experimental Design
von: Choudhury, Deepro, et al.
Veröffentlicht: (2025)
von: Choudhury, Deepro, et al.
Veröffentlicht: (2025)
Deep Bayesian Active Learning for Preference Modeling in Large Language Models
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2024)
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2024)
Estimating the Hallucination Rate of Generative AI
von: Jesson, Andrew, et al.
Veröffentlicht: (2024)
von: Jesson, Andrew, et al.
Veröffentlicht: (2024)
Reducing Large Language Model Safety Risks in Women's Health using Semantic Entropy
von: Penny-Dimri, Jahan C., et al.
Veröffentlicht: (2025)
von: Penny-Dimri, Jahan C., et al.
Veröffentlicht: (2025)
MADE: Benchmark Environments for Closed-Loop Materials Discovery
von: Malik, Shreshth A, et al.
Veröffentlicht: (2026)
von: Malik, Shreshth A, et al.
Veröffentlicht: (2026)
Towards a Neural Debugger for Python
von: Beck, Maximilian, et al.
Veröffentlicht: (2026)
von: Beck, Maximilian, et al.
Veröffentlicht: (2026)
Beyond Bayesian Model Averaging over Paths in Probabilistic Programs with Stochastic Support
von: Reichelt, Tim, et al.
Veröffentlicht: (2023)
von: Reichelt, Tim, et al.
Veröffentlicht: (2023)
On the Expected Size of Conformal Prediction Sets
von: Dhillon, Guneet S., et al.
Veröffentlicht: (2023)
von: Dhillon, Guneet S., et al.
Veröffentlicht: (2023)
Leveraging Deep Learning for Physical Model Bias of Global Air Quality Estimates
von: Doerksen, Kelsey, et al.
Veröffentlicht: (2025)
von: Doerksen, Kelsey, et al.
Veröffentlicht: (2025)
A Geometric Approach to Optimal Experimental Design
von: Kerrigan, Gavin, et al.
Veröffentlicht: (2025)
von: Kerrigan, Gavin, et al.
Veröffentlicht: (2025)
Detecting LLM Hallucination Through Layer-wise Information Deficiency: Analysis of Ambiguous Prompts and Unanswerable Questions
von: Kim, Hazel, et al.
Veröffentlicht: (2024)
von: Kim, Hazel, et al.
Veröffentlicht: (2024)
Simple Baselines are Competitive with Code Evolution
von: Gideoni, Yonatan, et al.
Veröffentlicht: (2026)
von: Gideoni, Yonatan, et al.
Veröffentlicht: (2026)
Uncertainty Quantification for Surface Ozone Emulators using Deep Learning
von: Doerksen, Kelsey, et al.
Veröffentlicht: (2025)
von: Doerksen, Kelsey, et al.
Veröffentlicht: (2025)
Step-DAD: Semi-Amortized Policy-Based Bayesian Experimental Design
von: Hedman, Marcel, et al.
Veröffentlicht: (2025)
von: Hedman, Marcel, et al.
Veröffentlicht: (2025)
Do Multilingual LLMs Think In English?
von: Schut, Lisa, et al.
Veröffentlicht: (2025)
von: Schut, Lisa, et al.
Veröffentlicht: (2025)
Stabilizing Policy Gradients for Sample-Efficient Reinforcement Learning in LLM Reasoning
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2025)
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2025)
Temporal-Difference Variational Continual Learning
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2024)
von: Melo, Luckeciano C., et al.
Veröffentlicht: (2024)
Existing Large Language Model Unlearning Evaluations Are Inconclusive
von: Feng, Zhili, et al.
Veröffentlicht: (2025)
von: Feng, Zhili, et al.
Veröffentlicht: (2025)
Incorporating Unlabelled Data into Bayesian Neural Networks
von: Sharma, Mrinank, et al.
Veröffentlicht: (2023)
von: Sharma, Mrinank, et al.
Veröffentlicht: (2023)
Evaluating & Reducing Deceptive Dialogue From Language Models with Multi-turn RL
von: Abdulhai, Marwa, et al.
Veröffentlicht: (2025)
von: Abdulhai, Marwa, et al.
Veröffentlicht: (2025)
Testing For Distribution Shifts with Conditional Conformal Test Martingales
von: Shaer, Shalev, et al.
Veröffentlicht: (2026)
von: Shaer, Shalev, et al.
Veröffentlicht: (2026)
Training Transformers for KV Cache Compressibility
von: Gelberg, Yoav, et al.
Veröffentlicht: (2026)
von: Gelberg, Yoav, et al.
Veröffentlicht: (2026)
Protected Test-Time Adaptation via Online Entropy Matching: A Betting Approach
von: Bar, Yarin, et al.
Veröffentlicht: (2024)
von: Bar, Yarin, et al.
Veröffentlicht: (2024)
Generative Flows on Discrete State-Spaces: Enabling Multimodal Flows with Applications to Protein Co-Design
von: Campbell, Andrew, et al.
Veröffentlicht: (2024)
von: Campbell, Andrew, et al.
Veröffentlicht: (2024)
Variational Inference Failures Under Model Symmetries: Permutation Invariant Posteriors for Bayesian Neural Networks
von: Gelberg, Yoav, et al.
Veröffentlicht: (2024)
von: Gelberg, Yoav, et al.
Veröffentlicht: (2024)
KoopAGRU: A Koopman-based Anomaly Detection in Time-Series using Gated Recurrent Units
von: Yahia, Issam Ait, et al.
Veröffentlicht: (2025)
von: Yahia, Issam Ait, et al.
Veröffentlicht: (2025)
Boundary Point Jailbreaking of Black-Box LLMs
von: Davies, Xander, et al.
Veröffentlicht: (2026)
von: Davies, Xander, et al.
Veröffentlicht: (2026)
Selective Safety Steering via Value-Filtered Decoding
von: Einbinder, Bat-Sheva, et al.
Veröffentlicht: (2026)
von: Einbinder, Bat-Sheva, et al.
Veröffentlicht: (2026)
TextCAVs: Debugging vision models using text
von: Nicolson, Angus, et al.
Veröffentlicht: (2024)
von: Nicolson, Angus, et al.
Veröffentlicht: (2024)
Fine-tuning can cripple your foundation model; preserving features may be the solution
von: Mukhoti, Jishnu, et al.
Veröffentlicht: (2023)
von: Mukhoti, Jishnu, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
In-Context Learning Learns Label Relationships but Is Not Conventional Learning
von: Kossen, Jannik, et al.
Veröffentlicht: (2023) -
Fine-Tuning Large Language Models to Appropriately Abstain with Semantic Entropy
von: Tjandra, Benedict Aaron, et al.
Veröffentlicht: (2024) -
Loss-Driven Bayesian Active Learning
von: Huang, Zhuoyue, et al.
Veröffentlicht: (2026) -
Making Better Use of Unlabelled Data in Bayesian Active Learning
von: Smith, Freddie Bickford, et al.
Veröffentlicht: (2024) -
Rethinking Aleatoric and Epistemic Uncertainty
von: Smith, Freddie Bickford, et al.
Veröffentlicht: (2024)