Saved in:
| Main Authors: | Ye, Zikun, Yoganarasimhan, Hema |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2604.17267 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Fair Document Valuation in LLM Summaries via Shapley Values
by: Ye, Zikun, et al.
Published: (2025)
by: Ye, Zikun, et al.
Published: (2025)
Impact of AI Search Summaries on Website Traffic: Evidence from Google AI Overviews and Wikipedia
by: Khosravi, Mehrzad, et al.
Published: (2026)
by: Khosravi, Mehrzad, et al.
Published: (2026)
LOLA: LLM-Assisted Online Learning Algorithm for Content Experiments
by: Ye, Zikun, et al.
Published: (2024)
by: Ye, Zikun, et al.
Published: (2024)
TextBO: Bayesian Optimization in Language Space for Eval-Efficient Self-Improving AI
by: Kang, Enoch Hyunwook, et al.
Published: (2025)
by: Kang, Enoch Hyunwook, et al.
Published: (2025)
Estimating Item Difficulty with Large Language Models as Experts
by: Kolesnikova, Diana, et al.
Published: (2026)
by: Kolesnikova, Diana, et al.
Published: (2026)
Scalable Stewardship of an LLM-Assisted Clinical Benchmark with Physician Oversight
by: Ye, Junze, et al.
Published: (2025)
by: Ye, Junze, et al.
Published: (2025)
An Empirical Risk Minimization Approach for Offline Inverse RL and Dynamic Discrete Choice Model
by: Kang, Enoch H., et al.
Published: (2025)
by: Kang, Enoch H., et al.
Published: (2025)
StatLLM: A Dataset for Evaluating the Performance of Large Language Models in Statistical Analysis
by: Song, Xinyi, et al.
Published: (2025)
by: Song, Xinyi, et al.
Published: (2025)
Lightweight Adaptation for LLM-based Technical Service Agent: Latent Logic Augmentation and Robust Noise Reduction
by: Yu, Yi, et al.
Published: (2026)
by: Yu, Yi, et al.
Published: (2026)
Incorporating LLM Embeddings for Variation Across the Human Genome
by: Niu, Hongqian, et al.
Published: (2025)
by: Niu, Hongqian, et al.
Published: (2025)
Efficient Inference Using Large Language Models with Limited Human Data: Fine-Tuning then Rectification
by: Wang, Lei, et al.
Published: (2025)
by: Wang, Lei, et al.
Published: (2025)
Improving LLM Leaderboards with Psychometrical Methodology
by: Federiakin, Denis
Published: (2025)
by: Federiakin, Denis
Published: (2025)
Unleashing The Power of Pre-Trained Language Models for Irregularly Sampled Time Series
by: Zhang, Weijia, et al.
Published: (2024)
by: Zhang, Weijia, et al.
Published: (2024)
Towards Reliable LLM Evaluation: Correcting the Winner's Curse in Adaptive Benchmarking
by: Xu, Yang, et al.
Published: (2026)
by: Xu, Yang, et al.
Published: (2026)
Prune 'n Predict: Optimizing LLM Decision-making with Conformal Prediction
by: Vishwakarma, Harit, et al.
Published: (2024)
by: Vishwakarma, Harit, et al.
Published: (2024)
ImplicitRM: Unbiased Reward Modeling from Implicit Preference Data for LLM alignment
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
Wafer-Level Etch Spatial Profiling for Process Monitoring from Time-Series with Time-LLM
by: Kim, Hyunwoo, et al.
Published: (2026)
by: Kim, Hyunwoo, et al.
Published: (2026)
Binary Gaussian Copula Synthesis: A Novel Data Augmentation Technique to Advance ML-based Clinical Decision Support Systems for Early Prediction of Dialysis Among CKD Patients
by: Khosravi, Hamed, et al.
Published: (2024)
by: Khosravi, Hamed, et al.
Published: (2024)
Crowdsourced Adaptive Surveys
by: Velez, Yamil
Published: (2024)
by: Velez, Yamil
Published: (2024)
Optimal design of experiments to identify latent behavioral types
by: Balietti, Stefano, et al.
Published: (2018)
by: Balietti, Stefano, et al.
Published: (2018)
Finding the Sweet Spot: Optimal Data Augmentation Ratio for Imbalanced Credit Scoring Using ADASYN
by: Chia, Luis H.
Published: (2025)
by: Chia, Luis H.
Published: (2025)
Decision Quality Evaluation Framework at Pinterest
by: Tian, Yuqi, et al.
Published: (2026)
by: Tian, Yuqi, et al.
Published: (2026)
Surrogate-Based Prevalence Measurement for Large-Scale A/B Testing
by: Xu, Zehao, et al.
Published: (2026)
by: Xu, Zehao, et al.
Published: (2026)
CERES: A Probabilistic Early Warning System for Acute Food Insecurity
by: Pedersen, Tom Danny S.
Published: (2026)
by: Pedersen, Tom Danny S.
Published: (2026)
SEED-SET: Scalable Evolving Experimental Design for System-level Ethical Testing
by: Parashar, Anjali, et al.
Published: (2026)
by: Parashar, Anjali, et al.
Published: (2026)
From Passive Metric to Active Signal: The Evolving Role of Uncertainty Quantification in Large Language Models
by: Zhang, Jiaxin, et al.
Published: (2026)
by: Zhang, Jiaxin, et al.
Published: (2026)
Efficient Detection of Bad Benchmark Items with Novel Scalability Coefficients
by: Hardy, Michael, et al.
Published: (2026)
by: Hardy, Michael, et al.
Published: (2026)
On the Mechanistic Interpretability of Neural Networks for Causality in Bio-statistics
by: Conan, Jean-Baptiste A.
Published: (2025)
by: Conan, Jean-Baptiste A.
Published: (2025)
Quantitative Technology Forecasting: a Review of Trend Extrapolation Methods
by: Tsai, Peng-Hung, et al.
Published: (2024)
by: Tsai, Peng-Hung, et al.
Published: (2024)
HiBayES: A Hierarchical Bayesian Modeling Framework for AI Evaluation Statistics
by: Luettgau, Lennart, et al.
Published: (2025)
by: Luettgau, Lennart, et al.
Published: (2025)
Data-Driven Bayesian Network Models of Hurricane Evacuation Decision Making
by: Wang, Hui Sophie, et al.
Published: (2023)
by: Wang, Hui Sophie, et al.
Published: (2023)
Decade-long Emission Forecasting with an Ensemble Model in Taiwan
by: Hung, Gordon, et al.
Published: (2025)
by: Hung, Gordon, et al.
Published: (2025)
ChatGPT and post-test probability
by: Weisenthal, Samuel J.
Published: (2023)
by: Weisenthal, Samuel J.
Published: (2023)
Process-Aware Analysis of Treatment Paths in Heart Failure Patients: A Case Study
by: Beyel, Harry H., et al.
Published: (2024)
by: Beyel, Harry H., et al.
Published: (2024)
Calculating Customer Lifetime Value and Churn using Beta Geometric Negative Binomial and Gamma-Gamma Distribution in a NFT based setting
by: Das, Sagarnil
Published: (2025)
by: Das, Sagarnil
Published: (2025)
Unlocking the Potential of Past Research: Using Generative AI to Reconstruct Healthcare Simulation Models
by: Monks, Thomas, et al.
Published: (2025)
by: Monks, Thomas, et al.
Published: (2025)
The Advancement of Personalized Learning Potentially Accelerated by Generative AI
by: Wei, Yuang, et al.
Published: (2024)
by: Wei, Yuang, et al.
Published: (2024)
Bridging the Data Gap in AI Reliability Research and Establishing DR-AIR, a Comprehensive Data Repository for AI Reliability
by: Zheng, Simin, et al.
Published: (2025)
by: Zheng, Simin, et al.
Published: (2025)
Performance Evaluation of Large Language Models in Statistical Programming
by: Song, Xinyi, et al.
Published: (2025)
by: Song, Xinyi, et al.
Published: (2025)
TCKAN:A Novel Integrated Network Model for Predicting Mortality Risk in Sepsis Patients
by: Dong, Fanglin
Published: (2024)
by: Dong, Fanglin
Published: (2024)
Similar Items
-
Fair Document Valuation in LLM Summaries via Shapley Values
by: Ye, Zikun, et al.
Published: (2025) -
Impact of AI Search Summaries on Website Traffic: Evidence from Google AI Overviews and Wikipedia
by: Khosravi, Mehrzad, et al.
Published: (2026) -
LOLA: LLM-Assisted Online Learning Algorithm for Content Experiments
by: Ye, Zikun, et al.
Published: (2024) -
TextBO: Bayesian Optimization in Language Space for Eval-Efficient Self-Improving AI
by: Kang, Enoch Hyunwook, et al.
Published: (2025) -
Estimating Item Difficulty with Large Language Models as Experts
by: Kolesnikova, Diana, et al.
Published: (2026)