Predicting Language Models' Success at Zero-Shot Probabilistic Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Ren, Kevin, Cortes-Gomez, Santiago, Patiño, Carlos Miguel, Joshi, Ananya, Lyu, Ruiqi, Tang, Jingjing, Turcan, Alistair, Yamin, Khurram, Wu, Steven, Wilder, Bryan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Explaining Concept Shift with Interpretable Feature Attribution
by: Lyu, Ruiqi, et al.
Published: (2025)
by: Lyu, Ruiqi, et al.
Published: (2025)
Combining digital data streams and epidemic networks for real time outbreak detection
by: Lyu, Ruiqi, et al.
Published: (2025)
by: Lyu, Ruiqi, et al.
Published: (2025)
SpatialEpiBench: Benchmarking Spatial Information and Epidemic Priors in Forecasting
by: Lyu, Ruiqi, et al.
Published: (2026)
by: Lyu, Ruiqi, et al.
Published: (2026)
Improving constraint-based discovery with robust propagation and reliable LLM priors
by: Lyu, Ruiqi, et al.
Published: (2025)
by: Lyu, Ruiqi, et al.
Published: (2025)
Can Revealed Preferences Clarify LLM Alignment and Steering?
by: Yamin, Khurram, et al.
Published: (2026)
by: Yamin, Khurram, et al.
Published: (2026)
When Agents Say One Thing and Do Another: Validating Elicited Beliefs from LLMs
by: Yamin, Khurram, et al.
Published: (2026)
by: Yamin, Khurram, et al.
Published: (2026)
Dependent Randomized Rounding for Budget Constrained Experimental Design
by: Yamin, Khurram, et al.
Published: (2025)
by: Yamin, Khurram, et al.
Published: (2025)
Can LLMs Reconcile Knowledge Conflicts in Counterfactual Reasoning
by: Yamin, Khurram, et al.
Published: (2025)
by: Yamin, Khurram, et al.
Published: (2025)
Accounting for Missing Covariates in Heterogeneous Treatment Estimation
by: Yamin, Khurram, et al.
Published: (2024)
by: Yamin, Khurram, et al.
Published: (2024)
Utility-Directed Conformal Prediction: A Decision-Aware Framework for Actionable Uncertainty Quantification
by: Cortes-Gomez, Santiago, et al.
Published: (2024)
by: Cortes-Gomez, Santiago, et al.
Published: (2024)
Federated Epidemic Surveillance
by: Lyu, Ruiqi, et al.
Published: (2023)
by: Lyu, Ruiqi, et al.
Published: (2023)
The Limits of AI-Driven Allocation: Optimal Screening under Aleatoric Uncertainty
by: Cortes-Gomez, Santiago, et al.
Published: (2026)
by: Cortes-Gomez, Santiago, et al.
Published: (2026)
Failure Modes of LLMs for Causal Reasoning on Narratives
by: Yamin, Khurram, et al.
Published: (2024)
by: Yamin, Khurram, et al.
Published: (2024)
Computationally Assisted Quality Control for Public Health Data Streams
by: Joshi, Ananya, et al.
Published: (2023)
by: Joshi, Ananya, et al.
Published: (2023)
Zero-Shot Performance Prediction for Probabilistic Scaling Laws
by: Schram, Viktoria, et al.
Published: (2025)
by: Schram, Viktoria, et al.
Published: (2025)
Auditing Fairness by Betting
by: Chugg, Ben, et al.
Published: (2023)
by: Chugg, Ben, et al.
Published: (2023)
Data-driven Design of Randomized Control Trials with Guaranteed Treatment Effects
by: Cortes-Gomez, Santiago, et al.
Published: (2024)
by: Cortes-Gomez, Santiago, et al.
Published: (2024)
Decision-Focused Evaluation of Worst-Case Distribution Shift
by: Ren, Kevin, et al.
Published: (2024)
by: Ren, Kevin, et al.
Published: (2024)
Will It Zero-Shot?: Predicting Zero-Shot Classification Performance For Arbitrary Queries
by: Robbins, Kevin, et al.
Published: (2026)
by: Robbins, Kevin, et al.
Published: (2026)
Outlier Ranking in Large-Scale Public Health Streams
by: Joshi, Ananya, et al.
Published: (2024)
by: Joshi, Ananya, et al.
Published: (2024)
An AI-Based Public Health Data Monitoring System
by: Joshi, Ananya, et al.
Published: (2025)
by: Joshi, Ananya, et al.
Published: (2025)
Healthcare LLM Benchmarks Are Only as Good as Their Explicit Assumptions
by: Raman, Naveen, et al.
Published: (2026)
by: Raman, Naveen, et al.
Published: (2026)
Leaving the Nest: Going Beyond Local Loss Functions for Predict-Then-Optimize
by: Shah, Sanket, et al.
Published: (2023)
by: Shah, Sanket, et al.
Published: (2023)
TusoAI: Agentic Optimization for Scientific Methods
by: Turcan, Alistair, et al.
Published: (2025)
by: Turcan, Alistair, et al.
Published: (2025)
Conformal Prediction for Zero-Shot Models
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
by: Silva-Rodríguez, Julio, et al.
Published: (2025)
SmartWay: Enhanced Waypoint Prediction and Backtracking for Zero-Shot Vision-and-Language Navigation
by: Shi, Xiangyu, et al.
Published: (2025)
by: Shi, Xiangyu, et al.
Published: (2025)
A Generalization Theory for Zero-Shot Prediction
by: Mehta, Ronak, et al.
Published: (2025)
by: Mehta, Ronak, et al.
Published: (2025)
Zero Shot Health Trajectory Prediction Using Transformer
by: Renc, Pawel, et al.
Published: (2024)
by: Renc, Pawel, et al.
Published: (2024)
Learning treatment effects while treating those in need
by: Wilder, Bryan, et al.
Published: (2024)
by: Wilder, Bryan, et al.
Published: (2024)
Fostering the Ecosystem of AI for Social Impact Requires Expanding and Strengthening Evaluation Standards
by: Wilder, Bryan, et al.
Published: (2025)
by: Wilder, Bryan, et al.
Published: (2025)
Comparing Targeting Strategies for Maximizing Social Welfare with Limited Resources
by: Sharma, Vibhhu, et al.
Published: (2024)
by: Sharma, Vibhhu, et al.
Published: (2024)
Multimodal Analytics of Cybersecurity Crisis Preparation Exercises: What Predicts Success?
by: Borchers, Conrad, et al.
Published: (2026)
by: Borchers, Conrad, et al.
Published: (2026)
One-Shot Badminton Shuttle Detection for Mobile Robots
by: Dipner, Florentin, et al.
Published: (2026)
by: Dipner, Florentin, et al.
Published: (2026)
Inducing Causal World Models in LLMs for Zero-Shot Physical Reasoning
by: Sharma, Aditya, et al.
Published: (2025)
by: Sharma, Aditya, et al.
Published: (2025)
Zero-Shot Function Encoder-Based Differentiable Predictive Control
by: Iqbal, Hassan, et al.
Published: (2025)
by: Iqbal, Hassan, et al.
Published: (2025)
Zero-Shot and Efficient Clarification Need Prediction in Conversational Search
by: Lu, Lili, et al.
Published: (2025)
by: Lu, Lili, et al.
Published: (2025)
Predicting VBAC Outcomes from U.S. Natality Data using Deep and Classical Machine Learning Models
by: Anand, Ananya
Published: (2025)
by: Anand, Ananya
Published: (2025)
Zero-Shot Surgical Tool Segmentation in Monocular Video Using Segment Anything Model 2
by: Lou, Ange, et al.
Published: (2024)
by: Lou, Ange, et al.
Published: (2024)
Decomposing Theory of Mind: How Emotional Processing Mediates ToM Abilities in LLMs
by: Chulo, Ivan, et al.
Published: (2025)
by: Chulo, Ivan, et al.
Published: (2025)
Context-Aware Hospitalization Forecasting Evaluations for Decision Support using LLMs
by: Makkuni, Rhea, et al.
Published: (2026)
by: Makkuni, Rhea, et al.
Published: (2026)
Similar Items
-
Explaining Concept Shift with Interpretable Feature Attribution
by: Lyu, Ruiqi, et al.
Published: (2025) -
Combining digital data streams and epidemic networks for real time outbreak detection
by: Lyu, Ruiqi, et al.
Published: (2025) -
SpatialEpiBench: Benchmarking Spatial Information and Epidemic Priors in Forecasting
by: Lyu, Ruiqi, et al.
Published: (2026) -
Improving constraint-based discovery with robust propagation and reliable LLM priors
by: Lyu, Ruiqi, et al.
Published: (2025) -
Can Revealed Preferences Clarify LLM Alignment and Steering?
by: Yamin, Khurram, et al.
Published: (2026)