Weak Supervision Performance Evaluation via Partial Identification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Polo, Felipe Maia, Maity, Subha, Yurochkin, Mikhail, Banerjee, Moulinath, Sun, Yuekai |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Microfoundation Inference for Strategic Prediction
von: Bracale, Daniele, et al.
Veröffentlicht: (2024)
von: Bracale, Daniele, et al.
Veröffentlicht: (2024)
Bridging Human and LLM Judgments: Understanding and Narrowing the Gap
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2025)
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2025)
Sloth: scaling laws for LLM skills to predict multi-benchmark performance across families
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024)
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024)
Learning the Distribution Map in Reverse Causal Performative Prediction
von: Bracale, Daniele, et al.
Veröffentlicht: (2024)
von: Bracale, Daniele, et al.
Veröffentlicht: (2024)
Aligners: Decoupling LLMs and Alignment
von: Ngweta, Lilian, et al.
Veröffentlicht: (2024)
von: Ngweta, Lilian, et al.
Veröffentlicht: (2024)
Likelihood-Free Estimation for Spatiotemporal Hawkes processes with missing data and application to predictive policing
von: Das, Pramit, et al.
Veröffentlicht: (2025)
von: Das, Pramit, et al.
Veröffentlicht: (2025)
A transfer learning framework for weak-to-strong generalization
von: Somerstep, Seamus, et al.
Veröffentlicht: (2024)
von: Somerstep, Seamus, et al.
Veröffentlicht: (2024)
tinyBenchmarks: evaluating LLMs with fewer examples
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024)
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024)
Limitations of refinement methods for weak to strong generalization
von: Somerstep, Seamus, et al.
Veröffentlicht: (2025)
von: Somerstep, Seamus, et al.
Veröffentlicht: (2025)
Partial Identification Approach to Counterfactual Fairness Assessment
von: Rho, Saeyoung, et al.
Veröffentlicht: (2025)
von: Rho, Saeyoung, et al.
Veröffentlicht: (2025)
Efficient multi-prompt evaluation of LLMs
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024)
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024)
Optimal Intervention for Self-triggering Spatial Networks with Application to Urban Crime Analytics
von: Das, Pramit, et al.
Veröffentlicht: (2025)
von: Das, Pramit, et al.
Veröffentlicht: (2025)
Distribution Matching for Self-Supervised Transfer Learning
von: Jiao, Yuling, et al.
Veröffentlicht: (2025)
von: Jiao, Yuling, et al.
Veröffentlicht: (2025)
Partially Observed Structural Causal Models
von: Orujlu, Turan, et al.
Veröffentlicht: (2026)
von: Orujlu, Turan, et al.
Veröffentlicht: (2026)
Robust inference for risk heterogeneity under group imbalance
von: Xu, Mengqi, et al.
Veröffentlicht: (2026)
von: Xu, Mengqi, et al.
Veröffentlicht: (2026)
Fusing Models with Complementary Expertise
von: Wang, Hongyi, et al.
Veröffentlicht: (2023)
von: Wang, Hongyi, et al.
Veröffentlicht: (2023)
A Latent Variable Framework for Scaling Laws in Large Language Models
von: Cai, Peiyao, et al.
Veröffentlicht: (2025)
von: Cai, Peiyao, et al.
Veröffentlicht: (2025)
Uplift Modeling Under Limited Supervision
von: Panagopoulos, George, et al.
Veröffentlicht: (2024)
von: Panagopoulos, George, et al.
Veröffentlicht: (2024)
Causal Identification in Time Series Models
von: Jahn, Erik, et al.
Veröffentlicht: (2025)
von: Jahn, Erik, et al.
Veröffentlicht: (2025)
s-ID: Causal Effect Identification in a Sub-Population
von: Abouei, Amir Mohammad, et al.
Veröffentlicht: (2023)
von: Abouei, Amir Mohammad, et al.
Veröffentlicht: (2023)
Fast Proxy Experiment Design for Causal Effect Identification
von: Elahi, Sepehr, et al.
Veröffentlicht: (2024)
von: Elahi, Sepehr, et al.
Veröffentlicht: (2024)
Towards a holistic understanding of Selection Bias for Causal Effect Identification
von: Qiu, Yiwen, et al.
Veröffentlicht: (2026)
von: Qiu, Yiwen, et al.
Veröffentlicht: (2026)
Learning Identifiable Structures Helps Avoid Bias in DNN-based Supervised Causal Learning
von: Zhang, Jiaru, et al.
Veröffentlicht: (2025)
von: Zhang, Jiaru, et al.
Veröffentlicht: (2025)
Data-Augmented Few-Shot Neural Emulator for Computer-Model System Identification
von: Jantre, Sanket, et al.
Veröffentlicht: (2025)
von: Jantre, Sanket, et al.
Veröffentlicht: (2025)
De-confounding Representation Learning for Counterfactual Inference on Continuous Treatment via Generative Adversarial Network
von: Zhao, Yonghe, et al.
Veröffentlicht: (2023)
von: Zhao, Yonghe, et al.
Veröffentlicht: (2023)
Time Series Domain Adaptation via Latent Invariant Causal Mechanism
von: Cai, Ruichu, et al.
Veröffentlicht: (2025)
von: Cai, Ruichu, et al.
Veröffentlicht: (2025)
Structure-Adaptive Conformal Inference for Large-Scale Out-of-Distribution Testing
von: Sun, Rongyi, et al.
Veröffentlicht: (2026)
von: Sun, Rongyi, et al.
Veröffentlicht: (2026)
Evaluating the Effectiveness of Index-Based Treatment Allocation
von: Boehmer, Niclas, et al.
Veröffentlicht: (2024)
von: Boehmer, Niclas, et al.
Veröffentlicht: (2024)
Maximin Relative Improvement: Fair Learning as a Bargaining Problem
von: Han, Jiwoo, et al.
Veröffentlicht: (2026)
von: Han, Jiwoo, et al.
Veröffentlicht: (2026)
A Causal Framework for Evaluating ICU Discharge Strategies
von: Simha, Sagar Nagaraj, et al.
Veröffentlicht: (2026)
von: Simha, Sagar Nagaraj, et al.
Veröffentlicht: (2026)
Enhancing the Performance of Neural Networks Through Causal Discovery and Integration of Domain Knowledge
von: Zhang, Xiaoge, et al.
Veröffentlicht: (2023)
von: Zhang, Xiaoge, et al.
Veröffentlicht: (2023)
Learning from Double Positive and Unlabeled Data for Potential-Customer Identification
von: Kato, Masahiro, et al.
Veröffentlicht: (2025)
von: Kato, Masahiro, et al.
Veröffentlicht: (2025)
Generalized Neyman Allocation for Locally Minimax Optimal Best-Arm Identification
von: Kato, Masahiro
Veröffentlicht: (2024)
von: Kato, Masahiro
Veröffentlicht: (2024)
Off-Policy Evaluation and Learning for Survival Outcomes under Censoring
von: Kubota, Kohsuke, et al.
Veröffentlicht: (2026)
von: Kubota, Kohsuke, et al.
Veröffentlicht: (2026)
A Consequentialist Critique of Binary Classification Evaluation: Theory, Practice, and Tools
von: Flores, Gerardo, et al.
Veröffentlicht: (2025)
von: Flores, Gerardo, et al.
Veröffentlicht: (2025)
M$^3$TN: Multi-gate Mixture-of-Experts based Multi-valued Treatment Network for Uplift Modeling
von: Sun, Zexu, et al.
Veröffentlicht: (2024)
von: Sun, Zexu, et al.
Veröffentlicht: (2024)
A General Causal Inference Framework for Cross-Sectional Observational Data
von: Zhao, Yonghe, et al.
Veröffentlicht: (2024)
von: Zhao, Yonghe, et al.
Veröffentlicht: (2024)
Estimation with missing not at random binary outcomes via exponential tilts
von: Maity, Subha
Veröffentlicht: (2025)
von: Maity, Subha
Veröffentlicht: (2025)
CausalCompass: Evaluating the Robustness of Time-Series Causal Discovery in Misspecified Scenarios
von: Yi, Huiyang, et al.
Veröffentlicht: (2026)
von: Yi, Huiyang, et al.
Veröffentlicht: (2026)
ProCause: Generating Counterfactual Outcomes to Evaluate Prescriptive Process Monitoring Methods
von: De Moor, Jakob, et al.
Veröffentlicht: (2025)
von: De Moor, Jakob, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Microfoundation Inference for Strategic Prediction
von: Bracale, Daniele, et al.
Veröffentlicht: (2024) -
Bridging Human and LLM Judgments: Understanding and Narrowing the Gap
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2025) -
Sloth: scaling laws for LLM skills to predict multi-benchmark performance across families
von: Polo, Felipe Maia, et al.
Veröffentlicht: (2024) -
Learning the Distribution Map in Reverse Causal Performative Prediction
von: Bracale, Daniele, et al.
Veröffentlicht: (2024) -
Aligners: Decoupling LLMs and Alignment
von: Ngweta, Lilian, et al.
Veröffentlicht: (2024)