LLM Evaluation as Tensor Completion: Low Rank Structure and Semiparametric Efficiency
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Jiachun, Simchi-Levi, David, Sun, Will Wei |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Beyond ATE: Multi-Criteria Design for A/B Testing
by: Li, Jiachun, et al.
Published: (2025)
by: Li, Jiachun, et al.
Published: (2025)
Privacy Preserving Adaptive Experiment Design
by: Li, Jiachun, et al.
Published: (2024)
by: Li, Jiachun, et al.
Published: (2024)
Low-Rank Robust Subspace Tensor Clustering for Metro Passenger Flow Modeling
by: Hu, Jiuyun, et al.
Published: (2024)
by: Hu, Jiuyun, et al.
Published: (2024)
Beyond Covariance Matrix: The Statistical Complexity of Private Linear Regression
by: Chen, Fan, et al.
Published: (2025)
by: Chen, Fan, et al.
Published: (2025)
Beyond Pairwise: Empowering LLM Alignment With Ranked Choice Modeling
by: Tang, Yuxuan, et al.
Published: (2025)
by: Tang, Yuxuan, et al.
Published: (2025)
Design and Analysis of Switchback Experiments
by: Bojinov, Iavor, et al.
Published: (2020)
by: Bojinov, Iavor, et al.
Published: (2020)
High-Dimensional Tensor Classification with CP Low-Rank Discriminant Structure
by: Chen, Elynn, et al.
Published: (2024)
by: Chen, Elynn, et al.
Published: (2024)
Structure-Adaptive Conformal Inference for Large-Scale Out-of-Distribution Testing
by: Sun, Rongyi, et al.
Published: (2026)
by: Sun, Rongyi, et al.
Published: (2026)
ShapE-GRPO: Shapley-Enhanced Reward Allocation for Multi-Candidate LLM Training
by: Ai, Rui, et al.
Published: (2026)
by: Ai, Rui, et al.
Published: (2026)
Partial Identification under Missing Data Using Weak Shadow Variables from Pretrained Models
by: Chen, Hongyu, et al.
Published: (2026)
by: Chen, Hongyu, et al.
Published: (2026)
OptiRepair: Closed-Loop Diagnosis and Repair of Supply Chain Optimization Models with LLM Agents
by: Ao, Ruicheng, et al.
Published: (2026)
by: Ao, Ruicheng, et al.
Published: (2026)
Regret Distribution in Stochastic Bandits: Optimal Trade-off between Expectation and Tail Risk
by: Simchi-Levi, David, et al.
Published: (2023)
by: Simchi-Levi, David, et al.
Published: (2023)
Improving the Estimation of Lifetime Effects in A/B Testing via Treatment Locality
by: Chen, Shuze, et al.
Published: (2024)
by: Chen, Shuze, et al.
Published: (2024)
Fourier Low-rank and Sparse Tensor for Efficient Tensor Completion
by: Li, Jingyang, et al.
Published: (2025)
by: Li, Jingyang, et al.
Published: (2025)
Segmenting Human-LLM Co-authored Text via Change Point Detection
by: Li, Mengchu, et al.
Published: (2026)
by: Li, Mengchu, et al.
Published: (2026)
Weak Supervision Performance Evaluation via Partial Identification
by: Polo, Felipe Maia, et al.
Published: (2023)
by: Polo, Felipe Maia, et al.
Published: (2023)
When Should an AI Workflow Release? Always-Valid Inference for Black-Box Generate-Verify Systems
by: Cho, Young Hyun, et al.
Published: (2026)
by: Cho, Young Hyun, et al.
Published: (2026)
Adaptive Prediction-Powered AutoEval with Reliability and Efficiency Guarantees
by: Park, Sangwoo, et al.
Published: (2025)
by: Park, Sangwoo, et al.
Published: (2025)
Recidivism and Peer Influence with LLM Text Embeddings in Low Security Correctional Facilities
by: Nath, Shanjukta, et al.
Published: (2025)
by: Nath, Shanjukta, et al.
Published: (2025)
Detecting Structural Heart Disease from Electrocardiograms via a Generalized Additive Model of Interpretable Foundation-Model Predictors
by: Zhou, Ya, et al.
Published: (2026)
by: Zhou, Ya, et al.
Published: (2026)
First-Order Efficiency for Probabilistic Value Estimation via A Statistical Viewpoint
by: Liu, Ziqi, et al.
Published: (2026)
by: Liu, Ziqi, et al.
Published: (2026)
DeepVARwT: Deep Learning for a VAR Model with Trend
by: Li, Xixi, et al.
Published: (2022)
by: Li, Xixi, et al.
Published: (2022)
Choosing with unknown causal information: Action-outcome probabilities for decision making can be grounded in causal models
by: Soto, Mauricio Gonzalez, et al.
Published: (2019)
by: Soto, Mauricio Gonzalez, et al.
Published: (2019)
Bounding Causal Effects with Leaky Instruments
by: Watson, David S., et al.
Published: (2024)
by: Watson, David S., et al.
Published: (2024)
Recover Experimental Data with Selection Bias using Counterfactual Logic
by: He, Jingyang, et al.
Published: (2025)
by: He, Jingyang, et al.
Published: (2025)
Causal Concept Graphs in LLM Latent Space for Stepwise Reasoning
by: Meherab, Md Muntaqim, et al.
Published: (2026)
by: Meherab, Md Muntaqim, et al.
Published: (2026)
General sample size analysis for probabilities of causation: a delta method approach
by: Cheng, Tianyuan, et al.
Published: (2026)
by: Cheng, Tianyuan, et al.
Published: (2026)
Partially Observed Structural Causal Models
by: Orujlu, Turan, et al.
Published: (2026)
by: Orujlu, Turan, et al.
Published: (2026)
Causal Temporal Regime Structure Learning
by: Rahmani, Abdellah, et al.
Published: (2023)
by: Rahmani, Abdellah, et al.
Published: (2023)
Delving Into the Psychology of Machines: Exploring the Structure of Self-Regulated Learning via LLM-Generated Survey Responses
by: Vogelsmeier, Leonie V. D. E., et al.
Published: (2025)
by: Vogelsmeier, Leonie V. D. E., et al.
Published: (2025)
CrowdLLM: Building LLM-Based Digital Populations Augmented with Generative Models
by: Lin, Ryan Feng, et al.
Published: (2025)
by: Lin, Ryan Feng, et al.
Published: (2025)
Semiparametric Causal Discovery and Inference with Invalid Instruments
by: Zou, Jing, et al.
Published: (2025)
by: Zou, Jing, et al.
Published: (2025)
Exploring Multi-Modal Data with Tool-Augmented LLM Agents for Precise Causal Discovery
by: Shen, ChengAo, et al.
Published: (2024)
by: Shen, ChengAo, et al.
Published: (2024)
Evaluating the Effectiveness of Index-Based Treatment Allocation
by: Boehmer, Niclas, et al.
Published: (2024)
by: Boehmer, Niclas, et al.
Published: (2024)
Conformalized Tensor Completion with Riemannian Optimization
by: Sun, Hu, et al.
Published: (2024)
by: Sun, Hu, et al.
Published: (2024)
Learning Causal Abstractions of Linear Structural Causal Models
by: Massidda, Riccardo, et al.
Published: (2024)
by: Massidda, Riccardo, et al.
Published: (2024)
A Causal Framework for Evaluating ICU Discharge Strategies
by: Simha, Sagar Nagaraj, et al.
Published: (2026)
by: Simha, Sagar Nagaraj, et al.
Published: (2026)
Testing for LLM response differences: the case of a composite null consisting of semantically irrelevant query perturbations
by: Acharyya, Aranyak, et al.
Published: (2025)
by: Acharyya, Aranyak, et al.
Published: (2025)
Off-Policy Evaluation and Learning for Survival Outcomes under Censoring
by: Kubota, Kohsuke, et al.
Published: (2026)
by: Kubota, Kohsuke, et al.
Published: (2026)
Effective Bayesian Causal Inference via Structural Marginalisation and Autoregressive Orders
by: Toth, Christian, et al.
Published: (2024)
by: Toth, Christian, et al.
Published: (2024)
Similar Items
-
Beyond ATE: Multi-Criteria Design for A/B Testing
by: Li, Jiachun, et al.
Published: (2025) -
Privacy Preserving Adaptive Experiment Design
by: Li, Jiachun, et al.
Published: (2024) -
Low-Rank Robust Subspace Tensor Clustering for Metro Passenger Flow Modeling
by: Hu, Jiuyun, et al.
Published: (2024) -
Beyond Covariance Matrix: The Statistical Complexity of Private Linear Regression
by: Chen, Fan, et al.
Published: (2025) -
Beyond Pairwise: Empowering LLM Alignment With Ranked Choice Modeling
by: Tang, Yuxuan, et al.
Published: (2025)