Saved in:
| Main Author: | Ding, Kaihua |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2509.22751 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Quantifiable Visual Explanations Without Ground-Truth
by: Singh, Amritpal, et al.
Published: (2026)
by: Singh, Amritpal, et al.
Published: (2026)
Evaluating Model Explanations without Ground Truth
by: Rawal, Kaivalya, et al.
Published: (2025)
by: Rawal, Kaivalya, et al.
Published: (2025)
Fairness Evaluation for Uplift Modeling in the Absence of Ground Truth
by: Kadioglu, Serdar, et al.
Published: (2024)
by: Kadioglu, Serdar, et al.
Published: (2024)
Confidence Calibration under Ambiguous Ground Truth
by: Tao, Linwei, et al.
Published: (2026)
by: Tao, Linwei, et al.
Published: (2026)
Iterative Causal Segmentation: Filling the Gap between Market Segmentation and Marketing Strategy
by: Ding, Kaihua, et al.
Published: (2024)
by: Ding, Kaihua, et al.
Published: (2024)
From Ground Truth to Measurement: A Statistical Framework for Human Labeling
by: Chew, Robert, et al.
Published: (2026)
by: Chew, Robert, et al.
Published: (2026)
xaitimesynth: A Python Package for Evaluating Attribution Methods for Time Series with Synthetic Ground Truth
by: Baer, Gregor
Published: (2026)
by: Baer, Gregor
Published: (2026)
Synthetic Data and the Shifting Ground of Truth
by: Offenhuber, Dietmar
Published: (2025)
by: Offenhuber, Dietmar
Published: (2025)
Position: AI Evaluations Should be Grounded on a Theory of Capability
by: Jo, Nathanael, et al.
Published: (2025)
by: Jo, Nathanael, et al.
Published: (2025)
Variance-Dependent Regret Lower Bounds for Contextual Bandits
by: He, Jiafan, et al.
Published: (2025)
by: He, Jiafan, et al.
Published: (2025)
Designing AI-Resilient Assessments Using Interconnected Problems: A Theoretically Grounded and Empirically Validated Framework
by: Ding, Kaihua
Published: (2025)
by: Ding, Kaihua
Published: (2025)
Quantifying Variance in Evaluation Benchmarks
by: Madaan, Lovish, et al.
Published: (2024)
by: Madaan, Lovish, et al.
Published: (2024)
Variance-Dependent Regret Bounds for Non-stationary Linear Bandits
by: Wang, Zhiyong, et al.
Published: (2024)
by: Wang, Zhiyong, et al.
Published: (2024)
Feature-Centric Unsupervised Node Representation Learning Without Homophily Assumption
by: Kim, Sunwoo, et al.
Published: (2025)
by: Kim, Sunwoo, et al.
Published: (2025)
The Scaling Law of Evaluation Failure: Why Simple Averaging Collapses Under Data Sparsity and Item Difficulty Gaps, and How Item Response Theory Recovers Ground Truth Across Domains
by: Kang, Jung Min
Published: (2026)
by: Kang, Jung Min
Published: (2026)
Ranking Large Language Models without Ground Truth
by: Dhurandhar, Amit, et al.
Published: (2024)
by: Dhurandhar, Amit, et al.
Published: (2024)
Ground-Truth Subgraphs for Better Training and Evaluation of Knowledge Graph Augmented LLMs
by: Cattaneo, Alberto, et al.
Published: (2025)
by: Cattaneo, Alberto, et al.
Published: (2025)
Grounded Object Centric Learning
by: Kori, Avinash, et al.
Published: (2023)
by: Kori, Avinash, et al.
Published: (2023)
A Confidence-Variance Theory for Pseudo-Label Selection in Semi-Supervised Learning
by: Liu, Jinshi, et al.
Published: (2026)
by: Liu, Jinshi, et al.
Published: (2026)
Conflict-Aware Pseudo Labeling via Optimal Transport for Entity Alignment
by: Ding, Qijie, et al.
Published: (2022)
by: Ding, Qijie, et al.
Published: (2022)
GT-Space: Enhancing Heterogeneous Collaborative Perception with Ground Truth Feature Space
by: Wang, Wentao, et al.
Published: (2026)
by: Wang, Wentao, et al.
Published: (2026)
A Framework for Fair Evaluation of Variance-Aware Bandit Algorithms
by: Wolf, Elise
Published: (2025)
by: Wolf, Elise
Published: (2025)
Minimax Optimal Variance-Aware Regret Bounds for Multinomial Logistic MDPs
by: Boudart, Pierre, et al.
Published: (2026)
by: Boudart, Pierre, et al.
Published: (2026)
Decision-Centric Design for LLM Systems
by: Sun, Wei
Published: (2026)
by: Sun, Wei
Published: (2026)
DataMaster: Data-Centric Autonomous AI Research
by: Du, Yaxin, et al.
Published: (2026)
by: Du, Yaxin, et al.
Published: (2026)
Towards Unified Attribution in Explainable AI, Data-Centric AI, and Mechanistic Interpretability
by: Zhang, Shichang, et al.
Published: (2025)
by: Zhang, Shichang, et al.
Published: (2025)
Low Variance Off-policy Evaluation with State-based Importance Sampling
by: Bossens, David M., et al.
Published: (2022)
by: Bossens, David M., et al.
Published: (2022)
Mastering Chinese Chess AI (Xiangqi) Without Search
by: Chen, Yu, et al.
Published: (2024)
by: Chen, Yu, et al.
Published: (2024)
Can Generative AI Support Patients' & Caregivers' Informational Needs? Towards Task-Centric Evaluation Of AI Systems
by: Rajagopal, Shreya, et al.
Published: (2024)
by: Rajagopal, Shreya, et al.
Published: (2024)
Emergent Visual Grounding in Large Multimodal Models Without Grounding Supervision
by: Cao, Shengcao, et al.
Published: (2024)
by: Cao, Shengcao, et al.
Published: (2024)
DSAI: Unbiased and Interpretable Latent Feature Extraction for Data-Centric AI
by: Cho, Hyowon, et al.
Published: (2024)
by: Cho, Hyowon, et al.
Published: (2024)
Reducing Credit Assignment Variance via Counterfactual Reasoning Paths
by: Ding, Fei, et al.
Published: (2026)
by: Ding, Fei, et al.
Published: (2026)
A Survey on Data-Centric AI: Tabular Learning from Reinforcement Learning and Generative AI Perspective
by: Ying, Wangyang, et al.
Published: (2025)
by: Ying, Wangyang, et al.
Published: (2025)
TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning
by: Wei, Zhepei, et al.
Published: (2025)
by: Wei, Zhepei, et al.
Published: (2025)
Toward an Evaluation Science for Generative AI Systems
by: Weidinger, Laura, et al.
Published: (2025)
by: Weidinger, Laura, et al.
Published: (2025)
BEVDiffuser: Plug-and-Play Diffusion Model for BEV Denoising with Ground-Truth Guidance
by: Ye, Xin, et al.
Published: (2025)
by: Ye, Xin, et al.
Published: (2025)
Data-Centric Foundation Models in Computational Healthcare: A Survey
by: Zhang, Yunkun, et al.
Published: (2024)
by: Zhang, Yunkun, et al.
Published: (2024)
Measuring What AI Systems Might Do: Towards A Measurement Science in AI
by: Voudouris, Konstantinos, et al.
Published: (2026)
by: Voudouris, Konstantinos, et al.
Published: (2026)
TruthFlow: Truthful LLM Generation via Representation Flow Correction
by: Wang, Hanyu, et al.
Published: (2025)
by: Wang, Hanyu, et al.
Published: (2025)
Incentivizing Truthfulness and Collaborative Fairness in Bayesian Learning
by: Sim, Rachael Hwee Ling, et al.
Published: (2026)
by: Sim, Rachael Hwee Ling, et al.
Published: (2026)
Similar Items
-
Learning Quantifiable Visual Explanations Without Ground-Truth
by: Singh, Amritpal, et al.
Published: (2026) -
Evaluating Model Explanations without Ground Truth
by: Rawal, Kaivalya, et al.
Published: (2025) -
Fairness Evaluation for Uplift Modeling in the Absence of Ground Truth
by: Kadioglu, Serdar, et al.
Published: (2024) -
Confidence Calibration under Ambiguous Ground Truth
by: Tao, Linwei, et al.
Published: (2026) -
Iterative Causal Segmentation: Filling the Gap between Market Segmentation and Marketing Strategy
by: Ding, Kaihua, et al.
Published: (2024)