Promises and Pitfalls of Threshold-based Auto-labeling
Fuente:
arXiv
Saved in:
| Main Authors: | Vishwakarma, Harit, Lin, Heguang, Sala, Frederic, Vinayak, Ramya Korlakai |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Taming False Positives in Out-of-Distribution Detection with Human Feedback
by: Vishwakarma, Harit, et al.
Published: (2024)
by: Vishwakarma, Harit, et al.
Published: (2024)
Adaptive Scoring and Thresholding with Human Feedback for Robust Out-of-Distribution Detection
by: Yamada, Daisuke, et al.
Published: (2025)
by: Yamada, Daisuke, et al.
Published: (2025)
Pearls from Pebbles: Improved Confidence Functions for Auto-labeling
by: Vishwakarma, Harit, et al.
Published: (2024)
by: Vishwakarma, Harit, et al.
Published: (2024)
Time To Impeach LLM-as-a-Judge: Programs are the Future of Evaluation
by: Huang, Tzu-Heng, et al.
Published: (2025)
by: Huang, Tzu-Heng, et al.
Published: (2025)
OTTER: Effortless Label Distribution Adaptation of Zero-shot Models
by: Shin, Changho, et al.
Published: (2024)
by: Shin, Changho, et al.
Published: (2024)
Metric Learning in an RKHS
by: Tatli, Gokcan, et al.
Published: (2025)
by: Tatli, Gokcan, et al.
Published: (2025)
Learning from Less: Measuring the Effectiveness of RLVR in Low Data and Compute Regimes
by: Bauer, Justin, et al.
Published: (2026)
by: Bauer, Justin, et al.
Published: (2026)
Is Conformal Factuality for RAG-based LLMs Robust? Novel Metrics and Systematic Insights
by: Chen, Yi, et al.
Published: (2026)
by: Chen, Yi, et al.
Published: (2026)
Thresholded Lexicographic Ordered Multiobjective Reinforcement Learning
by: Tercan, Alperen, et al.
Published: (2024)
by: Tercan, Alperen, et al.
Published: (2024)
Bridging Lifelong and Multi-Task Representation Learning via Algorithm and Complexity Measure
by: Wang, Zhi, et al.
Published: (2025)
by: Wang, Zhi, et al.
Published: (2025)
Metric Learning from Limited Pairwise Preference Comparisons
by: Wang, Zhi, et al.
Published: (2024)
by: Wang, Zhi, et al.
Published: (2024)
GRASP: Graph Agentic Search over Propositions for Multi-hop Question Answering
by: Jenkins, Stockton, et al.
Published: (2026)
by: Jenkins, Stockton, et al.
Published: (2026)
Prune 'n Predict: Optimizing LLM Decision-making with Conformal Prediction
by: Vishwakarma, Harit, et al.
Published: (2024)
by: Vishwakarma, Harit, et al.
Published: (2024)
Causal Spherical Hypergraph Networks for Modelling Social Uncertainty
by: Harit, Anoushka, et al.
Published: (2025)
by: Harit, Anoushka, et al.
Published: (2025)
PAL: Pluralistic Alignment Framework for Learning from Heterogeneous Preferences
by: Chen, Daiwei, et al.
Published: (2024)
by: Chen, Daiwei, et al.
Published: (2024)
Actionable Interpretability via Causal Hypergraphs: Unravelling Batch Size Effects in Deep Learning
by: Sun, Zhongtian, et al.
Published: (2025)
by: Sun, Zhongtian, et al.
Published: (2025)
RicciFlowRec: A Geometric Root Cause Recommender Using Ricci Curvature on Financial Graphs
by: Sun, Zhongtian, et al.
Published: (2025)
by: Sun, Zhongtian, et al.
Published: (2025)
COSMOS: Predictable and Cost-Effective Adaptation of LLMs
by: Wang, Jiayu, et al.
Published: (2025)
by: Wang, Jiayu, et al.
Published: (2025)
Weak-to-Strong Generalization Through the Data-Centric Lens
by: Shin, Changho, et al.
Published: (2024)
by: Shin, Changho, et al.
Published: (2024)
Hyperparameter Importance Analysis for Multi-Objective AutoML
by: Theodorakopoulos, Daphne, et al.
Published: (2024)
by: Theodorakopoulos, Daphne, et al.
Published: (2024)
ManifoldMind: Dynamic Hyperbolic Reasoning for Trustworthy Recommendations
by: Harit, Anoushka, et al.
Published: (2025)
by: Harit, Anoushka, et al.
Published: (2025)
Breaking Down Financial News Impact: A Novel AI Approach with Geometric Hypergraphs
by: Harit, Anoushka, et al.
Published: (2024)
by: Harit, Anoushka, et al.
Published: (2024)
ScriptoriumWS: A Code Generation Assistant for Weak Supervision
by: Huang, Tzu-Heng, et al.
Published: (2025)
by: Huang, Tzu-Heng, et al.
Published: (2025)
Grounded Token Initialization for New Vocabulary in LMs for Generative Recommendation
by: Chen, Daiwei, et al.
Published: (2026)
by: Chen, Daiwei, et al.
Published: (2026)
Optimal partition of feature using Bayesian classifier
by: Vishwakarma, Sanjay, et al.
Published: (2023)
by: Vishwakarma, Sanjay, et al.
Published: (2023)
Zero-Shot Robustification of Zero-Shot Models
by: Adila, Dyah, et al.
Published: (2023)
by: Adila, Dyah, et al.
Published: (2023)
Quantifying Structure in CLIP Embeddings: A Statistical Framework for Concept Interpretation
by: Zhao, Jitian, et al.
Published: (2025)
by: Zhao, Jitian, et al.
Published: (2025)
The Pitfalls of KV Cache Compression
by: Chen, Alex, et al.
Published: (2025)
by: Chen, Alex, et al.
Published: (2025)
Deep Learning Inference on Heterogeneous Mobile Processors: Potentials and Pitfalls
by: Liu, Sicong, et al.
Published: (2024)
by: Liu, Sicong, et al.
Published: (2024)
GLANCE: Graph Logic Attention Network with Cluster Enhancement for Heterophilous Graph Representation Learning
by: Sun, Zhongtian, et al.
Published: (2025)
by: Sun, Zhongtian, et al.
Published: (2025)
Demystifying Synthetic Data in LLM Pre-training: A Systematic Study of Scaling Laws, Benefits, and Pitfalls
by: Kang, Feiyang, et al.
Published: (2025)
by: Kang, Feiyang, et al.
Published: (2025)
The Pitfalls of Memorization: When Memorization Hurts Generalization
by: Bayat, Reza, et al.
Published: (2024)
by: Bayat, Reza, et al.
Published: (2024)
Evaluating Sample Utility for Efficient Data Selection by Mimicking Model Weights
by: Huang, Tzu-Heng, et al.
Published: (2025)
by: Huang, Tzu-Heng, et al.
Published: (2025)
Statistical Guarantees in Synthetic Data through Conformal Adversarial Generation
by: Vishwakarma, Rahul, et al.
Published: (2025)
by: Vishwakarma, Rahul, et al.
Published: (2025)
Uncertainty-Aware Hardware Trojan Detection Using Multimodal Deep Learning
by: Vishwakarma, Rahul, et al.
Published: (2024)
by: Vishwakarma, Rahul, et al.
Published: (2024)
SkillOrchestra: Learning to Route Agents via Skill Transfer
by: Wang, Jiayu, et al.
Published: (2026)
by: Wang, Jiayu, et al.
Published: (2026)
Auto-Prompt Ensemble for LLM Judge
by: Li, Jiajie, et al.
Published: (2025)
by: Li, Jiajie, et al.
Published: (2025)
Addressing Pitfalls in the Evaluation of Uncertainty Estimation Methods for Natural Language Generation
by: Ielanskyi, Mykyta, et al.
Published: (2025)
by: Ielanskyi, Mykyta, et al.
Published: (2025)
From Challenges and Pitfalls to Recommendations and Opportunities: Implementing Federated Learning in Healthcare
by: Li, Ming, et al.
Published: (2024)
by: Li, Ming, et al.
Published: (2024)
Evaluating Supervised Machine Learning Models: Principles, Pitfalls, and Metric Selection
by: Liu, Xuanyan, et al.
Published: (2026)
by: Liu, Xuanyan, et al.
Published: (2026)
Similar Items
-
Taming False Positives in Out-of-Distribution Detection with Human Feedback
by: Vishwakarma, Harit, et al.
Published: (2024) -
Adaptive Scoring and Thresholding with Human Feedback for Robust Out-of-Distribution Detection
by: Yamada, Daisuke, et al.
Published: (2025) -
Pearls from Pebbles: Improved Confidence Functions for Auto-labeling
by: Vishwakarma, Harit, et al.
Published: (2024) -
Time To Impeach LLM-as-a-Judge: Programs are the Future of Evaluation
by: Huang, Tzu-Heng, et al.
Published: (2025) -
OTTER: Effortless Label Distribution Adaptation of Zero-shot Models
by: Shin, Changho, et al.
Published: (2024)