Guardado en:
| Autores principales: | Huang, Tzu-Heng, Cao, Catherine, Schoenberg, Spencer, Vishwakarma, Harit, Roberts, Nicholas, Sala, Frederic |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2502.12366 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Time To Impeach LLM-as-a-Judge: Programs are the Future of Evaluation
por: Huang, Tzu-Heng, et al.
Publicado: (2025)
por: Huang, Tzu-Heng, et al.
Publicado: (2025)
Promises and Pitfalls of Threshold-based Auto-labeling
por: Vishwakarma, Harit, et al.
Publicado: (2022)
por: Vishwakarma, Harit, et al.
Publicado: (2022)
OTTER: Effortless Label Distribution Adaptation of Zero-shot Models
por: Shin, Changho, et al.
Publicado: (2024)
por: Shin, Changho, et al.
Publicado: (2024)
The ALCHEmist: Automated Labeling 500x CHEaper Than LLM Data Annotators
por: Huang, Tzu-Heng, et al.
Publicado: (2024)
por: Huang, Tzu-Heng, et al.
Publicado: (2024)
Learning from Less: Measuring the Effectiveness of RLVR in Low Data and Compute Regimes
por: Bauer, Justin, et al.
Publicado: (2026)
por: Bauer, Justin, et al.
Publicado: (2026)
Stronger Than You Think: Benchmarking Weak Supervision on Realistic Tasks
por: Zhang, Tianyi, et al.
Publicado: (2025)
por: Zhang, Tianyi, et al.
Publicado: (2025)
Pearls from Pebbles: Improved Confidence Functions for Auto-labeling
por: Vishwakarma, Harit, et al.
Publicado: (2024)
por: Vishwakarma, Harit, et al.
Publicado: (2024)
Adaptive Scoring and Thresholding with Human Feedback for Robust Out-of-Distribution Detection
por: Yamada, Daisuke, et al.
Publicado: (2025)
por: Yamada, Daisuke, et al.
Publicado: (2025)
Taming False Positives in Out-of-Distribution Detection with Human Feedback
por: Vishwakarma, Harit, et al.
Publicado: (2024)
por: Vishwakarma, Harit, et al.
Publicado: (2024)
MoRe Fine-Tuning with 10x Fewer Parameters
por: Tan, Wenxuan, et al.
Publicado: (2024)
por: Tan, Wenxuan, et al.
Publicado: (2024)
Automating Benchmark Design
por: Dsouza, Amanda, et al.
Publicado: (2025)
por: Dsouza, Amanda, et al.
Publicado: (2025)
Evaluating Sample Utility for Efficient Data Selection by Mimicking Model Weights
por: Huang, Tzu-Heng, et al.
Publicado: (2025)
por: Huang, Tzu-Heng, et al.
Publicado: (2025)
Weak-to-Strong Generalization Through the Data-Centric Lens
por: Shin, Changho, et al.
Publicado: (2024)
por: Shin, Changho, et al.
Publicado: (2024)
RubiCap: Rubric-Guided Reinforcement Learning for Dense Image Captioning
por: Huang, Tzu-Heng, et al.
Publicado: (2026)
por: Huang, Tzu-Heng, et al.
Publicado: (2026)
Test-Time Scaling Makes Overtraining Compute-Optimal
por: Roberts, Nicholas, et al.
Publicado: (2026)
por: Roberts, Nicholas, et al.
Publicado: (2026)
Multimodal Data Curation via Object Detection and Filter Ensembles
por: Huang, Tzu-Heng, et al.
Publicado: (2024)
por: Huang, Tzu-Heng, et al.
Publicado: (2024)
CARE: Confounder-Aware Aggregation for Reliable LLM Evaluation
por: Zhao, Jitian, et al.
Publicado: (2026)
por: Zhao, Jitian, et al.
Publicado: (2026)
WS-GRPO: Weakly-Supervised Group-Relative Policy Optimization for Rollout-Efficient Reasoning
por: Mundada, Gagan, et al.
Publicado: (2026)
por: Mundada, Gagan, et al.
Publicado: (2026)
Tabby: A Language Model Architecture for Tabular and Structured Data Synthesis
por: Cromp, Sonia, et al.
Publicado: (2025)
por: Cromp, Sonia, et al.
Publicado: (2025)
Prune 'n Predict: Optimizing LLM Decision-making with Conformal Prediction
por: Vishwakarma, Harit, et al.
Publicado: (2024)
por: Vishwakarma, Harit, et al.
Publicado: (2024)
R&B: Domain Regrouping and Data Mixture Balancing for Efficient Foundation Model Training
por: Ge, Albert, et al.
Publicado: (2025)
por: Ge, Albert, et al.
Publicado: (2025)
Causal Spherical Hypergraph Networks for Modelling Social Uncertainty
por: Harit, Anoushka, et al.
Publicado: (2025)
por: Harit, Anoushka, et al.
Publicado: (2025)
RicciFlowRec: A Geometric Root Cause Recommender Using Ricci Curvature on Financial Graphs
por: Sun, Zhongtian, et al.
Publicado: (2025)
por: Sun, Zhongtian, et al.
Publicado: (2025)
From News to Returns: A Granger-Causal Hypergraph Transformer on the Sphere
por: Harit, Anoushka, et al.
Publicado: (2025)
por: Harit, Anoushka, et al.
Publicado: (2025)
Actionable Interpretability via Causal Hypergraphs: Unravelling Batch Size Effects in Deep Learning
por: Sun, Zhongtian, et al.
Publicado: (2025)
por: Sun, Zhongtian, et al.
Publicado: (2025)
Weakly Supervised Label Learning Flows
por: Lu, You, et al.
Publicado: (2023)
por: Lu, You, et al.
Publicado: (2023)
Pareto Optimal Code Generation
por: Orlanski, Gabriel, et al.
Publicado: (2025)
por: Orlanski, Gabriel, et al.
Publicado: (2025)
ManifoldMind: Dynamic Hyperbolic Reasoning for Trustworthy Recommendations
por: Harit, Anoushka, et al.
Publicado: (2025)
por: Harit, Anoushka, et al.
Publicado: (2025)
A General Framework for Learning from Weak Supervision
por: Chen, Hao, et al.
Publicado: (2024)
por: Chen, Hao, et al.
Publicado: (2024)
Scriptorium
Publicado: (2019)
Publicado: (2019)
COSMOS: Predictable and Cost-Effective Adaptation of LLMs
por: Wang, Jiayu, et al.
Publicado: (2025)
por: Wang, Jiayu, et al.
Publicado: (2025)
INDOTABVQA: A Benchmark for Cross-Lingual Table Understanding in Bahasa Indonesia Documents
por: Gautam, Somraj, et al.
Publicado: (2026)
por: Gautam, Somraj, et al.
Publicado: (2026)
Breaking Down Financial News Impact: A Novel AI Approach with Geometric Hypergraphs
por: Harit, Anoushka, et al.
Publicado: (2024)
por: Harit, Anoushka, et al.
Publicado: (2024)
Personalize Your LLM: Fake it then Align it
por: Zhang, Yijing, et al.
Publicado: (2025)
por: Zhang, Yijing, et al.
Publicado: (2025)
Expressivity-Efficiency Tradeoffs for Hybrid Sequence Models
por: Cooper, John, et al.
Publicado: (2026)
por: Cooper, John, et al.
Publicado: (2026)
Quantifying Structure in CLIP Embeddings: A Statistical Framework for Concept Interpretation
por: Zhao, Jitian, et al.
Publicado: (2025)
por: Zhao, Jitian, et al.
Publicado: (2025)
A Generic Self-Supervised Framework of Learning Invariant Discriminative Features
por: Ntelemis, Foivos, et al.
Publicado: (2022)
por: Ntelemis, Foivos, et al.
Publicado: (2022)
Table Detection with Active Learning
por: Gautam, Somraj, et al.
Publicado: (2025)
por: Gautam, Somraj, et al.
Publicado: (2025)
Pretrained Hybrids with MAD Skills
por: Roberts, Nicholas, et al.
Publicado: (2024)
por: Roberts, Nicholas, et al.
Publicado: (2024)
A Unified Empirical Risk Minimization Framework for Flexible N-Tuples Weak Supervision
por: Huang, Shuying, et al.
Publicado: (2025)
por: Huang, Shuying, et al.
Publicado: (2025)
Ejemplares similares
-
Time To Impeach LLM-as-a-Judge: Programs are the Future of Evaluation
por: Huang, Tzu-Heng, et al.
Publicado: (2025) -
Promises and Pitfalls of Threshold-based Auto-labeling
por: Vishwakarma, Harit, et al.
Publicado: (2022) -
OTTER: Effortless Label Distribution Adaptation of Zero-shot Models
por: Shin, Changho, et al.
Publicado: (2024) -
The ALCHEmist: Automated Labeling 500x CHEaper Than LLM Data Annotators
por: Huang, Tzu-Heng, et al.
Publicado: (2024) -
Learning from Less: Measuring the Effectiveness of RLVR in Low Data and Compute Regimes
por: Bauer, Justin, et al.
Publicado: (2026)