Stronger Than You Think: Benchmarking Weak Supervision on Realistic Tasks
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Zhang, Tianyi, Cai, Linrong, Li, Jeffrey, Roberts, Nicholas, Guha, Neel, Lee, Jinoh, Sala, Frederic |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
The Weak Form Is Stronger Than You Think
par: Messenger, Daniel A., et autres
Publié: (2024)
par: Messenger, Daniel A., et autres
Publié: (2024)
Zero-Shot Robustification of Zero-Shot Models
par: Adila, Dyah, et autres
Publié: (2023)
par: Adila, Dyah, et autres
Publié: (2023)
ScriptoriumWS: A Code Generation Assistant for Weak Supervision
par: Huang, Tzu-Heng, et autres
Publié: (2025)
par: Huang, Tzu-Heng, et autres
Publié: (2025)
ResNets Are Deeper Than You Think
par: Mehmeti-Göpel, Christian H. X. Ali, et autres
Publié: (2025)
par: Mehmeti-Göpel, Christian H. X. Ali, et autres
Publié: (2025)
LoRA Is Slower Than You Think
par: Ko, Seokmin
Publié: (2025)
par: Ko, Seokmin
Publié: (2025)
Aligning Text to Image in Diffusion Models is Easier Than You Think
par: Lee, Jaa-Yeon, et autres
Publié: (2025)
par: Lee, Jaa-Yeon, et autres
Publié: (2025)
VLA Models Are More Generalizable Than You Think: Revisiting Physical and Spatial Modeling
par: Li, Weiqi, et autres
Publié: (2025)
par: Li, Weiqi, et autres
Publié: (2025)
Fine-tuning MLLMs Without Forgetting Is Easier Than You Think
par: Li, He, et autres
Publié: (2026)
par: Li, He, et autres
Publié: (2026)
LLM Unlearning Reveals a Stronger-Than-Expected Coreset Effect in Current Benchmarks
par: Pal, Soumyadeep, et autres
Publié: (2025)
par: Pal, Soumyadeep, et autres
Publié: (2025)
Weak-to-Strong Generalization Through the Data-Centric Lens
par: Shin, Changho, et autres
Publié: (2024)
par: Shin, Changho, et autres
Publié: (2024)
Intermediate Outputs Are More Sensitive Than You Think
par: Huang, Tao, et autres
Publié: (2024)
par: Huang, Tao, et autres
Publié: (2024)
Reverse Thinking Makes LLMs Stronger Reasoners
par: Chen, Justin Chih-Yao, et autres
Publié: (2024)
par: Chen, Justin Chih-Yao, et autres
Publié: (2024)
MoEs Are Stronger than You Think: Hyper-Parallel Inference Scaling with RoE
par: Zibakhsh, Soheil, et autres
Publié: (2025)
par: Zibakhsh, Soheil, et autres
Publié: (2025)
Reasoning with Sampling: Your Base Model is Smarter Than You Think
par: Karan, Aayush, et autres
Publié: (2025)
par: Karan, Aayush, et autres
Publié: (2025)
A Realistic Protocol for Evaluation of Weakly Supervised Object Localization
par: Murtaza, Shakeeb, et autres
Publié: (2024)
par: Murtaza, Shakeeb, et autres
Publié: (2024)
CRAFT: Aligning Diffusion Models with Fine-Tuning Is Easier Than You Think
par: Sun, Zening, et autres
Publié: (2026)
par: Sun, Zening, et autres
Publié: (2026)
Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think
par: Yu, Sihyun, et autres
Publié: (2024)
par: Yu, Sihyun, et autres
Publié: (2024)
The ALCHEmist: Automated Labeling 500x CHEaper Than LLM Data Annotators
par: Huang, Tzu-Heng, et autres
Publié: (2024)
par: Huang, Tzu-Heng, et autres
Publié: (2024)
CoDAR: Continuous Diffusion Language Models are More Powerful Than You Think
par: Shen, Junzhe, et autres
Publié: (2026)
par: Shen, Junzhe, et autres
Publié: (2026)
Unlearning Works Better Than You Think: Local Reinforcement-Based Selection of Auxiliary Objectives
par: Bendahi, Abderrahim, et autres
Publié: (2025)
par: Bendahi, Abderrahim, et autres
Publié: (2025)
Benchmarking and Building Long-Context Retrieval Models with LoCo and M2-BERT
par: Saad-Falcon, Jon, et autres
Publié: (2024)
par: Saad-Falcon, Jon, et autres
Publié: (2024)
Odysseys: Benchmarking Web Agents on Realistic Long Horizon Tasks
par: Jang, Lawrence Keunho, et autres
Publié: (2026)
par: Jang, Lawrence Keunho, et autres
Publié: (2026)
Debate Helps Weak Judges Reward Stronger Models
par: Elasky, Ethan, et autres
Publié: (2026)
par: Elasky, Ethan, et autres
Publié: (2026)
Model Checking for Reinforcement Learning in Autonomous Driving: One Can Do More Than You Think!
par: Gu, Rong
Publié: (2024)
par: Gu, Rong
Publié: (2024)
ProgAgent:A Continual RL Agent with Progress-Aware Rewards
par: Tan, Jinzhou, et autres
Publié: (2026)
par: Tan, Jinzhou, et autres
Publié: (2026)
Rethinking Weak Supervision in Anomaly Detection: A Comprehensive Benchmark
par: Yao, Xu, et autres
Publié: (2026)
par: Yao, Xu, et autres
Publié: (2026)
More Than a Score: Probing the Impact of Prompt Specificity on LLM Code Generation
par: Zi, Yangtian, et autres
Publié: (2025)
par: Zi, Yangtian, et autres
Publié: (2025)
Confident or Seek Stronger: Exploring Uncertainty-Based On-device LLM Routing From Benchmarking to Generalization
par: Chuang, Yu-Neng, et autres
Publié: (2025)
par: Chuang, Yu-Neng, et autres
Publié: (2025)
Visual Latents Know More Than They Say: Unsilencing Latent Reasoning in MLLMs
par: Zhang, Xin, et autres
Publié: (2026)
par: Zhang, Xin, et autres
Publié: (2026)
RiskWebWorld: A Realistic Interactive Benchmark for GUI Agents in E-commerce Risk Management
par: Chen, Renqi, et autres
Publié: (2026)
par: Chen, Renqi, et autres
Publié: (2026)
Towards Understanding Why FixMatch Generalizes Better Than Supervised Learning
par: Li, Jingyang, et autres
Publié: (2024)
par: Li, Jingyang, et autres
Publié: (2024)
Personalize Your LLM: Fake it then Align it
par: Zhang, Yijing, et autres
Publié: (2025)
par: Zhang, Yijing, et autres
Publié: (2025)
Weakly Supervised Anomaly Detection via Knowledge-Data Alignment
par: Zhao, Haihong, et autres
Publié: (2024)
par: Zhao, Haihong, et autres
Publié: (2024)
MoRe Fine-Tuning with 10x Fewer Parameters
par: Tan, Wenxuan, et autres
Publié: (2024)
par: Tan, Wenxuan, et autres
Publié: (2024)
SALMUBench: A Benchmark for Sensitive Association-Level Multimodal Unlearning
par: Selvas-Sala, Cai, et autres
Publié: (2026)
par: Selvas-Sala, Cai, et autres
Publié: (2026)
Unified Approach for Weakly Supervised Multicalibration
par: Futami, Futoshi, et autres
Publié: (2026)
par: Futami, Futoshi, et autres
Publié: (2026)
Weakly Supervised Label Learning Flows
par: Lu, You, et autres
Publié: (2023)
par: Lu, You, et autres
Publié: (2023)
Weakly-Supervised Contrastive Learning for Imprecise Class Labels
par: Zhou, Zi-Hao, et autres
Publié: (2025)
par: Zhou, Zi-Hao, et autres
Publié: (2025)
Smoothie: Label Free Language Model Routing
par: Guha, Neel, et autres
Publié: (2024)
par: Guha, Neel, et autres
Publié: (2024)
Understanding Reasoning in Thinking Language Models via Steering Vectors
par: Venhoff, Constantin, et autres
Publié: (2025)
par: Venhoff, Constantin, et autres
Publié: (2025)
Documents similaires
-
The Weak Form Is Stronger Than You Think
par: Messenger, Daniel A., et autres
Publié: (2024) -
Zero-Shot Robustification of Zero-Shot Models
par: Adila, Dyah, et autres
Publié: (2023) -
ScriptoriumWS: A Code Generation Assistant for Weak Supervision
par: Huang, Tzu-Heng, et autres
Publié: (2025) -
ResNets Are Deeper Than You Think
par: Mehmeti-Göpel, Christian H. X. Ali, et autres
Publié: (2025) -
LoRA Is Slower Than You Think
par: Ko, Seokmin
Publié: (2025)