Meta-Harness: End-to-End Optimization of Model Harnesses
Fuente:
arXiv
Salvato in:
| Autori principali: | Lee, Yoonho, Nair, Roshen, Zhang, Qizheng, Lee, Kangwook, Khattab, Omar, Finn, Chelsea |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Self-Guided Masked Autoencoders for Domain-Agnostic Self-Supervised Learning
di: Xie, Johnathan, et al.
Pubblicazione: (2024)
di: Xie, Johnathan, et al.
Pubblicazione: (2024)
Calibrating Language Models with Adaptive Temperature Scaling
di: Xie, Johnathan, et al.
Pubblicazione: (2024)
di: Xie, Johnathan, et al.
Pubblicazione: (2024)
Clarify: Improving Model Robustness With Natural Language Corrections
di: Lee, Yoonho, et al.
Pubblicazione: (2024)
di: Lee, Yoonho, et al.
Pubblicazione: (2024)
Grounding by Trying: LLMs with Reinforcement Learning-Enhanced Retrieval
di: Hsu, Sheryl, et al.
Pubblicazione: (2024)
di: Hsu, Sheryl, et al.
Pubblicazione: (2024)
NORA: A Harness-Engineered Autonomous Research Agent for End-to-End Spatial Data Science
di: Zhou, Bing, et al.
Pubblicazione: (2026)
di: Zhou, Bing, et al.
Pubblicazione: (2026)
Conservative Prediction via Data-Driven Confidence Minimization
di: Choi, Caroline, et al.
Pubblicazione: (2023)
di: Choi, Caroline, et al.
Pubblicazione: (2023)
PEEK: Context Map as an Orientation Cache for Long-Context LLM Agents
di: Gu, Zhuohan, et al.
Pubblicazione: (2026)
di: Gu, Zhuohan, et al.
Pubblicazione: (2026)
Bidirectional Decoding: Improving Action Chunking via Guided Test-Time Sampling
di: Liu, Yuejiang, et al.
Pubblicazione: (2024)
di: Liu, Yuejiang, et al.
Pubblicazione: (2024)
RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems
di: Qu, Yuxiao, et al.
Pubblicazione: (2025)
di: Qu, Yuxiao, et al.
Pubblicazione: (2025)
Feedback Descent: Open-Ended Text Optimization via Pairwise Comparison
di: Lee, Yoonho, et al.
Pubblicazione: (2025)
di: Lee, Yoonho, et al.
Pubblicazione: (2025)
Causality-Aware End-to-End Autonomous Driving via Ego-Centric Joint Scene Modeling
di: Moon, Seokha, et al.
Pubblicazione: (2026)
di: Moon, Seokha, et al.
Pubblicazione: (2026)
TAPE: Tool-Guided Adaptive Planning and Constrained Execution in Language Model Agents
di: Jeong, Jongwon, et al.
Pubblicazione: (2026)
di: Jeong, Jongwon, et al.
Pubblicazione: (2026)
Harness-Bench: Measuring Harness Effects across Models in Realistic Agent Workflows
di: Yao, Yilun, et al.
Pubblicazione: (2026)
di: Yao, Yilun, et al.
Pubblicazione: (2026)
SynAD: Enhancing Real-World End-to-End Autonomous Driving Models through Synthetic Data Integration
di: Kim, Jongsuk, et al.
Pubblicazione: (2025)
di: Kim, Jongsuk, et al.
Pubblicazione: (2025)
HARBOR: Automated Harness Optimization
di: Sengupta, Biswa, et al.
Pubblicazione: (2026)
di: Sengupta, Biswa, et al.
Pubblicazione: (2026)
GSQA: An End-to-End Model for Generative Spoken Question Answering
di: Shih, Min-Han, et al.
Pubblicazione: (2023)
di: Shih, Min-Han, et al.
Pubblicazione: (2023)
EXAONE Path 2.0: Pathology Foundation Model with End-to-End Supervision
di: Pyeon, Myeongjang, et al.
Pubblicazione: (2025)
di: Pyeon, Myeongjang, et al.
Pubblicazione: (2025)
End-to-End Learning for Fair Multiobjective Optimization Under Uncertainty
di: Dinh, My H, et al.
Pubblicazione: (2024)
di: Dinh, My H, et al.
Pubblicazione: (2024)
Harnessing Optimization Dynamics for Curvature-Informed Model Merging
di: Mahdavinia, Pouria, et al.
Pubblicazione: (2025)
di: Mahdavinia, Pouria, et al.
Pubblicazione: (2025)
The Expressive Power of Low-Rank Adaptation
di: Zeng, Yuchen, et al.
Pubblicazione: (2023)
di: Zeng, Yuchen, et al.
Pubblicazione: (2023)
Identifying Functionally Important Features with End-to-End Sparse Dictionary Learning
di: Braun, Dan, et al.
Pubblicazione: (2024)
di: Braun, Dan, et al.
Pubblicazione: (2024)
TriGen: NPU Architecture for End-to-End Acceleration of Large Language Models based on SW-HW Co-Design
di: Lee, Jonghun, et al.
Pubblicazione: (2026)
di: Lee, Jonghun, et al.
Pubblicazione: (2026)
Weak-for-Strong: Training Weak Meta-Agent to Harness Strong Executors
di: Nie, Fan, et al.
Pubblicazione: (2025)
di: Nie, Fan, et al.
Pubblicazione: (2025)
vla-eval: A Unified Evaluation Harness for Vision-Language-Action Models
di: Choi, Suhwan, et al.
Pubblicazione: (2026)
di: Choi, Suhwan, et al.
Pubblicazione: (2026)
ORCA: An End-to-End Interactive Copilot for Optimized Root Cause Analysis
di: Xuan, Phi Nguyen, et al.
Pubblicazione: (2026)
di: Xuan, Phi Nguyen, et al.
Pubblicazione: (2026)
Harness Updating Is Not Harness Benefit: Disentangling Evolution Capabilities in Self-Evolving LLM Agents
di: Lin, Minhua, et al.
Pubblicazione: (2026)
di: Lin, Minhua, et al.
Pubblicazione: (2026)
MetaAgent-X : Breaking the Ceiling of Automatic Multi-Agent Systems via End-to-End Reinforcement Learning
di: Zhang, Yaolun, et al.
Pubblicazione: (2026)
di: Zhang, Yaolun, et al.
Pubblicazione: (2026)
Recursive Language Models
di: Zhang, Alex L., et al.
Pubblicazione: (2025)
di: Zhang, Alex L., et al.
Pubblicazione: (2025)
Large Language Models as End-to-end Combinatorial Optimization Solvers
di: Jiang, Xia, et al.
Pubblicazione: (2025)
di: Jiang, Xia, et al.
Pubblicazione: (2025)
Wave-U-Mamba: An End-To-End Framework For High-Quality And Efficient Speech Super Resolution
di: Lee, Yongjoon, et al.
Pubblicazione: (2024)
di: Lee, Yongjoon, et al.
Pubblicazione: (2024)
Evolving-RL: End-to-End Optimization of Experience-Driven Self-Evolving Capability within Agents
di: Fan, Zhiyuan, et al.
Pubblicazione: (2026)
di: Fan, Zhiyuan, et al.
Pubblicazione: (2026)
EndToEndML: An Open-Source End-to-End Pipeline for Machine Learning Applications
di: Pillai, Nisha, et al.
Pubblicazione: (2024)
di: Pillai, Nisha, et al.
Pubblicazione: (2024)
Self-Evolving Recommendation System: End-To-End Autonomous Model Optimization With LLM Agents
di: Wang, Haochen, et al.
Pubblicazione: (2026)
di: Wang, Haochen, et al.
Pubblicazione: (2026)
End-to-End Optimized Image Compression with the Frequency-Oriented Transform
di: Zhang, Yuefeng, et al.
Pubblicazione: (2024)
di: Zhang, Yuefeng, et al.
Pubblicazione: (2024)
The End of Manual Decoding: Towards Truly End-to-End Language Models
di: Wang, Zhichao, et al.
Pubblicazione: (2025)
di: Wang, Zhichao, et al.
Pubblicazione: (2025)
MEMERAG: A Multilingual End-to-End Meta-Evaluation Benchmark for Retrieval Augmented Generation
di: Blandón, María Andrea Cruz, et al.
Pubblicazione: (2025)
di: Blandón, María Andrea Cruz, et al.
Pubblicazione: (2025)
Towards Direct Evaluation of Harness Optimizers via Priority Ranking
di: Ong, Kai Tzu-iunn, et al.
Pubblicazione: (2026)
di: Ong, Kai Tzu-iunn, et al.
Pubblicazione: (2026)
Fine-Tuning and Prompt Optimization: Two Great Steps that Work Better Together
di: Soylu, Dilara, et al.
Pubblicazione: (2024)
di: Soylu, Dilara, et al.
Pubblicazione: (2024)
ClickDiffusion: Harnessing LLMs for Interactive Precise Image Editing
di: Helbling, Alec, et al.
Pubblicazione: (2024)
di: Helbling, Alec, et al.
Pubblicazione: (2024)
HARIVO: Harnessing Text-to-Image Models for Video Generation
di: Kwon, Mingi, et al.
Pubblicazione: (2024)
di: Kwon, Mingi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Self-Guided Masked Autoencoders for Domain-Agnostic Self-Supervised Learning
di: Xie, Johnathan, et al.
Pubblicazione: (2024) -
Calibrating Language Models with Adaptive Temperature Scaling
di: Xie, Johnathan, et al.
Pubblicazione: (2024) -
Clarify: Improving Model Robustness With Natural Language Corrections
di: Lee, Yoonho, et al.
Pubblicazione: (2024) -
Grounding by Trying: LLMs with Reinforcement Learning-Enhanced Retrieval
di: Hsu, Sheryl, et al.
Pubblicazione: (2024) -
NORA: A Harness-Engineered Autonomous Research Agent for End-to-End Spatial Data Science
di: Zhou, Bing, et al.
Pubblicazione: (2026)