AdaEvolve: Adaptive LLM Driven Zeroth-Order Optimization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cemri, Mert, Agrawal, Shubham, Gupta, Akshat, Liu, Shu, Cheng, Audrey, Mang, Qiuyang, Naren, Ashwin, Erdogan, Lutfi Eren, Sen, Koushik, Zaharia, Matei, Dimakis, Alex, Stoica, Ion |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EvoX: Meta-Evolution for Automated Discovery
von: Liu, Shu, et al.
Veröffentlicht: (2026)
von: Liu, Shu, et al.
Veröffentlicht: (2026)
The Time is Here for Just-in-Time Systems: Challenges and Opportunities
von: Liu, Shu, et al.
Veröffentlicht: (2026)
von: Liu, Shu, et al.
Veröffentlicht: (2026)
Inductive Deductive Synthesis: Enabling AI to Generate Formally Verified Systems
von: Agarwal, Shubham, et al.
Veröffentlicht: (2026)
von: Agarwal, Shubham, et al.
Veröffentlicht: (2026)
SIEVE: Sample-Efficient Parametric Learning from Natural Language
von: Asawa, Parth, et al.
Veröffentlicht: (2026)
von: Asawa, Parth, et al.
Veröffentlicht: (2026)
The Price Reversal Phenomenon: When Cheaper Reasoning Models Cost More
von: Chen, Lingjiao, et al.
Veröffentlicht: (2026)
von: Chen, Lingjiao, et al.
Veröffentlicht: (2026)
DeepScholar-Bench: A Live Benchmark and Automated Evaluation for Generative Research Synthesis
von: Patel, Liana, et al.
Veröffentlicht: (2025)
von: Patel, Liana, et al.
Veröffentlicht: (2025)
Let the Barbarians In: How AI Can Accelerate Systems Performance Research
von: Cheng, Audrey, et al.
Veröffentlicht: (2025)
von: Cheng, Audrey, et al.
Veröffentlicht: (2025)
RAFT: Adapting Language Model to Domain Specific RAG
von: Zhang, Tianjun, et al.
Veröffentlicht: (2024)
von: Zhang, Tianjun, et al.
Veröffentlicht: (2024)
Are More LLM Calls All You Need? Towards Scaling Laws of Compound Inference Systems
von: Chen, Lingjiao, et al.
Veröffentlicht: (2024)
von: Chen, Lingjiao, et al.
Veröffentlicht: (2024)
Optimizing Model Selection for Compound AI Systems
von: Chen, Lingjiao, et al.
Veröffentlicht: (2025)
von: Chen, Lingjiao, et al.
Veröffentlicht: (2025)
Specifications: The missing link to making the development of LLM systems an engineering discipline
von: Stoica, Ion, et al.
Veröffentlicht: (2024)
von: Stoica, Ion, et al.
Veröffentlicht: (2024)
optimize_anything: A Universal API for Optimizing any Text Parameter
von: Agrawal, Lakshya A, et al.
Veröffentlicht: (2026)
von: Agrawal, Lakshya A, et al.
Veröffentlicht: (2026)
$\texttt{SPECS}$: Faster Test-Time Scaling through Speculative Drafts
von: Cemri, Mert, et al.
Veröffentlicht: (2025)
von: Cemri, Mert, et al.
Veröffentlicht: (2025)
DSPy Assertions: Computational Constraints for Self-Refining Language Model Pipelines
von: Singhvi, Arnav, et al.
Veröffentlicht: (2023)
von: Singhvi, Arnav, et al.
Veröffentlicht: (2023)
Barbarians at the Gate: How AI is Upending Systems Research
von: Cheng, Audrey, et al.
Veröffentlicht: (2025)
von: Cheng, Audrey, et al.
Veröffentlicht: (2025)
Characterizing Prompt Compression Methods for Long Context Inference
von: Jha, Siddharth, et al.
Veröffentlicht: (2024)
von: Jha, Siddharth, et al.
Veröffentlicht: (2024)
Stochastic Communication Avoidance for Recommendation Systems
von: Erdogan, Lutfi Eren, et al.
Veröffentlicht: (2024)
von: Erdogan, Lutfi Eren, et al.
Veröffentlicht: (2024)
R2E-Gym: Procedural Environments and Hybrid Verifiers for Scaling Open-Weights SWE Agents
von: Jain, Naman, et al.
Veröffentlicht: (2025)
von: Jain, Naman, et al.
Veröffentlicht: (2025)
Delta Fair Sharing: Performance Isolation for Multi-Tenant Storage Systems
von: Griggs, Tyler, et al.
Veröffentlicht: (2026)
von: Griggs, Tyler, et al.
Veröffentlicht: (2026)
BARE: Leveraging Base Language Models for Few-Shot Synthetic Data Generation
von: Zhu, Alan, et al.
Veröffentlicht: (2025)
von: Zhu, Alan, et al.
Veröffentlicht: (2025)
GSO: Challenging Software Optimization Tasks for Evaluating SWE-Agents
von: Shetty, Manish, et al.
Veröffentlicht: (2025)
von: Shetty, Manish, et al.
Veröffentlicht: (2025)
LangProBe: a Language Programs Benchmark
von: Tan, Shangyin, et al.
Veröffentlicht: (2025)
von: Tan, Shangyin, et al.
Veröffentlicht: (2025)
How to Train Your Advisor: Steering Black-Box LLMs with Advisor Models
von: Asawa, Parth, et al.
Veröffentlicht: (2025)
von: Asawa, Parth, et al.
Veröffentlicht: (2025)
GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
von: Agrawal, Lakshya A, et al.
Veröffentlicht: (2025)
von: Agrawal, Lakshya A, et al.
Veröffentlicht: (2025)
MoE-Lightning: High-Throughput MoE Inference on Memory-constrained GPUs
von: Cao, Shiyi, et al.
Veröffentlicht: (2024)
von: Cao, Shiyi, et al.
Veröffentlicht: (2024)
SVG-EAR: Parameter-Free Linear Compensation for Sparse Video Generation via Error-aware Routing
von: Zhou, Xuanyi, et al.
Veröffentlicht: (2026)
von: Zhou, Xuanyi, et al.
Veröffentlicht: (2026)
Can QPP Choose the Right Query Variant? Evaluating Query Variant Selection for RAG Pipelines
von: Arabzadeh, Negar, et al.
Veröffentlicht: (2026)
von: Arabzadeh, Negar, et al.
Veröffentlicht: (2026)
RedunCut: Measurement-Driven Sampling and Accuracy Performance Modeling for Low-Cost Live Video Analytics
von: Sela, Gur-Eyal, et al.
Veröffentlicht: (2025)
von: Sela, Gur-Eyal, et al.
Veröffentlicht: (2025)
Deep Image Composition Meets Image Forgery
von: Tahir, Eren, et al.
Veröffentlicht: (2024)
von: Tahir, Eren, et al.
Veröffentlicht: (2024)
Deep Image Restoration For Image Anti-Forensics
von: Tahir, Eren, et al.
Veröffentlicht: (2024)
von: Tahir, Eren, et al.
Veröffentlicht: (2024)
Plan-and-Act: Improving Planning of Agents for Long-Horizon Tasks
von: Erdogan, Lutfi Eren, et al.
Veröffentlicht: (2025)
von: Erdogan, Lutfi Eren, et al.
Veröffentlicht: (2025)
Why Do Multi-Agent LLM Systems Fail?
von: Cemri, Mert, et al.
Veröffentlicht: (2025)
von: Cemri, Mert, et al.
Veröffentlicht: (2025)
Efficient and Scalable Estimation of Tool Representations in Vector Space
von: Moon, Suhong, et al.
Veröffentlicht: (2024)
von: Moon, Suhong, et al.
Veröffentlicht: (2024)
RAG over Thinking Traces Can Improve Reasoning Tasks
von: Arabzadeh, Negar, et al.
Veröffentlicht: (2026)
von: Arabzadeh, Negar, et al.
Veröffentlicht: (2026)
AI-Driven Research for Databases
von: Cheng, Audrey, et al.
Veröffentlicht: (2026)
von: Cheng, Audrey, et al.
Veröffentlicht: (2026)
Agentic Test-Time Scaling for WebAgents
von: Lee, Nicholas, et al.
Veröffentlicht: (2026)
von: Lee, Nicholas, et al.
Veröffentlicht: (2026)
ARES: An Automated Evaluation Framework for Retrieval-Augmented Generation Systems
von: Saad-Falcon, Jon, et al.
Veröffentlicht: (2023)
von: Saad-Falcon, Jon, et al.
Veröffentlicht: (2023)
Combee: Scaling Prompt Learning for Self-Improving Language Model Agents
von: Li, Hanchen, et al.
Veröffentlicht: (2026)
von: Li, Hanchen, et al.
Veröffentlicht: (2026)
Long Context RAG Performance of Large Language Models
von: Leng, Quinn, et al.
Veröffentlicht: (2024)
von: Leng, Quinn, et al.
Veröffentlicht: (2024)
HashAttention: Semantic Sparsity for Faster Inference
von: Desai, Aditya, et al.
Veröffentlicht: (2024)
von: Desai, Aditya, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
EvoX: Meta-Evolution for Automated Discovery
von: Liu, Shu, et al.
Veröffentlicht: (2026) -
The Time is Here for Just-in-Time Systems: Challenges and Opportunities
von: Liu, Shu, et al.
Veröffentlicht: (2026) -
Inductive Deductive Synthesis: Enabling AI to Generate Formally Verified Systems
von: Agarwal, Shubham, et al.
Veröffentlicht: (2026) -
SIEVE: Sample-Efficient Parametric Learning from Natural Language
von: Asawa, Parth, et al.
Veröffentlicht: (2026) -
The Price Reversal Phenomenon: When Cheaper Reasoning Models Cost More
von: Chen, Lingjiao, et al.
Veröffentlicht: (2026)