Saved in:
| Main Authors: | Saad-Falcon, Jon, Buchanan, E. Kelly, Chen, Mayee F., Huang, Tzu-Heng, McLaughlin, Brendan, Bhathal, Tanvir, Zhu, Shang, Athiwaratkun, Ben, Sala, Frederic, Linderman, Scott, Mirhoseini, Azalia, Ré, Christopher |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2506.18203 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Archon: An Architecture Search Framework for Inference-Time Techniques
by: Saad-Falcon, Jon, et al.
Published: (2024)
by: Saad-Falcon, Jon, et al.
Published: (2024)
TRACE: Capability-Targeted Agentic Training
by: Kang, Hangoo, et al.
Published: (2026)
by: Kang, Hangoo, et al.
Published: (2026)
OpenJarvis: Personal AI, On Personal Devices
by: Saad-Falcon, Jon, et al.
Published: (2026)
by: Saad-Falcon, Jon, et al.
Published: (2026)
WebSight: A Vision-First Architecture for Robust Web Agents
by: Bhathal, Tanvir, et al.
Published: (2025)
by: Bhathal, Tanvir, et al.
Published: (2025)
EmissionNet: Air Quality Pollution Forecasting for Agriculture
by: Saligram, Prady, et al.
Published: (2025)
by: Saligram, Prady, et al.
Published: (2025)
On the Role of Temperature Sampling in Test-Time Scaling
by: Wu, Yuheng, et al.
Published: (2025)
by: Wu, Yuheng, et al.
Published: (2025)
That Chip Has Sailed: A Critique of Unfounded Skepticism Around AI for Chip Design
by: Goldie, Anna, et al.
Published: (2024)
by: Goldie, Anna, et al.
Published: (2024)
CodeMonkeys: Scaling Test-Time Compute for Software Engineering
by: Ehrlich, Ryan, et al.
Published: (2025)
by: Ehrlich, Ryan, et al.
Published: (2025)
Hydragen: High-Throughput LLM Inference with Shared Prefixes
by: Juravsky, Jordan, et al.
Published: (2024)
by: Juravsky, Jordan, et al.
Published: (2024)
2048: Reinforcement Learning in a Delayed Reward Environment
by: Saligram, Prady, et al.
Published: (2025)
by: Saligram, Prady, et al.
Published: (2025)
Think, Prune, Train, Improve: Scaling Reasoning without Scaling Models
by: Costello, Caia, et al.
Published: (2025)
by: Costello, Caia, et al.
Published: (2025)
Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
by: Brown, Bradley, et al.
Published: (2024)
by: Brown, Bradley, et al.
Published: (2024)
KernelBench: Can LLMs Write Efficient GPU Kernels?
by: Ouyang, Anne, et al.
Published: (2025)
by: Ouyang, Anne, et al.
Published: (2025)
Scaling Verification Can Be More Effective than Scaling Policy Learning for Vision-Language-Action Alignment
by: Kwok, Jacky, et al.
Published: (2026)
by: Kwok, Jacky, et al.
Published: (2026)
Time To Impeach LLM-as-a-Judge: Programs are the Future of Evaluation
by: Huang, Tzu-Heng, et al.
Published: (2025)
by: Huang, Tzu-Heng, et al.
Published: (2025)
Think Deep, Think Fast: Investigating Efficiency of Verifier-free Inference-time-scaling Methods
by: Wang, Junlin, et al.
Published: (2025)
by: Wang, Junlin, et al.
Published: (2025)
Federation of Experts: Communication Efficient Distributed Inference for Large Language Models
by: Abdurrahman, Muhammad Shahir, et al.
Published: (2026)
by: Abdurrahman, Muhammad Shahir, et al.
Published: (2026)
ForTIFAI: Fending Off Recursive Training Induced Failure for AI Model Collapse
by: Shabgahi, Soheil Zibakhsh, et al.
Published: (2025)
by: Shabgahi, Soheil Zibakhsh, et al.
Published: (2025)
Agent JIT Compilation for Latency-Optimizing Web Agent Planning and Scheduling
by: Winston, Caleb, et al.
Published: (2026)
by: Winston, Caleb, et al.
Published: (2026)
ScriptoriumWS: A Code Generation Assistant for Weak Supervision
by: Huang, Tzu-Heng, et al.
Published: (2025)
by: Huang, Tzu-Heng, et al.
Published: (2025)
RoboMonkey: Scaling Test-Time Sampling and Verification for Vision-Language-Action Models
by: Kwok, Jacky, et al.
Published: (2025)
by: Kwok, Jacky, et al.
Published: (2025)
Stellar astrophysics
by: Bhathal, Ragbir S
Published: (2009)
by: Bhathal, Ragbir S
Published: (2009)
Benchmarking and Building Long-Context Retrieval Models with LoCo and M2-BERT
by: Saad-Falcon, Jon, et al.
Published: (2024)
by: Saad-Falcon, Jon, et al.
Published: (2024)
Synthetic Data Generation & Multi-Step RL for Reasoning & Tool Use
by: Goldie, Anna, et al.
Published: (2025)
by: Goldie, Anna, et al.
Published: (2025)
CHESS: Contextual Harnessing for Efficient SQL Synthesis
by: Talaei, Shayan, et al.
Published: (2024)
by: Talaei, Shayan, et al.
Published: (2024)
CATS: Contextually-Aware Thresholding for Sparsity in Large Language Models
by: Lee, Donghyun, et al.
Published: (2024)
by: Lee, Donghyun, et al.
Published: (2024)
The ALCHEmist: Automated Labeling 500x CHEaper Than LLM Data Annotators
by: Huang, Tzu-Heng, et al.
Published: (2024)
by: Huang, Tzu-Heng, et al.
Published: (2024)
Cookbook: A framework for improving LLM generative abilities via programmatic data generating templates
by: Narayan, Avanika, et al.
Published: (2024)
by: Narayan, Avanika, et al.
Published: (2024)
Mind the Gap: Detecting Black-box Adversarial Attacks in the Making through Query Update Analysis
by: Park, Jeonghwan, et al.
Published: (2025)
by: Park, Jeonghwan, et al.
Published: (2025)
The War Correspondent - Second Edition
by: McLaughlin, Greg
Published: (2016)
by: McLaughlin, Greg
Published: (2016)
Edge of Catastrophe: Erich Fromm, Fascism, and the HolocaustBy Roger Frie, Oxford University Press, 2024. 216 pp. $35 (cloth). ISBN: 9780197748770.
by: Neil McLaughlin
Published: (2025)
by: Neil McLaughlin
Published: (2025)
Priority Projects: Constituent Spending Demand and the Benefits of Congressional Credit Claiming
by: Peter T. McLaughlin
Published: (2025)
by: Peter T. McLaughlin
Published: (2025)
La era de Klein / Patricia McLaughlin
by: McLaughlin, Patricia
by: McLaughlin, Patricia
Using Play‐Doh to teach creative data collection
by: Sarah McLaughlin
Published: (2024)
by: Sarah McLaughlin
Published: (2024)
Visionaries and Crackpots, Maniacs and Saints: Existential Risk and the Politics of Longtermism
by: Alex McLaughlin
Published: (2025)
by: Alex McLaughlin
Published: (2025)
Managing Computer Software Collections. ERIC Digest.
by: McLaughlin, Pamela
Published: (1984)
by: McLaughlin, Pamela
Published: (1984)
Media Specialists and the Curriculum: A Selected ERIC Bibliography.
by: McLaughlin, Pamela
Published: (1985)
by: McLaughlin, Pamela
Published: (1985)
Kids, Kulture, and Us
by: McLaughlin, Frank
Published: (1972)
by: McLaughlin, Frank
Published: (1972)
Test-Time Scaling Makes Overtraining Compute-Optimal
by: Roberts, Nicholas, et al.
Published: (2026)
by: Roberts, Nicholas, et al.
Published: (2026)
SPRINT: Enabling Interleaved Planning and Parallelized Execution in Reasoning Models
by: Biju, Emil, et al.
Published: (2025)
by: Biju, Emil, et al.
Published: (2025)
Similar Items
-
Archon: An Architecture Search Framework for Inference-Time Techniques
by: Saad-Falcon, Jon, et al.
Published: (2024) -
TRACE: Capability-Targeted Agentic Training
by: Kang, Hangoo, et al.
Published: (2026) -
OpenJarvis: Personal AI, On Personal Devices
by: Saad-Falcon, Jon, et al.
Published: (2026) -
WebSight: A Vision-First Architecture for Robust Web Agents
by: Bhathal, Tanvir, et al.
Published: (2025) -
EmissionNet: Air Quality Pollution Forecasting for Agriculture
by: Saligram, Prady, et al.
Published: (2025)