SPRINT: Enabling Interleaved Planning and Parallelized Execution in Reasoning Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Biju, Emil, Talaei, Shayan, Huang, Zhemin, Pourreza, Mohammadreza, Mirhoseini, Azalia, Saberi, Amin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
CHESS: Contextual Harnessing for Efficient SQL Synthesis
von: Talaei, Shayan, et al.
Veröffentlicht: (2024)
von: Talaei, Shayan, et al.
Veröffentlicht: (2024)
Reasoning-SQL: Reinforcement Learning with SQL Tailored Partial Rewards for Reasoning-Enhanced Text-to-SQL
von: Pourreza, Mohammadreza, et al.
Veröffentlicht: (2025)
von: Pourreza, Mohammadreza, et al.
Veröffentlicht: (2025)
CHASE-SQL: Multi-Path Reasoning and Preference Optimized Candidate Selection in Text-to-SQL
von: Pourreza, Mohammadreza, et al.
Veröffentlicht: (2024)
von: Pourreza, Mohammadreza, et al.
Veröffentlicht: (2024)
That Chip Has Sailed: A Critique of Unfounded Skepticism Around AI for Chip Design
von: Goldie, Anna, et al.
Veröffentlicht: (2024)
von: Goldie, Anna, et al.
Veröffentlicht: (2024)
Agent JIT Compilation for Latency-Optimizing Web Agent Planning and Scheduling
von: Winston, Caleb, et al.
Veröffentlicht: (2026)
von: Winston, Caleb, et al.
Veröffentlicht: (2026)
On the Role of Temperature Sampling in Test-Time Scaling
von: Wu, Yuheng, et al.
Veröffentlicht: (2025)
von: Wu, Yuheng, et al.
Veröffentlicht: (2025)
ForTIFAI: Fending Off Recursive Training Induced Failure for AI Model Collapse
von: Shabgahi, Soheil Zibakhsh, et al.
Veröffentlicht: (2025)
von: Shabgahi, Soheil Zibakhsh, et al.
Veröffentlicht: (2025)
Synthetic Data Generation & Multi-Step RL for Reasoning & Tool Use
von: Goldie, Anna, et al.
Veröffentlicht: (2025)
von: Goldie, Anna, et al.
Veröffentlicht: (2025)
Plantain: Plan-Answer Interleaved Reasoning
von: Liang, Anthony, et al.
Veröffentlicht: (2025)
von: Liang, Anthony, et al.
Veröffentlicht: (2025)
Large Language Monkeys: Scaling Inference Compute with Repeated Sampling
von: Brown, Bradley, et al.
Veröffentlicht: (2024)
von: Brown, Bradley, et al.
Veröffentlicht: (2024)
Think, Prune, Train, Improve: Scaling Reasoning without Scaling Models
von: Costello, Caia, et al.
Veröffentlicht: (2025)
von: Costello, Caia, et al.
Veröffentlicht: (2025)
How Do Large Language Monkeys Get Their Power (Laws)?
von: Schaeffer, Rylan, et al.
Veröffentlicht: (2025)
von: Schaeffer, Rylan, et al.
Veröffentlicht: (2025)
KernelBench: Can LLMs Write Efficient GPU Kernels?
von: Ouyang, Anne, et al.
Veröffentlicht: (2025)
von: Ouyang, Anne, et al.
Veröffentlicht: (2025)
SQL-GEN: Bridging the Dialect Gap for Text-to-SQL Via Synthetic Data And Model Merging
von: Pourreza, Mohammadreza, et al.
Veröffentlicht: (2024)
von: Pourreza, Mohammadreza, et al.
Veröffentlicht: (2024)
TransZero: Parallel Tree Expansion in MuZero using Transformer Networks
von: Malmsten, Emil, et al.
Veröffentlicht: (2025)
von: Malmsten, Emil, et al.
Veröffentlicht: (2025)
SPRINT: Robust Model Attribution of Generated Images via Secret Pixel Reconstruction
von: Yao, Kai, et al.
Veröffentlicht: (2025)
von: Yao, Kai, et al.
Veröffentlicht: (2025)
SPRINT: Scalable Policy Pre-Training via Language Instruction Relabeling
von: Zhang, Jesse, et al.
Veröffentlicht: (2023)
von: Zhang, Jesse, et al.
Veröffentlicht: (2023)
Cartridges: Lightweight and general-purpose long context representations via self-study
von: Eyuboglu, Sabri, et al.
Veröffentlicht: (2025)
von: Eyuboglu, Sabri, et al.
Veröffentlicht: (2025)
BitPipe: Bidirectional Interleaved Pipeline Parallelism for Accelerating Large Models Training
von: Wu, Houming, et al.
Veröffentlicht: (2024)
von: Wu, Houming, et al.
Veröffentlicht: (2024)
When to Trust the Cheap Check: Weak and Strong Verification for Reasoning
von: Kiyani, Shayan, et al.
Veröffentlicht: (2026)
von: Kiyani, Shayan, et al.
Veröffentlicht: (2026)
Exploring Diffusion Transformer Designs via Grafting
von: Chandrasegaran, Keshigeyan, et al.
Veröffentlicht: (2025)
von: Chandrasegaran, Keshigeyan, et al.
Veröffentlicht: (2025)
AdaReasoner: Adaptive Reasoning Enables More Flexible Thinking in Large Language Models
von: Wang, Xiangqi, et al.
Veröffentlicht: (2025)
von: Wang, Xiangqi, et al.
Veröffentlicht: (2025)
StorySage: Conversational Autobiography Writing Powered by a Multi-Agent Framework
von: Talaei, Shayan, et al.
Veröffentlicht: (2025)
von: Talaei, Shayan, et al.
Veröffentlicht: (2025)
Archon: An Architecture Search Framework for Inference-Time Techniques
von: Saad-Falcon, Jon, et al.
Veröffentlicht: (2024)
von: Saad-Falcon, Jon, et al.
Veröffentlicht: (2024)
Optimizing Data Curation through Spectral Analysis and Joint Batch Selection (SALN)
von: Sharifi, Mohammadreza
Veröffentlicht: (2024)
von: Sharifi, Mohammadreza
Veröffentlicht: (2024)
Model Parallelism With Subnetwork Data Parallelism
von: Singh, Vaibhav, et al.
Veröffentlicht: (2025)
von: Singh, Vaibhav, et al.
Veröffentlicht: (2025)
Federation of Experts: Communication Efficient Distributed Inference for Large Language Models
von: Abdurrahman, Muhammad Shahir, et al.
Veröffentlicht: (2026)
von: Abdurrahman, Muhammad Shahir, et al.
Veröffentlicht: (2026)
Interleaving Reasoning for Better Text-to-Image Generation
von: Huang, Wenxuan, et al.
Veröffentlicht: (2025)
von: Huang, Wenxuan, et al.
Veröffentlicht: (2025)
Astra: A Multi-Agent System for GPU Kernel Performance Optimization
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning
von: Hwang, Jaebak, et al.
Veröffentlicht: (2025)
von: Hwang, Jaebak, et al.
Veröffentlicht: (2025)
SMOGAN: Synthetic Minority Oversampling with GAN Refinement for Imbalanced Regression
von: Alahyari, Shayan, et al.
Veröffentlicht: (2025)
von: Alahyari, Shayan, et al.
Veröffentlicht: (2025)
Parallel Test-Time Scaling for Latent Reasoning Models
von: You, Runyang, et al.
Veröffentlicht: (2025)
von: You, Runyang, et al.
Veröffentlicht: (2025)
Drone-Based Multispectral Imaging and Deep Learning for Timely Detection of Branched Broomrape in Tomato Farms
von: Narimani, Mohammadreza, et al.
Veröffentlicht: (2025)
von: Narimani, Mohammadreza, et al.
Veröffentlicht: (2025)
Branched Broomrape Detection in Tomato Farms Using Satellite Imagery and Time-Series Analysis
von: Narimani, Mohammadreza, et al.
Veröffentlicht: (2025)
von: Narimani, Mohammadreza, et al.
Veröffentlicht: (2025)
ProSpec RL: Plan Ahead, then Execute
von: Liu, Liangliang, et al.
Veröffentlicht: (2024)
von: Liu, Liangliang, et al.
Veröffentlicht: (2024)
TRACE: Capability-Targeted Agentic Training
von: Kang, Hangoo, et al.
Veröffentlicht: (2026)
von: Kang, Hangoo, et al.
Veröffentlicht: (2026)
Architectural Flaw Detection in Civil Engineering Using GPT-4
von: Kumar, Saket, et al.
Veröffentlicht: (2024)
von: Kumar, Saket, et al.
Veröffentlicht: (2024)
Reasoning Planning for Language Models
von: Nguyen, Bao, et al.
Veröffentlicht: (2025)
von: Nguyen, Bao, et al.
Veröffentlicht: (2025)
Reinforced Reasoning for Embodied Planning
von: Wu, Di, et al.
Veröffentlicht: (2025)
von: Wu, Di, et al.
Veröffentlicht: (2025)
Plan, Verify and Fill: A Structured Parallel Decoding Approach for Diffusion Language Models
von: Li, Miao, et al.
Veröffentlicht: (2026)
von: Li, Miao, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
CHESS: Contextual Harnessing for Efficient SQL Synthesis
von: Talaei, Shayan, et al.
Veröffentlicht: (2024) -
Reasoning-SQL: Reinforcement Learning with SQL Tailored Partial Rewards for Reasoning-Enhanced Text-to-SQL
von: Pourreza, Mohammadreza, et al.
Veröffentlicht: (2025) -
CHASE-SQL: Multi-Path Reasoning and Preference Optimized Candidate Selection in Text-to-SQL
von: Pourreza, Mohammadreza, et al.
Veröffentlicht: (2024) -
That Chip Has Sailed: A Critique of Unfounded Skepticism Around AI for Chip Design
von: Goldie, Anna, et al.
Veröffentlicht: (2024) -
Agent JIT Compilation for Latency-Optimizing Web Agent Planning and Scheduling
von: Winston, Caleb, et al.
Veröffentlicht: (2026)