Plantain: Plan-Answer Interleaved Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Liang, Anthony, Berant, Jonathan, Fisch, Adam, Goyal, Abhimanyu, Krishna, Kalpesh, Eisenstein, Jacob |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MT-PingEval: Evaluating Multi-Turn Collaboration with Private Information Games
by: Eisenstein, Jacob, et al.
Published: (2026)
by: Eisenstein, Jacob, et al.
Published: (2026)
Cost-Optimal Active AI Model Evaluation
by: Angelopoulos, Anastasios N., et al.
Published: (2025)
by: Angelopoulos, Anastasios N., et al.
Published: (2025)
ALTA: Compiler-Based Analysis of Transformers
by: Shaw, Peter, et al.
Published: (2024)
by: Shaw, Peter, et al.
Published: (2024)
SPRINT: Enabling Interleaved Planning and Parallelized Execution in Reasoning Models
by: Biju, Emil, et al.
Published: (2025)
by: Biju, Emil, et al.
Published: (2025)
Learning Steerable Clarification Policies with Collaborative Self-play
by: Berant, Jonathan, et al.
Published: (2025)
by: Berant, Jonathan, et al.
Published: (2025)
Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning
by: Setlur, Amrith, et al.
Published: (2024)
by: Setlur, Amrith, et al.
Published: (2024)
Don't lie to your friends: Learning what you know from collaborative self-play
by: Eisenstein, Jacob, et al.
Published: (2025)
by: Eisenstein, Jacob, et al.
Published: (2025)
Robust Preference Optimization through Reward Model Distillation
by: Fisch, Adam, et al.
Published: (2024)
by: Fisch, Adam, et al.
Published: (2024)
Mitigating Preference Hacking in Policy Optimization with Pessimism
by: Gupta, Dhawal, et al.
Published: (2025)
by: Gupta, Dhawal, et al.
Published: (2025)
Beyond the Answer: Decoding the Behavior of LLMs as Scientific Reasoners
by: Pandey, Rohan, et al.
Published: (2026)
by: Pandey, Rohan, et al.
Published: (2026)
Reasoning Model Unlearning: Forgetting Traces, Not Just Answers, While Preserving Reasoning Skills
by: Wang, Changsheng, et al.
Published: (2025)
by: Wang, Changsheng, et al.
Published: (2025)
Rich Insights from Cheap Signals: Efficient Evaluations via Tensor Factorization
by: Polo, Felipe Maia, et al.
Published: (2026)
by: Polo, Felipe Maia, et al.
Published: (2026)
SEMQA: Semi-Extractive Multi-Source Question Answering
by: Schuster, Tal, et al.
Published: (2023)
by: Schuster, Tal, et al.
Published: (2023)
PostMark: A Robust Blackbox Watermark for Large Language Models
by: Chang, Yapei, et al.
Published: (2024)
by: Chang, Yapei, et al.
Published: (2024)
Reasoned Safety Alignment: Ensuring Jailbreak Defense via Answer-Then-Check
by: Cao, Chentao, et al.
Published: (2025)
by: Cao, Chentao, et al.
Published: (2025)
Foundational Autoraters: Taming Large Language Models for Better Automatic Evaluation
by: Vu, Tu, et al.
Published: (2024)
by: Vu, Tu, et al.
Published: (2024)
Bridging Kolmogorov Complexity and Deep Learning: Asymptotically Optimal Description Length Objectives for Transformers
by: Shaw, Peter, et al.
Published: (2025)
by: Shaw, Peter, et al.
Published: (2025)
Reasoning-Aware Training for Time Series Forecasting
by: Ahamed, Md Atik, et al.
Published: (2026)
by: Ahamed, Md Atik, et al.
Published: (2026)
Hindsight Hint Distillation: Scaffolded Reasoning for SWE Agents from CoT-free Answers
by: Wang, Shengjie, et al.
Published: (2026)
by: Wang, Shengjie, et al.
Published: (2026)
From Stochastic Answers to Verifiable Reasoning: Interpretable Decision-Making with LLM-Generated Code
by: Mahesh, Anirudh Jaidev, et al.
Published: (2026)
by: Mahesh, Anirudh Jaidev, et al.
Published: (2026)
Metacognitive Reuse: Turning Recurring LLM Reasoning Into Concise Behaviors
by: Didolkar, Aniket, et al.
Published: (2025)
by: Didolkar, Aniket, et al.
Published: (2025)
OOWM: Structuring Embodied Reasoning and Planning via Object-Oriented Programmatic World Modeling
by: Chen, Hongyu, et al.
Published: (2026)
by: Chen, Hongyu, et al.
Published: (2026)
Learning What Reinforcement Learning Can't: Interleaved Online Fine-Tuning for Hardest Questions
by: Ma, Lu, et al.
Published: (2025)
by: Ma, Lu, et al.
Published: (2025)
Enabling Unstructured Sparse Acceleration on Structured Sparse Accelerators
by: Jeong, Geonhwa, et al.
Published: (2024)
by: Jeong, Geonhwa, et al.
Published: (2024)
Reinforced Reasoning for Embodied Planning
by: Wu, Di, et al.
Published: (2025)
by: Wu, Di, et al.
Published: (2025)
Memorization vs. Reasoning: Updating LLMs with New Knowledge
by: Li, Aochong Oliver, et al.
Published: (2025)
by: Li, Aochong Oliver, et al.
Published: (2025)
From Answers to Arguments: Toward Trustworthy Clinical Diagnostic Reasoning with Toulmin-Guided Curriculum Goal-Conditioned Learning
by: Zhan, Chen, et al.
Published: (2026)
by: Zhan, Chen, et al.
Published: (2026)
Simultaneous Multi-Robot Motion Planning with Projected Diffusion Models
by: Liang, Jinhao, et al.
Published: (2025)
by: Liang, Jinhao, et al.
Published: (2025)
ASAG2024: A Combined Benchmark for Short Answer Grading
by: Meyer, Gérôme, et al.
Published: (2024)
by: Meyer, Gérôme, et al.
Published: (2024)
Explaining How Quantization Disparately Skews a Model
by: Bellam, Abhimanyu, et al.
Published: (2025)
by: Bellam, Abhimanyu, et al.
Published: (2025)
Interleaving Reasoning for Better Text-to-Image Generation
by: Huang, Wenxuan, et al.
Published: (2025)
by: Huang, Wenxuan, et al.
Published: (2025)
Statistical Comparative Analysis of Semantic Similarities and Model Transferability Across Datasets for Short Answer Grading
by: Bonthu, Sridevi, et al.
Published: (2025)
by: Bonthu, Sridevi, et al.
Published: (2025)
On the Empirical Complexity of Reasoning and Planning in LLMs
by: Kang, Liwei, et al.
Published: (2024)
by: Kang, Liwei, et al.
Published: (2024)
Interleaved Gibbs Diffusion: Generating Discrete-Continuous Data with Implicit Constraints
by: Anil, Gautham Govind, et al.
Published: (2025)
by: Anil, Gautham Govind, et al.
Published: (2025)
Interleaved Multitask Learning with Energy Modulated Learning Progress
by: Say, Hanne, et al.
Published: (2025)
by: Say, Hanne, et al.
Published: (2025)
Learning to Answer from Correct Demonstrations
by: Joshi, Nirmit, et al.
Published: (2025)
by: Joshi, Nirmit, et al.
Published: (2025)
Incentivizing LLMs to Self-Verify Their Answers
by: Zhang, Fuxiang, et al.
Published: (2025)
by: Zhang, Fuxiang, et al.
Published: (2025)
Graph-enhanced Large Language Models in Asynchronous Plan Reasoning
by: Lin, Fangru, et al.
Published: (2024)
by: Lin, Fangru, et al.
Published: (2024)
Spatial Reasoning and Planning for Deep Embodied Agents
by: Ishida, Shu
Published: (2024)
by: Ishida, Shu
Published: (2024)
Conformal Risk Control
by: Angelopoulos, Anastasios N., et al.
Published: (2022)
by: Angelopoulos, Anastasios N., et al.
Published: (2022)
Similar Items
-
MT-PingEval: Evaluating Multi-Turn Collaboration with Private Information Games
by: Eisenstein, Jacob, et al.
Published: (2026) -
Cost-Optimal Active AI Model Evaluation
by: Angelopoulos, Anastasios N., et al.
Published: (2025) -
ALTA: Compiler-Based Analysis of Transformers
by: Shaw, Peter, et al.
Published: (2024) -
SPRINT: Enabling Interleaved Planning and Parallelized Execution in Reasoning Models
by: Biju, Emil, et al.
Published: (2025) -
Learning Steerable Clarification Policies with Collaborative Self-play
by: Berant, Jonathan, et al.
Published: (2025)