CoSPlay: Cooperative Self-Play at Test-Time with Self-Generated Code and Unit Test
Fuente:
arXiv
Saved in:
| Main Authors: | Hu, Zhangyi, Liu, Chenhui, Huang, Tian, Li, Jindong, Yang, Yang, Wu, Jiemin, Zhong, Zining, Yang, Menglin, Yue, Yutao |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TurboEvolve: Towards Fast and Robust LLM-Driven Program Evolution
by: Yang, Yang, et al.
Published: (2026)
by: Yang, Yang, et al.
Published: (2026)
Hypothesis Generation via LLM-Automated Language Bias for ILP
by: Yang, Yang, et al.
Published: (2025)
by: Yang, Yang, et al.
Published: (2025)
RLIE: Rule Generation with Logistic Regression, Iterative Refinement, and Evaluation for Large Language Models
by: Yang, Yang, et al.
Published: (2025)
by: Yang, Yang, et al.
Published: (2025)
Self-Reflective Generation at Test Time
by: Mu, Jian, et al.
Published: (2025)
by: Mu, Jian, et al.
Published: (2025)
IMTS is Worth Time $\times$ Channel Patches: Visual Masked Autoencoders for Irregular Multivariate Time Series Prediction
by: Hu, Zhangyi, et al.
Published: (2025)
by: Hu, Zhangyi, et al.
Published: (2025)
Self-Harmony: Learning to Harmonize Self-Supervision and Self-Play in Test-Time Reinforcement Learning
by: Wang, Ru, et al.
Published: (2025)
by: Wang, Ru, et al.
Published: (2025)
Learning to Solve and Verify: A Self-Play Framework for Code and Test Generation
by: Lin, Zi, et al.
Published: (2025)
by: Lin, Zi, et al.
Published: (2025)
CasModaTest: A Cascaded and Model-agnostic Self-directed Framework for Unit Test Generation
by: Ni, Chao, et al.
Published: (2024)
by: Ni, Chao, et al.
Published: (2024)
ACE: Self-Evolving LLM Coding Framework via Adversarial Unit Test Generation and Preference Optimization
by: Huang, Yixu, et al.
Published: (2026)
by: Huang, Yixu, et al.
Published: (2026)
Revisit Self-Debugging with Self-Generated Tests for Code Generation
by: Chen, Xiancai, et al.
Published: (2025)
by: Chen, Xiancai, et al.
Published: (2025)
UTMath: Math Evaluation with Unit Test via Reasoning-to-Coding Thoughts
by: Yang, Bo, et al.
Published: (2024)
by: Yang, Bo, et al.
Published: (2024)
Superhuman AI for Stratego Using Self-Play Reinforcement Learning and Test-Time Search
by: Sokota, Samuel, et al.
Published: (2025)
by: Sokota, Samuel, et al.
Published: (2025)
Automated Unit Test Refactoring
by: Gao, Yi, et al.
Published: (2024)
by: Gao, Yi, et al.
Published: (2024)
PETS: A Principled Framework Towards Optimal Trajectory Allocation for Efficient Test-Time Self-Consistency
by: Liu, Zhangyi, et al.
Published: (2026)
by: Liu, Zhangyi, et al.
Published: (2026)
DSTC: Direct Preference Learning with Only Self-Generated Tests and Code to Improve Code LMs
by: Liu, Zhihan, et al.
Published: (2024)
by: Liu, Zhihan, et al.
Published: (2024)
Requirements-Based Test Generation: A Comprehensive Survey
by: Yang, Zhenzhen, et al.
Published: (2025)
by: Yang, Zhenzhen, et al.
Published: (2025)
What You See Is What You Get: Attention-based Self-guided Automatic Unit Test Generation
by: Yin, Xin, et al.
Published: (2024)
by: Yin, Xin, et al.
Published: (2024)
TTCS: Test-Time Curriculum Synthesis for Self-Evolving
by: Yang, Chengyi, et al.
Published: (2026)
by: Yang, Chengyi, et al.
Published: (2026)
Cogito, ergo sum: A Neurobiologically-Inspired Cognition-Memory-Growth System for Code Generation
by: Li, Yanlong, et al.
Published: (2025)
by: Li, Yanlong, et al.
Published: (2025)
EvoTest: Evolutionary Test-Time Learning for Self-Improving Agentic Systems
by: He, Yufei, et al.
Published: (2025)
by: He, Yufei, et al.
Published: (2025)
Clarifying Semantics of In-Context Examples for Unit Test Generation
by: Yang, Chen, et al.
Published: (2025)
by: Yang, Chen, et al.
Published: (2025)
DRIVE: Dual-Robustness via Information Variability and Entropic Consistency in Source-Free Unsupervised Domain Adaptation
by: Xiao, Ruiqiang, et al.
Published: (2024)
by: Xiao, Ruiqiang, et al.
Published: (2024)
Maintaining Informative Coherence: Migrating Hallucinations in Large Language Models via Absorbing Markov Chains
by: Wu, Jiemin, et al.
Published: (2024)
by: Wu, Jiemin, et al.
Published: (2024)
Synthesizing File-Level Data for Unit Test Generation with Chain-of-Thoughts via Self-Debugging
by: Hua, Ziyue, et al.
Published: (2026)
by: Hua, Ziyue, et al.
Published: (2026)
SRTJ: Self-Evolving Rule-Driven Training-Free LLM Jailbreaking
by: Li, Jindong, et al.
Published: (2026)
by: Li, Jindong, et al.
Published: (2026)
Testing Calibration in Nearly-Linear Time
by: Hu, Lunjia, et al.
Published: (2024)
by: Hu, Lunjia, et al.
Published: (2024)
Self-Trained Verification for Training- and Test-Time Self-Improvement
by: Wu, Chen Henry, et al.
Published: (2026)
by: Wu, Chen Henry, et al.
Published: (2026)
SETS: Leveraging Self-Verification and Self-Correction for Improved Test-Time Scaling
by: Chen, Jiefeng, et al.
Published: (2025)
by: Chen, Jiefeng, et al.
Published: (2025)
Self-Play Preference Optimization for Language Model Alignment
by: Wu, Yue, et al.
Published: (2024)
by: Wu, Yue, et al.
Published: (2024)
TTSR: Test-Time Self-Reflection for Continual Reasoning Improvement
by: He, Haoyang, et al.
Published: (2026)
by: He, Haoyang, et al.
Published: (2026)
Efficient Test-Time Scaling via Self-Calibration
by: Huang, Chengsong, et al.
Published: (2025)
by: Huang, Chengsong, et al.
Published: (2025)
Unify and Triumph: Polyglot, Diverse, and Self-Consistent Generation of Unit Tests with LLMs
by: Khelladi, Djamel Eddine, et al.
Published: (2025)
by: Khelladi, Djamel Eddine, et al.
Published: (2025)
Text2Weight: Bridging Natural Language and Neural Network Weight Spaces
by: Tian, Bowen, et al.
Published: (2025)
by: Tian, Bowen, et al.
Published: (2025)
Beyond Task Vectors: Selective Task Arithmetic Based on Importance Metrics
by: Bowen, Tian, et al.
Published: (2024)
by: Bowen, Tian, et al.
Published: (2024)
On the Evaluation of Large Language Models in Unit Test Generation
by: Yang, Lin, et al.
Published: (2024)
by: Yang, Lin, et al.
Published: (2024)
HyperbolicRAG: Enhancing Retrieval-Augmented Generation with Hyperbolic Representations
by: Cao, Linxiao, et al.
Published: (2025)
by: Cao, Linxiao, et al.
Published: (2025)
Sampling-Efficient Test-Time Scaling: Self-Estimating the Best-of-N Sampling in Early Decoding
by: Wang, Yiming, et al.
Published: (2025)
by: Wang, Yiming, et al.
Published: (2025)
TestART: Improving LLM-based Unit Testing via Co-evolution of Automated Generation and Repair Iteration
by: Gu, Siqi, et al.
Published: (2024)
by: Gu, Siqi, et al.
Published: (2024)
Uncovering Business Logic Bugs via Semantics-Driven Unit Test Generation
by: Yang, Chen, et al.
Published: (2026)
by: Yang, Chen, et al.
Published: (2026)
AL-GNN: Privacy-Preserving and Replay-Free Continual Graph Learning via Analytic Learning
by: Zhang, Xuling, et al.
Published: (2025)
by: Zhang, Xuling, et al.
Published: (2025)
Similar Items
-
TurboEvolve: Towards Fast and Robust LLM-Driven Program Evolution
by: Yang, Yang, et al.
Published: (2026) -
Hypothesis Generation via LLM-Automated Language Bias for ILP
by: Yang, Yang, et al.
Published: (2025) -
RLIE: Rule Generation with Logistic Regression, Iterative Refinement, and Evaluation for Large Language Models
by: Yang, Yang, et al.
Published: (2025) -
Self-Reflective Generation at Test Time
by: Mu, Jian, et al.
Published: (2025) -
IMTS is Worth Time $\times$ Channel Patches: Visual Masked Autoencoders for Irregular Multivariate Time Series Prediction
by: Hu, Zhangyi, et al.
Published: (2025)