CoSPlay: Cooperative Self-Play at Test-Time with Self-Generated Code and Unit Test
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Hu, Zhangyi, Liu, Chenhui, Huang, Tian, Li, Jindong, Yang, Yang, Wu, Jiemin, Zhong, Zining, Yang, Menglin, Yue, Yutao |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
TurboEvolve: Towards Fast and Robust LLM-Driven Program Evolution
par: Yang, Yang, et autres
Publié: (2026)
par: Yang, Yang, et autres
Publié: (2026)
Hypothesis Generation via LLM-Automated Language Bias for ILP
par: Yang, Yang, et autres
Publié: (2025)
par: Yang, Yang, et autres
Publié: (2025)
RLIE: Rule Generation with Logistic Regression, Iterative Refinement, and Evaluation for Large Language Models
par: Yang, Yang, et autres
Publié: (2025)
par: Yang, Yang, et autres
Publié: (2025)
Self-Reflective Generation at Test Time
par: Mu, Jian, et autres
Publié: (2025)
par: Mu, Jian, et autres
Publié: (2025)
IMTS is Worth Time $\times$ Channel Patches: Visual Masked Autoencoders for Irregular Multivariate Time Series Prediction
par: Hu, Zhangyi, et autres
Publié: (2025)
par: Hu, Zhangyi, et autres
Publié: (2025)
Self-Harmony: Learning to Harmonize Self-Supervision and Self-Play in Test-Time Reinforcement Learning
par: Wang, Ru, et autres
Publié: (2025)
par: Wang, Ru, et autres
Publié: (2025)
Learning to Solve and Verify: A Self-Play Framework for Code and Test Generation
par: Lin, Zi, et autres
Publié: (2025)
par: Lin, Zi, et autres
Publié: (2025)
CasModaTest: A Cascaded and Model-agnostic Self-directed Framework for Unit Test Generation
par: Ni, Chao, et autres
Publié: (2024)
par: Ni, Chao, et autres
Publié: (2024)
ACE: Self-Evolving LLM Coding Framework via Adversarial Unit Test Generation and Preference Optimization
par: Huang, Yixu, et autres
Publié: (2026)
par: Huang, Yixu, et autres
Publié: (2026)
Revisit Self-Debugging with Self-Generated Tests for Code Generation
par: Chen, Xiancai, et autres
Publié: (2025)
par: Chen, Xiancai, et autres
Publié: (2025)
UTMath: Math Evaluation with Unit Test via Reasoning-to-Coding Thoughts
par: Yang, Bo, et autres
Publié: (2024)
par: Yang, Bo, et autres
Publié: (2024)
Superhuman AI for Stratego Using Self-Play Reinforcement Learning and Test-Time Search
par: Sokota, Samuel, et autres
Publié: (2025)
par: Sokota, Samuel, et autres
Publié: (2025)
Automated Unit Test Refactoring
par: Gao, Yi, et autres
Publié: (2024)
par: Gao, Yi, et autres
Publié: (2024)
PETS: A Principled Framework Towards Optimal Trajectory Allocation for Efficient Test-Time Self-Consistency
par: Liu, Zhangyi, et autres
Publié: (2026)
par: Liu, Zhangyi, et autres
Publié: (2026)
DSTC: Direct Preference Learning with Only Self-Generated Tests and Code to Improve Code LMs
par: Liu, Zhihan, et autres
Publié: (2024)
par: Liu, Zhihan, et autres
Publié: (2024)
Requirements-Based Test Generation: A Comprehensive Survey
par: Yang, Zhenzhen, et autres
Publié: (2025)
par: Yang, Zhenzhen, et autres
Publié: (2025)
What You See Is What You Get: Attention-based Self-guided Automatic Unit Test Generation
par: Yin, Xin, et autres
Publié: (2024)
par: Yin, Xin, et autres
Publié: (2024)
TTCS: Test-Time Curriculum Synthesis for Self-Evolving
par: Yang, Chengyi, et autres
Publié: (2026)
par: Yang, Chengyi, et autres
Publié: (2026)
Cogito, ergo sum: A Neurobiologically-Inspired Cognition-Memory-Growth System for Code Generation
par: Li, Yanlong, et autres
Publié: (2025)
par: Li, Yanlong, et autres
Publié: (2025)
EvoTest: Evolutionary Test-Time Learning for Self-Improving Agentic Systems
par: He, Yufei, et autres
Publié: (2025)
par: He, Yufei, et autres
Publié: (2025)
Clarifying Semantics of In-Context Examples for Unit Test Generation
par: Yang, Chen, et autres
Publié: (2025)
par: Yang, Chen, et autres
Publié: (2025)
DRIVE: Dual-Robustness via Information Variability and Entropic Consistency in Source-Free Unsupervised Domain Adaptation
par: Xiao, Ruiqiang, et autres
Publié: (2024)
par: Xiao, Ruiqiang, et autres
Publié: (2024)
Maintaining Informative Coherence: Migrating Hallucinations in Large Language Models via Absorbing Markov Chains
par: Wu, Jiemin, et autres
Publié: (2024)
par: Wu, Jiemin, et autres
Publié: (2024)
Synthesizing File-Level Data for Unit Test Generation with Chain-of-Thoughts via Self-Debugging
par: Hua, Ziyue, et autres
Publié: (2026)
par: Hua, Ziyue, et autres
Publié: (2026)
SRTJ: Self-Evolving Rule-Driven Training-Free LLM Jailbreaking
par: Li, Jindong, et autres
Publié: (2026)
par: Li, Jindong, et autres
Publié: (2026)
Testing Calibration in Nearly-Linear Time
par: Hu, Lunjia, et autres
Publié: (2024)
par: Hu, Lunjia, et autres
Publié: (2024)
Self-Trained Verification for Training- and Test-Time Self-Improvement
par: Wu, Chen Henry, et autres
Publié: (2026)
par: Wu, Chen Henry, et autres
Publié: (2026)
SETS: Leveraging Self-Verification and Self-Correction for Improved Test-Time Scaling
par: Chen, Jiefeng, et autres
Publié: (2025)
par: Chen, Jiefeng, et autres
Publié: (2025)
Self-Play Preference Optimization for Language Model Alignment
par: Wu, Yue, et autres
Publié: (2024)
par: Wu, Yue, et autres
Publié: (2024)
TTSR: Test-Time Self-Reflection for Continual Reasoning Improvement
par: He, Haoyang, et autres
Publié: (2026)
par: He, Haoyang, et autres
Publié: (2026)
Efficient Test-Time Scaling via Self-Calibration
par: Huang, Chengsong, et autres
Publié: (2025)
par: Huang, Chengsong, et autres
Publié: (2025)
Unify and Triumph: Polyglot, Diverse, and Self-Consistent Generation of Unit Tests with LLMs
par: Khelladi, Djamel Eddine, et autres
Publié: (2025)
par: Khelladi, Djamel Eddine, et autres
Publié: (2025)
Text2Weight: Bridging Natural Language and Neural Network Weight Spaces
par: Tian, Bowen, et autres
Publié: (2025)
par: Tian, Bowen, et autres
Publié: (2025)
Beyond Task Vectors: Selective Task Arithmetic Based on Importance Metrics
par: Bowen, Tian, et autres
Publié: (2024)
par: Bowen, Tian, et autres
Publié: (2024)
On the Evaluation of Large Language Models in Unit Test Generation
par: Yang, Lin, et autres
Publié: (2024)
par: Yang, Lin, et autres
Publié: (2024)
HyperbolicRAG: Enhancing Retrieval-Augmented Generation with Hyperbolic Representations
par: Cao, Linxiao, et autres
Publié: (2025)
par: Cao, Linxiao, et autres
Publié: (2025)
Sampling-Efficient Test-Time Scaling: Self-Estimating the Best-of-N Sampling in Early Decoding
par: Wang, Yiming, et autres
Publié: (2025)
par: Wang, Yiming, et autres
Publié: (2025)
TestART: Improving LLM-based Unit Testing via Co-evolution of Automated Generation and Repair Iteration
par: Gu, Siqi, et autres
Publié: (2024)
par: Gu, Siqi, et autres
Publié: (2024)
Uncovering Business Logic Bugs via Semantics-Driven Unit Test Generation
par: Yang, Chen, et autres
Publié: (2026)
par: Yang, Chen, et autres
Publié: (2026)
AL-GNN: Privacy-Preserving and Replay-Free Continual Graph Learning via Analytic Learning
par: Zhang, Xuling, et autres
Publié: (2025)
par: Zhang, Xuling, et autres
Publié: (2025)
Documents similaires
-
TurboEvolve: Towards Fast and Robust LLM-Driven Program Evolution
par: Yang, Yang, et autres
Publié: (2026) -
Hypothesis Generation via LLM-Automated Language Bias for ILP
par: Yang, Yang, et autres
Publié: (2025) -
RLIE: Rule Generation with Logistic Regression, Iterative Refinement, and Evaluation for Large Language Models
par: Yang, Yang, et autres
Publié: (2025) -
Self-Reflective Generation at Test Time
par: Mu, Jian, et autres
Publié: (2025) -
IMTS is Worth Time $\times$ Channel Patches: Visual Masked Autoencoders for Irregular Multivariate Time Series Prediction
par: Hu, Zhangyi, et autres
Publié: (2025)