VeriScale: Adversarial Test-Suite Scaling for Verifiable Code Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Bai, Yifan, Liu, Xiaoyang, Mou, Zihao, Wang, Guihong, Yu, Jian, Xie, Shuhan, Li, Yantao, Zhang, Yangyu, Liang, Jingwei, Luo, Tao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
ATLAS: Automated Toolkit for Large-Scale Verified Code Synthesis
von: Baksys, Mantas, et al.
Veröffentlicht: (2025)
von: Baksys, Mantas, et al.
Veröffentlicht: (2025)
Efficient Incremental Code Coverage Analysis for Regression Test Suites
von: Wang, Jiale Amber, et al.
Veröffentlicht: (2024)
von: Wang, Jiale Amber, et al.
Veröffentlicht: (2024)
Scaling Test-Driven Code Generation from Functions to Classes: An Empirical Study
von: Liang, Yunhao, et al.
Veröffentlicht: (2026)
von: Liang, Yunhao, et al.
Veröffentlicht: (2026)
VeriContest: A Competitive-Programming Benchmark for Verifiable Code Generation
von: Xie, Zichen, et al.
Veröffentlicht: (2026)
von: Xie, Zichen, et al.
Veröffentlicht: (2026)
Dynamic Scaling of Unit Tests for Code Reward Modeling
von: Ma, Zeyao, et al.
Veröffentlicht: (2025)
von: Ma, Zeyao, et al.
Veröffentlicht: (2025)
Scaling Laws Behind Code Understanding Model
von: Lin, Jiayi, et al.
Veröffentlicht: (2024)
von: Lin, Jiayi, et al.
Veröffentlicht: (2024)
VeriFix: Verifying Your Fix Towards An Atomicity Violation
von: Li, Zhuang, et al.
Veröffentlicht: (2025)
von: Li, Zhuang, et al.
Veröffentlicht: (2025)
On the Effectiveness of Modular Testing in EvoSuite
von: Dinella, Elizabeth
Veröffentlicht: (2026)
von: Dinella, Elizabeth
Veröffentlicht: (2026)
Reusable Test Suites for Reinforcement Learning
von: Betten, Jørn Eirik, et al.
Veröffentlicht: (2025)
von: Betten, Jørn Eirik, et al.
Veröffentlicht: (2025)
TestForge: Feedback-Driven, Agentic Test Suite Generation
von: Jain, Kush, et al.
Veröffentlicht: (2025)
von: Jain, Kush, et al.
Veröffentlicht: (2025)
AlgoVeri: An Aligned Benchmark for Verified Code Generation on Classical Algorithms
von: Zhao, Haoyu, et al.
Veröffentlicht: (2026)
von: Zhao, Haoyu, et al.
Veröffentlicht: (2026)
CodeChemist: Functional Knowledge Transfer for Low-Resource Code Generation via Test-Time Scaling
von: Wang, Kaixin, et al.
Veröffentlicht: (2025)
von: Wang, Kaixin, et al.
Veröffentlicht: (2025)
Empirical Derivations from an Evolving Test Suite
von: Ruohonen, Jukka, et al.
Veröffentlicht: (2025)
von: Ruohonen, Jukka, et al.
Veröffentlicht: (2025)
Learning to Solve and Verify: A Self-Play Framework for Code and Test Generation
von: Lin, Zi, et al.
Veröffentlicht: (2025)
von: Lin, Zi, et al.
Veröffentlicht: (2025)
Automatically Removing Unnecessary Stubbings from Test Suites
von: Li, Mengzhen, et al.
Veröffentlicht: (2024)
von: Li, Mengzhen, et al.
Veröffentlicht: (2024)
Interval Analysis in Industrial-Scale BMC Software Verifiers: A Case Study
von: Menezes, Rafael Sá, et al.
Veröffentlicht: (2024)
von: Menezes, Rafael Sá, et al.
Veröffentlicht: (2024)
FormalRTL: Verified RTL Synthesis at Scale
von: Li, Kezhi, et al.
Veröffentlicht: (2026)
von: Li, Kezhi, et al.
Veröffentlicht: (2026)
Scalable Similarity-Aware Test Suite Minimization with Reinforcement Learning
von: Gu, Sijia, et al.
Veröffentlicht: (2024)
von: Gu, Sijia, et al.
Veröffentlicht: (2024)
Ever-Improving Test Suite by Leveraging Large Language Models
von: Qiu, Ketai
Veröffentlicht: (2025)
von: Qiu, Ketai
Veröffentlicht: (2025)
Regression Test Suite for Payment Switch using jPOS
von: Sardesai, Atharv, et al.
Veröffentlicht: (2022)
von: Sardesai, Atharv, et al.
Veröffentlicht: (2022)
AOCI: Symbolic-Semantic Indexing for Practical Repository-Scale Code Understanding with LLMs
von: Liu, Jinshi, et al.
Veröffentlicht: (2026)
von: Liu, Jinshi, et al.
Veröffentlicht: (2026)
Adversarial Attack Classification and Robustness Testing for Large Language Models for Code
von: Liu, Yang, et al.
Veröffentlicht: (2025)
von: Liu, Yang, et al.
Veröffentlicht: (2025)
TestDecision: Sequential Test Suite Generation via Greedy Optimization and Reinforcement Learning
von: Wang, Guoqing, et al.
Veröffentlicht: (2026)
von: Wang, Guoqing, et al.
Veröffentlicht: (2026)
Evaluating Small-Scale Code Models for Code Clone Detection
von: Martinez-Gil, Jorge
Veröffentlicht: (2025)
von: Martinez-Gil, Jorge
Veröffentlicht: (2025)
Scaling Mobile Chaos Testing with AI-Driven Test Execution
von: Marcano, Juan, et al.
Veröffentlicht: (2026)
von: Marcano, Juan, et al.
Veröffentlicht: (2026)
Temporal Modeling of Change History for Black-Box Test Suite Minimization
von: Asif, Kamruzzaman, et al.
Veröffentlicht: (2026)
von: Asif, Kamruzzaman, et al.
Veröffentlicht: (2026)
A Test Suite for Efficient Robustness Evaluation of Face Recognition Systems
von: Zhang, Ruihan, et al.
Veröffentlicht: (2025)
von: Zhang, Ruihan, et al.
Veröffentlicht: (2025)
CodeDPO: Aligning Code Models with Self Generated and Verified Source Code
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
Code Generation by Differential Test Time Scaling
von: He, Yifeng, et al.
Veröffentlicht: (2026)
von: He, Yifeng, et al.
Veröffentlicht: (2026)
AutoICE: Automatically Synthesizing Verifiable C Code via LLM-driven Evolution
von: Luo, Weilin, et al.
Veröffentlicht: (2025)
von: Luo, Weilin, et al.
Veröffentlicht: (2025)
ExecVerify: White-Box RL with Verifiable Stepwise Rewards for Code Execution Reasoning
von: Tang, Lingxiao, et al.
Veröffentlicht: (2026)
von: Tang, Lingxiao, et al.
Veröffentlicht: (2026)
SWE-TRACE: Optimizing Long-Horizon SWE Agents Through Rubric Process Reward Models and Heuristic Test-Time Scaling
von: Han, Hao, et al.
Veröffentlicht: (2026)
von: Han, Hao, et al.
Veröffentlicht: (2026)
Are Benchmark Tests Strong Enough? Mutation-Guided Diagnosis and Augmentation of Regression Suites
von: Li, Chenglin, et al.
Veröffentlicht: (2026)
von: Li, Chenglin, et al.
Veröffentlicht: (2026)
Improving Spectrum-Based Localization of Multiple Faults by Iterative Test Suite Reduction
von: Callaghan, Dylan, et al.
Veröffentlicht: (2023)
von: Callaghan, Dylan, et al.
Veröffentlicht: (2023)
Governing the Commons: Code Ownership and Code-Clones in Large-Scale Software Development
von: Sundelin, Anders, et al.
Veröffentlicht: (2024)
von: Sundelin, Anders, et al.
Veröffentlicht: (2024)
Write Your Own CodeChecker: An Automated Test-Driven Checker Development Approach with LLMs
von: Liu, Jun, et al.
Veröffentlicht: (2024)
von: Liu, Jun, et al.
Veröffentlicht: (2024)
Scaling Coding Agents via Atomic Skills
von: Ma, Yingwei, et al.
Veröffentlicht: (2026)
von: Ma, Yingwei, et al.
Veröffentlicht: (2026)
SWE-Universe: Scale Real-World Verifiable Environments to Millions
von: Chen, Mouxiang, et al.
Veröffentlicht: (2026)
von: Chen, Mouxiang, et al.
Veröffentlicht: (2026)
VeriSBOM: Secure and Verifiable SBOM Sharing Via Zero-Knowledge Proofs
von: Castiglione, Gianpietro, et al.
Veröffentlicht: (2026)
von: Castiglione, Gianpietro, et al.
Veröffentlicht: (2026)
VeriAct: Beyond Verifiability -- Agentic Synthesis of Correct and Complete Formal Specifications
von: Misu, Md Rakib Hossain, et al.
Veröffentlicht: (2026)
von: Misu, Md Rakib Hossain, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
ATLAS: Automated Toolkit for Large-Scale Verified Code Synthesis
von: Baksys, Mantas, et al.
Veröffentlicht: (2025) -
Efficient Incremental Code Coverage Analysis for Regression Test Suites
von: Wang, Jiale Amber, et al.
Veröffentlicht: (2024) -
Scaling Test-Driven Code Generation from Functions to Classes: An Empirical Study
von: Liang, Yunhao, et al.
Veröffentlicht: (2026) -
VeriContest: A Competitive-Programming Benchmark for Verifiable Code Generation
von: Xie, Zichen, et al.
Veröffentlicht: (2026) -
Dynamic Scaling of Unit Tests for Code Reward Modeling
von: Ma, Zeyao, et al.
Veröffentlicht: (2025)