Verify Implementation Equivalence of Large Models
Fuente:
arXiv
Saved in:
| Main Authors: | Zhan, Qi, Hu, Xing, Xia, Xin, Li, Shanping |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PS$^3$: Precise Patch Presence Test based on Semantic Symbolic Signature
by: Zhan, Qi, et al.
Published: (2023)
by: Zhan, Qi, et al.
Published: (2023)
Mitigating Implicit Inconsistencies in Patch Porting
by: Pan, Shengyi, et al.
Published: (2026)
by: Pan, Shengyi, et al.
Published: (2026)
Automating Zero-Shot Patch Porting for Hard Forks
by: Pan, Shengyi, et al.
Published: (2024)
by: Pan, Shengyi, et al.
Published: (2024)
SelfPiCo: Self-Guided Partial Code Execution with LLMs
by: Xue, Zhipeng, et al.
Published: (2024)
by: Xue, Zhipeng, et al.
Published: (2024)
Actionable Warning Is Not Enough: Recommending Valid Actionable Warnings with Weak Supervision
by: Xue, Zhipeng, et al.
Published: (2025)
by: Xue, Zhipeng, et al.
Published: (2025)
Clean Code, Better Models: Enhancing LLM Performance with Smell-Cleaned Dataset
by: Xue, Zhipeng, et al.
Published: (2025)
by: Xue, Zhipeng, et al.
Published: (2025)
Less is More: On the Importance of Data Quality for Unit Test Generation
by: Zhang, Junwei, et al.
Published: (2025)
by: Zhang, Junwei, et al.
Published: (2025)
FGIT: Fault-Guided Fine-Tuning for Code Generation
by: Fan, Lishui, et al.
Published: (2025)
by: Fan, Lishui, et al.
Published: (2025)
Exploring the Capabilities of LLMs for Code Change Related Tasks
by: Fan, Lishui, et al.
Published: (2024)
by: Fan, Lishui, et al.
Published: (2024)
A Large-scale Empirical Study on the Generalizability of Disclosed Java Library Vulnerability Exploits
by: Chen, Zirui, et al.
Published: (2026)
by: Chen, Zirui, et al.
Published: (2026)
Towards Explainable Vulnerability Detection with Large Language Models
by: Mao, Qiheng, et al.
Published: (2024)
by: Mao, Qiheng, et al.
Published: (2024)
The Current Challenges of Software Engineering in the Era of Large Language Models
by: Gao, Cuiyun, et al.
Published: (2024)
by: Gao, Cuiyun, et al.
Published: (2024)
ZeroCoder: Can LLMs Improve Code Generation Without Ground-Truth Supervision?
by: Fan, Lishui, et al.
Published: (2026)
by: Fan, Lishui, et al.
Published: (2026)
Where Is Self-admitted Code Generated by Large Language Models on GitHub?
by: Yu, Xiao, et al.
Published: (2024)
by: Yu, Xiao, et al.
Published: (2024)
VERT: Verified Equivalent Rust Transpilation with Large Language Models as Few-Shot Learners
by: Yang, Aidan Z. H., et al.
Published: (2024)
by: Yang, Aidan Z. H., et al.
Published: (2024)
Towards Understanding Bugs in Distributed Training and Inference Frameworks for Large Language Models
by: Yu, Xiao, et al.
Published: (2025)
by: Yu, Xiao, et al.
Published: (2025)
HFuzzer: Testing Large Language Models for Package Hallucinations via Phrase-based Fuzzing
by: Zhao, Yukai, et al.
Published: (2025)
by: Zhao, Yukai, et al.
Published: (2025)
Assessing and Advancing Benchmarks for Evaluating Large Language Models in Software Engineering Tasks
by: Hu, Xing, et al.
Published: (2025)
by: Hu, Xing, et al.
Published: (2025)
An Empirical Study of Speculative Decoding on Software Engineering Tasks
by: Li, Yijia, et al.
Published: (2026)
by: Li, Yijia, et al.
Published: (2026)
KBX: Verified Model Synchronization via Formal Bidirectional Transformation
by: Zhao, Jianhong, et al.
Published: (2024)
by: Zhao, Jianhong, et al.
Published: (2024)
Every Maintenance Has Its Exemplar: The Future of Software Maintenance through Migration
by: Chen, Zirui, et al.
Published: (2026)
by: Chen, Zirui, et al.
Published: (2026)
Automating TODO-missed Methods Detection and Patching
by: Gao, Zhipeng, et al.
Published: (2024)
by: Gao, Zhipeng, et al.
Published: (2024)
Automated Unit Test Refactoring
by: Gao, Yi, et al.
Published: (2024)
by: Gao, Yi, et al.
Published: (2024)
ActRef: Enhancing the Understanding of Python Code Refactoring with Action-Based Analysis
by: Wang, Siqi, et al.
Published: (2025)
by: Wang, Siqi, et al.
Published: (2025)
Learning in the Wild: Towards Leveraging Unlabeled Data for Effectively Tuning Pre-trained Code Models
by: Gao, Shuzheng, et al.
Published: (2024)
by: Gao, Shuzheng, et al.
Published: (2024)
Large Language Models for Equivalent Mutant Detection: How Far Are We?
by: Tian, Zhao, et al.
Published: (2024)
by: Tian, Zhao, et al.
Published: (2024)
CREME: Robustness Enhancement of Code LLMs via Layer-Aware Model Editing
by: Liu, Shuhan, et al.
Published: (2025)
by: Liu, Shuhan, et al.
Published: (2025)
Lightweight Model Editing for LLMs to Correct Deprecated API Recommendations
by: Lin, Guancheng, et al.
Published: (2025)
by: Lin, Guancheng, et al.
Published: (2025)
PPT4J: Patch Presence Test for Java Binaries
by: Pan, Zhiyuan, et al.
Published: (2023)
by: Pan, Zhiyuan, et al.
Published: (2023)
Assessing Large Language Models in Comprehending and Verifying Concurrent Programs across Memory Models
by: Jain, Ridhi, et al.
Published: (2025)
by: Jain, Ridhi, et al.
Published: (2025)
Understanding Practitioners' Expectations on Clear Code Review Comments
by: Chen, Junkai, et al.
Published: (2024)
by: Chen, Junkai, et al.
Published: (2024)
Easy over Hard: A Simple Baseline for Test Failures Causes Prediction
by: Gao, Zhipeng, et al.
Published: (2024)
by: Gao, Zhipeng, et al.
Published: (2024)
Generating Mitigations for Downstream Projects to Neutralize Upstream Library Vulnerability
by: Chen, Zirui, et al.
Published: (2025)
by: Chen, Zirui, et al.
Published: (2025)
Ensemble Fuzzing with Dynamic Resource Scheduling and Multidimensional Seed Evaluation
by: Zhao, Yukai, et al.
Published: (2025)
by: Zhao, Yukai, et al.
Published: (2025)
A Rule-Based Approach for UI Migration from Android to iOS
by: Gao, Yi, et al.
Published: (2024)
by: Gao, Yi, et al.
Published: (2024)
NLPerturbator: Studying the Robustness of Code LLMs to Natural Language Variations
by: Chen, Junkai, et al.
Published: (2024)
by: Chen, Junkai, et al.
Published: (2024)
Possible Value Analysis based on Symbolic Lattice
by: Zhan, Qi
Published: (2024)
by: Zhan, Qi
Published: (2024)
VeCoGen: Automating Generation of Formally Verified C Code with Large Language Models
by: Sevenhuijsen, Merlijn, et al.
Published: (2024)
by: Sevenhuijsen, Merlijn, et al.
Published: (2024)
ATLAS: Automated Toolkit for Large-Scale Verified Code Synthesis
by: Baksys, Mantas, et al.
Published: (2025)
by: Baksys, Mantas, et al.
Published: (2025)
An Empirical Study of Retrieval-Augmented Code Generation: Challenges and Opportunities
by: Yang, Zezhou, et al.
Published: (2025)
by: Yang, Zezhou, et al.
Published: (2025)
Similar Items
-
PS$^3$: Precise Patch Presence Test based on Semantic Symbolic Signature
by: Zhan, Qi, et al.
Published: (2023) -
Mitigating Implicit Inconsistencies in Patch Porting
by: Pan, Shengyi, et al.
Published: (2026) -
Automating Zero-Shot Patch Porting for Hard Forks
by: Pan, Shengyi, et al.
Published: (2024) -
SelfPiCo: Self-Guided Partial Code Execution with LLMs
by: Xue, Zhipeng, et al.
Published: (2024) -
Actionable Warning Is Not Enough: Recommending Valid Actionable Warnings with Weak Supervision
by: Xue, Zhipeng, et al.
Published: (2025)