Symbol Preference Aware Generative Models for Recovering Variable Names from Stripped Binary
Fuente:
arXiv
Saved in:
| Main Authors: | Xu, Xiangzhe, Zhang, Zhuo, Su, Zian, Huang, Ziyang, Feng, Shiwei, Ye, Yapeng, Jiang, Nan, Xie, Danning, Cheng, Siyuan, Tan, Lin, Zhang, Xiangyu |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CodeArt: Better Code Models by Attention Regularization When Symbols Are Lacking
by: Su, Zian, et al.
Published: (2024)
by: Su, Zian, et al.
Published: (2024)
Source Code Foundation Models are Transferable Binary Analysis Knowledge Bases
by: Su, Zian, et al.
Published: (2024)
by: Su, Zian, et al.
Published: (2024)
Position: Intelligent Coding Systems Should Write Programs with Justifications
by: Xu, Xiangzhe, et al.
Published: (2025)
by: Xu, Xiangzhe, et al.
Published: (2025)
ROCAS: Root Cause Analysis of Autonomous Driving Accidents via Cyber-Physical Co-mutation
by: Feng, Shiwei, et al.
Published: (2024)
by: Feng, Shiwei, et al.
Published: (2024)
TAI3: Testing Agent Integrity in Interpreting User Intent
by: Feng, Shiwei, et al.
Published: (2025)
by: Feng, Shiwei, et al.
Published: (2025)
ProSec: Fortifying Code LLMs with Proactive Security Alignment
by: Xu, Xiangzhe, et al.
Published: (2024)
by: Xu, Xiangzhe, et al.
Published: (2024)
Large Language Models for Validating Network Protocol Parsers
by: Zheng, Mingwei, et al.
Published: (2025)
by: Zheng, Mingwei, et al.
Published: (2025)
How Effective are Large Language Models in Generating Software Specifications?
by: Xie, Danning, et al.
Published: (2023)
by: Xie, Danning, et al.
Published: (2023)
LLMDFA: Analyzing Dataflow in Code with Large Language Models
by: Wang, Chengpeng, et al.
Published: (2024)
by: Wang, Chengpeng, et al.
Published: (2024)
RepoAudit: An Autonomous LLM-Agent for Repository-Level Code Auditing
by: Guo, Jinyao, et al.
Published: (2025)
by: Guo, Jinyao, et al.
Published: (2025)
ASTRA: Autonomous Spatial-Temporal Red-teaming for AI Software Assistants
by: Xu, Xiangzhe, et al.
Published: (2025)
by: Xu, Xiangzhe, et al.
Published: (2025)
Cross-modal Retrieval Models for Stripped Binary Analysis
by: Chen, Guoqiang, et al.
Published: (2025)
by: Chen, Guoqiang, et al.
Published: (2025)
Extracting Protocol Format as State Machine via Controlled Static Loop Analysis
by: Shi, Qingkai, et al.
Published: (2023)
by: Shi, Qingkai, et al.
Published: (2023)
Identifying Adversary Tactics and Techniques in Malware Binaries with an LLM Agent
by: Xuan, Zhou, et al.
Published: (2026)
by: Xuan, Zhou, et al.
Published: (2026)
Nova: Generative Language Models for Assembly Code with Hierarchical Attention and Contrastive Learning
by: Jiang, Nan, et al.
Published: (2023)
by: Jiang, Nan, et al.
Published: (2023)
Validating Network Protocol Parsers with Traceable RFC Document Interpretation
by: Zheng, Mingwei, et al.
Published: (2025)
by: Zheng, Mingwei, et al.
Published: (2025)
Beyond C/C++: Probabilistic and LLM Methods for Next-Generation Software Reverse Engineering
by: Zhuo, Zhuo, et al.
Published: (2025)
by: Zhuo, Zhuo, et al.
Published: (2025)
Can LLMs Recover Program Semantics? A Systematic Evaluation with Symbolic Execution
by: Feng, Rong, et al.
Published: (2025)
by: Feng, Rong, et al.
Published: (2025)
CoRe: Benchmarking LLMs Code Reasoning Capabilities through Static Analysis Tasks
by: Xie, Danning, et al.
Published: (2025)
by: Xie, Danning, et al.
Published: (2025)
REBENCH: A Procedural, Fair-by-Construction Benchmark for LLMs on Stripped-Binary Types and Names (Extended Version)
by: Won, Jun Yeon, et al.
Published: (2026)
by: Won, Jun Yeon, et al.
Published: (2026)
DocTer: Documentation Guided Fuzzing for Testing Deep Learning API Functions
by: Xie, Danning, et al.
Published: (2021)
by: Xie, Danning, et al.
Published: (2021)
Recommending Variable Names for Extract Local Variable Refactorings
by: Wang, Taiming, et al.
Published: (2025)
by: Wang, Taiming, et al.
Published: (2025)
RFCAudit: An LLM Agent for Functional Bug Detection in Network Protocols
by: Zheng, Mingwei, et al.
Published: (2025)
by: Zheng, Mingwei, et al.
Published: (2025)
SpecOps: A Fully Automated AI Agent Testing Framework in Real-World GUI Environments
by: Ahmed, Syed Yusuf, et al.
Published: (2026)
by: Ahmed, Syed Yusuf, et al.
Published: (2026)
NESA: Relational Neuro-Symbolic Static Program Analysis
by: Wang, Chengpeng, et al.
Published: (2024)
by: Wang, Chengpeng, et al.
Published: (2024)
Adaptive Proof Refinement with LLM-Guided Strategy Selection
by: Lu, Minghai, et al.
Published: (2025)
by: Lu, Minghai, et al.
Published: (2025)
Neural Variable Name Repair: Learning to Rename Identifiers for Readability
by: Yousuf, Muhammad, et al.
Published: (2025)
by: Yousuf, Muhammad, et al.
Published: (2025)
BugScope: Learn to Find Bugs Like Human
by: Guo, Jinyao, et al.
Published: (2025)
by: Guo, Jinyao, et al.
Published: (2025)
Enhancing Function Name Prediction using Votes-Based Name Tokenization and Multi-Task Learning
by: Zhang, Xiaoling, et al.
Published: (2024)
by: Zhang, Xiaoling, et al.
Published: (2024)
KEENHash: Hashing Programs into Function-Aware Embeddings for Large-Scale Binary Code Similarity Analysis
by: Liu, Zhijie, et al.
Published: (2025)
by: Liu, Zhijie, et al.
Published: (2025)
Decompile-Bench: Million-Scale Binary-Source Function Pairs for Real-World Binary Decompilation
by: Tan, Hanzhuo, et al.
Published: (2025)
by: Tan, Hanzhuo, et al.
Published: (2025)
Neuro-Symbolic Generation and Validation of Memory-Aware Formal Function Specifications
by: Zhang, Liao, et al.
Published: (2026)
by: Zhang, Liao, et al.
Published: (2026)
Static Code Analyzer Recommendation via Preference Mining
by: Ge, Xiuting, et al.
Published: (2024)
by: Ge, Xiuting, et al.
Published: (2024)
AOCI: Symbolic-Semantic Indexing for Practical Repository-Scale Code Understanding with LLMs
by: Liu, Jinshi, et al.
Published: (2026)
by: Liu, Jinshi, et al.
Published: (2026)
LLM as an Execution Estimator: Recovering Missing Dependency for Practical Time-travelling Debugging
by: Pei, Yunrui, et al.
Published: (2025)
by: Pei, Yunrui, et al.
Published: (2025)
From Poisoned to Aware: Fostering Backdoor Self-Awareness in LLMs
by: Shen, Guangyu, et al.
Published: (2025)
by: Shen, Guangyu, et al.
Published: (2025)
An Empirical Study of False Negatives and Positives of Static Code Analyzers From the Perspective of Historical Issues
by: Cui, Han, et al.
Published: (2024)
by: Cui, Han, et al.
Published: (2024)
Chain-of-Thought in Neural Code Generation: From and For Lightweight Language Models
by: Yang, Guang, et al.
Published: (2023)
by: Yang, Guang, et al.
Published: (2023)
Context-Aware Functional Test Generation via Business Logic Extraction and Adaptation
by: Zhang, Yakun, et al.
Published: (2026)
by: Zhang, Yakun, et al.
Published: (2026)
BinaryAI: Binary Software Composition Analysis via Intelligent Binary Source Code Matching
by: Jiang, Ling, et al.
Published: (2024)
by: Jiang, Ling, et al.
Published: (2024)
Similar Items
-
CodeArt: Better Code Models by Attention Regularization When Symbols Are Lacking
by: Su, Zian, et al.
Published: (2024) -
Source Code Foundation Models are Transferable Binary Analysis Knowledge Bases
by: Su, Zian, et al.
Published: (2024) -
Position: Intelligent Coding Systems Should Write Programs with Justifications
by: Xu, Xiangzhe, et al.
Published: (2025) -
ROCAS: Root Cause Analysis of Autonomous Driving Accidents via Cyber-Physical Co-mutation
by: Feng, Shiwei, et al.
Published: (2024) -
TAI3: Testing Agent Integrity in Interpreting User Intent
by: Feng, Shiwei, et al.
Published: (2025)