A Large Scale Study of AI-based Binary Function Similarity Detection Techniques for Security Researchers and Practitioners
Fuente:
arXiv
Saved in:
| Main Authors: | Shi, Jingyi, Chen, Yufeng, Xiao, Yang, Li, Yuekang, Xu, Zhengzi, Qiu, Sihao, Zhang, Chi, Qi, Keyu, Li, Yeting, Chen, Xingchu, Zou, Yanyan, Liu, Yang, Huo, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Vulnerability-Affected Versions Identification: How Far Are We?
by: Chen, Xingchu, et al.
Published: (2025)
by: Chen, Xingchu, et al.
Published: (2025)
Enhancing Function Name Prediction using Votes-Based Name Tokenization and Multi-Task Learning
by: Zhang, Xiaoling, et al.
Published: (2024)
by: Zhang, Xiaoling, et al.
Published: (2024)
Decompile-Bench: Million-Scale Binary-Source Function Pairs for Real-World Binary Decompilation
by: Tan, Hanzhuo, et al.
Published: (2025)
by: Tan, Hanzhuo, et al.
Published: (2025)
Practitioners' Expectations on Log Anomaly Detection
by: Ma, Xiaoxue, et al.
Published: (2024)
by: Ma, Xiaoxue, et al.
Published: (2024)
ORCAS: Obfuscation-Resilient Binary Code Similarity Analysis using Dominance Enhanced Semantic Graph
by: Wang, Yufeng, et al.
Published: (2025)
by: Wang, Yufeng, et al.
Published: (2025)
Jailbreaking ChatGPT via Prompt Engineering: An Empirical Study
by: Liu, Yi, et al.
Published: (2023)
by: Liu, Yi, et al.
Published: (2023)
Cross-Inlining Binary Function Similarity Detection
by: Jia, Ang, et al.
Published: (2024)
by: Jia, Ang, et al.
Published: (2024)
CEBin: A Cost-Effective Framework for Large-Scale Binary Code Similarity Detection
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
KEENHash: Hashing Programs into Function-Aware Embeddings for Large-Scale Binary Code Similarity Analysis
by: Liu, Zhijie, et al.
Published: (2025)
by: Liu, Zhijie, et al.
Published: (2025)
Binary Code Similarity Detection via Graph Contrastive Learning on Intermediate Representations
by: Shang, Xiuwei, et al.
Published: (2024)
by: Shang, Xiuwei, et al.
Published: (2024)
Challenges of Using Pre-trained Models: the Practitioners' Perspective
by: Tan, Xin, et al.
Published: (2024)
by: Tan, Xin, et al.
Published: (2024)
A Large-Scale Evaluation for Log Parsing Techniques: How Far Are We?
by: Jiang, Zhihan, et al.
Published: (2023)
by: Jiang, Zhihan, et al.
Published: (2023)
Uncovering and Mitigating the Impact of Frozen Package Versions for Fixed-Release Linux
by: Tang, Wei, et al.
Published: (2024)
by: Tang, Wei, et al.
Published: (2024)
Understanding the AI-powered Binary Code Similarity Detection
by: Fu, Lirong, et al.
Published: (2024)
by: Fu, Lirong, et al.
Published: (2024)
Doctor: Optimizing Container Rebuild Efficiency by Instruction Re-Orchestration
by: Zhu, Zhiling, et al.
Published: (2025)
by: Zhu, Zhiling, et al.
Published: (2025)
Agent Skills in the Wild: An Empirical Study of Security Vulnerabilities at Scale
by: Liu, Yi, et al.
Published: (2026)
by: Liu, Yi, et al.
Published: (2026)
Security Debt in Practice: Nuanced Insights from Practitioners
by: Boufaied, Chaima, et al.
Published: (2025)
by: Boufaied, Chaima, et al.
Published: (2025)
TransferFuzz: Fuzzing with Historical Trace for Verifying Propagated Vulnerability Code
by: Li, Siyuan, et al.
Published: (2024)
by: Li, Siyuan, et al.
Published: (2024)
Managing Security Issues in Software Containers: From Practitioners Perspective
by: Sroor, Maha, et al.
Published: (2025)
by: Sroor, Maha, et al.
Published: (2025)
Understanding Practitioners' Expectations on Clear Code Review Comments
by: Chen, Junkai, et al.
Published: (2024)
by: Chen, Junkai, et al.
Published: (2024)
Drop the Golden Apples: Identifying Third-Party Reuse by DB-Less Software Composition Analysis
by: Zhang, Lyuye, et al.
Published: (2025)
by: Zhang, Lyuye, et al.
Published: (2025)
Industry Practitioners Perspectives on AI Model Quality: Perceptions, Challenges, and Solutions
by: Wang, Chenyu, et al.
Published: (2024)
by: Wang, Chenyu, et al.
Published: (2024)
What Makes a Good LLM Agent for Real-world Penetration Testing?
by: Deng, Gelei, et al.
Published: (2026)
by: Deng, Gelei, et al.
Published: (2026)
Incorporating Verification Standards for Security Requirements Generation from Functional Specifications
by: Lian, Xiaoli, et al.
Published: (2025)
by: Lian, Xiaoli, et al.
Published: (2025)
Low-Code Paradox in DevOps: Security and Governance Insights from Practitioners
by: Akbar, Muhammad Azeem, et al.
Published: (2026)
by: Akbar, Muhammad Azeem, et al.
Published: (2026)
IntelliRadar: A Comprehensive Platform to Pinpoint Malicious Package Information from Cyber Intelligence
by: Guo, Wenbo, et al.
Published: (2024)
by: Guo, Wenbo, et al.
Published: (2024)
Where Code Meets Natural Language: Taxonomy-Driven Information Flow Analysis for LLM-Integrated Applications
by: Xu, Zihao, et al.
Published: (2026)
by: Xu, Zihao, et al.
Published: (2026)
MiniScope: Automated UI Exploration and Privacy Inconsistency Detection of MiniApps via Two-phase Iterative Hybrid Analysis
by: Wang, Shenao, et al.
Published: (2024)
by: Wang, Shenao, et al.
Published: (2024)
Drowzee: Metamorphic Testing for Fact-Conflicting Hallucination Detection in Large Language Models
by: Li, Ningke, et al.
Published: (2024)
by: Li, Ningke, et al.
Published: (2024)
It Only Gets Worse: Revisiting DL-Based Vulnerability Detectors from a Practical Perspective
by: Wang, Yunqian, et al.
Published: (2025)
by: Wang, Yunqian, et al.
Published: (2025)
Multi-Agent Systems for Dataset Adaptation in Software Engineering: Capabilities, Limitations, and Future Directions
by: Chen, Jingyi, et al.
Published: (2025)
by: Chen, Jingyi, et al.
Published: (2025)
When LLMs Meet API Documentation: Can Retrieval Augmentation Aid Code Generation Just as It Helps Developers?
by: Chen, Jingyi, et al.
Published: (2025)
by: Chen, Jingyi, et al.
Published: (2025)
MeTMaP: Metamorphic Testing for Detecting False Vector Matching Problems in LLM Augmented Generation
by: Wang, Guanyu, et al.
Published: (2024)
by: Wang, Guanyu, et al.
Published: (2024)
Secure-Instruct: An Automated Pipeline for Synthesizing Instruction-Tuning Datasets Using LLMs for Secure Code Generation
by: Li, Junjie, et al.
Published: (2025)
by: Li, Junjie, et al.
Published: (2025)
CrossPL: Evaluating Large Language Models on Cross Programming Language Code Generation
by: Xiong, Zhanhang, et al.
Published: (2025)
by: Xiong, Zhanhang, et al.
Published: (2025)
Understanding NPM Malicious Package Detection: A Benchmark-Driven Empirical Analysis
by: Guo, Wenbo, et al.
Published: (2026)
by: Guo, Wenbo, et al.
Published: (2026)
Smart Contract and DeFi Security Tools: Do They Meet the Needs of Practitioners?
by: Chaliasos, Stefanos, et al.
Published: (2023)
by: Chaliasos, Stefanos, et al.
Published: (2023)
Efficient Function Orchestration for Large Language Models
by: Liu, Xiaoxia, et al.
Published: (2025)
by: Liu, Xiaoxia, et al.
Published: (2025)
Unveiling Dynamic Binary Instrumentation Techniques
by: Llorente-Vazquez, Oscar, et al.
Published: (2025)
by: Llorente-Vazquez, Oscar, et al.
Published: (2025)
JC-Finder: Detecting Java Clone-based Third-Party Library by Class-level Tree Analysis
by: Zhao, Lida, et al.
Published: (2025)
by: Zhao, Lida, et al.
Published: (2025)
Similar Items
-
Vulnerability-Affected Versions Identification: How Far Are We?
by: Chen, Xingchu, et al.
Published: (2025) -
Enhancing Function Name Prediction using Votes-Based Name Tokenization and Multi-Task Learning
by: Zhang, Xiaoling, et al.
Published: (2024) -
Decompile-Bench: Million-Scale Binary-Source Function Pairs for Real-World Binary Decompilation
by: Tan, Hanzhuo, et al.
Published: (2025) -
Practitioners' Expectations on Log Anomaly Detection
by: Ma, Xiaoxue, et al.
Published: (2024) -
ORCAS: Obfuscation-Resilient Binary Code Similarity Analysis using Dominance Enhanced Semantic Graph
by: Wang, Yufeng, et al.
Published: (2025)