Advanced Detection of Source Code Clones via an Ensemble of Unsupervised Similarity Measures
Fuente:
arXiv
Guardado en:
| Autor principal: | Martinez-Gil, Jorge |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Source Code Clone Detection Using Unsupervised Similarity Measures
por: Martinez-Gil, Jorge
Publicado: (2024)
por: Martinez-Gil, Jorge
Publicado: (2024)
Evaluating Small-Scale Code Models for Code Clone Detection
por: Martinez-Gil, Jorge
Publicado: (2025)
por: Martinez-Gil, Jorge
Publicado: (2025)
Assessing the Code Clone Detection Capability of Large Language Models
por: Zhang, Zixian, et al.
Publicado: (2024)
por: Zhang, Zixian, et al.
Publicado: (2024)
SimClone: Detecting Tabular Data Clones using Value Similarity
por: Yang, Xu, et al.
Publicado: (2024)
por: Yang, Xu, et al.
Publicado: (2024)
Improving Source Code Similarity Detection Through GraphCodeBERT and Integration of Additional Features
por: Martinez-Gil, Jorge
Publicado: (2024)
por: Martinez-Gil, Jorge
Publicado: (2024)
Semantic Similarity Loss for Neural Source Code Summarization
por: Su, Chia-Yi, et al.
Publicado: (2023)
por: Su, Chia-Yi, et al.
Publicado: (2023)
MAGNET: A Multi-Graph Attentional Network for Code Clone Detection
por: Zhang, Zixian, et al.
Publicado: (2025)
por: Zhang, Zixian, et al.
Publicado: (2025)
Selecting and Combining Large Language Models for Scalable Code Clone Detection
por: Chochlov, Muslim, et al.
Publicado: (2025)
por: Chochlov, Muslim, et al.
Publicado: (2025)
Unraveling Code Clone Dynamics in Deep Learning Frameworks
por: Assi, Maram, et al.
Publicado: (2024)
por: Assi, Maram, et al.
Publicado: (2024)
AST-Enhanced or AST-Overloaded? The Surprising Impact of Hybrid Graph Representations on Code Clone Detection
por: Zhang, Zixian, et al.
Publicado: (2025)
por: Zhang, Zixian, et al.
Publicado: (2025)
The Struggles of LLMs in Cross-lingual Code Clone Detection
por: Moumoula, Micheline Bénédicte, et al.
Publicado: (2024)
por: Moumoula, Micheline Bénédicte, et al.
Publicado: (2024)
SourceP: Detecting Ponzi Schemes on Ethereum with Source Code
por: Lu, Pengcheng, et al.
Publicado: (2023)
por: Lu, Pengcheng, et al.
Publicado: (2023)
Machine Learning Techniques for Python Source Code Vulnerability Detection
por: Farasat, Talaya, et al.
Publicado: (2024)
por: Farasat, Talaya, et al.
Publicado: (2024)
Standing on the Shoulders of Giants: Stabilized Knowledge Distillation for Cross--Language Code Clone Detection
por: Khajezade, Mohamad, et al.
Publicado: (2026)
por: Khajezade, Mohamad, et al.
Publicado: (2026)
Investigating the Efficacy of Large Language Models for Code Clone Detection
por: Khajezade, Mohamad, et al.
Publicado: (2024)
por: Khajezade, Mohamad, et al.
Publicado: (2024)
Do Machines and Humans Focus on Similar Code? Exploring Explainability of Large Language Models in Code Summarization
por: Li, Jiliang, et al.
Publicado: (2024)
por: Li, Jiliang, et al.
Publicado: (2024)
Distilled GPT for Source Code Summarization
por: Su, Chia-Yi, et al.
Publicado: (2023)
por: Su, Chia-Yi, et al.
Publicado: (2023)
Generating High-Quality Datasets for Code Editing via Open-Source Language Models
por: Zhang, Zekai, et al.
Publicado: (2025)
por: Zhang, Zekai, et al.
Publicado: (2025)
What Makes Code Generation Ethically Sourced?
por: Xu, Zhuolin, et al.
Publicado: (2025)
por: Xu, Zhuolin, et al.
Publicado: (2025)
A Controlled Experiment on the Energy Efficiency of the Source Code Generated by Code Llama
por: Cursaru, Vlad-Andrei, et al.
Publicado: (2024)
por: Cursaru, Vlad-Andrei, et al.
Publicado: (2024)
HGAdapter: Hypergraph-based Adapters in Language Models for Code Summarization and Clone Detection
por: Yang, Guang, et al.
Publicado: (2025)
por: Yang, Guang, et al.
Publicado: (2025)
Harnessing the Power of LLMs in Source Code Vulnerability Detection
por: Mahyari, Andrew A
Publicado: (2024)
por: Mahyari, Andrew A
Publicado: (2024)
LibreLog: Accurate and Efficient Unsupervised Log Parsing Using Open-Source Large Language Models
por: Ma, Zeyang, et al.
Publicado: (2024)
por: Ma, Zeyang, et al.
Publicado: (2024)
Will It Survive? Deciphering the Fate of AI-Generated Code in Open Source
por: Rahman, Musfiqur, et al.
Publicado: (2026)
por: Rahman, Musfiqur, et al.
Publicado: (2026)
Detecting and Correcting Hallucinations in LLM-Generated Code via Deterministic AST Analysis
por: Khati, Dipin, et al.
Publicado: (2026)
por: Khati, Dipin, et al.
Publicado: (2026)
Code2Bench: Scaling Source and Rigor for Dynamic Benchmark Construction
por: Zhang, Zhe, et al.
Publicado: (2025)
por: Zhang, Zhe, et al.
Publicado: (2025)
Towards Leveraging Large Language Model Summaries for Topic Modeling in Source Code
por: Carissimi, Michele, et al.
Publicado: (2025)
por: Carissimi, Michele, et al.
Publicado: (2025)
Boosting Source Code Learning with Text-Oriented Data Augmentation: An Empirical Study
por: Dong, Zeming, et al.
Publicado: (2023)
por: Dong, Zeming, et al.
Publicado: (2023)
Automatic Identification of Parallelizable Loops Using Transformer-Based Source Code Representations
por: Correia, Izavan dos S., et al.
Publicado: (2026)
por: Correia, Izavan dos S., et al.
Publicado: (2026)
Specification and Detection of LLM Code Smells
por: Mahmoudi, Brahim, et al.
Publicado: (2025)
por: Mahmoudi, Brahim, et al.
Publicado: (2025)
AI Code in the Wild: Measuring Security Risks and Ecosystem Shifts of AI-Generated Code in Modern Software
por: Wang, Bin, et al.
Publicado: (2025)
por: Wang, Bin, et al.
Publicado: (2025)
REINFOREST: Reinforcing Semantic Code Similarity for Cross-Lingual Code Search Models
por: Saieva, Anthony, et al.
Publicado: (2023)
por: Saieva, Anthony, et al.
Publicado: (2023)
An Effective Approach to Embedding Source Code by Combining Large Language and Sentence Embedding Models
por: Xian, Zixiang, et al.
Publicado: (2024)
por: Xian, Zixiang, et al.
Publicado: (2024)
Automated Classification of Source Code Changes Based on Metrics Clustering in the Software Development Process
por: Kniazev, Evgenii
Publicado: (2026)
por: Kniazev, Evgenii
Publicado: (2026)
On the Limitations of Embedding Based Methods for Measuring Functional Correctness for Code Generation
por: Naik, Atharva
Publicado: (2024)
por: Naik, Atharva
Publicado: (2024)
XSearch: Explainable Code Search via Concept-to-Code Alignment
por: Liu, Yiming, et al.
Publicado: (2026)
por: Liu, Yiming, et al.
Publicado: (2026)
ML Code Smells: From Specification to Detection
por: Mahmoudi, Brahim, et al.
Publicado: (2025)
por: Mahmoudi, Brahim, et al.
Publicado: (2025)
Towards Advancing Code Generation with Large Language Models: A Research Roadmap
por: Jin, Haolin, et al.
Publicado: (2025)
por: Jin, Haolin, et al.
Publicado: (2025)
Asm2SrcEval: Evaluating Large Language Models for Assembly-to-Source Code Translation
por: Hamedi, Parisa, et al.
Publicado: (2025)
por: Hamedi, Parisa, et al.
Publicado: (2025)
Beyond Embeddings: Interpretable Feature Extraction for Binary Code Similarity
por: Gagnon, Charles E., et al.
Publicado: (2025)
por: Gagnon, Charles E., et al.
Publicado: (2025)
Ejemplares similares
-
Source Code Clone Detection Using Unsupervised Similarity Measures
por: Martinez-Gil, Jorge
Publicado: (2024) -
Evaluating Small-Scale Code Models for Code Clone Detection
por: Martinez-Gil, Jorge
Publicado: (2025) -
Assessing the Code Clone Detection Capability of Large Language Models
por: Zhang, Zixian, et al.
Publicado: (2024) -
SimClone: Detecting Tabular Data Clones using Value Similarity
por: Yang, Xu, et al.
Publicado: (2024) -
Improving Source Code Similarity Detection Through GraphCodeBERT and Integration of Additional Features
por: Martinez-Gil, Jorge
Publicado: (2024)