SimClone: Detecting Tabular Data Clones using Value Similarity
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Xu, Rajbahadur, Gopi Krishnan, Lin, Dayi, Wang, Shaowei, Ming, Zhen, Jiang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Keeping Deep Learning Models in Check: A History-Based Approach to Mitigate Overfitting
di: Li, Hao, et al.
Pubblicazione: (2024)
di: Li, Hao, et al.
Pubblicazione: (2024)
Data Quality Antipatterns for Software Analytics
di: Bhatia, Aaditya, et al.
Pubblicazione: (2024)
di: Bhatia, Aaditya, et al.
Pubblicazione: (2024)
From Cool Demos to Production-Ready FMware: Core Challenges and a Technology Roadmap
di: Rajbahadur, Gopi Krishnan, et al.
Pubblicazione: (2024)
di: Rajbahadur, Gopi Krishnan, et al.
Pubblicazione: (2024)
Edit, But Verify: An Empirical Audit of Instructed Code-Editing Benchmarks
di: Ebrahimi, Amir M., et al.
Pubblicazione: (2026)
di: Ebrahimi, Amir M., et al.
Pubblicazione: (2026)
Studying the Impact of TensorFlow and PyTorch Bindings on Machine Learning Software Quality
di: Li, Hao, et al.
Pubblicazione: (2024)
di: Li, Hao, et al.
Pubblicazione: (2024)
Quality Assessment of Tabular Data using Large Language Models and Code Generation
di: Akella, Ashlesha, et al.
Pubblicazione: (2025)
di: Akella, Ashlesha, et al.
Pubblicazione: (2025)
Advanced Detection of Source Code Clones via an Ensemble of Unsupervised Similarity Measures
di: Martinez-Gil, Jorge
Pubblicazione: (2024)
di: Martinez-Gil, Jorge
Pubblicazione: (2024)
Rethinking Software Engineering in the Foundation Model Era: A Curated Catalogue of Challenges in the Development of Trustworthy FMware
di: Hassan, Ahmed E., et al.
Pubblicazione: (2024)
di: Hassan, Ahmed E., et al.
Pubblicazione: (2024)
From Hugging Face to GitHub: Tracing License Drift in the Open-Source AI Ecosystem
di: Jewitt, James, et al.
Pubblicazione: (2025)
di: Jewitt, James, et al.
Pubblicazione: (2025)
SPICE: An Automated SWE-Bench Labeling Pipeline for Issue Clarity, Test Coverage, and Effort Estimation
di: Oliva, Gustavo A., et al.
Pubblicazione: (2025)
di: Oliva, Gustavo A., et al.
Pubblicazione: (2025)
Assessing the Code Clone Detection Capability of Large Language Models
di: Zhang, Zixian, et al.
Pubblicazione: (2024)
di: Zhang, Zixian, et al.
Pubblicazione: (2024)
MAGNET: A Multi-Graph Attentional Network for Code Clone Detection
di: Zhang, Zixian, et al.
Pubblicazione: (2025)
di: Zhang, Zixian, et al.
Pubblicazione: (2025)
Selecting and Combining Large Language Models for Scalable Code Clone Detection
di: Chochlov, Muslim, et al.
Pubblicazione: (2025)
di: Chochlov, Muslim, et al.
Pubblicazione: (2025)
The Hitchhikers Guide to Production-ready Trustworthy Foundation Model powered Software (FMware)
di: Vasilevski, Kirill, et al.
Pubblicazione: (2025)
di: Vasilevski, Kirill, et al.
Pubblicazione: (2025)
LicenseGPT: A Fine-tuned Foundation Model for Publicly Available Dataset License Compliance
di: Tan, Jingwen, et al.
Pubblicazione: (2024)
di: Tan, Jingwen, et al.
Pubblicazione: (2024)
Adaptive Data Quality Scoring Operations Framework using Drift-Aware Mechanism for Industrial Applications
di: Bayram, Firas, et al.
Pubblicazione: (2024)
di: Bayram, Firas, et al.
Pubblicazione: (2024)
Unraveling Code Clone Dynamics in Deep Learning Frameworks
di: Assi, Maram, et al.
Pubblicazione: (2024)
di: Assi, Maram, et al.
Pubblicazione: (2024)
The Struggles of LLMs in Cross-lingual Code Clone Detection
di: Moumoula, Micheline Bénédicte, et al.
Pubblicazione: (2024)
di: Moumoula, Micheline Bénédicte, et al.
Pubblicazione: (2024)
Detect, Localize, and Explain: Interactive Hierarchical Log Anomaly Analytics with LLM Augmentation
di: Ma, Lei, et al.
Pubblicazione: (2026)
di: Ma, Lei, et al.
Pubblicazione: (2026)
AST-Enhanced or AST-Overloaded? The Surprising Impact of Hybrid Graph Representations on Code Clone Detection
di: Zhang, Zixian, et al.
Pubblicazione: (2025)
di: Zhang, Zixian, et al.
Pubblicazione: (2025)
Declarative Techniques for NL Queries over Heterogeneous Data
di: Khabiri, Elham, et al.
Pubblicazione: (2025)
di: Khabiri, Elham, et al.
Pubblicazione: (2025)
KRONE: Scalable LLM-Augmented Log Anomaly Detection via Hierarchical Abstraction
di: Ma, Lei, et al.
Pubblicazione: (2026)
di: Ma, Lei, et al.
Pubblicazione: (2026)
A Pythonic Functional Approach for Semantic Data Harmonisation in the ILIAD Project
di: Nystad, Erik Johan, et al.
Pubblicazione: (2026)
di: Nystad, Erik Johan, et al.
Pubblicazione: (2026)
Permissive-Washing in the Open AI Supply Chain: A Large-Scale Audit of License Integrity
di: Jewitt, James, et al.
Pubblicazione: (2026)
di: Jewitt, James, et al.
Pubblicazione: (2026)
Generating Reliable Adverse event Profiles for Health through Automated Integrated Data (GRAPH-AID): A Semi-Automated Ontology Building Approach
di: Gadusu, Srikar Reddy, et al.
Pubblicazione: (2025)
di: Gadusu, Srikar Reddy, et al.
Pubblicazione: (2025)
Rethinking Software Engineering in the Foundation Model Era: From Task-Driven AI Copilots to Goal-Driven AI Pair Programmers
di: Hassan, Ahmed E., et al.
Pubblicazione: (2024)
di: Hassan, Ahmed E., et al.
Pubblicazione: (2024)
Towards AI-Native Software Engineering (SE 3.0): A Vision and a Challenge Roadmap
di: Hassan, Ahmed E., et al.
Pubblicazione: (2024)
di: Hassan, Ahmed E., et al.
Pubblicazione: (2024)
SkillClone: Multi-Modal Clone Detection and Clone Propagation Analysis in the Agent Skill Ecosystem
di: Zhu, Jiaying, et al.
Pubblicazione: (2026)
di: Zhu, Jiaying, et al.
Pubblicazione: (2026)
Standing on the Shoulders of Giants: Stabilized Knowledge Distillation for Cross--Language Code Clone Detection
di: Khajezade, Mohamad, et al.
Pubblicazione: (2026)
di: Khajezade, Mohamad, et al.
Pubblicazione: (2026)
HGAdapter: Hypergraph-based Adapters in Language Models for Code Summarization and Clone Detection
di: Yang, Guang, et al.
Pubblicazione: (2025)
di: Yang, Guang, et al.
Pubblicazione: (2025)
Investigating the Efficacy of Large Language Models for Code Clone Detection
di: Khajezade, Mohamad, et al.
Pubblicazione: (2024)
di: Khajezade, Mohamad, et al.
Pubblicazione: (2024)
Prompt Migration: Stabilizing GenAI Applications with Evolving Large Language Models
di: Tripathi, Shivani, et al.
Pubblicazione: (2025)
di: Tripathi, Shivani, et al.
Pubblicazione: (2025)
GEE-OPs: An Operator Knowledge Base for Geospatial Code Generation on the Google Earth Engine Platform Powered by Large Language Models
di: Hou, Shuyang, et al.
Pubblicazione: (2024)
di: Hou, Shuyang, et al.
Pubblicazione: (2024)
Tabularis Formatus: Predictive Formatting for Tables
di: Singh, Mukul, et al.
Pubblicazione: (2025)
di: Singh, Mukul, et al.
Pubblicazione: (2025)
From Natural Language to PromQL: A Catalog-Driven Framework with Dynamic Temporal Resolution for Cloud-Native Observability
di: Sisodia, Twinkll
Pubblicazione: (2026)
di: Sisodia, Twinkll
Pubblicazione: (2026)
LeGo-Code: Can Modular Curriculum Learning Advance Complex Code Generation? Insights from Text-to-SQL
di: Chafik, Salmane, et al.
Pubblicazione: (2026)
di: Chafik, Salmane, et al.
Pubblicazione: (2026)
Enhancing LLM Fine-tuning for Text-to-SQLs by SQL Quality Measurement
di: Sarker, Shouvon, et al.
Pubblicazione: (2024)
di: Sarker, Shouvon, et al.
Pubblicazione: (2024)
Control-flow Reconstruction Attacks on Business Process Models
di: Kirchmann, Henrik, et al.
Pubblicazione: (2024)
di: Kirchmann, Henrik, et al.
Pubblicazione: (2024)
Unraveling the Never-Ending Story of Lifecycles and Vitalizing Processes
di: Fahrenkrog-Petersen, Stephan A., et al.
Pubblicazione: (2024)
di: Fahrenkrog-Petersen, Stephan A., et al.
Pubblicazione: (2024)
Geo-FuB: A Method for Constructing an Operator-Function Knowledge Base for Geospatial Code Generation Tasks Using Large Language Models
di: Hou, Shuyang, et al.
Pubblicazione: (2024)
di: Hou, Shuyang, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Keeping Deep Learning Models in Check: A History-Based Approach to Mitigate Overfitting
di: Li, Hao, et al.
Pubblicazione: (2024) -
Data Quality Antipatterns for Software Analytics
di: Bhatia, Aaditya, et al.
Pubblicazione: (2024) -
From Cool Demos to Production-Ready FMware: Core Challenges and a Technology Roadmap
di: Rajbahadur, Gopi Krishnan, et al.
Pubblicazione: (2024) -
Edit, But Verify: An Empirical Audit of Instructed Code-Editing Benchmarks
di: Ebrahimi, Amir M., et al.
Pubblicazione: (2026) -
Studying the Impact of TensorFlow and PyTorch Bindings on Machine Learning Software Quality
di: Li, Hao, et al.
Pubblicazione: (2024)