Searching by Code: a New SearchBySnippet Dataset and SnippeR Retrieval Model for Searching by Code Snippets
Fuente:
arXiv
Saved in:
| Main Authors: | Sedykh, Ivan, Abulkhanov, Dmitry, Sorokin, Nikita, Nikolenko, Sergey, Malykh, Valentin |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CCT-Code: Cross-Consistency Training for Multilingual Clone Detection and Code Search
by: Tikhonov, Anton, et al.
Published: (2023)
by: Tikhonov, Anton, et al.
Published: (2023)
Iterative Self-Training for Code Generation via Reinforced Re-Ranking
by: Sorokin, Nikita, et al.
Published: (2025)
by: Sorokin, Nikita, et al.
Published: (2025)
Hierarchical Embedding Fusion for Retrieval-Augmented Code Generation
by: Sorokin, Nikita, et al.
Published: (2026)
by: Sorokin, Nikita, et al.
Published: (2026)
ArkTS-CodeSearch: A Open-Source ArkTS Dataset for Code Retrieval
by: He, Yulong, et al.
Published: (2026)
by: He, Yulong, et al.
Published: (2026)
Can We Identify Stack Overflow Questions Requiring Code Snippets? Investigating the Cause & Effect of Missing Code Snippets
by: Mondal, Saikat, et al.
Published: (2024)
by: Mondal, Saikat, et al.
Published: (2024)
StRuCom: A Novel Dataset of Structured Code Comments in Russian
by: Dziuba, Maria, et al.
Published: (2025)
by: Dziuba, Maria, et al.
Published: (2025)
Towards Summarizing Code Snippets Using Pre-Trained Transformers
by: Mastropaolo, Antonio, et al.
Published: (2024)
by: Mastropaolo, Antonio, et al.
Published: (2024)
GENCNIPPET: Automated Generation of Code Snippets for Supporting Programming Questions
by: Mondal, Saikat, et al.
Published: (2025)
by: Mondal, Saikat, et al.
Published: (2025)
Automated Snippet-Alignment Data Augmentation for Code Translation
by: Zhang, Zhiming, et al.
Published: (2025)
by: Zhang, Zhiming, et al.
Published: (2025)
Readability and Understandability of Snippets Recommended by General-purpose Web Search Engines: a Comparative Study
by: Dantas, Carlos Eduardo C., et al.
Published: (2021)
by: Dantas, Carlos Eduardo C., et al.
Published: (2021)
Unmasking the Genuine Type Inference Capabilities of LLMs for Java Code Snippets
by: Dong, Yiwen, et al.
Published: (2025)
by: Dong, Yiwen, et al.
Published: (2025)
Not All Denoising Steps Are Equal: Model Scheduling for Faster Masked Diffusion Language Models
by: Sedykh, Ivan, et al.
Published: (2026)
by: Sedykh, Ivan, et al.
Published: (2026)
Identifying Smart Contract Security Issues in Code Snippets from Stack Overflow
by: Chen, Jiachi, et al.
Published: (2024)
by: Chen, Jiachi, et al.
Published: (2024)
Code2API: A Tool for Generating Reusable APIs from Stack Overflow Code Snippets
by: Mai, Yubo, et al.
Published: (2025)
by: Mai, Yubo, et al.
Published: (2025)
Uncovering Intention through LLM-Driven Code Snippet Description Generation
by: Nugroho, Yusuf Sulistyo, et al.
Published: (2025)
by: Nugroho, Yusuf Sulistyo, et al.
Published: (2025)
Beyond Code Snippets: Benchmarking LLMs on Repository-Level Question Answering
by: Alebachew, Yoseph Berhanu, et al.
Published: (2026)
by: Alebachew, Yoseph Berhanu, et al.
Published: (2026)
CIDRe: A Reference-Free Multi-Aspect Criterion for Code Comment Quality Measurement
by: Dziuba, Maria, et al.
Published: (2025)
by: Dziuba, Maria, et al.
Published: (2025)
Demystifying Code Snippets in Code Reviews: A Study of the OpenStack and Qt Communities and A Practitioner Survey
by: Zhang, Beiqi, et al.
Published: (2023)
by: Zhang, Beiqi, et al.
Published: (2023)
Efficient and Scalable Provenance Tracking for LLM-Generated Code Snippets
by: Gurioli, Andrea, et al.
Published: (2026)
by: Gurioli, Andrea, et al.
Published: (2026)
AdaptEval: A Benchmark for Evaluating Large Language Models on Code Snippet Adaptation
by: Zhang, Tanghaoran, et al.
Published: (2026)
by: Zhang, Tanghaoran, et al.
Published: (2026)
Automated Code Editing with Search-Generate-Modify
by: Liu, Changshu, et al.
Published: (2023)
by: Liu, Changshu, et al.
Published: (2023)
ZS4C: Zero-Shot Synthesis of Compilable Code for Incomplete Code Snippets using LLMs
by: Kabir, Azmain, et al.
Published: (2024)
by: Kabir, Azmain, et al.
Published: (2024)
Structural Code Search using Natural Language Queries
by: Limpanukorn, Ben, et al.
Published: (2025)
by: Limpanukorn, Ben, et al.
Published: (2025)
Readability and Understandability Scores for Snippet Assessment: an Exploratory Study
by: Dantas, Carlos Eduardo C., et al.
Published: (2021)
by: Dantas, Carlos Eduardo C., et al.
Published: (2021)
Search-Based LLMs for Code Optimization
by: Gao, Shuzheng, et al.
Published: (2024)
by: Gao, Shuzheng, et al.
Published: (2024)
REINFOREST: Reinforcing Semantic Code Similarity for Cross-Lingual Code Search Models
by: Saieva, Anthony, et al.
Published: (2023)
by: Saieva, Anthony, et al.
Published: (2023)
AUTOGENICS: Automated Generation of Context-Aware Inline Comments for Code Snippets on Programming Q&A Sites Using LLM
by: Bappon, Suborno Deb, et al.
Published: (2024)
by: Bappon, Suborno Deb, et al.
Published: (2024)
CodeScout: An Effective Recipe for Reinforcement Learning of Code Search Agents
by: Sutawika, Lintang, et al.
Published: (2026)
by: Sutawika, Lintang, et al.
Published: (2026)
Zero-Shot Cross-Domain Code Search without Fine-Tuning
by: Liang, Keyu, et al.
Published: (2025)
by: Liang, Keyu, et al.
Published: (2025)
Symbolic Execution Meets Multi-LLM Orchestration: Detecting Memory Vulnerabilities in Incomplete Rust CVE Snippets
by: Abdelrazek, Zeyad, et al.
Published: (2026)
by: Abdelrazek, Zeyad, et al.
Published: (2026)
A Multi-Perspective Architecture for Semantic Code Search
by: Haldar, Rajarshi, et al.
Published: (2020)
by: Haldar, Rajarshi, et al.
Published: (2020)
RethinkMCTS: Refining Erroneous Thoughts in Monte Carlo Tree Search for Code Generation
by: Li, Qingyao, et al.
Published: (2024)
by: Li, Qingyao, et al.
Published: (2024)
STELLAR: A Search-Based Testing Framework for Large Language Model Applications
by: Sorokin, Lev, et al.
Published: (2026)
by: Sorokin, Lev, et al.
Published: (2026)
ProCQA: A Large-scale Community-based Programming Question Answering Dataset for Code Search
by: Li, Zehan, et al.
Published: (2024)
by: Li, Zehan, et al.
Published: (2024)
Instruct or Interact? Exploring and Eliciting LLMs' Capability in Code Snippet Adaptation Through Prompt Engineering
by: Zhang, Tanghaoran, et al.
Published: (2024)
by: Zhang, Tanghaoran, et al.
Published: (2024)
LLM Agents Improve Semantic Code Search
by: Jain, Sarthak, et al.
Published: (2024)
by: Jain, Sarthak, et al.
Published: (2024)
Rewriting the Code: A Simple Method for Large Language Model Augmented Code Search
by: Li, Haochen, et al.
Published: (2024)
by: Li, Haochen, et al.
Published: (2024)
Beyond Retrieval: A Multitask Benchmark and Model for Code Search
by: Xue, Siqiao, et al.
Published: (2026)
by: Xue, Siqiao, et al.
Published: (2026)
Issue Localization via LLM-Driven Iterative Code Graph Searching
by: Jiang, Zhonghao, et al.
Published: (2025)
by: Jiang, Zhonghao, et al.
Published: (2025)
From Completion to Editing: Unlocking Context-Aware Code Infilling via Search-and-Replace Instruction Tuning
by: Zhang, Jiajun, et al.
Published: (2026)
by: Zhang, Jiajun, et al.
Published: (2026)
Similar Items
-
CCT-Code: Cross-Consistency Training for Multilingual Clone Detection and Code Search
by: Tikhonov, Anton, et al.
Published: (2023) -
Iterative Self-Training for Code Generation via Reinforced Re-Ranking
by: Sorokin, Nikita, et al.
Published: (2025) -
Hierarchical Embedding Fusion for Retrieval-Augmented Code Generation
by: Sorokin, Nikita, et al.
Published: (2026) -
ArkTS-CodeSearch: A Open-Source ArkTS Dataset for Code Retrieval
by: He, Yulong, et al.
Published: (2026) -
Can We Identify Stack Overflow Questions Requiring Code Snippets? Investigating the Cause & Effect of Missing Code Snippets
by: Mondal, Saikat, et al.
Published: (2024)