STAB: Specification-driven Testing for Algorithmic Bottlenecks
Fuente:
arXiv
Saved in:
| Main Authors: | Lim, Soohan, Hahn, Joonghyuk, Jin, Hyundong, Han, Yo-Sub |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
MEC$^3$O: Multi-Expert Consensus for Code Time Complexity Prediction
by: Hahn, Joonghyuk, et al.
Published: (2025)
by: Hahn, Joonghyuk, et al.
Published: (2025)
TCProF: Time-Complexity Prediction SSL Framework
by: Hahn, Joonghyuk, et al.
Published: (2025)
by: Hahn, Joonghyuk, et al.
Published: (2025)
ContractEval: A Benchmark for Evaluating Contract-Satisfying Assertions in Code Generation
by: Lim, Soohan, et al.
Published: (2025)
by: Lim, Soohan, et al.
Published: (2025)
Repairing Regex Vulnerabilities via Localization-Guided Instructions
by: Sung, Sicheol, et al.
Published: (2025)
by: Sung, Sicheol, et al.
Published: (2025)
AmpleHate: Amplifying the Attention for Versatile Implicit Hate Detection
by: Lee, Yejin, et al.
Published: (2025)
by: Lee, Yejin, et al.
Published: (2025)
RegexPSPACE: A Benchmark for Evaluating LLM Reasoning on PSPACE-complete Regex Problems
by: Jin, Hyundong, et al.
Published: (2025)
by: Jin, Hyundong, et al.
Published: (2025)
Obfuscation Rules for Detecting and Detoxifying Korean Toxicity
by: Lee, Yejin, et al.
Published: (2025)
by: Lee, Yejin, et al.
Published: (2025)
Test-driven Software Experimentation with LASSO: an LLM Prompt Benchmarking Example
by: Kessel, Marcus
Published: (2024)
by: Kessel, Marcus
Published: (2024)
Active Context Compression: Autonomous Memory Management in LLM Agents
by: Verma, Nikhil
Published: (2026)
by: Verma, Nikhil
Published: (2026)
Test-Driven AI Agent Definition (TDAD): Compiling Tool-Using Agents from Behavioral Specifications
by: Rehan, Tzafrir
Published: (2026)
by: Rehan, Tzafrir
Published: (2026)
Addressing Data Leakage in HumanEval Using Combinatorial Test Design
by: Bradbury, Jeremy S., et al.
Published: (2024)
by: Bradbury, Jeremy S., et al.
Published: (2024)
RV-HATE: Reinforced Multi-Module Voting for Implicit Hate Speech Detection
by: Lee, Yejin, et al.
Published: (2025)
by: Lee, Yejin, et al.
Published: (2025)
Algorithmic Trading Strategy Development and Optimisation
by: Yuan, Owen Nyo Wei, et al.
Published: (2026)
by: Yuan, Owen Nyo Wei, et al.
Published: (2026)
Breaking Validity-Induced Boundaries to Expand Algorithm Search Space: A Two-Stage AST-Based Operator for LLM-Driven Automated Heuristic Evolution
by: Shengming, Sun, et al.
Published: (2026)
by: Shengming, Sun, et al.
Published: (2026)
CodeEvolve: LLM-Driven Evolutionary Optimization with Runtime-Enriched Target Selection for Multi-Language Code Enhancement
by: Borra, Ajay Krishna, et al.
Published: (2026)
by: Borra, Ajay Krishna, et al.
Published: (2026)
CortexCompile: Harnessing Cortical-Inspired Architectures for Enhanced Multi-Agent NLP Code Synthesis
by: Ramachandran, Gautham, et al.
Published: (2024)
by: Ramachandran, Gautham, et al.
Published: (2024)
DELTA: Variational Disentangled Learning for Privacy-Preserving Data Reprogramming
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
Solve it with EASE
by: Viktorin, Adam, et al.
Published: (2025)
by: Viktorin, Adam, et al.
Published: (2025)
Deriving Coding-Specific Sub-Models from LLMs using Resource-Efficient Pruning
by: Puccioni, Laura, et al.
Published: (2025)
by: Puccioni, Laura, et al.
Published: (2025)
4OPS: Structural Difficulty Modeling in Integer Arithmetic Puzzles
by: Zeytuncu, Yunus E.
Published: (2026)
by: Zeytuncu, Yunus E.
Published: (2026)
Vision-Guided Iterative Refinement for Frontend Code Generation
by: Sansford, Hannah, et al.
Published: (2026)
by: Sansford, Hannah, et al.
Published: (2026)
Utility Function is All You Need: LLM-based Congestion Control
by: Rozen-Schiff, Neta, et al.
Published: (2026)
by: Rozen-Schiff, Neta, et al.
Published: (2026)
Train to Defend: First Defense Against Cryptanalytic Neural Network Parameter Extraction Attacks
by: Kurian, Ashley, et al.
Published: (2025)
by: Kurian, Ashley, et al.
Published: (2025)
TRIM: Achieving Extreme Sparsity with Targeted Row-wise Iterative Metric-driven Pruning
by: Beck, Florentin, et al.
Published: (2025)
by: Beck, Florentin, et al.
Published: (2025)
Random-Key Algorithms for Optimizing Integrated Operating Room Scheduling
by: Vieira, Bruno Salezze, et al.
Published: (2025)
by: Vieira, Bruno Salezze, et al.
Published: (2025)
Examination of Code generated by Large Language Models
by: Beer, Robin, et al.
Published: (2024)
by: Beer, Robin, et al.
Published: (2024)
QiMeng-TensorOp: Automatically Generating High-Performance Tensor Operators with Hardware Primitives
by: Zhang, Xuzhi, et al.
Published: (2025)
by: Zhang, Xuzhi, et al.
Published: (2025)
MaPPing Your Model: Assessing the Impact of Adversarial Attacks on LLM-based Programming Assistants
by: Heibel, John, et al.
Published: (2024)
by: Heibel, John, et al.
Published: (2024)
Towards Platonic Representation for Table Reasoning: A Foundation for Permutation-Invariant Retrieval
by: Tchuitcheu, Willy Carlos, et al.
Published: (2026)
by: Tchuitcheu, Willy Carlos, et al.
Published: (2026)
Domain-Independent Dynamic Programming
by: Kuroiwa, Ryo, et al.
Published: (2024)
by: Kuroiwa, Ryo, et al.
Published: (2024)
An In-depth Study of LLM Contributions to the Bin Packing Problem
by: Herrmann, Julien, et al.
Published: (2025)
by: Herrmann, Julien, et al.
Published: (2025)
Large Language Model Interface for Home Energy Management Systems
by: Michelon, François, et al.
Published: (2025)
by: Michelon, François, et al.
Published: (2025)
Policy-Grounded Safety Evaluation of 20 Large Language Models
by: Contreras, Juan Manuel
Published: (2025)
by: Contreras, Juan Manuel
Published: (2025)
Evaluating the efficacy of LLM Safety Solutions : The Palit Benchmark Dataset
by: Palit, Sayon, et al.
Published: (2025)
by: Palit, Sayon, et al.
Published: (2025)
Evolving Programmatic Skill Networks
by: Shi, Haochen, et al.
Published: (2026)
by: Shi, Haochen, et al.
Published: (2026)
Interactive LLM-assisted Curriculum Learning for Multi-Task Evolutionary Policy Search
by: Sakallioglu, Berfin, et al.
Published: (2026)
by: Sakallioglu, Berfin, et al.
Published: (2026)
RoboGrind: Intuitive and Interactive Surface Treatment with Industrial Robots
by: Alt, Benjamin, et al.
Published: (2024)
by: Alt, Benjamin, et al.
Published: (2024)
Directed Domain Fine-Tuning: Tailoring Separate Modalities for Specific Training Tasks
by: Wen, Daniel, et al.
Published: (2024)
by: Wen, Daniel, et al.
Published: (2024)
Can ChatGPT Support Developers? An Empirical Evaluation of Large Language Models for Code Generation
by: Jin, Kailun, et al.
Published: (2024)
by: Jin, Kailun, et al.
Published: (2024)
Evaluating the Limitations of Local LLMs in Solving Complex Programming Challenges
by: Matotek, Kadin, et al.
Published: (2025)
by: Matotek, Kadin, et al.
Published: (2025)
Similar Items
-
MEC$^3$O: Multi-Expert Consensus for Code Time Complexity Prediction
by: Hahn, Joonghyuk, et al.
Published: (2025) -
TCProF: Time-Complexity Prediction SSL Framework
by: Hahn, Joonghyuk, et al.
Published: (2025) -
ContractEval: A Benchmark for Evaluating Contract-Satisfying Assertions in Code Generation
by: Lim, Soohan, et al.
Published: (2025) -
Repairing Regex Vulnerabilities via Localization-Guided Instructions
by: Sung, Sicheol, et al.
Published: (2025) -
AmpleHate: Amplifying the Attention for Versatile Implicit Hate Detection
by: Lee, Yejin, et al.
Published: (2025)