Towards Robust Agentic CUDA Kernel Benchmarking, Verification, and Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Lange, Robert Tjarko, Sun, Qi, Prasad, Aaditya, Faldor, Maxence, Tang, Yujin, Ha, David |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ProofWright: Towards Agentic Formal Verification of CUDA
by: Chatterjee, Bodhisatwa, et al.
Published: (2025)
by: Chatterjee, Bodhisatwa, et al.
Published: (2025)
Kevin: Multi-Turn RL for Generating CUDA Kernels
by: Baronio, Carlo, et al.
Published: (2025)
by: Baronio, Carlo, et al.
Published: (2025)
Towards Agentic Runtime Healing
by: Sun, Zhensu, et al.
Published: (2024)
by: Sun, Zhensu, et al.
Published: (2024)
Vision2Web: A Hierarchical Benchmark for Visual Website Development with Agent Verification
by: He, Zehai, et al.
Published: (2026)
by: He, Zehai, et al.
Published: (2026)
RefEvo: Agentic Design with Co-Evolutionary Verification for Agile Reference Model Generation
by: Zhang, Yifan, et al.
Published: (2026)
by: Zhang, Yifan, et al.
Published: (2026)
Toward Reliable Design of LLM-Enabled Agentic Workflows: Optimizing Latency-Reliability-Cost Tradeoffs
by: Yang, Ya-Ting, et al.
Published: (2026)
by: Yang, Ya-Ting, et al.
Published: (2026)
cuVSLAM: CUDA accelerated visual odometry and mapping
by: Korovko, Alexander, et al.
Published: (2025)
by: Korovko, Alexander, et al.
Published: (2025)
Evolution of Kernels: Automated RISC-V Kernel Optimization with Large Language Models
by: Chen, Siyuan, et al.
Published: (2025)
by: Chen, Siyuan, et al.
Published: (2025)
FeatureBench: Benchmarking Agentic Coding for Complex Feature Development
by: Zhou, Qixing, et al.
Published: (2026)
by: Zhou, Qixing, et al.
Published: (2026)
Towards Automated Formal Verification of Backend Systems with LLMs
by: Xu, Kangping, et al.
Published: (2025)
by: Xu, Kangping, et al.
Published: (2025)
Benchmarks are Not Enough: RAMP for Runtime Assessing of Agentic Models in Production Systems
by: Ouyang, Yipeng, et al.
Published: (2026)
by: Ouyang, Yipeng, et al.
Published: (2026)
LLM-Based Agentic Systems for Software Engineering: Challenges and Opportunities
by: Tang, Yongjian, et al.
Published: (2026)
by: Tang, Yongjian, et al.
Published: (2026)
GitGoodBench: A Novel Benchmark For Evaluating Agentic Performance On Git
by: Lindenbauer, Tobias, et al.
Published: (2025)
by: Lindenbauer, Tobias, et al.
Published: (2025)
VeriStruct: AI-assisted Automated Verification of Data-Structure Modules in Verus
by: Sun, Chuyue, et al.
Published: (2025)
by: Sun, Chuyue, et al.
Published: (2025)
OptiML: An End-to-End Framework for Program Synthesis and CUDA Kernel Optimization
by: Bhattacharjee, Arijit, et al.
Published: (2026)
by: Bhattacharjee, Arijit, et al.
Published: (2026)
AgenticTCAD: A LLM-based Multi-Agent Framework for Automated TCAD Code Generation and Device Optimization
by: Fan, Guangxi, et al.
Published: (2025)
by: Fan, Guangxi, et al.
Published: (2025)
ToolMisuseBench: An Offline Deterministic Benchmark for Tool Misuse and Recovery in Agentic Systems
by: Sigdel, Akshey, et al.
Published: (2026)
by: Sigdel, Akshey, et al.
Published: (2026)
From Laboratory to Real-World Applications: Benchmarking Agentic Code Reasoning at the Repository Level
by: Li, Jia, et al.
Published: (2026)
by: Li, Jia, et al.
Published: (2026)
SWE-Compass: Towards Unified Evaluation of Agentic Coding Abilities for Large Language Models
by: Xu, Jingxuan, et al.
Published: (2025)
by: Xu, Jingxuan, et al.
Published: (2025)
The Rise of Agentic Testing: Multi-Agent Systems for Robust Software Quality Assurance
by: Naqvi, Saba, et al.
Published: (2026)
by: Naqvi, Saba, et al.
Published: (2026)
Agentic Business Process Management Systems
by: Dumas, Marlon, et al.
Published: (2026)
by: Dumas, Marlon, et al.
Published: (2026)
Verification-Guided Context Optimization for Tool Calling via Hierarchical LLMs-as-Editors
by: Li, Henger, et al.
Published: (2025)
by: Li, Henger, et al.
Published: (2025)
Towards a Declarative Agentic Layer for Intelligent Agents in MCP-Based Server Ecosystems
by: Rodriguez-Sanchez, Maria Jesus, et al.
Published: (2026)
by: Rodriguez-Sanchez, Maria Jesus, et al.
Published: (2026)
On the Effectiveness of LLMs for Manual Test Verifications
by: Peixoto, Myron David Lucena Campos, et al.
Published: (2024)
by: Peixoto, Myron David Lucena Campos, et al.
Published: (2024)
The Semi-Executable Stack: Agentic Software Engineering and the Expanding Scope of SE
by: Feldt, Robert, et al.
Published: (2026)
by: Feldt, Robert, et al.
Published: (2026)
A Dual-Helix Governance Approach Towards Reliable Agentic AI for WebGIS Development
by: Boyuan, et al.
Published: (2026)
by: Boyuan, et al.
Published: (2026)
FormulaCode: Evaluating Agentic Optimization on Large Codebases
by: Sehgal, Atharva, et al.
Published: (2026)
by: Sehgal, Atharva, et al.
Published: (2026)
DockSmith: Scaling Reliable Coding Environments via an Agentic Docker Builder
by: Zhang, Jiaran, et al.
Published: (2026)
by: Zhang, Jiaran, et al.
Published: (2026)
Beyond the 'Diff': Addressing Agentic Entropy in Agentic Software Development
by: Casserini, Matteo, et al.
Published: (2026)
by: Casserini, Matteo, et al.
Published: (2026)
Beyond Function-Level Search: Repository-Aware Dual-Encoder Code Retrieval with Adversarial Verification
by: Liu, Aofan, et al.
Published: (2025)
by: Liu, Aofan, et al.
Published: (2025)
Towards Comprehensive Benchmarking Infrastructure for LLMs In Software Engineering
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2026)
by: Rodriguez-Cardenas, Daniel, et al.
Published: (2026)
LLM-Driven Kernel Evolution: Automating Driver Updates in Linux
by: Kharlamova, Arina, et al.
Published: (2025)
by: Kharlamova, Arina, et al.
Published: (2025)
Agentic Software Issue Resolution with Large Language Models: A Survey
by: Jiang, Zhonghao, et al.
Published: (2025)
by: Jiang, Zhonghao, et al.
Published: (2025)
Toward an Agentic Infused Software Ecosystem
by: Marron, Mark
Published: (2026)
by: Marron, Mark
Published: (2026)
Agentic Bug Reproduction for Effective Automated Program Repair at Google
by: Cheng, Runxiang, et al.
Published: (2025)
by: Cheng, Runxiang, et al.
Published: (2025)
On Generalization in Agentic Tool Calling: CoreThink Agentic Reasoner and MAVEN Dataset
by: Bhat, Vishvesh, et al.
Published: (2025)
by: Bhat, Vishvesh, et al.
Published: (2025)
Towards Precise Observations of Neural Model Robustness in Classification
by: Mu, Wenchuan, et al.
Published: (2024)
by: Mu, Wenchuan, et al.
Published: (2024)
From Natural Language to Verified Code: Toward AI Assisted Problem-to-Code Generation with Dafny-Based Formal Verification
by: Erfan, Md, et al.
Published: (2026)
by: Erfan, Md, et al.
Published: (2026)
Querying Large Automotive Software Models: Agentic vs. Direct LLM Approaches
by: Mazur, Lukasz, et al.
Published: (2025)
by: Mazur, Lukasz, et al.
Published: (2025)
Agentic Proving for Program Verification
by: Sosso, Alessandro, et al.
Published: (2026)
by: Sosso, Alessandro, et al.
Published: (2026)
Similar Items
-
ProofWright: Towards Agentic Formal Verification of CUDA
by: Chatterjee, Bodhisatwa, et al.
Published: (2025) -
Kevin: Multi-Turn RL for Generating CUDA Kernels
by: Baronio, Carlo, et al.
Published: (2025) -
Towards Agentic Runtime Healing
by: Sun, Zhensu, et al.
Published: (2024) -
Vision2Web: A Hierarchical Benchmark for Visual Website Development with Agent Verification
by: He, Zehai, et al.
Published: (2026) -
RefEvo: Agentic Design with Co-Evolutionary Verification for Agile Reference Model Generation
by: Zhang, Yifan, et al.
Published: (2026)