s2n-bignum-bench: A practical benchmark for evaluating low-level code reasoning of LLMs

Fuente: arXiv
Saved in:
Bibliographic Details
Main Authors: Rao, Balaji, Harrison, John, Kong, Soonho, Lee, Juneyoung, Lipizzi, Carlo
Format: Preprint
Published: 2026
Subjects:
Online Access:
Tags: Add Tag
No Tags, Be the first to tag this record!