C2RUST-BENCH: A Minimized, Representative Dataset for C-to-Rust Transpilation Evaluation
Fuente:
arXiv
Saved in:
| Main Authors: | Sirlanci, Melih, Yagemann, Carter, Lin, Zhiqiang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Practical Framework for Evaluating Medical AI Security: Reproducible Assessment of Jailbreaking and Privacy Vulnerabilities Across Clinical Specialties
by: Wang, Jinghao, et al.
Published: (2025)
by: Wang, Jinghao, et al.
Published: (2025)
PAuth - Precise Task-Scoped Authorization For Agents
by: Sharma, Reshabh K, et al.
Published: (2026)
by: Sharma, Reshabh K, et al.
Published: (2026)
EXHIB: A Benchmark for Realistic and Diverse Evaluation of Function Similarity in the Wild
by: Fan, Yiming, et al.
Published: (2026)
by: Fan, Yiming, et al.
Published: (2026)
PermRust: A Token-based Permission System for Rust
by: Gehring, Lukas, et al.
Published: (2025)
by: Gehring, Lukas, et al.
Published: (2025)
RustMC: Extending the GenMC stateless model checker to Rust
by: Pearce, Oliver, et al.
Published: (2025)
by: Pearce, Oliver, et al.
Published: (2025)
AC4A: Access Control for Agents
by: Sharma, Reshabh K, et al.
Published: (2026)
by: Sharma, Reshabh K, et al.
Published: (2026)
Translating C To Rust: Lessons from a User Study
by: Li, Ruishi, et al.
Published: (2024)
by: Li, Ruishi, et al.
Published: (2024)
How secure is AI-generated Code: A Large-Scale Comparison of Large Language Models
by: Tihanyi, Norbert, et al.
Published: (2024)
by: Tihanyi, Norbert, et al.
Published: (2024)
CrypTorch: PyTorch-based Auto-tuning Compiler for Machine Learning with Multi-party Computation
by: Liu, Jinyu, et al.
Published: (2025)
by: Liu, Jinyu, et al.
Published: (2025)
Language-Based Agent Control
by: Zhou, Timothy, et al.
Published: (2026)
by: Zhou, Timothy, et al.
Published: (2026)
Semia: Auditing Agent Skills via Constraint-Guided Representation Synthesis
by: Wen, Hongbo, et al.
Published: (2026)
by: Wen, Hongbo, et al.
Published: (2026)
TYPEPULSE: Detecting Type Confusion Bugs in Rust Programs
by: Chen, Hung-Mao, et al.
Published: (2025)
by: Chen, Hung-Mao, et al.
Published: (2025)
Filament: Denning-Style Information Flow Control for Rust
by: Ching, Jeffrey C., et al.
Published: (2026)
by: Ching, Jeffrey C., et al.
Published: (2026)
SafeTrans: LLM-assisted Transpilation from C to Rust
by: Farrukh, Muhammad, et al.
Published: (2025)
by: Farrukh, Muhammad, et al.
Published: (2025)
R1-Fuzz: Specializing Language Models for Textual Fuzzing via Reinforcement Learning
by: Lin, Jiayi, et al.
Published: (2025)
by: Lin, Jiayi, et al.
Published: (2025)
Friend or Foe Inside? Exploring In-Process Isolation to Maintain Memory Safety for Unsafe Rust
by: Gülmez, Merve, et al.
Published: (2023)
by: Gülmez, Merve, et al.
Published: (2023)
Text2VLM: Adapting Text-Only Datasets to Evaluate Alignment Training in Visual Language Models
by: Downer, Gabriel, et al.
Published: (2025)
by: Downer, Gabriel, et al.
Published: (2025)
SafeFFI: Efficient Sanitization at the Boundary Between Safe and Unsafe Code in Rust and Mixed-Language Applications
by: Braunsdorf, Oliver, et al.
Published: (2025)
by: Braunsdorf, Oliver, et al.
Published: (2025)
Static Deadlock Detection for Rust Programs
by: Zhang, Yu, et al.
Published: (2024)
by: Zhang, Yu, et al.
Published: (2024)
LLM for SoC Security: A Paradigm Shift
by: Saha, Dipayan, et al.
Published: (2023)
by: Saha, Dipayan, et al.
Published: (2023)
A Fast, Reliable, and Secure Programming Language for LLM Agents with Code Actions
by: Mell, Stephen, et al.
Published: (2025)
by: Mell, Stephen, et al.
Published: (2025)
s2n-bignum-bench: A practical benchmark for evaluating low-level code reasoning of LLMs
by: Rao, Balaji, et al.
Published: (2026)
by: Rao, Balaji, et al.
Published: (2026)
Hound: Relation-First Knowledge Graphs for Complex-System Reasoning in Security Audits
by: Mueller, Bernhard
Published: (2025)
by: Mueller, Bernhard
Published: (2025)
Fuzzing: Randomness? Reasoning! Efficient Directed Fuzzing via Large Language Models
by: Feng, Xiaotao, et al.
Published: (2025)
by: Feng, Xiaotao, et al.
Published: (2025)
BaxBench: Can LLMs Generate Correct and Secure Backends?
by: Vero, Mark, et al.
Published: (2025)
by: Vero, Mark, et al.
Published: (2025)
Agentic Specification Generator for Move Programs
by: Fu, Yu-Fu, et al.
Published: (2025)
by: Fu, Yu-Fu, et al.
Published: (2025)
AutoBaxBuilder: Bootstrapping Code Security Benchmarking
by: von Arx, Tobias, et al.
Published: (2025)
by: von Arx, Tobias, et al.
Published: (2025)
Arbiter: Detecting Interference in LLM Agent System Prompts
by: Mason, Tony
Published: (2026)
by: Mason, Tony
Published: (2026)
ForensicsData: A Digital Forensics Dataset for Large Language Models
by: Chakir, Youssef, et al.
Published: (2025)
by: Chakir, Youssef, et al.
Published: (2025)
Primus: A Pioneering Collection of Open-Source Datasets for Cybersecurity LLM Training
by: Yu, Yao-Ching, et al.
Published: (2025)
by: Yu, Yao-Ching, et al.
Published: (2025)
An Independent Safety Evaluation of Kimi K2.5
by: Yong, Zheng-Xin, et al.
Published: (2026)
by: Yong, Zheng-Xin, et al.
Published: (2026)
ai.txt: A Domain-Specific Language for Guiding AI Interactions with the Internet
by: Li, Yuekang, et al.
Published: (2025)
by: Li, Yuekang, et al.
Published: (2025)
Agent Tools Orchestration Leaks More: Dataset, Benchmark, and Mitigation
by: Qiao, Yuxuan, et al.
Published: (2025)
by: Qiao, Yuxuan, et al.
Published: (2025)
Unleashing the Unseen: Harnessing Benign Datasets for Jailbreaking Large Language Models
by: Zhao, Wei, et al.
Published: (2024)
by: Zhao, Wei, et al.
Published: (2024)
Acquiring Clean Language Models from Backdoor Poisoned Datasets by Downscaling Frequency Space
by: Wu, Zongru, et al.
Published: (2024)
by: Wu, Zongru, et al.
Published: (2024)
SECOMP: Formally Secure Compilation of Compartmentalized C Programs
by: Thibault, Jérémy, et al.
Published: (2024)
by: Thibault, Jérémy, et al.
Published: (2024)
The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs
by: Liu, Songyang, et al.
Published: (2025)
by: Liu, Songyang, et al.
Published: (2025)
From Description to Score: Can LLMs Quantify Vulnerabilities?
by: Jafarikhah, Sima, et al.
Published: (2025)
by: Jafarikhah, Sima, et al.
Published: (2025)
BELLS: A Framework Towards Future Proof Benchmarks for the Evaluation of LLM Safeguards
by: Dorn, Diego, et al.
Published: (2024)
by: Dorn, Diego, et al.
Published: (2024)
Supporting Artifact Evaluation with LLMs: A Study with Published Security Research Papers
by: Heye, David, et al.
Published: (2026)
by: Heye, David, et al.
Published: (2026)
Similar Items
-
A Practical Framework for Evaluating Medical AI Security: Reproducible Assessment of Jailbreaking and Privacy Vulnerabilities Across Clinical Specialties
by: Wang, Jinghao, et al.
Published: (2025) -
PAuth - Precise Task-Scoped Authorization For Agents
by: Sharma, Reshabh K, et al.
Published: (2026) -
EXHIB: A Benchmark for Realistic and Diverse Evaluation of Function Similarity in the Wild
by: Fan, Yiming, et al.
Published: (2026) -
PermRust: A Token-based Permission System for Rust
by: Gehring, Lukas, et al.
Published: (2025) -
RustMC: Extending the GenMC stateless model checker to Rust
by: Pearce, Oliver, et al.
Published: (2025)