Instance-Level Safety-Aware Fidelity of Synthetic Data and Its Calibration
Fuente:
arXiv
Salvato in:
| Autori principali: | Cheng, Chih-Hong, Stöckel, Paul, Zhao, Xingyu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
What, Indeed, is an Achievable Provable Guarantee for Learning-Enabled Safety Critical Systems
di: Bensalem, Saddek, et al.
Pubblicazione: (2023)
di: Bensalem, Saddek, et al.
Pubblicazione: (2023)
SMARTCAL: An Approach to Self-Aware Tool-Use Evaluation and Calibration
di: Shen, Yuanhao, et al.
Pubblicazione: (2024)
di: Shen, Yuanhao, et al.
Pubblicazione: (2024)
Beyond Synthetic Benchmarks: Evaluating LLM Performance on Real-World Class-Level Code Generation
di: Rahman, Musfiqur, et al.
Pubblicazione: (2025)
di: Rahman, Musfiqur, et al.
Pubblicazione: (2025)
Runtime Monitoring and Enforcement of Conditional Fairness in Generative AIs
di: Cheng, Chih-Hong, et al.
Pubblicazione: (2024)
di: Cheng, Chih-Hong, et al.
Pubblicazione: (2024)
From Hazard Identification to Controller Design: Proactive and LLM-Supported Safety Engineering for ML-Powered Systems
di: Hong, Yining, et al.
Pubblicazione: (2025)
di: Hong, Yining, et al.
Pubblicazione: (2025)
Operational Robustness of LLMs on Code Generation
di: Paul, Debalina Ghosh, et al.
Pubblicazione: (2026)
di: Paul, Debalina Ghosh, et al.
Pubblicazione: (2026)
SynthTools: A Framework for Scaling Synthetic Tools for Agent Development
di: Castellani, Tommaso, et al.
Pubblicazione: (2025)
di: Castellani, Tommaso, et al.
Pubblicazione: (2025)
Standardization Trends on Safety and Trustworthiness Technology for Advanced AI
di: Jeon, Jonghong
Pubblicazione: (2024)
di: Jeon, Jonghong
Pubblicazione: (2024)
The Causal Impact of Tool Affordance on Safety Alignment in LLM Agents
di: Yu, Shasha, et al.
Pubblicazione: (2026)
di: Yu, Shasha, et al.
Pubblicazione: (2026)
SMARLA: A Safety Monitoring Approach for Deep Reinforcement Learning Agents
di: Zolfagharian, Amirhossein, et al.
Pubblicazione: (2023)
di: Zolfagharian, Amirhossein, et al.
Pubblicazione: (2023)
On The Importance of Reasoning for Context Retrieval in Repository-Level Code Editing
di: Kovrigin, Alexander, et al.
Pubblicazione: (2024)
di: Kovrigin, Alexander, et al.
Pubblicazione: (2024)
Insights Generator: Systematic Corpus-Level Trace Diagnostics for LLM Agents
di: Manglik, Akshay, et al.
Pubblicazione: (2026)
di: Manglik, Akshay, et al.
Pubblicazione: (2026)
CodeTaste: Can LLMs Generate Human-Level Code Refactorings?
di: Thillen, Alex, et al.
Pubblicazione: (2026)
di: Thillen, Alex, et al.
Pubblicazione: (2026)
ProToken: Token-Level Attribution for Federated Large Language Models
di: Gill, Waris, et al.
Pubblicazione: (2026)
di: Gill, Waris, et al.
Pubblicazione: (2026)
GREPO: A Benchmark for Graph Neural Networks on Repository-Level Bug Localization
di: Wang, Juntong, et al.
Pubblicazione: (2026)
di: Wang, Juntong, et al.
Pubblicazione: (2026)
MobiFlow: Real-World Mobile Agent Benchmarking through Trajectory Fusion
di: Feng, Yunfei, et al.
Pubblicazione: (2026)
di: Feng, Yunfei, et al.
Pubblicazione: (2026)
Teaching an Online Multi-Institutional Research Level Software Engineering Course with Industry -- an Experience Report
di: Jalote, Pankaj, et al.
Pubblicazione: (2025)
di: Jalote, Pankaj, et al.
Pubblicazione: (2025)
MASTEST: A LLM-Based Multi-Agent System For RESTful API Tests
di: Han, Xiaoke, et al.
Pubblicazione: (2025)
di: Han, Xiaoke, et al.
Pubblicazione: (2025)
Keeping Code-Aware LLMs Fresh: Full Refresh, In-Context Deltas, and Incremental Fine-Tuning
di: Sharma, Pradeep Kumar, et al.
Pubblicazione: (2025)
di: Sharma, Pradeep Kumar, et al.
Pubblicazione: (2025)
ZnTrack -- Data as Code
di: Zills, Fabian, et al.
Pubblicazione: (2024)
di: Zills, Fabian, et al.
Pubblicazione: (2024)
Hammer: Robust Function-Calling for On-Device Language Models via Function Masking
di: Lin, Qiqiang, et al.
Pubblicazione: (2024)
di: Lin, Qiqiang, et al.
Pubblicazione: (2024)
Assuring the Safety of Reinforcement Learning Components: AMLAS-RL
di: Imrie, Calum Corrie, et al.
Pubblicazione: (2025)
di: Imrie, Calum Corrie, et al.
Pubblicazione: (2025)
Relative Positioning Based Code Chunking Method For Rich Context Retrieval In Repository Level Code Completion Task With Code Language Model
di: Rahman, Imranur, et al.
Pubblicazione: (2025)
di: Rahman, Imranur, et al.
Pubblicazione: (2025)
Automated File-Level Logging Generation for Machine Learning Applications using LLMs: A Case Study using GPT-4o Mini
di: Rodriguez, Mayra Sofia Ruiz, et al.
Pubblicazione: (2025)
di: Rodriguez, Mayra Sofia Ruiz, et al.
Pubblicazione: (2025)
On the Need for a Statistical Foundation in Scenario-Based Testing of Autonomous Vehicles
di: Zhao, Xingyu, et al.
Pubblicazione: (2025)
di: Zhao, Xingyu, et al.
Pubblicazione: (2025)
Generative AI to Generate Test Data Generators
di: Baudry, Benoit, et al.
Pubblicazione: (2024)
di: Baudry, Benoit, et al.
Pubblicazione: (2024)
System Safety Monitoring of Learned Components Using Temporal Metric Forecasting
di: Sharifi, Sepehr, et al.
Pubblicazione: (2024)
di: Sharifi, Sepehr, et al.
Pubblicazione: (2024)
Leveraging XP and CRISP-DM for Agile Data Science Projects
di: Shimaoka, Andre Massahiro, et al.
Pubblicazione: (2025)
di: Shimaoka, Andre Massahiro, et al.
Pubblicazione: (2025)
SWE-MiniSandbox: Container-Free Reinforcement Learning for Building Software Engineering Agents
di: Yuan, Danlong, et al.
Pubblicazione: (2026)
di: Yuan, Danlong, et al.
Pubblicazione: (2026)
Deep Learning and Data Augmentation for Detecting Self-Admitted Technical Debt
di: Sutoyo, Edi, et al.
Pubblicazione: (2024)
di: Sutoyo, Edi, et al.
Pubblicazione: (2024)
daVinci-Agency: Unlocking Long-Horizon Agency Data-Efficiently
di: Jiang, Mohan, et al.
Pubblicazione: (2026)
di: Jiang, Mohan, et al.
Pubblicazione: (2026)
Workflow-Level Design Principles for Trustworthy GenAI in Automotive System Engineering
di: Cheng, Chih-Hong, et al.
Pubblicazione: (2026)
di: Cheng, Chih-Hong, et al.
Pubblicazione: (2026)
Machine Learning Systems: A Survey from a Data-Oriented Perspective
di: Cabrera, Christian, et al.
Pubblicazione: (2023)
di: Cabrera, Christian, et al.
Pubblicazione: (2023)
SLIM: a Scalable Light-weight Root Cause Analysis for Imbalanced Data in Microservice
di: Ren, Rui, et al.
Pubblicazione: (2024)
di: Ren, Rui, et al.
Pubblicazione: (2024)
Function+Data Flow: A Framework to Specify Machine Learning Pipelines for Digital Twinning
di: de Conto, Eduardo, et al.
Pubblicazione: (2024)
di: de Conto, Eduardo, et al.
Pubblicazione: (2024)
Beyond Output Correctness: Benchmarking and Evaluating Large Language Model Reasoning in Coding Tasks
di: Li, Yuangang, et al.
Pubblicazione: (2026)
di: Li, Yuangang, et al.
Pubblicazione: (2026)
Data-Driven Methods and AI in Engineering Design: A Systematic Literature Review Focusing on Challenges and Opportunities
di: Afifi, Nehal, et al.
Pubblicazione: (2025)
di: Afifi, Nehal, et al.
Pubblicazione: (2025)
ADReFT: Adaptive Decision Repair for Safe Autonomous Driving via Reinforcement Fine-Tuning
di: Cheng, Mingfei, et al.
Pubblicazione: (2025)
di: Cheng, Mingfei, et al.
Pubblicazione: (2025)
Gradient-Based Model Fingerprinting for LLM Similarity Detection and Family Classification
di: Wu, Zehao, et al.
Pubblicazione: (2025)
di: Wu, Zehao, et al.
Pubblicazione: (2025)
Scoring Verifiers: Evaluating Synthetic Verification for Code and Reasoning
di: Ficek, Aleksander, et al.
Pubblicazione: (2025)
di: Ficek, Aleksander, et al.
Pubblicazione: (2025)
Documenti analoghi
-
What, Indeed, is an Achievable Provable Guarantee for Learning-Enabled Safety Critical Systems
di: Bensalem, Saddek, et al.
Pubblicazione: (2023) -
SMARTCAL: An Approach to Self-Aware Tool-Use Evaluation and Calibration
di: Shen, Yuanhao, et al.
Pubblicazione: (2024) -
Beyond Synthetic Benchmarks: Evaluating LLM Performance on Real-World Class-Level Code Generation
di: Rahman, Musfiqur, et al.
Pubblicazione: (2025) -
Runtime Monitoring and Enforcement of Conditional Fairness in Generative AIs
di: Cheng, Chih-Hong, et al.
Pubblicazione: (2024) -
From Hazard Identification to Controller Design: Proactive and LLM-Supported Safety Engineering for ML-Powered Systems
di: Hong, Yining, et al.
Pubblicazione: (2025)