Rigorous Assessment of Model Inference Accuracy using Language Cardinality
Fuente:
arXiv
Guardado en:
| Autores principales: | Clun, Donato, Shin, Donghwan, Filieri, Antonio, Bianculli, Domenico |
|---|---|
| Formato: | Preprint |
| Publicado: |
2022
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Towards a Taxonomy of Software Log Smells
por: Saarimäki, Nyyti, et al.
Publicado: (2024)
por: Saarimäki, Nyyti, et al.
Publicado: (2024)
Impact of Log Parsing on Deep Learning-Based Anomaly Detection
por: Khan, Zanis Ali, et al.
Publicado: (2023)
por: Khan, Zanis Ali, et al.
Publicado: (2023)
Systematic Evaluation of Deep Learning Models for Log-based Failure Prediction
por: Hadadi, Fatemeh, et al.
Publicado: (2023)
por: Hadadi, Fatemeh, et al.
Publicado: (2023)
Towards Generating Executable Metamorphic Relations Using Large Language Models
por: Shin, Seung Yeob, et al.
Publicado: (2024)
por: Shin, Seung Yeob, et al.
Publicado: (2024)
Quantum Program Linting with LLMs: Emerging Results from a Comparative Study
por: Shin, Seung Yeob, et al.
Publicado: (2025)
por: Shin, Seung Yeob, et al.
Publicado: (2025)
A Piece of QAICCC: Towards a Countermeasure Against Crosstalk Attacks in Quantum Servers
por: Marquer, Yoann, et al.
Publicado: (2025)
por: Marquer, Yoann, et al.
Publicado: (2025)
Randomized and Diverse Input State Generation for Quantum Program Testing
por: Ernzer, Maryse, et al.
Publicado: (2026)
por: Ernzer, Maryse, et al.
Publicado: (2026)
Diagnosing Violations of State-based Specifications in iCFTL
por: Stratan, Cristina, et al.
Publicado: (2025)
por: Stratan, Cristina, et al.
Publicado: (2025)
Stress Testing Control Loops in Cyber-Physical Systems
por: Mandrioli, Claudio, et al.
Publicado: (2023)
por: Mandrioli, Claudio, et al.
Publicado: (2023)
Beyond Rules: LLM-Powered Linting for Quantum Programs
por: Cassieri, Pietro, et al.
Publicado: (2026)
por: Cassieri, Pietro, et al.
Publicado: (2026)
Testing CPS with Design Assumptions-Based Metamorphic Relations and Genetic Programming
por: Mandrioli, Claudio, et al.
Publicado: (2024)
por: Mandrioli, Claudio, et al.
Publicado: (2024)
Trace Diagnostics for Signal-based Temporal Properties
por: Boufaied, Chaima, et al.
Publicado: (2022)
por: Boufaied, Chaima, et al.
Publicado: (2022)
LLM meets ML: Data-efficient Anomaly Detection on Unstable Logs
por: Hadadi, Fatemeh, et al.
Publicado: (2024)
por: Hadadi, Fatemeh, et al.
Publicado: (2024)
Can LLMs Solve Science or Just Write Code? Evaluating Quantum Solver Generation
por: Baresi, Luciano, et al.
Publicado: (2026)
por: Baresi, Luciano, et al.
Publicado: (2026)
Lachesis: Predicting LLM Inference Accuracy using Structural Properties of Reasoning Paths
por: Kim, Naryeong, et al.
Publicado: (2024)
por: Kim, Naryeong, et al.
Publicado: (2024)
A Machine Learning Approach for Automated Filling of Categorical Fields in Data Entry Forms
por: Belgacem, Hichem, et al.
Publicado: (2022)
por: Belgacem, Hichem, et al.
Publicado: (2022)
Learning-Based Relaxation of Completeness Requirements for Data Entry Forms
por: Belgacem, Hichem, et al.
Publicado: (2023)
por: Belgacem, Hichem, et al.
Publicado: (2023)
Using Causal Inference to Test Systems with Hidden and Interacting Variables: An Evaluative Case Study
por: Foster, Michael, et al.
Publicado: (2025)
por: Foster, Michael, et al.
Publicado: (2025)
SAGA: Detecting Security Vulnerabilities Using Static Aspect Analysis
por: Marquer, Yoann, et al.
Publicado: (2026)
por: Marquer, Yoann, et al.
Publicado: (2026)
Detecting Multiple Semantic Concerns in Tangled Code Commits
por: Koh, Beomsu, et al.
Publicado: (2026)
por: Koh, Beomsu, et al.
Publicado: (2026)
Mutation-based Consistency Testing for Evaluating the Code Understanding Capability of LLMs
por: Li, Ziyu, et al.
Publicado: (2024)
por: Li, Ziyu, et al.
Publicado: (2024)
Software Testing for Extended Reality Applications: A Systematic Mapping Study
por: Gu, Ruizhen, et al.
Publicado: (2025)
por: Gu, Ruizhen, et al.
Publicado: (2025)
Automated Testing of Prevalent 3D User Interactions in Virtual Reality Applications
por: Gu, Ruizhen, et al.
Publicado: (2026)
por: Gu, Ruizhen, et al.
Publicado: (2026)
GDPR-Relevant Privacy Concerns in Mobile Apps Research: A Systematic Literature Review
por: Cejas, Orlando Amaral, et al.
Publicado: (2024)
por: Cejas, Orlando Amaral, et al.
Publicado: (2024)
A Systematic Mapping Study on the Debugging of Autonomous Driving Systems
por: Shaw, Nathan, et al.
Publicado: (2026)
por: Shaw, Nathan, et al.
Publicado: (2026)
A Comprehensive Study of Machine Learning Techniques for Log-Based Anomaly Detection
por: Ali, Shan, et al.
Publicado: (2023)
por: Ali, Shan, et al.
Publicado: (2023)
Pricing4APIs: A Rigorous Model for RESTful API Pricings
por: Fresno-Aranda, Rafael, et al.
Publicado: (2023)
por: Fresno-Aranda, Rafael, et al.
Publicado: (2023)
VulStamp: Vulnerability Assessment using Large Language Model
por: Shen, Hao, et al.
Publicado: (2025)
por: Shen, Hao, et al.
Publicado: (2025)
Language Models in Software Development Tasks: An Experimental Analysis of Energy and Accuracy
por: Alizadeh, Negar, et al.
Publicado: (2024)
por: Alizadeh, Negar, et al.
Publicado: (2024)
Mapping Cardinality-based Feature Models to Weighted Automata over Featured Multiset Semirings (Extended Version)
por: Müller, Robert, et al.
Publicado: (2024)
por: Müller, Robert, et al.
Publicado: (2024)
PLACIDUS: Engineering Product Lines of Rigorous Assurance Cases
por: Murphy, Logan, et al.
Publicado: (2024)
por: Murphy, Logan, et al.
Publicado: (2024)
DeVAIC: A Tool for Security Assessment of AI-generated Code
por: Cotroneo, Domenico, et al.
Publicado: (2024)
por: Cotroneo, Domenico, et al.
Publicado: (2024)
Neural Fault Injection: Generating Software Faults from Natural Language
por: Cotroneo, Domenico, et al.
Publicado: (2024)
por: Cotroneo, Domenico, et al.
Publicado: (2024)
PatchGuru: Patch Oracle Inference from Natural Language Artifacts with Large Language Models
por: Le-Cong, Thanh, et al.
Publicado: (2026)
por: Le-Cong, Thanh, et al.
Publicado: (2026)
Smart but Costly? Benchmarking LLMs on Functional Accuracy and Energy Efficiency
por: Mehditabar, Mohammadjavad, et al.
Publicado: (2025)
por: Mehditabar, Mohammadjavad, et al.
Publicado: (2025)
Human-Aligned Code Readability Assessment with Large Language Models
por: Ouédraogo, Wendkûuni C., et al.
Publicado: (2025)
por: Ouédraogo, Wendkûuni C., et al.
Publicado: (2025)
Quality Assessment of Python Tests Generated by Large Language Models
por: Alves, Victor, et al.
Publicado: (2025)
por: Alves, Victor, et al.
Publicado: (2025)
Characterizing the Usefulness of Code Review Comments in Scientific Software for Software Quality and Scientific Rigor
por: Ahmed, Sharif, et al.
Publicado: (2026)
por: Ahmed, Sharif, et al.
Publicado: (2026)
CERT: Finding Performance Issues in Database Systems Through the Lens of Cardinality Estimation
por: Ba, Jinsheng, et al.
Publicado: (2023)
por: Ba, Jinsheng, et al.
Publicado: (2023)
Towards Understanding Bugs in Distributed Training and Inference Frameworks for Large Language Models
por: Yu, Xiao, et al.
Publicado: (2025)
por: Yu, Xiao, et al.
Publicado: (2025)
Ejemplares similares
-
Towards a Taxonomy of Software Log Smells
por: Saarimäki, Nyyti, et al.
Publicado: (2024) -
Impact of Log Parsing on Deep Learning-Based Anomaly Detection
por: Khan, Zanis Ali, et al.
Publicado: (2023) -
Systematic Evaluation of Deep Learning Models for Log-based Failure Prediction
por: Hadadi, Fatemeh, et al.
Publicado: (2023) -
Towards Generating Executable Metamorphic Relations Using Large Language Models
por: Shin, Seung Yeob, et al.
Publicado: (2024) -
Quantum Program Linting with LLMs: Emerging Results from a Comparative Study
por: Shin, Seung Yeob, et al.
Publicado: (2025)