Overhead Measurement Noise in Different Runtime Environments
Fuente:
arXiv
Guardado en:
| Autores principales: | Reichelt, David Georg, Jung, Reiner, van Hoorn, André |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Detection of Performance Changes in MooBench Results Using Nyrkiö on GitHub Actions
por: Yang, Shinhyung, et al.
Publicado: (2025)
por: Yang, Shinhyung, et al.
Publicado: (2025)
OODEval: Evaluating Large Language Models on Object-Oriented Design
por: Xiao, Bingxu, et al.
Publicado: (2026)
por: Xiao, Bingxu, et al.
Publicado: (2026)
How Quickly Do Development Teams Update Their Vulnerable Dependencies?
por: Rahman, Imranur, et al.
Publicado: (2024)
por: Rahman, Imranur, et al.
Publicado: (2024)
Comprehensive Evaluation of Large Language Models on Software Engineering Tasks: A Multi-Task Benchmark
por: Gunawan, Go Frendi, et al.
Publicado: (2026)
por: Gunawan, Go Frendi, et al.
Publicado: (2026)
Reliability of AI Bots Footprints in GitHub Actions CI/CD Workflows
por: Shah, Syed Muhammad Ashhar, et al.
Publicado: (2026)
por: Shah, Syed Muhammad Ashhar, et al.
Publicado: (2026)
Analyzing the Adoption of Database Management Systems Throughout the History of Open Source Projects
por: Paiva, Camila A., et al.
Publicado: (2026)
por: Paiva, Camila A., et al.
Publicado: (2026)
Evaluating the Overhead of the Performance Profiler Cloudprofiler With MooBench
por: Yang, Shinhyung, et al.
Publicado: (2024)
por: Yang, Shinhyung, et al.
Publicado: (2024)
Token Arena: A Continuous Benchmark Unifying Energy and Cognition in AI Inference
por: Gao, Yuxuan, et al.
Publicado: (2026)
por: Gao, Yuxuan, et al.
Publicado: (2026)
The Kieker Observability Framework Version 2
por: Yang, Shinhyung, et al.
Publicado: (2025)
por: Yang, Shinhyung, et al.
Publicado: (2025)
Source Code Hotspots: A Diagnostic Method for Quality Issues
por: Muzammil, Saleha, et al.
Publicado: (2026)
por: Muzammil, Saleha, et al.
Publicado: (2026)
Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study
por: Alshaikh, Moaath, et al.
Publicado: (2026)
por: Alshaikh, Moaath, et al.
Publicado: (2026)
ConfProBench: A Confidence Evaluation Benchmark for MLLM-Based Process Judges
por: Zhou, Yue, et al.
Publicado: (2025)
por: Zhou, Yue, et al.
Publicado: (2025)
Interoperability From Kieker to OpenTelemetry: Demonstrated as Export to ExplorViz
por: Reichelt, David Georg, et al.
Publicado: (2024)
por: Reichelt, David Georg, et al.
Publicado: (2024)
The Single-File Test: A Longitudinal Public-Interface Evaluation of First-Output LLM Web Generation with Social Reach Tracking
por: Palacios, Diego Cabezas
Publicado: (2026)
por: Palacios, Diego Cabezas
Publicado: (2026)
Factors that Contribute to the Success of a Software Organisation's DevOps Environment: A Systematic Review
por: Gwangwadza, Ashley, et al.
Publicado: (2022)
por: Gwangwadza, Ashley, et al.
Publicado: (2022)
GEML: A Grammar-based Evolutionary Machine Learning Approach for Design-Pattern Detection
por: Barbudo, Rafael, et al.
Publicado: (2024)
por: Barbudo, Rafael, et al.
Publicado: (2024)
Feedback-Normalized Developer Memory for Reinforcement-Learning Coding Agents: A Safety-Gated MCP Architecture
por: Iscan, Mehmet
Publicado: (2026)
por: Iscan, Mehmet
Publicado: (2026)
Cost-Aware Logging: Measuring the Financial Impact of Excessive Log Retention in Small-Scale Cloud Deployments
por: Putra, Jody Almaida
Publicado: (2026)
por: Putra, Jody Almaida
Publicado: (2026)
Structured Prompting and Feedback-Guided Reasoning with LLMs for Data Interpretation
por: Rath, Amit
Publicado: (2025)
por: Rath, Amit
Publicado: (2025)
How Generation Architecture Shapes Code Complexity in Multi-Agent LLM Systems: A Paired Study on HumanEval
por: Ashrafi, Nazmus
Publicado: (2026)
por: Ashrafi, Nazmus
Publicado: (2026)
Binary BPE: A Family of Cross-Platform Tokenizers for Binary Analysis
por: Bommarito II, Michael J.
Publicado: (2025)
por: Bommarito II, Michael J.
Publicado: (2025)
Predictive Analytics for Collaborators Answers, Code Quality, and Dropout on Stack Overflow
por: Zolduoarrati, Elijah, et al.
Publicado: (2025)
por: Zolduoarrati, Elijah, et al.
Publicado: (2025)
AutoBench: Automating LLM Evaluation through Reciprocal Peer Assessment
por: Loi, Dario, et al.
Publicado: (2025)
por: Loi, Dario, et al.
Publicado: (2025)
Rust vs. C for Python Libraries: Evaluating Rust-Compatible Bindings Toolchains
por: Amaral, Isabella Basso do, et al.
Publicado: (2025)
por: Amaral, Isabella Basso do, et al.
Publicado: (2025)
An Analysis of XML Compression Efficiency
por: Augeri, Christopher James, et al.
Publicado: (2024)
por: Augeri, Christopher James, et al.
Publicado: (2024)
VulCPE: Context-Aware Cybersecurity Vulnerability Retrieval and Management
por: Jiang, Yuning, et al.
Publicado: (2025)
por: Jiang, Yuning, et al.
Publicado: (2025)
Spark-LLM-Eval: A Distributed Framework for Statistically Rigorous Large Language Model Evaluation
por: Mitra, Subhadip
Publicado: (2026)
por: Mitra, Subhadip
Publicado: (2026)
ChatGPT4PCG 2 Competition: Prompt Engineering for Science Birds Level Generation
por: Taveekitworachai, Pittawat, et al.
Publicado: (2024)
por: Taveekitworachai, Pittawat, et al.
Publicado: (2024)
HELEA: Hard-Negative Benchmark and LLM-based Reranking for Robust Entity Alignment
por: Jang, Yoonjin, et al.
Publicado: (2026)
por: Jang, Yoonjin, et al.
Publicado: (2026)
TrafficRAG: A Multimodal RAG Framework for Traffic Accident Liability Determination
por: Li, Xu, et al.
Publicado: (2026)
por: Li, Xu, et al.
Publicado: (2026)
LLMs as Idiomatic Decompilers: Recovering High-Level Code from x86-64 Assembly for Dart
por: Abualazm, Raafat, et al.
Publicado: (2026)
por: Abualazm, Raafat, et al.
Publicado: (2026)
How Clued up are LLMs? Evaluating Multi-Step Deductive Reasoning in a Text-Based Game Environment
por: Ansell, Rebecca, et al.
Publicado: (2026)
por: Ansell, Rebecca, et al.
Publicado: (2026)
Beyond Greenfield: The D3 Framework for AI-Driven Productivity in Brownfield Engineering
por: Sharma, Krishna Kumaar
Publicado: (2025)
por: Sharma, Krishna Kumaar
Publicado: (2025)
Generative AI and the Transformation of Software Development Practices
por: Acharya, Vivek
Publicado: (2025)
por: Acharya, Vivek
Publicado: (2025)
Stabilization Without Simplification: A Two-Dimensional Model of Software Evolution
por: Furukawa, Masaru
Publicado: (2026)
por: Furukawa, Masaru
Publicado: (2026)
From Monolith to Microservices: A Comparative Evaluation of Decomposition Frameworks
por: Weerasinghe, Mineth, et al.
Publicado: (2026)
por: Weerasinghe, Mineth, et al.
Publicado: (2026)
SAGAI-MID: A Generative AI-Driven Middleware for Dynamic Runtime Interoperability
por: Larsen, Oliver Aleksander, et al.
Publicado: (2026)
por: Larsen, Oliver Aleksander, et al.
Publicado: (2026)
Energy-Aware Decision Making in Software Stack Upgrades
por: Stocker, Mirko, et al.
Publicado: (2026)
por: Stocker, Mirko, et al.
Publicado: (2026)
GitHub Copilot and Developer Productivity: An Observational Dose-Response Analysis
por: Heilman, Alex, et al.
Publicado: (2026)
por: Heilman, Alex, et al.
Publicado: (2026)
Comparing Human and LLM Generated Code: The Jury is Still Out!
por: Licorish, Sherlock A., et al.
Publicado: (2025)
por: Licorish, Sherlock A., et al.
Publicado: (2025)
Ejemplares similares
-
Detection of Performance Changes in MooBench Results Using Nyrkiö on GitHub Actions
por: Yang, Shinhyung, et al.
Publicado: (2025) -
OODEval: Evaluating Large Language Models on Object-Oriented Design
por: Xiao, Bingxu, et al.
Publicado: (2026) -
How Quickly Do Development Teams Update Their Vulnerable Dependencies?
por: Rahman, Imranur, et al.
Publicado: (2024) -
Comprehensive Evaluation of Large Language Models on Software Engineering Tasks: A Multi-Task Benchmark
por: Gunawan, Go Frendi, et al.
Publicado: (2026) -
Reliability of AI Bots Footprints in GitHub Actions CI/CD Workflows
por: Shah, Syed Muhammad Ashhar, et al.
Publicado: (2026)