On the Impact of Black-box Deployment Strategies for Edge AI on Latency and Model Performance
Fuente:
arXiv
Salvato in:
| Autori principali: | Singh, Jaskirat, Fallahzadeh, Emad, Adams, Bram, Hassan, Ahmed E. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
On the Impact of White-box Deployment Strategies for Edge AI on Latency and Model Performance
di: Singh, Jaskirat, et al.
Pubblicazione: (2024)
di: Singh, Jaskirat, et al.
Pubblicazione: (2024)
A State-of-the-practice Release-readiness Checklist for Generative AI-based Software Products
di: Patel, Harsh, et al.
Pubblicazione: (2024)
di: Patel, Harsh, et al.
Pubblicazione: (2024)
Predicting the First Response Latency of Maintainers and Contributors in Pull Requests
di: Khatoonabadi, SayedHassan, et al.
Pubblicazione: (2023)
di: Khatoonabadi, SayedHassan, et al.
Pubblicazione: (2023)
Towards Evaluation Engineering: An Empirical Study of ML Evaluation Harnesses in the Wild
di: Zhao, Zhimin, et al.
Pubblicazione: (2026)
di: Zhao, Zhimin, et al.
Pubblicazione: (2026)
Permissive-Washing in the Open AI Supply Chain: A Large-Scale Audit of License Integrity
di: Jewitt, James, et al.
Pubblicazione: (2026)
di: Jewitt, James, et al.
Pubblicazione: (2026)
The Impact of Environment Configurations on the Stability of AI-Enabled Systems
di: Rahman, Musfiqur, et al.
Pubblicazione: (2024)
di: Rahman, Musfiqur, et al.
Pubblicazione: (2024)
Beyond Synthetic Benchmarks: Evaluating LLM Performance on Real-World Class-Level Code Generation
di: Rahman, Musfiqur, et al.
Pubblicazione: (2025)
di: Rahman, Musfiqur, et al.
Pubblicazione: (2025)
HAFix: History-Augmented Large Language Models for Bug Fixing
di: Shi, Yu, et al.
Pubblicazione: (2025)
di: Shi, Yu, et al.
Pubblicazione: (2025)
Automatic Detection of LLM-Generated Code: A Comparative Case Study of Contemporary Models Across Function and Class Granularities
di: Rahman, Musfiqur, et al.
Pubblicazione: (2024)
di: Rahman, Musfiqur, et al.
Pubblicazione: (2024)
Evaluating the Use of LLMs for Documentation to Code Traceability
di: Alor, Ebube, et al.
Pubblicazione: (2025)
di: Alor, Ebube, et al.
Pubblicazione: (2025)
How Robust are LLM-Generated Library Imports? An Empirical Study using Stack Overflow
di: Latendresse, Jasmine, et al.
Pubblicazione: (2025)
di: Latendresse, Jasmine, et al.
Pubblicazione: (2025)
OpenClassGen: A Large-Scale Corpus of Real-World Python Classes for LLM Research
di: Rahman, Musfiqur, et al.
Pubblicazione: (2025)
di: Rahman, Musfiqur, et al.
Pubblicazione: (2025)
Assessing and Improving the Representativeness of Code Generation Benchmarks Using Knowledge Units (KUs) of Programming Languages -- An Empirical Study
di: Ahasanuzzaman, Md, et al.
Pubblicazione: (2026)
di: Ahasanuzzaman, Md, et al.
Pubblicazione: (2026)
AIPC: Agent-Based Automation for AI Model Deployment with Qualcomm AI Runtime
di: Su, Jianhao, et al.
Pubblicazione: (2026)
di: Su, Jianhao, et al.
Pubblicazione: (2026)
FuzzTheREST: An Intelligent Automated Black-box RESTful API Fuzzer
di: Dias, Tiago, et al.
Pubblicazione: (2024)
di: Dias, Tiago, et al.
Pubblicazione: (2024)
Automated File-Level Logging Generation for Machine Learning Applications using LLMs: A Case Study using GPT-4o Mini
di: Rodriguez, Mayra Sofia Ruiz, et al.
Pubblicazione: (2025)
di: Rodriguez, Mayra Sofia Ruiz, et al.
Pubblicazione: (2025)
Is ChatGPT a Good Software Librarian? An Exploratory Study on the Use of ChatGPT for Software Library Recommendations
di: Latendresse, Jasmine, et al.
Pubblicazione: (2024)
di: Latendresse, Jasmine, et al.
Pubblicazione: (2024)
On Wasted Contributions: Understanding the Dynamics of Contributor-Abandoned Pull Requests
di: Khatoonabadi, SayedHassan, et al.
Pubblicazione: (2021)
di: Khatoonabadi, SayedHassan, et al.
Pubblicazione: (2021)
Understanding the Helpfulness of Stale Bot for Pull-based Development: An Empirical Study of 20 Large Open-Source Projects
di: Khatoonabadi, SayedHassan, et al.
Pubblicazione: (2023)
di: Khatoonabadi, SayedHassan, et al.
Pubblicazione: (2023)
HAFixAgent: History-Aware Program Repair Agent
di: Shi, Yu, et al.
Pubblicazione: (2025)
di: Shi, Yu, et al.
Pubblicazione: (2025)
From Hugging Face to GitHub: Tracing License Drift in the Open-Source AI Ecosystem
di: Jewitt, James, et al.
Pubblicazione: (2025)
di: Jewitt, James, et al.
Pubblicazione: (2025)
OmniLLP: Enhancing LLM-based Log Level Prediction with Context-Aware Retrieval
di: Ouatiti, Youssef Esseddiq, et al.
Pubblicazione: (2025)
di: Ouatiti, Youssef Esseddiq, et al.
Pubblicazione: (2025)
Synergizing LLMs and Knowledge Graphs: A Novel Approach to Software Repository-Related Question Answering
di: Abedu, Samuel, et al.
Pubblicazione: (2024)
di: Abedu, Samuel, et al.
Pubblicazione: (2024)
The Rise of AI Teammates in Software Engineering (SE) 3.0: How Autonomous Coding Agents Are Reshaping Software Engineering
di: Li, Hao, et al.
Pubblicazione: (2025)
di: Li, Hao, et al.
Pubblicazione: (2025)
An Empirical Study of Testing Practices in Open Source AI Agent Frameworks and Agentic Applications
di: Hasan, Mohammed Mehedi, et al.
Pubblicazione: (2025)
di: Hasan, Mohammed Mehedi, et al.
Pubblicazione: (2025)
Deploying Geospatial Foundation Models in the Real World: Lessons from WorldCereal
di: Butsko, Christina, et al.
Pubblicazione: (2025)
di: Butsko, Christina, et al.
Pubblicazione: (2025)
Data Quality Antipatterns for Software Analytics
di: Bhatia, Aaditya, et al.
Pubblicazione: (2024)
di: Bhatia, Aaditya, et al.
Pubblicazione: (2024)
An Empirical Study of Challenges in Machine Learning Asset Management
di: Zhao, Zhimin, et al.
Pubblicazione: (2024)
di: Zhao, Zhimin, et al.
Pubblicazione: (2024)
An Approach for Auto Generation of Labeling Functions for Software Engineering Chatbots
di: Alor, Ebube, et al.
Pubblicazione: (2024)
di: Alor, Ebube, et al.
Pubblicazione: (2024)
On the Workflows and Smells of Leaderboard Operations (LBOps): An Exploratory Study of Foundation Model Leaderboards
di: Zhao, Zhimin, et al.
Pubblicazione: (2024)
di: Zhao, Zhimin, et al.
Pubblicazione: (2024)
Model Context Protocol (MCP) at First Glance: Studying the Security and Maintainability of MCP Servers
di: Hasan, Mohammed Mehedi, et al.
Pubblicazione: (2025)
di: Hasan, Mohammed Mehedi, et al.
Pubblicazione: (2025)
PCBSchemaGen: Constraint-Guided Schematic Design via LLM for Printed Circuit Boards (PCB)
di: Zou, Huanghaohe, et al.
Pubblicazione: (2026)
di: Zou, Huanghaohe, et al.
Pubblicazione: (2026)
Intuition to Evidence: Measuring AI's True Impact on Developer Productivity
di: Kumar, Anand, et al.
Pubblicazione: (2025)
di: Kumar, Anand, et al.
Pubblicazione: (2025)
An Empirical Evaluation of Locally Deployed LLMs for Bug Detection in Python Code
di: Vulićević, Jelena Ilić
Pubblicazione: (2026)
di: Vulićević, Jelena Ilić
Pubblicazione: (2026)
A Regression Framework for Understanding Prompt Component Impact on LLM Performance
di: Lauziere, Andrew, et al.
Pubblicazione: (2026)
di: Lauziere, Andrew, et al.
Pubblicazione: (2026)
CSR-Bench: Benchmarking LLM Agents in Deployment of Computer Science Research Repositories
di: Xiao, Yijia, et al.
Pubblicazione: (2025)
di: Xiao, Yijia, et al.
Pubblicazione: (2025)
AgentTrace: Causal Graph Tracing for Root Cause Analysis in Deployed Multi-Agent Systems
di: Wang, Zhaohui Geoffrey
Pubblicazione: (2026)
di: Wang, Zhaohui Geoffrey
Pubblicazione: (2026)
Understanding Prompt Management in GitHub Repositories: A Call for Best Practices
di: Li, Hao, et al.
Pubblicazione: (2025)
di: Li, Hao, et al.
Pubblicazione: (2025)
The Hitchhikers Guide to Production-ready Trustworthy Foundation Model powered Software (FMware)
di: Vasilevski, Kirill, et al.
Pubblicazione: (2025)
di: Vasilevski, Kirill, et al.
Pubblicazione: (2025)
What's documented in AI? Systematic Analysis of 32K AI Model Cards
di: Liang, Weixin, et al.
Pubblicazione: (2024)
di: Liang, Weixin, et al.
Pubblicazione: (2024)
Documenti analoghi
-
On the Impact of White-box Deployment Strategies for Edge AI on Latency and Model Performance
di: Singh, Jaskirat, et al.
Pubblicazione: (2024) -
A State-of-the-practice Release-readiness Checklist for Generative AI-based Software Products
di: Patel, Harsh, et al.
Pubblicazione: (2024) -
Predicting the First Response Latency of Maintainers and Contributors in Pull Requests
di: Khatoonabadi, SayedHassan, et al.
Pubblicazione: (2023) -
Towards Evaluation Engineering: An Empirical Study of ML Evaluation Harnesses in the Wild
di: Zhao, Zhimin, et al.
Pubblicazione: (2026) -
Permissive-Washing in the Open AI Supply Chain: A Large-Scale Audit of License Integrity
di: Jewitt, James, et al.
Pubblicazione: (2026)