A Framework for Effective Invocation Methods of Various LLM Services
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Can, Sui, Dianbo, Zhang, Bolin, Liu, Xiaoyu, Kang, Jiabao, Qiao, Zhidong, Tu, Zhiying |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Carbon-aware Software Services
di: Forti, Stefano, et al.
Pubblicazione: (2024)
di: Forti, Stefano, et al.
Pubblicazione: (2024)
CARISMA: CAR-Integrated Service Mesh Architecture
di: Klein, Kevin, et al.
Pubblicazione: (2024)
di: Klein, Kevin, et al.
Pubblicazione: (2024)
GitFarm: Git as a Service for Large-Scale Monorepos
di: Dwivedi, Preetam, et al.
Pubblicazione: (2026)
di: Dwivedi, Preetam, et al.
Pubblicazione: (2026)
LLM-HPC++: Evaluating LLM-Generated Modern C++ and MPI+OpenMP Codes for Scalable Mandelbrot Set Computation
di: Diehl, Patrick, et al.
Pubblicazione: (2025)
di: Diehl, Patrick, et al.
Pubblicazione: (2025)
Cost-Effective Big Data Orchestration Using Dagster: A Multi-Platform Approach
di: Picatto, Hernan, et al.
Pubblicazione: (2024)
di: Picatto, Hernan, et al.
Pubblicazione: (2024)
LLM4FaaS: No-Code Application Development using LLMs and FaaS
di: Wang, Minghe, et al.
Pubblicazione: (2025)
di: Wang, Minghe, et al.
Pubblicazione: (2025)
An SLO Driven and Cost-Aware Autoscaling Framework for Kubernetes
di: Punniyamoorthy, Vinoth, et al.
Pubblicazione: (2025)
di: Punniyamoorthy, Vinoth, et al.
Pubblicazione: (2025)
L4: Diagnosing Large-scale LLM Training Failures via Automated Log Analysis
di: Jiang, Zhihan, et al.
Pubblicazione: (2025)
di: Jiang, Zhihan, et al.
Pubblicazione: (2025)
AdaptiFlow: An Extensible Framework for Event-Driven Autonomy in Cloud Microservices
di: Ndadji, Brice Arléon Zemtsop, et al.
Pubblicazione: (2025)
di: Ndadji, Brice Arléon Zemtsop, et al.
Pubblicazione: (2025)
A Comprehensive Benchmarking Analysis of Fault Recovery in Stream Processing Frameworks
di: Vogel, Adriano, et al.
Pubblicazione: (2024)
di: Vogel, Adriano, et al.
Pubblicazione: (2024)
A Unifying Framework to Enable Artificial Intelligence in High Performance Computing Workflows
di: Domke, Jens, et al.
Pubblicazione: (2025)
di: Domke, Jens, et al.
Pubblicazione: (2025)
FMI Meets SystemC: A Framework for Cross-Tool Virtual Prototyping
di: Bosbach, Nils, et al.
Pubblicazione: (2025)
di: Bosbach, Nils, et al.
Pubblicazione: (2025)
A Comprehensive Experimentation Framework for Energy-Efficient Design of Cloud-Native Applications
di: Werner, Sebastian, et al.
Pubblicazione: (2025)
di: Werner, Sebastian, et al.
Pubblicazione: (2025)
ShuffleBench: A Benchmark for Large-Scale Data Shuffling Operations with Distributed Stream Processing Frameworks
di: Henning, Sören, et al.
Pubblicazione: (2024)
di: Henning, Sören, et al.
Pubblicazione: (2024)
CLAID: Closing the Loop on AI & Data Collection -- A Cross-Platform Transparent Computing Middleware Framework for Smart Edge-Cloud and Digital Biomarker Applications
di: Langer, Patrick, et al.
Pubblicazione: (2023)
di: Langer, Patrick, et al.
Pubblicazione: (2023)
Supporting Long-term Transactions in Smart Contracts Generated from Business Process Model and Notation (BPMN) Models
di: Liu, Christian Gang
Pubblicazione: (2025)
di: Liu, Christian Gang
Pubblicazione: (2025)
FSM Modeling For Off-Blockchain Computation
di: Liu, Christian Gang
Pubblicazione: (2025)
di: Liu, Christian Gang
Pubblicazione: (2025)
CSnake: Detecting Self-Sustaining Cascading Failure via Causal Stitching of Fault Propagations
di: Qian, Shangshu, et al.
Pubblicazione: (2025)
di: Qian, Shangshu, et al.
Pubblicazione: (2025)
A Reference Architecture for Governance of Cloud Native Applications
di: Pourmajidi, William, et al.
Pubblicazione: (2023)
di: Pourmajidi, William, et al.
Pubblicazione: (2023)
MegaFlow: Large-Scale Distributed Orchestration System for the Agentic Era
di: Zhang, Lei, et al.
Pubblicazione: (2026)
di: Zhang, Lei, et al.
Pubblicazione: (2026)
Wherefore Art Thou? Provenance-Guided Automatic Online Debugging with Lumos
di: Chen, Jingyuan, et al.
Pubblicazione: (2026)
di: Chen, Jingyuan, et al.
Pubblicazione: (2026)
Supercharging Federated Learning with Flower and NVIDIA FLARE
di: Roth, Holger R., et al.
Pubblicazione: (2024)
di: Roth, Holger R., et al.
Pubblicazione: (2024)
SeBS-Flow: Benchmarking Serverless Cloud Function Workflows
di: Schmid, Larissa, et al.
Pubblicazione: (2024)
di: Schmid, Larissa, et al.
Pubblicazione: (2024)
CloudHeatMap: Heatmap-Based Monitoring for Large-Scale Cloud Systems
di: Sohana, Sarah, et al.
Pubblicazione: (2024)
di: Sohana, Sarah, et al.
Pubblicazione: (2024)
$μ$OpTime: Statically Reducing the Execution Time of Microbenchmark Suites Using Stability Metrics
di: Japke, Nils, et al.
Pubblicazione: (2025)
di: Japke, Nils, et al.
Pubblicazione: (2025)
Adaptable TeaStore
di: Bliudze, Simon, et al.
Pubblicazione: (2024)
di: Bliudze, Simon, et al.
Pubblicazione: (2024)
A Test Taxonomy and Continuous Integration Ecosystem for Dynamic Resource Management in HPC
di: Sandås, Petter, et al.
Pubblicazione: (2026)
di: Sandås, Petter, et al.
Pubblicazione: (2026)
Building Castles in the Cloud: Architecting Resilient and Scalable Infrastructure
di: Gundla, Naresh Kumar
Pubblicazione: (2024)
di: Gundla, Naresh Kumar
Pubblicazione: (2024)
Histrio: a Serverless Actor System
di: Buttiglieri, Giorgio Natale, et al.
Pubblicazione: (2024)
di: Buttiglieri, Giorgio Natale, et al.
Pubblicazione: (2024)
Predictive Autoscaling for Node.js on Kubernetes: Lower Latency, Right-Sized Capacity
di: Tymoshenko, Ivan, et al.
Pubblicazione: (2026)
di: Tymoshenko, Ivan, et al.
Pubblicazione: (2026)
Efficiently Reproducing Distributed Workflows in Notebook-based Systems
di: Azaz, Talha, et al.
Pubblicazione: (2026)
di: Azaz, Talha, et al.
Pubblicazione: (2026)
AlertGuardian: Intelligent Alert Life-Cycle Management for Large-scale Cloud Systems
di: Yu, Guangba, et al.
Pubblicazione: (2026)
di: Yu, Guangba, et al.
Pubblicazione: (2026)
Do Large Language Models Understand Performance Optimization?
di: Cui, Bowen, et al.
Pubblicazione: (2025)
di: Cui, Bowen, et al.
Pubblicazione: (2025)
Umbilical Choir: Automated Live Testing for Edge-To-Cloud FaaS Applications
di: Malekabbasi, Mohammadreza, et al.
Pubblicazione: (2025)
di: Malekabbasi, Mohammadreza, et al.
Pubblicazione: (2025)
Specx: a C++ task-based runtime system for heterogeneous distributed architectures
di: Cardosi, Paul, et al.
Pubblicazione: (2023)
di: Cardosi, Paul, et al.
Pubblicazione: (2023)
ATOM: Asynchronous Training of Massive Models for Deep Learning in a Decentralized Environment
di: Wu, Xiaofeng, et al.
Pubblicazione: (2024)
di: Wu, Xiaofeng, et al.
Pubblicazione: (2024)
Container-level Energy Observability in Kubernetes Clusters
di: Pijnacker, Bjorn, et al.
Pubblicazione: (2025)
di: Pijnacker, Bjorn, et al.
Pubblicazione: (2025)
FlowUnits: Extending Dataflow for the Edge-to-Cloud Computing Continuum
di: Chini, Fabio, et al.
Pubblicazione: (2025)
di: Chini, Fabio, et al.
Pubblicazione: (2025)
Learning Recovery Strategies for Dynamic Self-healing in Reactive Systems
di: Sanabria, Mateo, et al.
Pubblicazione: (2024)
di: Sanabria, Mateo, et al.
Pubblicazione: (2024)
SoK: Microservice Architectures from a Dependability Perspective
di: Kažemaks, Dāvis, et al.
Pubblicazione: (2025)
di: Kažemaks, Dāvis, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Carbon-aware Software Services
di: Forti, Stefano, et al.
Pubblicazione: (2024) -
CARISMA: CAR-Integrated Service Mesh Architecture
di: Klein, Kevin, et al.
Pubblicazione: (2024) -
GitFarm: Git as a Service for Large-Scale Monorepos
di: Dwivedi, Preetam, et al.
Pubblicazione: (2026) -
LLM-HPC++: Evaluating LLM-Generated Modern C++ and MPI+OpenMP Codes for Scalable Mandelbrot Set Computation
di: Diehl, Patrick, et al.
Pubblicazione: (2025) -
Cost-Effective Big Data Orchestration Using Dagster: A Multi-Platform Approach
di: Picatto, Hernan, et al.
Pubblicazione: (2024)