MARCO: Multi-Agent Code Optimization with Real-Time Knowledge Integration for High-Performance Computing
Fuente:
arXiv
Guardado en:
| Autores principales: | Rahman, Asif, Cvetkovic, Veljko, Reece, Kathleen, Walters, Aidan, Hassan, Yasir, Tummeti, Aneesh, Torres, Bryan, Cooney, Denise, Ellis, Margaret, Nikolopoulos, Dimitrios S. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
FACT: Compositional Kernel Synthesis with a Three-Stage Agentic Workflow
por: Heidari, Sina, et al.
Publicado: (2026)
por: Heidari, Sina, et al.
Publicado: (2026)
Taming the Memory Footprint Crisis: System Design for Production Diffusion LLM Serving
por: Fan, Jiakun, et al.
Publicado: (2025)
por: Fan, Jiakun, et al.
Publicado: (2025)
APEX: Asynchronous Parallel CPU-GPU Execution for Online LLM Inference on Constrained GPUs
por: Fan, Jiakun, et al.
Publicado: (2025)
por: Fan, Jiakun, et al.
Publicado: (2025)
Modality Inflation: Energy Characterization and Optimization Opportunities for MLLM Inference
por: Moghadampanah, Mona, et al.
Publicado: (2025)
por: Moghadampanah, Mona, et al.
Publicado: (2025)
Task queue implementation for edge computing platform
por: Maksimovic, Veljko, et al.
Publicado: (2024)
por: Maksimovic, Veljko, et al.
Publicado: (2024)
Configuration management in the distributed cloud
por: Ranković, Tamara, et al.
Publicado: (2024)
por: Ranković, Tamara, et al.
Publicado: (2024)
ParvaGPU: Efficient Spatial GPU Sharing for Large-Scale DNN Inference in Cloud Environments
por: Lee, Munkyu, et al.
Publicado: (2024)
por: Lee, Munkyu, et al.
Publicado: (2024)
Autonomous Electrochemistry Platform with Real-Time Normality Testing of Voltammetry Measurements Using ML
por: Al-Najjar, Anees, et al.
Publicado: (2025)
por: Al-Najjar, Anees, et al.
Publicado: (2025)
Trustworthy Scheduling for Big Data Applications
por: Tomaras, Dimitrios, et al.
Publicado: (2026)
por: Tomaras, Dimitrios, et al.
Publicado: (2026)
TIMBER: On supporting data pipelines in Mobile Cloud Environments
por: Tomaras, Dimitrios, et al.
Publicado: (2024)
por: Tomaras, Dimitrios, et al.
Publicado: (2024)
Where is the Testbed for my Federated Learning Research?
por: Božič, Janez, et al.
Publicado: (2024)
por: Božič, Janez, et al.
Publicado: (2024)
Dirigent: Lightweight Serverless Orchestration
por: Cvetković, Lazar, et al.
Publicado: (2024)
por: Cvetković, Lazar, et al.
Publicado: (2024)
Towards a real-time distributed feedback system for the transportation assistance of PwD
por: Polenakis, Iosif, et al.
Publicado: (2024)
por: Polenakis, Iosif, et al.
Publicado: (2024)
Fused Breadth-First Probabilistic Traversals on Distributed GPU Systems
por: Neff, Reece, et al.
Publicado: (2023)
por: Neff, Reece, et al.
Publicado: (2023)
QPART: Adaptive Model Quantization and Dynamic Workload Balancing for Accuracy-aware Edge Inference
por: Li, Xiangchen, et al.
Publicado: (2025)
por: Li, Xiangchen, et al.
Publicado: (2025)
Leveraging Public Cloud Infrastructure for Real-time Connected Vehicle Speed Advisory at a Signalized Corridor
por: Deng, Hsien-Wen, et al.
Publicado: (2024)
por: Deng, Hsien-Wen, et al.
Publicado: (2024)
CG-Kit: Code Generation Toolkit for Performant and Maintainable Variants of Source Code Applied to Flash-X Hydrodynamics Simulations
por: Rudi, Johann, et al.
Publicado: (2024)
por: Rudi, Johann, et al.
Publicado: (2024)
Picasso: Memory-Efficient Graph Coloring Using Palettes With Applications in Quantum Computing
por: Ferdous, S M, et al.
Publicado: (2024)
por: Ferdous, S M, et al.
Publicado: (2024)
Generalized Data Placement Strategies for Racetrack Memories
por: Khan, Asif Ali, et al.
Publicado: (2019)
por: Khan, Asif Ali, et al.
Publicado: (2019)
SHEATH: Defending Horizontal Collaboration for Distributed CNNs against Adversarial Noise
por: Asif, Muneeba, et al.
Publicado: (2024)
por: Asif, Muneeba, et al.
Publicado: (2024)
Prediction-driven resource provisioning for serverless container runtimes
por: Tomaras, Dimitrios, et al.
Publicado: (2024)
por: Tomaras, Dimitrios, et al.
Publicado: (2024)
Leveraging Core and Uncore Frequency Scaling for Power-Efficient Serverless Workflows
por: Tzenetopoulos, Achilleas, et al.
Publicado: (2024)
por: Tzenetopoulos, Achilleas, et al.
Publicado: (2024)
Stream-K Optimization and Exploration
por: Rackley, Nick, et al.
Publicado: (2024)
por: Rackley, Nick, et al.
Publicado: (2024)
Ichnos: A Carbon Footprint Estimator for Scientific Workflows
por: West, Kathleen, et al.
Publicado: (2024)
por: West, Kathleen, et al.
Publicado: (2024)
SuperSFL: Resource-Heterogeneous Federated Split Learning with Weight-Sharing Super-Networks
por: Asif, Abdullah Al, et al.
Publicado: (2026)
por: Asif, Abdullah Al, et al.
Publicado: (2026)
Cooperative Gradient Coding
por: Weng, Shudi, et al.
Publicado: (2025)
por: Weng, Shudi, et al.
Publicado: (2025)
Code once, Run Green: Automated Green Code Translation in Serverless Computing
por: Werner, Sebastian, et al.
Publicado: (2025)
por: Werner, Sebastian, et al.
Publicado: (2025)
A Survey on Scheduling Techniques in the Edge Cloud: Issues, Challenges and Future Directions
por: Asghar, Hassan, et al.
Publicado: (2022)
por: Asghar, Hassan, et al.
Publicado: (2022)
Design and Implementation of Code Completion System Based on LLM and CodeBERT Hybrid Subsystem
por: Zhang, Bingbing, et al.
Publicado: (2025)
por: Zhang, Bingbing, et al.
Publicado: (2025)
Rethinking Knowledge Distillation in Collaborative Machine Learning: Memory, Knowledge, and Their Interactions
por: Han, Pengchao, et al.
Publicado: (2025)
por: Han, Pengchao, et al.
Publicado: (2025)
Melding the Serverless Control Plane with the Conventional Cluster Manager for Speed and Resource Efficiency
por: Kondrashov, Leonid, et al.
Publicado: (2025)
por: Kondrashov, Leonid, et al.
Publicado: (2025)
HeLoCo: Efficient asynchronous low-communication training under data and device heterogeneity
por: Asif, Abdullah Al, et al.
Publicado: (2026)
por: Asif, Abdullah Al, et al.
Publicado: (2026)
SmartPQ: An Adaptive Concurrent Priority Queue for NUMA Architectures
por: Giannoula, Christina, et al.
Publicado: (2024)
por: Giannoula, Christina, et al.
Publicado: (2024)
On Similarity of Computational Kernels in our Codes and Proxies
por: McKinsey, Michael, et al.
Publicado: (2026)
por: McKinsey, Michael, et al.
Publicado: (2026)
Biased Compression in Gradient Coding for Distributed Learning
por: Li, Chengxi, et al.
Publicado: (2026)
por: Li, Chengxi, et al.
Publicado: (2026)
SynergAI: Edge-to-Cloud Synergy for Architecture-Driven High-Performance Orchestration for AI Inference
por: Stathopoulou, Foteini, et al.
Publicado: (2025)
por: Stathopoulou, Foteini, et al.
Publicado: (2025)
Mobile Edge Computing
por: Ahmed, Sohaib, et al.
Publicado: (2024)
por: Ahmed, Sohaib, et al.
Publicado: (2024)
A High-throughput and Secure Coded Blockchain for IoT
por: Taherpour, Amirhossein, et al.
Publicado: (2023)
por: Taherpour, Amirhossein, et al.
Publicado: (2023)
New Wide Locally Recoverable Codes with Unified Locality
por: Xu, Liangliang, et al.
Publicado: (2025)
por: Xu, Liangliang, et al.
Publicado: (2025)
Augur: Pre-Execution Energy Prediction for Workflow Tasks in Heterogeneous Clusters
por: West, Kathleen, et al.
Publicado: (2026)
por: West, Kathleen, et al.
Publicado: (2026)
Ejemplares similares
-
FACT: Compositional Kernel Synthesis with a Three-Stage Agentic Workflow
por: Heidari, Sina, et al.
Publicado: (2026) -
Taming the Memory Footprint Crisis: System Design for Production Diffusion LLM Serving
por: Fan, Jiakun, et al.
Publicado: (2025) -
APEX: Asynchronous Parallel CPU-GPU Execution for Online LLM Inference on Constrained GPUs
por: Fan, Jiakun, et al.
Publicado: (2025) -
Modality Inflation: Energy Characterization and Optimization Opportunities for MLLM Inference
por: Moghadampanah, Mona, et al.
Publicado: (2025) -
Task queue implementation for edge computing platform
por: Maksimovic, Veljko, et al.
Publicado: (2024)