Guardado en:
| Autores principales: | Srinivasan, Sriram, Alabsi, Hamdan, Obeidat, Rand, Ponnala, Nithisha, Zenebe, Azene |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2509.13703 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A GPU-accelerated Molecular Docking Workflow with Kubernetes and Apache Airflow
por: Medeiros, Daniel, et al.
Publicado: (2024)
por: Medeiros, Daniel, et al.
Publicado: (2024)
Cultivating Multidisciplinary AI Workforce Development on iTiger GPU Cluster: Practices and Challenges
por: Sharif, Mayira, et al.
Publicado: (2025)
por: Sharif, Mayira, et al.
Publicado: (2025)
SageSched: Efficient LLM Scheduling Confronting Demand Uncertainty and Hybridity
por: Gan, Zhenghao, et al.
Publicado: (2026)
por: Gan, Zhenghao, et al.
Publicado: (2026)
VDCores: Resource Decoupled Programming and Execution for Asynchronous GPU
por: He, Zijian, et al.
Publicado: (2026)
por: He, Zijian, et al.
Publicado: (2026)
AI-coupled HPC Workflow Applications, Middleware and Performance
por: Brewer, Wes, et al.
Publicado: (2024)
por: Brewer, Wes, et al.
Publicado: (2024)
RHAPSODY: Execution of Hybrid AI-HPC Workflows at Scale
por: Alsaadi, Aymen, et al.
Publicado: (2025)
por: Alsaadi, Aymen, et al.
Publicado: (2025)
Evaluation of Programming Models and Performance for Stencil Computation on Current GPU Architectures
por: Shan, Baodi, et al.
Publicado: (2024)
por: Shan, Baodi, et al.
Publicado: (2024)
Concurrent Scheduling of High-Level Parallel Programs on Multi-GPU Systems
por: Knorr, Fabian, et al.
Publicado: (2025)
por: Knorr, Fabian, et al.
Publicado: (2025)
SageServe: Optimizing LLM Serving on Cloud Data Centers with Forecast Aware Auto-Scaling
por: Jaiswal, Shashwat, et al.
Publicado: (2025)
por: Jaiswal, Shashwat, et al.
Publicado: (2025)
HAP: SPMD DNN Training on Heterogeneous GPU Clusters with Automated Program Synthesis
por: Zhang, Shiwei, et al.
Publicado: (2024)
por: Zhang, Shiwei, et al.
Publicado: (2024)
HiRace: Accurate and Fast Source-Level Race Checking of GPU Programs
por: Jacobson, John, et al.
Publicado: (2024)
por: Jacobson, John, et al.
Publicado: (2024)
A Study of Performance Programming of CPU, GPU accelerated Computers and SIMD Architecture
por: Yi, Xinyao
Publicado: (2024)
por: Yi, Xinyao
Publicado: (2024)
In-Transit Data Transport Strategies for Coupled AI-Simulation Workflow Patterns
por: Tummalapalli, Harikrishna, et al.
Publicado: (2025)
por: Tummalapalli, Harikrishna, et al.
Publicado: (2025)
Automated Market Makers for Cross-chain DeFi and Sharded Blockchains
por: Aanes, Jon Michael, et al.
Publicado: (2023)
por: Aanes, Jon Michael, et al.
Publicado: (2023)
Taking GPU Programming Models to Task for Performance Portability
por: Davis, Joshua H., et al.
Publicado: (2024)
por: Davis, Joshua H., et al.
Publicado: (2024)
KEET: Explaining Performance of GPU Kernels Using LLM Agents
por: Davis, Joshua H., et al.
Publicado: (2026)
por: Davis, Joshua H., et al.
Publicado: (2026)
BandPilot: Towards Performance- and Contention-Aware GPU Dispatching in AI Clusters
por: Zhang, Kunming, et al.
Publicado: (2025)
por: Zhang, Kunming, et al.
Publicado: (2025)
Workflows Community Summit 2024: Future Trends and Challenges in Scientific Workflows
por: da Silva, Rafael Ferreira, et al.
Publicado: (2024)
por: da Silva, Rafael Ferreira, et al.
Publicado: (2024)
The Fused Kernel Library: A C++ API to Develop Highly-Efficient GPU Libraries
por: Amoros, Oscar, et al.
Publicado: (2025)
por: Amoros, Oscar, et al.
Publicado: (2025)
exa-AMD: A Scalable Workflow for Accelerating AI-Assisted Materials Discovery and Design
por: Moraru, Maxim, et al.
Publicado: (2025)
por: Moraru, Maxim, et al.
Publicado: (2025)
Workflow-Driven Modeling for the Compute Continuum: An Optimization Approach to Automated System and Workload Scheduling
por: Sharma, Aasish Kumar, et al.
Publicado: (2025)
por: Sharma, Aasish Kumar, et al.
Publicado: (2025)
Towards singular optimality in the presence of local initial knowledge
por: Ji, Hongyan, et al.
Publicado: (2024)
por: Ji, Hongyan, et al.
Publicado: (2024)
Future-Proofing IoT: Unleashing the Power of AWS Greengrass in Propelling Smart Devices to New Heights
por: Kokkula, Sahasra, et al.
Publicado: (2024)
por: Kokkula, Sahasra, et al.
Publicado: (2024)
Workflow Mini-Apps: Portable, Scalable, Tunable & Faithful Representations of Scientific Workflows
por: Kilic, Ozgur Ozan, et al.
Publicado: (2024)
por: Kilic, Ozgur Ozan, et al.
Publicado: (2024)
WOW: Workflow-Aware Data Movement and Task Scheduling for Dynamic Scientific Workflows
por: Lehmann, Fabian, et al.
Publicado: (2025)
por: Lehmann, Fabian, et al.
Publicado: (2025)
A Deep Reinforcement Learning Approach for Cost Optimized Workflow Scheduling in Cloud Computing Environments
por: Jayanetti, Amanda, et al.
Publicado: (2024)
por: Jayanetti, Amanda, et al.
Publicado: (2024)
Improving GPU Multi-Tenancy Through Dynamic Multi-Instance GPU Reconfiguration
por: Wang, Tianyu, et al.
Publicado: (2024)
por: Wang, Tianyu, et al.
Publicado: (2024)
Serving Compound Inference Systems on Datacenter GPUs
por: Devata, Sriram, et al.
Publicado: (2026)
por: Devata, Sriram, et al.
Publicado: (2026)
Towards Advancing Research with Workflows: A perspective from the Workflows Community Summit -- Amsterdam, 2025
por: Bonati, Irene, et al.
Publicado: (2026)
por: Bonati, Irene, et al.
Publicado: (2026)
Accelerating Biclique Counting on GPU
por: Qiu, Linshan, et al.
Publicado: (2024)
por: Qiu, Linshan, et al.
Publicado: (2024)
GPU Sharing with Triples Mode
por: Byun, Chansup, et al.
Publicado: (2024)
por: Byun, Chansup, et al.
Publicado: (2024)
Syncopate: Efficient Multi-GPU AI Kernels via Automatic Chunk-Centric Compute-Communication Overlap
por: Qiang, Xinwei, et al.
Publicado: (2026)
por: Qiang, Xinwei, et al.
Publicado: (2026)
ParvaGPU: Efficient Spatial GPU Sharing for Large-Scale DNN Inference in Cloud Environments
por: Lee, Munkyu, et al.
Publicado: (2024)
por: Lee, Munkyu, et al.
Publicado: (2024)
Accelerating Intra-Node GPU-to-GPU Communication Through Multi-Path Transfers with CUDA Graphs
por: Sojoodi, Amirhossein, et al.
Publicado: (2026)
por: Sojoodi, Amirhossein, et al.
Publicado: (2026)
A Terminology for Scientific Workflow Systems
por: Suter, Frédéric, et al.
Publicado: (2025)
por: Suter, Frédéric, et al.
Publicado: (2025)
An Ecosystem of Services for FAIR Computational Workflows
por: Wilkinson, Sean R., et al.
Publicado: (2025)
por: Wilkinson, Sean R., et al.
Publicado: (2025)
AgentServe: Algorithm-System Co-Design for Efficient Agentic AI Serving on a Consumer-Grade GPU
por: Zhang, Yuning, et al.
Publicado: (2026)
por: Zhang, Yuning, et al.
Publicado: (2026)
DuaLip-GPU Technical Report
por: Dexter, Gregory, et al.
Publicado: (2026)
por: Dexter, Gregory, et al.
Publicado: (2026)
Incidence Constraints in Hypergraph Partitioning on GPU
por: Ronzani, Marco, et al.
Publicado: (2026)
por: Ronzani, Marco, et al.
Publicado: (2026)
Predictable LLM Serving on GPU Clusters
por: Darzi, Erfan, et al.
Publicado: (2025)
por: Darzi, Erfan, et al.
Publicado: (2025)
Ejemplares similares
-
A GPU-accelerated Molecular Docking Workflow with Kubernetes and Apache Airflow
por: Medeiros, Daniel, et al.
Publicado: (2024) -
Cultivating Multidisciplinary AI Workforce Development on iTiger GPU Cluster: Practices and Challenges
por: Sharif, Mayira, et al.
Publicado: (2025) -
SageSched: Efficient LLM Scheduling Confronting Demand Uncertainty and Hybridity
por: Gan, Zhenghao, et al.
Publicado: (2026) -
VDCores: Resource Decoupled Programming and Execution for Asynchronous GPU
por: He, Zijian, et al.
Publicado: (2026) -
AI-coupled HPC Workflow Applications, Middleware and Performance
por: Brewer, Wes, et al.
Publicado: (2024)