The Landscape and Challenges of HPC Research and LLMs
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Chen, Le, Ahmed, Nesreen K., Dutta, Akash, Bhattacharjee, Arijit, Yu, Sixing, Mahmud, Quazi Ishtiaque, Abebe, Waqwoya, Phan, Hung, Sarkar, Aishwarya, Butler, Branden, Hasabnis, Niranjan, Oren, Gal, Vo, Vy A., Munoz, Juan Pablo, Willke, Theodore L., Mattson, Tim, Jannesari, Ali |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
OMPGPT: A Generative Pre-trained Transformer Model for OpenMP
par: Chen, Le, et autres
Publié: (2024)
par: Chen, Le, et autres
Publié: (2024)
MonoCoder: Domain-Specific Code Language Model for HPC Codes and Tasks
par: Kadosh, Tal, et autres
Publié: (2023)
par: Kadosh, Tal, et autres
Publié: (2023)
Learning to Parallelize with OpenMP by Augmented Heterogeneous AST Representation
par: Chen, Le, et autres
Publié: (2023)
par: Chen, Le, et autres
Publié: (2023)
AutoParLLM: GNN-guided Context Generation for Zero-Shot Code Parallelization using LLMs
par: Mahmud, Quazi Ishtiaque, et autres
Publié: (2023)
par: Mahmud, Quazi Ishtiaque, et autres
Publié: (2023)
MPIrigen: MPI Code Generation through Domain-Specific Language Models
par: Schneider, Nadav, et autres
Publié: (2024)
par: Schneider, Nadav, et autres
Publié: (2024)
Data Race Satisfiability on Array Elements
par: Shim, Junhyung, et autres
Publié: (2025)
par: Shim, Junhyung, et autres
Publié: (2025)
OMPILOT: Harnessing Transformer Models for Auto Parallelization to Shared Memory Computing Paradigms
par: Bhattacharjee, Arijit, et autres
Publié: (2025)
par: Bhattacharjee, Arijit, et autres
Publié: (2025)
OMPar: Automatic Parallelization with AI-Driven Source-to-Source Compilation
par: Kadosh, Tal, et autres
Publié: (2024)
par: Kadosh, Tal, et autres
Publié: (2024)
PipeInfer: Accelerating LLM Inference using Asynchronous Pipelined Speculation
par: Butler, Branden, et autres
Publié: (2024)
par: Butler, Branden, et autres
Publié: (2024)
LEFL: Low Entropy Client Sampling in Federated Learning
par: Abebe, Waqwoya, et autres
Publié: (2023)
par: Abebe, Waqwoya, et autres
Publié: (2023)
SuperSAM: Crafting a SAM Supernetwork via Structured Pruning and Unstructured Parameter Prioritization
par: Abebe, Waqwoya, et autres
Publié: (2025)
par: Abebe, Waqwoya, et autres
Publié: (2025)
UniPar: A Unified LLM-Based Framework for Parallel and Accelerated Code Translation in HPC
par: Bitan, Tomer, et autres
Publié: (2025)
par: Bitan, Tomer, et autres
Publié: (2025)
OptiML: An End-to-End Framework for Program Synthesis and CUDA Kernel Optimization
par: Bhattacharjee, Arijit, et autres
Publié: (2026)
par: Bhattacharjee, Arijit, et autres
Publié: (2026)
CodeRosetta: Pushing the Boundaries of Unsupervised Code Translation for Parallel Programming
par: TehraniJamsaz, Ali, et autres
Publié: (2024)
par: TehraniJamsaz, Ali, et autres
Publié: (2024)
ParaCodex: A Profiling-Guided Autonomous Coding Agent for Reliable Parallel Code Generation and Translation
par: Kaplan, Erel, et autres
Publié: (2026)
par: Kaplan, Erel, et autres
Publié: (2026)
Structure Guided Prompt: Instructing Large Language Model in Multi-Step Reasoning by Exploring Graph Structure of the Text
par: Cheng, Kewei, et autres
Publié: (2024)
par: Cheng, Kewei, et autres
Publié: (2024)
Counting Without Running: Evaluating LLMs' Reasoning About Code Complexity
par: Bolet, Gregory, et autres
Publié: (2025)
par: Bolet, Gregory, et autres
Publié: (2025)
Can Large Language Models Predict Parallel Code Performance?
par: Bolet, Gregory, et autres
Publié: (2025)
par: Bolet, Gregory, et autres
Publié: (2025)
SAM-I-Am: Semantic Boosting for Zero-shot Atomic-Scale Electron Micrograph Segmentation
par: Abebe, Waqwoya, et autres
Publié: (2024)
par: Abebe, Waqwoya, et autres
Publié: (2024)
Federated Foundation Models: Privacy-Preserving and Collaborative Learning for Large Models
par: Yu, Sixing, et autres
Publié: (2023)
par: Yu, Sixing, et autres
Publié: (2023)
Resource-Aware Heterogeneous Federated Learning using Neural Architecture Search
par: Yu, Sixing, et autres
Publié: (2022)
par: Yu, Sixing, et autres
Publié: (2022)
MIREncoder: Multi-modal IR-based Pretrained Embeddings for Performance Optimizations
par: Dutta, Akash, et autres
Publié: (2024)
par: Dutta, Akash, et autres
Publié: (2024)
VeriMoA: A Mixture-of-Agents Framework for Spec-to-HDL Generation
par: Ping, Heng, et autres
Publié: (2025)
par: Ping, Heng, et autres
Publié: (2025)
Toward Optimal Search and Retrieval for RAG
par: Leto, Alexandria, et autres
Publié: (2024)
par: Leto, Alexandria, et autres
Publié: (2024)
Adaptive USVs Swarm Optimization for Target Tracking in Dynamic Environments
par: Gal, Oren
Publié: (2024)
par: Gal, Oren
Publié: (2024)
Static Generation of Efficient OpenMP Offload Data Mappings
par: Marzen, Luke, et autres
Publié: (2024)
par: Marzen, Luke, et autres
Publié: (2024)
Verified Code Transpilation with LLMs
par: Bhatia, Sahil, et autres
Publié: (2024)
par: Bhatia, Sahil, et autres
Publié: (2024)
Fast MoE Inference via Predictive Prefetching and Expert Replication
par: Jyothish, Ankit, et autres
Publié: (2026)
par: Jyothish, Ankit, et autres
Publié: (2026)
ProfilingAgent: Profiling-Guided Agentic Reasoning for Adaptive Model Optimization
par: Jafari, Sadegh, et autres
Publié: (2025)
par: Jafari, Sadegh, et autres
Publié: (2025)
Enhanced Soups for Graph Neural Networks
par: Zuber, Joseph, et autres
Publié: (2025)
par: Zuber, Joseph, et autres
Publié: (2025)
SuperSFL: Resource-Heterogeneous Federated Split Learning with Weight-Sharing Super-Networks
par: Asif, Abdullah Al, et autres
Publié: (2026)
par: Asif, Abdullah Al, et autres
Publié: (2026)
PerfMamba: Performance Analysis and Pruning of Selective State Space Models
par: Asif, Abdullah Al, et autres
Publié: (2025)
par: Asif, Abdullah Al, et autres
Publié: (2025)
Diabetic Retinopathy Classification from Retinal Images using Machine Learning Approaches
par: Bhattacharjee, Indronil, et autres
Publié: (2024)
par: Bhattacharjee, Indronil, et autres
Publié: (2024)
MassiveGNN: Efficient Training via Prefetching for Massively Connected Distributed Graphs
par: Sarkar, Aishwarya, et autres
Publié: (2024)
par: Sarkar, Aishwarya, et autres
Publié: (2024)
NOMAD: Generating Embeddings for Massive Distributed Graphs
par: Sarkar, Aishwarya, et autres
Publié: (2026)
par: Sarkar, Aishwarya, et autres
Publié: (2026)
Learning Bug Context for PyTorch-to-JAX Translation with LLMs
par: Phan, Hung, et autres
Publié: (2025)
par: Phan, Hung, et autres
Publié: (2025)
Long COVID and financial hardship: A disaggregated analysis at income and education levels
par: Biplab Kumar Datta, et autres
Publié: (2024)
par: Biplab Kumar Datta, et autres
Publié: (2024)
Tenspiler: A Verified Lifting-Based Compiler for Tensor Operations (Extended Version)
par: Qiu, Jie, et autres
Publié: (2024)
par: Qiu, Jie, et autres
Publié: (2024)
MOSAIC-Bench: Measuring Compositional Vulnerability Induction in Coding Agents
par: Steinberg, Jonathan, et autres
Publié: (2026)
par: Steinberg, Jonathan, et autres
Publié: (2026)
Semantic Denial of Service in LLM-controlled robots
par: Steinberg, Jonathan, et autres
Publié: (2026)
par: Steinberg, Jonathan, et autres
Publié: (2026)
Documents similaires
-
OMPGPT: A Generative Pre-trained Transformer Model for OpenMP
par: Chen, Le, et autres
Publié: (2024) -
MonoCoder: Domain-Specific Code Language Model for HPC Codes and Tasks
par: Kadosh, Tal, et autres
Publié: (2023) -
Learning to Parallelize with OpenMP by Augmented Heterogeneous AST Representation
par: Chen, Le, et autres
Publié: (2023) -
AutoParLLM: GNN-guided Context Generation for Zero-Shot Code Parallelization using LLMs
par: Mahmud, Quazi Ishtiaque, et autres
Publié: (2023) -
MPIrigen: MPI Code Generation through Domain-Specific Language Models
par: Schneider, Nadav, et autres
Publié: (2024)