Analyzing Latency Hiding and Parallelism in an MLIR-based AI Kernel Compiler
Fuente:
arXiv
Salvato in:
| Autori principali: | Absar, Javed, Narang, Samarth, Baskaran, Muthu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Tensor Evolution: A Framework for Fast Evaluation of Tensor Computations using Recurrences
di: Absar, Javed, et al.
Pubblicazione: (2025)
di: Absar, Javed, et al.
Pubblicazione: (2025)
Hexagon-MLIR: An AI Compilation Stack For Qualcomm's Neural Processing Units (NPUs)
di: Absar, Mohammed Javed, et al.
Pubblicazione: (2026)
di: Absar, Mohammed Javed, et al.
Pubblicazione: (2026)
Executing as You Generate: Hiding Execution Latency in LLM Code Generation
di: Sun, Zhensu, et al.
Pubblicazione: (2026)
di: Sun, Zhensu, et al.
Pubblicazione: (2026)
An LLM-Tool Compiler for Fused Parallel Function Calling
di: Singh, Simranjit, et al.
Pubblicazione: (2024)
di: Singh, Simranjit, et al.
Pubblicazione: (2024)
Towards a high-performance AI compiler with upstream MLIR
di: Golin, Renato, et al.
Pubblicazione: (2024)
di: Golin, Renato, et al.
Pubblicazione: (2024)
Meta Large Language Model Compiler: Foundation Models of Compiler Optimization
di: Cummins, Chris, et al.
Pubblicazione: (2024)
di: Cummins, Chris, et al.
Pubblicazione: (2024)
AwareCompiler: Agentic Context-Aware Compiler Optimization via a Synergistic Knowledge-Data Driven Framework
di: Lin, Hongyu, et al.
Pubblicazione: (2025)
di: Lin, Hongyu, et al.
Pubblicazione: (2025)
Algorithmic Language Models with Neurally Compiled Libraries
di: Saldyt, Lucas, et al.
Pubblicazione: (2024)
di: Saldyt, Lucas, et al.
Pubblicazione: (2024)
Model2Kernel: Model-Aware Symbolic Execution For Safe CUDA Kernels
di: He, Mengting, et al.
Pubblicazione: (2026)
di: He, Mengting, et al.
Pubblicazione: (2026)
Grammar-Aware Literate Generative Mathematical Programming with Compiler-in-the-Loop
di: Rossi, Roberto, et al.
Pubblicazione: (2026)
di: Rossi, Roberto, et al.
Pubblicazione: (2026)
The New Compiler Stack: A Survey on the Synergy of LLMs and Compilers
di: Zhang, Shuoming, et al.
Pubblicazione: (2026)
di: Zhang, Shuoming, et al.
Pubblicazione: (2026)
LEGO-Compiler: Enhancing Neural Compilation Through Translation Composability
di: Zhang, Shuoming, et al.
Pubblicazione: (2025)
di: Zhang, Shuoming, et al.
Pubblicazione: (2025)
AutoLALA: Automatic Loop Algebraic Locality Analysis for AI and HPC Kernels
di: Zhu, Yifan, et al.
Pubblicazione: (2026)
di: Zhu, Yifan, et al.
Pubblicazione: (2026)
AIOS Compiler: LLM as Interpreter for Natural Language Programming and Flow Programming of AI Agents
di: Xu, Shuyuan, et al.
Pubblicazione: (2024)
di: Xu, Shuyuan, et al.
Pubblicazione: (2024)
WAMI: Compilation to WebAssembly through MLIR without Losing Abstraction
di: Kang, Byeongjee, et al.
Pubblicazione: (2025)
di: Kang, Byeongjee, et al.
Pubblicazione: (2025)
Compiler-Guided Inference-Time Adaptation: Improving GPT-5 Programming Performance in Idris
di: Li, Minda, et al.
Pubblicazione: (2026)
di: Li, Minda, et al.
Pubblicazione: (2026)
CrypTorch: PyTorch-based Auto-tuning Compiler for Machine Learning with Multi-party Computation
di: Liu, Jinyu, et al.
Pubblicazione: (2025)
di: Liu, Jinyu, et al.
Pubblicazione: (2025)
Reinforcement Learning from Compiler and Language Server Feedback
di: Zhang, Yifan, et al.
Pubblicazione: (2025)
di: Zhang, Yifan, et al.
Pubblicazione: (2025)
LightCode: Compiling LLM Inference for Photonic-Electronic Systems
di: Tomich, Ryan, et al.
Pubblicazione: (2025)
di: Tomich, Ryan, et al.
Pubblicazione: (2025)
Analyzing the Effectiveness of Large Language Models on Text-to-SQL Synthesis
di: Roberson, Richard, et al.
Pubblicazione: (2024)
di: Roberson, Richard, et al.
Pubblicazione: (2024)
Demonstrating a Future for MLIR-native DSL Compilers on a NumPy-like Example
di: Friebel, Karl F. A., et al.
Pubblicazione: (2026)
di: Friebel, Karl F. A., et al.
Pubblicazione: (2026)
Magellan: Autonomous Discovery of Novel Compiler Optimization Heuristics with AlphaEvolve
di: Chen, Hongzheng, et al.
Pubblicazione: (2026)
di: Chen, Hongzheng, et al.
Pubblicazione: (2026)
MLIR-Smith: A Novel Random Program Generator for Evaluating Compiler Pipelines
di: Ates, Berke, et al.
Pubblicazione: (2026)
di: Ates, Berke, et al.
Pubblicazione: (2026)
Hexcute: A Compiler Framework for Automating Layout Synthesis in GPU Programs
di: Zhang, Xiao, et al.
Pubblicazione: (2025)
di: Zhang, Xiao, et al.
Pubblicazione: (2025)
PassNet: Scaling Large Language Models for Graph Compiler Pass Generation
di: Liu, Yiqun, et al.
Pubblicazione: (2026)
di: Liu, Yiqun, et al.
Pubblicazione: (2026)
depyf: Open the Opaque Box of PyTorch Compiler for Machine Learning Researchers
di: You, Kaichao, et al.
Pubblicazione: (2024)
di: You, Kaichao, et al.
Pubblicazione: (2024)
Evaluating adaptive and generative AI-based feedback and recommendations in a knowledge-graph-integrated programming learning system
di: Nongkhai, Lalita Na, et al.
Pubblicazione: (2026)
di: Nongkhai, Lalita Na, et al.
Pubblicazione: (2026)
Compile to Compress: Boosting Formal Theorem Provers by Compiler Outputs
di: Li, Guchan, et al.
Pubblicazione: (2026)
di: Li, Guchan, et al.
Pubblicazione: (2026)
ECCO: Evidence-Driven Causal Reasoning for Compiler Optimization
di: Pan, Haolin, et al.
Pubblicazione: (2026)
di: Pan, Haolin, et al.
Pubblicazione: (2026)
BODHI: Precise OS Kernel Specification Inference
di: Chang, Zhiming, et al.
Pubblicazione: (2026)
di: Chang, Zhiming, et al.
Pubblicazione: (2026)
BuildBench: Benchmarking LLM Agents on Compiling Real-World Open-Source Software
di: Zhang, Zehua, et al.
Pubblicazione: (2025)
di: Zhang, Zehua, et al.
Pubblicazione: (2025)
PopPy: Opportunistically Exploiting Parallelism in Python Compound AI Applications
di: Mell, Stephen, et al.
Pubblicazione: (2026)
di: Mell, Stephen, et al.
Pubblicazione: (2026)
Benchmarking Large Language Models for ABAP Code Generation: An Empirical Study on Iterative Improvement by Compiler Feedback
di: Wallraven, Stephan, et al.
Pubblicazione: (2026)
di: Wallraven, Stephan, et al.
Pubblicazione: (2026)
LiteCoOp: Lightweight Multi-LLM Shared-Tree Reasoning for Model-Serving Compiler Optimizations
di: Tang, Annabelle Sujun, et al.
Pubblicazione: (2026)
di: Tang, Annabelle Sujun, et al.
Pubblicazione: (2026)
Is It a Good Idea to Build an HLS Tool on Top of MLIR? Experience from Building the Dynamatic HLS Compiler
di: Xu, Jiahui, et al.
Pubblicazione: (2026)
di: Xu, Jiahui, et al.
Pubblicazione: (2026)
The why, what, and how of AI-based coding in scientific research
di: Zhuang, Tonghe, et al.
Pubblicazione: (2024)
di: Zhuang, Tonghe, et al.
Pubblicazione: (2024)
MTP: A Meaning-Typed Language Abstraction for AI-Integrated Programming
di: Dantanarayana, Jayanaka L., et al.
Pubblicazione: (2024)
di: Dantanarayana, Jayanaka L., et al.
Pubblicazione: (2024)
From Prompts to Performance: Evaluating LLMs for Task-based Parallel Code Generation
di: Bantel, Linus, et al.
Pubblicazione: (2026)
di: Bantel, Linus, et al.
Pubblicazione: (2026)
DSDL: Data Set Description Language for Bridging Modalities and Tasks in AI Data
di: Wang, Bin, et al.
Pubblicazione: (2024)
di: Wang, Bin, et al.
Pubblicazione: (2024)
Practical Formal Verification for MLIR Programs
di: Tucker, Emily, et al.
Pubblicazione: (2026)
di: Tucker, Emily, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Tensor Evolution: A Framework for Fast Evaluation of Tensor Computations using Recurrences
di: Absar, Javed, et al.
Pubblicazione: (2025) -
Hexagon-MLIR: An AI Compilation Stack For Qualcomm's Neural Processing Units (NPUs)
di: Absar, Mohammed Javed, et al.
Pubblicazione: (2026) -
Executing as You Generate: Hiding Execution Latency in LLM Code Generation
di: Sun, Zhensu, et al.
Pubblicazione: (2026) -
An LLM-Tool Compiler for Fused Parallel Function Calling
di: Singh, Simranjit, et al.
Pubblicazione: (2024) -
Towards a high-performance AI compiler with upstream MLIR
di: Golin, Renato, et al.
Pubblicazione: (2024)