Discovering Software Parallelization Points Using Deep Neural Networks
Fuente:
arXiv
Saved in:
| Main Authors: | Correia, Izavan dos S., Santos, Henrique C. T., Ferreira, Tiago A. E. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Reducing Data Bottlenecks in Distributed, Heterogeneous Neural Networks
by: Lin, Ruhai, et al.
Published: (2024)
by: Lin, Ruhai, et al.
Published: (2024)
A Fresh Approach to Evaluate Performance in Distributed Parallel Genetic Algorithms
by: Harada, Tomohiro, et al.
Published: (2021)
by: Harada, Tomohiro, et al.
Published: (2021)
Wireless Sensor Networks as Parallel and Distributed Hardware Platform for Artificial Neural Networks
by: Serpen, Gursel
Published: (2025)
by: Serpen, Gursel
Published: (2025)
Cyclic Data Parallelism for Efficient Parallelism of Deep Neural Networks
by: Fournier, Louis, et al.
Published: (2024)
by: Fournier, Louis, et al.
Published: (2024)
A Frequency-based Parent Selection for Reducing the Effect of Evaluation Time Bias in Asynchronous Parallel Multi-objective Evolutionary Algorithms
by: Harada, Tomohiro
Published: (2021)
by: Harada, Tomohiro
Published: (2021)
Asynchronous Evolution of Deep Neural Network Architectures
by: Liang, Jason, et al.
Published: (2023)
by: Liang, Jason, et al.
Published: (2023)
Sparse Spiking Neural-like Membrane Systems on Graphics Processing Units
by: Hernández-Tello, Javier, et al.
Published: (2024)
by: Hernández-Tello, Javier, et al.
Published: (2024)
Kernel Foundry: A Diagnosis-driven Evolutionary Kernel Optimizer with Multi-Experts
by: Huang, Zixuan, et al.
Published: (2026)
by: Huang, Zixuan, et al.
Published: (2026)
D-CODE: Data Colony Optimization for Dynamic Network Efficiency
by: Pandey, Tannu, et al.
Published: (2024)
by: Pandey, Tannu, et al.
Published: (2024)
Code generation and runtime techniques for enabling data-efficient deep learning training on GPUs
by: Wu, Kun
Published: (2024)
by: Wu, Kun
Published: (2024)
Neuromorphic Simulation of Drosophila Melanogaster Brain Connectome on Loihi 2
by: Wang, Felix, et al.
Published: (2025)
by: Wang, Felix, et al.
Published: (2025)
SSDTrain: An Activation Offloading Framework to SSDs for Faster Large Language Model Training
by: Wu, Kun, et al.
Published: (2024)
by: Wu, Kun, et al.
Published: (2024)
Evaluation and Efficiency Comparison of Evolutionary Algorithms for Service Placement Optimization in Fog Architectures
by: Guerrero, Carlos, et al.
Published: (2025)
by: Guerrero, Carlos, et al.
Published: (2025)
Trackable Agent-based Evolution Models at Wafer Scale
by: Moreno, Matthew Andres, et al.
Published: (2024)
by: Moreno, Matthew Andres, et al.
Published: (2024)
Trackable Island-model Genetic Algorithms at Wafer Scale
by: Moreno, Matthew Andres, et al.
Published: (2024)
by: Moreno, Matthew Andres, et al.
Published: (2024)
phys-MCP: A Control Plane for Heterogeneous Physical Neural Networks
by: Fischer, Stefan, et al.
Published: (2026)
by: Fischer, Stefan, et al.
Published: (2026)
Fray: An Efficient General-Purpose Concurrency Testing Platform for the JVM (Extended Version)
by: Li, Ao, et al.
Published: (2025)
by: Li, Ao, et al.
Published: (2025)
Determinacy with Priorities up to Clocks
by: Liquori, Luigi, et al.
Published: (2026)
by: Liquori, Luigi, et al.
Published: (2026)
NeuroRing: Scaling Spiking Neural Networks via Multi-FPGA Bidirectional Ring Topologies and Stream-Dataflow Architectures
by: Hafiz, Muhammad Ihsan Al, et al.
Published: (2026)
by: Hafiz, Muhammad Ihsan Al, et al.
Published: (2026)
Neuromorphic Programming: Emerging Directions for Brain-Inspired Hardware
by: Abreu, Steven, et al.
Published: (2024)
by: Abreu, Steven, et al.
Published: (2024)
HiAER-Spike: Hardware-Software Co-Design for Large-Scale Reconfigurable Event-Driven Neuromorphic Computing
by: Frank, Gwenevere, et al.
Published: (2025)
by: Frank, Gwenevere, et al.
Published: (2025)
Leveraging LLMs to Automate Energy-Aware Refactoring of Parallel Scientific Codes
by: Dearing, Matthew T., et al.
Published: (2025)
by: Dearing, Matthew T., et al.
Published: (2025)
LASSI: An LLM-based Automated Self-Correcting Pipeline for Translating Parallel Scientific Codes
by: Dearing, Matthew T., et al.
Published: (2024)
by: Dearing, Matthew T., et al.
Published: (2024)
MAC-DO: An Efficient Output-Stationary GEMM Accelerator for CNNs Using DRAM Technology
by: Jeong, Minki, et al.
Published: (2022)
by: Jeong, Minki, et al.
Published: (2022)
ParEVO: Synthesizing Code for Irregular Data: High-Performance Parallelism through Agentic Evolution
by: Yang, Liu, et al.
Published: (2026)
by: Yang, Liu, et al.
Published: (2026)
Comparative analysis of large data processing in Apache Spark using Java, Python and Scala
by: Borodii, Ivan, et al.
Published: (2025)
by: Borodii, Ivan, et al.
Published: (2025)
Scalable Construction of Spiking Neural Networks using up to thousands of GPUs
by: Golosio, Bruno, et al.
Published: (2025)
by: Golosio, Bruno, et al.
Published: (2025)
Edge Intelligence with Spiking Neural Networks
by: Deng, Shuiguang, et al.
Published: (2025)
by: Deng, Shuiguang, et al.
Published: (2025)
Dynamic hashtag recommendation in social media with trend shift detection and adaptation
by: Cantini, Riccardo, et al.
Published: (2025)
by: Cantini, Riccardo, et al.
Published: (2025)
STEMS: Spatial-Temporal Mapping For Spiking Neural Networks
by: Eissa, Sherif, et al.
Published: (2025)
by: Eissa, Sherif, et al.
Published: (2025)
Neuromorphic Computing: A Theoretical Framework for Time, Space, and Energy Scaling
by: Aimone, James B
Published: (2025)
by: Aimone, James B
Published: (2025)
Neuromorphic hardware for sustainable AI data centers
by: Vogginger, Bernhard, et al.
Published: (2024)
by: Vogginger, Bernhard, et al.
Published: (2024)
Federated Learning in Chemical Engineering: A Tutorial on a Framework for Privacy-Preserving Collaboration Across Distributed Data Sources
by: Dutta, Siddhant, et al.
Published: (2024)
by: Dutta, Siddhant, et al.
Published: (2024)
GAP2WSS: A Genetic Algorithm based on the Pareto Principle for Web Service Selection
by: Khatoonabadi, SayedHassan, et al.
Published: (2021)
by: Khatoonabadi, SayedHassan, et al.
Published: (2021)
The AI Shadow War: SaaS vs. Edge Computing Architectures
by: Marpu, Rhea Pritham, et al.
Published: (2025)
by: Marpu, Rhea Pritham, et al.
Published: (2025)
A Reinforced Evolution-Based Approach to Multi-Resource Load Balancing
by: Sliwko, Leszek
Published: (2025)
by: Sliwko, Leszek
Published: (2025)
Distributed genetic algorithm for application placement in the compute continuum leveraging infrastructure nodes for optimization
by: Guerrero, Carlos, et al.
Published: (2024)
by: Guerrero, Carlos, et al.
Published: (2024)
Enhancing Computational Efficiency in Multiscale Systems Using Deep Learning of Coordinates and Flow Maps
by: Hamid, Asif, et al.
Published: (2024)
by: Hamid, Asif, et al.
Published: (2024)
Carbon-aware Software Services
by: Forti, Stefano, et al.
Published: (2024)
by: Forti, Stefano, et al.
Published: (2024)
Online Continual Learning on Intel Loihi 2 via a Co-designed Spiking Neural Network
by: Hajizada, Elvin, et al.
Published: (2025)
by: Hajizada, Elvin, et al.
Published: (2025)
Similar Items
-
Reducing Data Bottlenecks in Distributed, Heterogeneous Neural Networks
by: Lin, Ruhai, et al.
Published: (2024) -
A Fresh Approach to Evaluate Performance in Distributed Parallel Genetic Algorithms
by: Harada, Tomohiro, et al.
Published: (2021) -
Wireless Sensor Networks as Parallel and Distributed Hardware Platform for Artificial Neural Networks
by: Serpen, Gursel
Published: (2025) -
Cyclic Data Parallelism for Efficient Parallelism of Deep Neural Networks
by: Fournier, Louis, et al.
Published: (2024) -
A Frequency-based Parent Selection for Reducing the Effect of Evaluation Time Bias in Asynchronous Parallel Multi-objective Evolutionary Algorithms
by: Harada, Tomohiro
Published: (2021)