Optimizing Branch Predictor for Graph Applications
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Upasna, Tavva, Venkata Kalyan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Efficient Page Migration in Hybrid Memory Systems
von: Upasna, et al.
Veröffentlicht: (2026)
von: Upasna, et al.
Veröffentlicht: (2026)
RowHammer Vulnerability Counter (RVC): Redefining RowHammer Detection with Victim-Centric Tracking
von: Jain, Lavi, et al.
Veröffentlicht: (2026)
von: Jain, Lavi, et al.
Veröffentlicht: (2026)
ReLMXEL: Adaptive RL-Based Memory Controller with Explainable Energy and Latency Optimization
von: Sai, Panuganti Chirag, et al.
Veröffentlicht: (2026)
von: Sai, Panuganti Chirag, et al.
Veröffentlicht: (2026)
Dissecting Conditional Branch Predictors of Apple Firestorm and Qualcomm Oryon for Software Optimization and Architectural Analysis
von: Chen, Jiajie, et al.
Veröffentlicht: (2024)
von: Chen, Jiajie, et al.
Veröffentlicht: (2024)
Exposing Shadow Branches
von: Pepi, Chrysanthos, et al.
Veröffentlicht: (2024)
von: Pepi, Chrysanthos, et al.
Veröffentlicht: (2024)
Workload Characterization for Branch Predictability
von: Vikas, FNU, et al.
Veröffentlicht: (2025)
von: Vikas, FNU, et al.
Veröffentlicht: (2025)
Branch Target Buffer Reverse Engineering on Arm
von: Wan, Junpeng
Veröffentlicht: (2024)
von: Wan, Junpeng
Veröffentlicht: (2024)
The Non-Predictability of Mispredicted Branches using Timing Information
von: Constantinou, Ioannis, et al.
Veröffentlicht: (2026)
von: Constantinou, Ioannis, et al.
Veröffentlicht: (2026)
ROVER: RTL Optimization via Verified E-Graph Rewriting
von: Coward, Samuel, et al.
Veröffentlicht: (2024)
von: Coward, Samuel, et al.
Veröffentlicht: (2024)
Branch Prediction in Hardcaml for a RISC-V 32im CPU
von: Saveau, Alex
Veröffentlicht: (2023)
von: Saveau, Alex
Veröffentlicht: (2023)
Analyzing and Exploiting Branch Mispredictions in Microcode
von: Mosier, Nicholas, et al.
Veröffentlicht: (2025)
von: Mosier, Nicholas, et al.
Veröffentlicht: (2025)
GDR-HGNN: A Heterogeneous Graph Neural Networks Accelerator Frontend with Graph Decoupling and Recoupling
von: Xue, Runzhen, et al.
Veröffentlicht: (2024)
von: Xue, Runzhen, et al.
Veröffentlicht: (2024)
Leveraging Recurrent Patterns in Graph Accelerators
von: Rahimi, Masoud, et al.
Veröffentlicht: (2025)
von: Rahimi, Masoud, et al.
Veröffentlicht: (2025)
Efficient and Accurate Graph Classification with Hyperdimensional Computing on FPGA
von: Arockiaraj, Jebacyril, et al.
Veröffentlicht: (2025)
von: Arockiaraj, Jebacyril, et al.
Veröffentlicht: (2025)
Exploiting Inaccurate Branch History in Side-Channel Attacks
von: Zhu, Yuhui, et al.
Veröffentlicht: (2025)
von: Zhu, Yuhui, et al.
Veröffentlicht: (2025)
Holistic Optimization Framework for FPGA Accelerators
von: Pouget, Stéphane, et al.
Veröffentlicht: (2025)
von: Pouget, Stéphane, et al.
Veröffentlicht: (2025)
RidgeWalker: Perfectly Pipelined Graph Random Walks on FPGAs
von: Tan, Hongshi, et al.
Veröffentlicht: (2026)
von: Tan, Hongshi, et al.
Veröffentlicht: (2026)
Flexible Bit-Truncation Memory for Approximate Applications on the Edge
von: Oswald, William, et al.
Veröffentlicht: (2025)
von: Oswald, William, et al.
Veröffentlicht: (2025)
CIS: Composable Instruction Set for Data Streaming Applications
von: Yang, Yu, et al.
Veröffentlicht: (2024)
von: Yang, Yu, et al.
Veröffentlicht: (2024)
LightningSimV2: Faster and Scalable Simulation for High-Level Synthesis via Graph Compilation and Optimization
von: Sarkar, Rishov, et al.
Veröffentlicht: (2024)
von: Sarkar, Rishov, et al.
Veröffentlicht: (2024)
CODO: An Automated Compiler for Comprehensive Dataflow Optimization
von: Zhang, Weichuang, et al.
Veröffentlicht: (2026)
von: Zhang, Weichuang, et al.
Veröffentlicht: (2026)
Convolutions Predictable Offloading to an Accelerator: Formalization and Optimization
von: Husson, Benjamin, et al.
Veröffentlicht: (2026)
von: Husson, Benjamin, et al.
Veröffentlicht: (2026)
Fast Cross-Operator Optimization of Attention Dataflow
von: Chang, Haodong, et al.
Veröffentlicht: (2026)
von: Chang, Haodong, et al.
Veröffentlicht: (2026)
Modeling and Optimizing Performance Bottlenecks for Neuromorphic Accelerators
von: Yik, Jason, et al.
Veröffentlicht: (2025)
von: Yik, Jason, et al.
Veröffentlicht: (2025)
ACS: Concurrent Kernel Execution on Irregular, Input-Dependent Computational Graphs
von: Durvasula, Sankeerth, et al.
Veröffentlicht: (2024)
von: Durvasula, Sankeerth, et al.
Veröffentlicht: (2024)
Affordable HPC: Leveraging Small Clusters for Big Data and Graph Computing
von: Wu, Ruilong, et al.
Veröffentlicht: (2024)
von: Wu, Ruilong, et al.
Veröffentlicht: (2024)
Image processing Application Development on Software Configurable Processor Array
von: Prabhu, Ganesh, et al.
Veröffentlicht: (2025)
von: Prabhu, Ganesh, et al.
Veröffentlicht: (2025)
A Mess of Memory System Benchmarking, Simulation and Application Profiling
von: Esmaili-Dokht, Pouya, et al.
Veröffentlicht: (2024)
von: Esmaili-Dokht, Pouya, et al.
Veröffentlicht: (2024)
CIBPU: A Conflict-Invisible Secure Branch Prediction Unit
von: Zhou, Zhe, et al.
Veröffentlicht: (2025)
von: Zhou, Zhe, et al.
Veröffentlicht: (2025)
Piccolo: Large-Scale Graph Processing with Fine-Grained In-Memory Scatter-Gather
von: Shin, Changmin, et al.
Veröffentlicht: (2025)
von: Shin, Changmin, et al.
Veröffentlicht: (2025)
Fast Graph Vector Search via Hardware Acceleration and Delayed-Synchronization Traversal
von: Jiang, Wenqi, et al.
Veröffentlicht: (2024)
von: Jiang, Wenqi, et al.
Veröffentlicht: (2024)
Swift: A Multi-FPGA Framework for Scaling Up Accelerated Graph Analytics
von: Jaiyeoba, Oluwole, et al.
Veröffentlicht: (2024)
von: Jaiyeoba, Oluwole, et al.
Veröffentlicht: (2024)
GUST: Graph Edge-Coloring Utilization for Accelerating Sparse Matrix Vector Multiplication
von: Gerami, Armin, et al.
Veröffentlicht: (2024)
von: Gerami, Armin, et al.
Veröffentlicht: (2024)
Optimizing GEMM for Energy and Performance on Versal ACAP Architectures
von: Papalamprou, Ilias, et al.
Veröffentlicht: (2025)
von: Papalamprou, Ilias, et al.
Veröffentlicht: (2025)
Optimized Spatial Architecture Mapping Flow for Transformer Accelerators
von: Xu, Haocheng, et al.
Veröffentlicht: (2024)
von: Xu, Haocheng, et al.
Veröffentlicht: (2024)
Combining Power and Arithmetic Optimization via Datapath Rewriting
von: Coward, Samuel, et al.
Veröffentlicht: (2024)
von: Coward, Samuel, et al.
Veröffentlicht: (2024)
Optimizing Energy Efficiency in Subthreshold RISC-V Cores
von: Djupdal, Asbjørn, et al.
Veröffentlicht: (2025)
von: Djupdal, Asbjørn, et al.
Veröffentlicht: (2025)
MENAGE: Mixed-Signal Event-Driven Neuromorphic Accelerator for Edge Applications
von: Abdollahi, Armin, et al.
Veröffentlicht: (2024)
von: Abdollahi, Armin, et al.
Veröffentlicht: (2024)
Evaluating the Effectiveness of Microarchitectural Hardware Fault Detection for Application-Specific Requirements
von: Papadopoulos, Konstantinos-Nikolaos, et al.
Veröffentlicht: (2024)
von: Papadopoulos, Konstantinos-Nikolaos, et al.
Veröffentlicht: (2024)
AutoGNN: End-to-End Hardware-Driven Graph Preprocessing for Enhanced GNN Performance
von: Kang, Seungkwan, et al.
Veröffentlicht: (2026)
von: Kang, Seungkwan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Efficient Page Migration in Hybrid Memory Systems
von: Upasna, et al.
Veröffentlicht: (2026) -
RowHammer Vulnerability Counter (RVC): Redefining RowHammer Detection with Victim-Centric Tracking
von: Jain, Lavi, et al.
Veröffentlicht: (2026) -
ReLMXEL: Adaptive RL-Based Memory Controller with Explainable Energy and Latency Optimization
von: Sai, Panuganti Chirag, et al.
Veröffentlicht: (2026) -
Dissecting Conditional Branch Predictors of Apple Firestorm and Qualcomm Oryon for Software Optimization and Architectural Analysis
von: Chen, Jiajie, et al.
Veröffentlicht: (2024) -
Exposing Shadow Branches
von: Pepi, Chrysanthos, et al.
Veröffentlicht: (2024)