An Open-Source Framework for Efficient Numerically-Tailored Computations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ledoux, Louis, Casas, Marc |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MATLAB Simulator of Level-Index Arithmetic
von: Mikaitis, Mantas
Veröffentlicht: (2024)
von: Mikaitis, Mantas
Veröffentlicht: (2024)
Mixed-precision finite element kernels and assembly: Rounding error analysis and hardware acceleration
von: Croci, M., et al.
Veröffentlicht: (2024)
von: Croci, M., et al.
Veröffentlicht: (2024)
Accurate Models of NVIDIA Tensor Cores
von: Khattak, Faizan A., et al.
Veröffentlicht: (2025)
von: Khattak, Faizan A., et al.
Veröffentlicht: (2025)
eXmY: A Data Type and Technique for Arbitrary Bit Precision Quantization
von: Agrawal, Aditya, et al.
Veröffentlicht: (2024)
von: Agrawal, Aditya, et al.
Veröffentlicht: (2024)
Bit-Accurate Modeling of GPU Matrix Multiply-Accumulate Units: Demystifying Numerical Discrepancy and Accuracy
von: Xie, Peichen, et al.
Veröffentlicht: (2025)
von: Xie, Peichen, et al.
Veröffentlicht: (2025)
QuArch: A Benchmark for Evaluating LLM Reasoning in Computer Architecture
von: Prakash, Shvetank, et al.
Veröffentlicht: (2025)
von: Prakash, Shvetank, et al.
Veröffentlicht: (2025)
Accurate Block Quantization in LLMs with Outliers
von: Trukhanov, Nikita, et al.
Veröffentlicht: (2024)
von: Trukhanov, Nikita, et al.
Veröffentlicht: (2024)
SARIS: Accelerating Stencil Computations on Energy-Efficient RISC-V Compute Clusters with Indirect Stream Registers
von: Scheffler, Paul, et al.
Veröffentlicht: (2024)
von: Scheffler, Paul, et al.
Veröffentlicht: (2024)
A Survey: Collaborative Hardware and Software Design in the Era of Large Language Models
von: Guo, Cong, et al.
Veröffentlicht: (2024)
von: Guo, Cong, et al.
Veröffentlicht: (2024)
David vs. Goliath: Can Small Models Win Big with Agentic AI in Hardware Design?
von: Shankar, Shashwat, et al.
Veröffentlicht: (2025)
von: Shankar, Shashwat, et al.
Veröffentlicht: (2025)
Generalized Methodology for Determining Numerical Features of Hardware Floating-Point Matrix Multipliers: Part I
von: Khattak, Faizan A, et al.
Veröffentlicht: (2025)
von: Khattak, Faizan A, et al.
Veröffentlicht: (2025)
UGrid: An Efficient-And-Rigorous Neural Multigrid Solver for Linear PDEs
von: Han, Xi, et al.
Veröffentlicht: (2024)
von: Han, Xi, et al.
Veröffentlicht: (2024)
FuzzWiz -- Fuzzing Framework for Efficient Hardware Coverage
von: Gadde, Deepak Narayan, et al.
Veröffentlicht: (2024)
von: Gadde, Deepak Narayan, et al.
Veröffentlicht: (2024)
Toward Capturing Genetic Epistasis From Multivariate Genome-Wide Association Studies Using Mixed-Precision Kernel Ridge Regression
von: Ltaief, Hatem, et al.
Veröffentlicht: (2024)
von: Ltaief, Hatem, et al.
Veröffentlicht: (2024)
ChipExpert: The Open-Source Integrated-Circuit-Design-Specific Large Language Model
von: Xu, Ning, et al.
Veröffentlicht: (2024)
von: Xu, Ning, et al.
Veröffentlicht: (2024)
rule4ml: An Open-Source Tool for Resource Utilization and Latency Estimation for ML Models on FPGA
von: Rahimifar, Mohammad Mehdi, et al.
Veröffentlicht: (2024)
von: Rahimifar, Mohammad Mehdi, et al.
Veröffentlicht: (2024)
On Stochastic Rounding with Few Random Bits
von: Fitzgibbon, Andrew, et al.
Veröffentlicht: (2025)
von: Fitzgibbon, Andrew, et al.
Veröffentlicht: (2025)
MEMHD: Memory-Efficient Multi-Centroid Hyperdimensional Computing for Fully-Utilized In-Memory Computing Architectures
von: Kang, Do Yeong, et al.
Veröffentlicht: (2025)
von: Kang, Do Yeong, et al.
Veröffentlicht: (2025)
AttentionLego: An Open-Source Building Block For Spatially-Scalable Large Language Model Accelerator With Processing-In-Memory Technology
von: Cong, Rongqing, et al.
Veröffentlicht: (2024)
von: Cong, Rongqing, et al.
Veröffentlicht: (2024)
Design and accuracy trade-offs in Computational Statistics
von: Xu, Tiancheng, et al.
Veröffentlicht: (2025)
von: Xu, Tiancheng, et al.
Veröffentlicht: (2025)
Efficient Calibration for RRAM-based In-Memory Computing using DoRA
von: Dong, Weirong, et al.
Veröffentlicht: (2025)
von: Dong, Weirong, et al.
Veröffentlicht: (2025)
Hawkeye: Reproducing GPU-Level Non-Determinism
von: Badash, Erez, et al.
Veröffentlicht: (2026)
von: Badash, Erez, et al.
Veröffentlicht: (2026)
ElasticAI: Creating and Deploying Energy-Efficient Deep Learning Accelerator for Pervasive Computing
von: Qian, Chao, et al.
Veröffentlicht: (2024)
von: Qian, Chao, et al.
Veröffentlicht: (2024)
NeuralMatrix: Compute the Entire Neural Networks with Linear Matrix Operations for Efficient Inference
von: Sun, Ruiqi, et al.
Veröffentlicht: (2023)
von: Sun, Ruiqi, et al.
Veröffentlicht: (2023)
Column-wise Quantization of Weights and Partial Sums for Accurate and Efficient Compute-In-Memory Accelerators
von: Kim, Jiyoon, et al.
Veröffentlicht: (2025)
von: Kim, Jiyoon, et al.
Veröffentlicht: (2025)
Inexactness and Correction of Floating-Point Reciprocal, Division and Square Root
von: Dutton, Lucas M., et al.
Veröffentlicht: (2024)
von: Dutton, Lucas M., et al.
Veröffentlicht: (2024)
Hardware-Accelerated Algorithm for Complex Function Roots Density Graph Plotting
von: Tang, Ruibai, et al.
Veröffentlicht: (2025)
von: Tang, Ruibai, et al.
Veröffentlicht: (2025)
Zero-Shot RTL Code Generation with Attention Sink Augmented Large Language Models
von: Sandal, Selim, et al.
Veröffentlicht: (2024)
von: Sandal, Selim, et al.
Veröffentlicht: (2024)
CoopetitiveV: Leveraging LLM-powered Coopetitive Multi-Agent Prompting for High-quality Verilog Generation
von: Mi, Zhendong, et al.
Veröffentlicht: (2024)
von: Mi, Zhendong, et al.
Veröffentlicht: (2024)
Efficient FRW Transitions via Stochastic Finite Differences for Handling Non-Stratified Dielectrics
von: Huang, Jiechen, et al.
Veröffentlicht: (2025)
von: Huang, Jiechen, et al.
Veröffentlicht: (2025)
VeriCoder: Enhancing LLM-Based RTL Code Generation through Functional Correctness Validation
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
von: Wei, Anjiang, et al.
Veröffentlicht: (2025)
A Hybrid Residue Floating Numerical Architecture for High Precision Arithmetic on FPGAs
von: Darvishi, Mostafa
Veröffentlicht: (2025)
von: Darvishi, Mostafa
Veröffentlicht: (2025)
The Cream Rises to the Top: Efficient Reranking Method for Verilog Code Generation
von: Yang, Guang, et al.
Veröffentlicht: (2025)
von: Yang, Guang, et al.
Veröffentlicht: (2025)
vTrain: A Simulation Framework for Evaluating Cost-effective and Compute-optimal Large Language Model Training
von: Bang, Jehyeon, et al.
Veröffentlicht: (2023)
von: Bang, Jehyeon, et al.
Veröffentlicht: (2023)
A Semi-Formal Verification Methodology for Efficient Configuration Coverage of Highly Configurable Digital Designs
von: Kumar, Aman, et al.
Veröffentlicht: (2024)
von: Kumar, Aman, et al.
Veröffentlicht: (2024)
Veri-Sure: A Contract-Aware Multi-Agent Framework with Temporal Tracing and Formal Verification for Correct RTL Code Generation
von: Liu, Jiale, et al.
Veröffentlicht: (2026)
von: Liu, Jiale, et al.
Veröffentlicht: (2026)
Hierarchical Source-to-Post-Route QoR Prediction in High-Level Synthesis with GNNs
von: Gao, Mingzhe, et al.
Veröffentlicht: (2024)
von: Gao, Mingzhe, et al.
Veröffentlicht: (2024)
HYLU: Hybrid Parallel Sparse LU Factorization
von: Chen, Xiaoming
Veröffentlicht: (2025)
von: Chen, Xiaoming
Veröffentlicht: (2025)
Design Rules for Extreme-Edge Scientific Computing on AI Engines
von: Ma, Zhenghua, et al.
Veröffentlicht: (2026)
von: Ma, Zhenghua, et al.
Veröffentlicht: (2026)
The Role of Advanced Computer Architectures in Accelerating Artificial Intelligence Workloads
von: Amin, Shahid, et al.
Veröffentlicht: (2025)
von: Amin, Shahid, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
MATLAB Simulator of Level-Index Arithmetic
von: Mikaitis, Mantas
Veröffentlicht: (2024) -
Mixed-precision finite element kernels and assembly: Rounding error analysis and hardware acceleration
von: Croci, M., et al.
Veröffentlicht: (2024) -
Accurate Models of NVIDIA Tensor Cores
von: Khattak, Faizan A., et al.
Veröffentlicht: (2025) -
eXmY: A Data Type and Technique for Arbitrary Bit Precision Quantization
von: Agrawal, Aditya, et al.
Veröffentlicht: (2024) -
Bit-Accurate Modeling of GPU Matrix Multiply-Accumulate Units: Demystifying Numerical Discrepancy and Accuracy
von: Xie, Peichen, et al.
Veröffentlicht: (2025)