Flexible Vector Integration in Embedded RISC-V SoCs for End to End CNN Inference Acceleration
Fuente:
arXiv
Saved in:
| Main Author: | Lyalikov, Dmitri |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GCV-Turbo: End-to-end Acceleration of GNN-based Computer Vision Tasks on FPGA
by: Zhang, Bingyi, et al.
Published: (2024)
by: Zhang, Bingyi, et al.
Published: (2024)
Divide and Save: Splitting Workload Among Containers in an Edge Device to Save Energy and Time
by: Khoshsirat, Aria, et al.
Published: (2023)
by: Khoshsirat, Aria, et al.
Published: (2023)
IPMN Risk Assessment under Federated Learning Paradigm
by: Pan, Hongyi, et al.
Published: (2024)
by: Pan, Hongyi, et al.
Published: (2024)
Decentralized Gossip Mutual Learning (GML) for automatic head and neck tumor segmentation
by: Chen, Jingyun, et al.
Published: (2024)
by: Chen, Jingyun, et al.
Published: (2024)
Cryo-RALib -- a modular library for accelerating alignment in cryo-EM
by: Chung, Szu-Chi, et al.
Published: (2020)
by: Chung, Szu-Chi, et al.
Published: (2020)
VcLLM: Video Codecs are Secretly Tensor Codecs
by: Xu, Ceyu, et al.
Published: (2024)
by: Xu, Ceyu, et al.
Published: (2024)
TREA: Low-precision Time-Multiplexed, Resource-Efficient Edge Accelerator for Object Detection and Classification
by: Sharma, Vijay Pratap, et al.
Published: (2026)
by: Sharma, Vijay Pratap, et al.
Published: (2026)
Multi-DNN Inference of Sparse Models on Edge SoCs
by: Luo, Jiawei, et al.
Published: (2026)
by: Luo, Jiawei, et al.
Published: (2026)
XDMA: A Distributed, Extensible DMA Architecture for Layout-Flexible Data Movements in Heterogeneous Multi-Accelerator SoCs
by: Kong, Fanchen, et al.
Published: (2025)
by: Kong, Fanchen, et al.
Published: (2025)
Adaptive Aggregation Weights for Federated Segmentation of Pancreas MRI
by: Pan, Hongyi, et al.
Published: (2024)
by: Pan, Hongyi, et al.
Published: (2024)
DMVC: Multi-Camera Video Compression Network aimed at Improving Deep Learning Accuracy
by: Cui, Huan, et al.
Published: (2024)
by: Cui, Huan, et al.
Published: (2024)
Deep Learning-based Assessment of the Relation Between the Third Molar and Mandibular Canal on Panoramic Radiographs using Local, Centralized, and Federated Learning
by: Rubak, Johan Andreas Balle, et al.
Published: (2026)
by: Rubak, Johan Andreas Balle, et al.
Published: (2026)
More is Different: Prototyping and Analyzing a New Form of Edge Server with Massive Mobile SoCs
by: Zhang, Li, et al.
Published: (2022)
by: Zhang, Li, et al.
Published: (2022)
Privacy Preserving Federated Learning in Medical Imaging with Uncertainty Estimation
by: Koutsoubis, Nikolas, et al.
Published: (2024)
by: Koutsoubis, Nikolas, et al.
Published: (2024)
FrankenSplit: Efficient Neural Feature Compression with Shallow Variational Bottleneck Injection for Mobile Edge Computing
by: Furutanpey, Alireza, et al.
Published: (2023)
by: Furutanpey, Alireza, et al.
Published: (2023)
FedD2S: Personalized Data-Free Federated Knowledge Distillation
by: Atapour, Kawa, et al.
Published: (2024)
by: Atapour, Kawa, et al.
Published: (2024)
Characterizing Mobile SoC for Accelerating Heterogeneous LLM Inference
by: Chen, Le, et al.
Published: (2025)
by: Chen, Le, et al.
Published: (2025)
Adversarial Robustness of Bottleneck Injected Deep Neural Networks for Task-Oriented Communication
by: Furutanpey, Alireza, et al.
Published: (2024)
by: Furutanpey, Alireza, et al.
Published: (2024)
Flex-PE: Flexible and SIMD Multi-Precision Processing Element for AI Workloads
by: Lokhande, Mukul, et al.
Published: (2024)
by: Lokhande, Mukul, et al.
Published: (2024)
Accelerating stencils on the Tenstorrent Grayskull RISC-V accelerator
by: Brown, Nick, et al.
Published: (2024)
by: Brown, Nick, et al.
Published: (2024)
MATCHA: Efficient Deployment of Deep Neural Networks on Multi-Accelerator Heterogeneous Edge SoCs
by: Russo, Enrico, et al.
Published: (2026)
by: Russo, Enrico, et al.
Published: (2026)
SynthPix: A lightspeed PIV image generator
by: Terpin, Antonio, et al.
Published: (2025)
by: Terpin, Antonio, et al.
Published: (2025)
CFL-SparseMed: Communication-Efficient Federated Learning for Medical Imaging with Top-k Sparse Updates
by: Habib, Gousia, et al.
Published: (2025)
by: Habib, Gousia, et al.
Published: (2025)
Adaptive and Robust Image Processing on CubeSats
by: Bayer, Robert, et al.
Published: (2025)
by: Bayer, Robert, et al.
Published: (2025)
ROIX-Comp: Optimizing X-ray Computed Tomography Imaging Strategy for Data Reduction and Reconstruction
by: Singh, Amarjit, et al.
Published: (2026)
by: Singh, Amarjit, et al.
Published: (2026)
Secure Federated Learning Approaches to Diagnosing COVID-19
by: Adhikari, Rittika, et al.
Published: (2024)
by: Adhikari, Rittika, et al.
Published: (2024)
FedMinds: Privacy-Preserving Personalized Brain Visual Decoding
by: Bao, Guangyin, et al.
Published: (2024)
by: Bao, Guangyin, et al.
Published: (2024)
Semantic Labeling of Large-Area Geographic Regions Using Multi-View and Multi-Date Satellite Images and Noisy OSM Training Labels
by: Comandur, Bharath, et al.
Published: (2020)
by: Comandur, Bharath, et al.
Published: (2020)
Closer in the Gap: Towards Portable Performance on RISC-V Vector Processors
by: Shi, Ruimin, et al.
Published: (2026)
by: Shi, Ruimin, et al.
Published: (2026)
HeRo: Adaptive Orchestration of Agentic RAG on Heterogeneous Mobile SoC
by: Li, Maoliang, et al.
Published: (2026)
by: Li, Maoliang, et al.
Published: (2026)
Accelerating End-Cloud Collaborative Inference via Near Bubble-free Pipeline Optimization
by: Gao, Luyao, et al.
Published: (2024)
by: Gao, Luyao, et al.
Published: (2024)
Fine-Grained Vectorized Merge Sorting on RISC-V: From Register to Cache
by: Zhang, Jin, et al.
Published: (2024)
by: Zhang, Jin, et al.
Published: (2024)
ORBIT: Oak Ridge Base Foundation Model for Earth System Predictability
by: Wang, Xiao, et al.
Published: (2024)
by: Wang, Xiao, et al.
Published: (2024)
Seamless Optical Cloud Computing across Edge-Metro Network for Generative AI
by: Xing, Sizhe, et al.
Published: (2024)
by: Xing, Sizhe, et al.
Published: (2024)
A Performance Analysis Modeling Framework for Extended Reality Applications in Edge-Assisted Wireless Networks
by: Mallik, Anik, et al.
Published: (2024)
by: Mallik, Anik, et al.
Published: (2024)
Kelvin v1.0: A Neural Pre-Encoder for H.264: A standards-compliant learned preprocessor with -27.62% BD-VMAF on UVG
by: Graziano, Marco
Published: (2026)
by: Graziano, Marco
Published: (2026)
PixRO: Pixel-Distributed Rotational Odometry with Gaussian Belief Propagation
by: Alzugaray, Ignacio, et al.
Published: (2024)
by: Alzugaray, Ignacio, et al.
Published: (2024)
Preparing for HPC on RISC-V: Examining Vectorization and Distributed Performance of an Astrophyiscs Application with HPX and Kokkos
by: Diehl, Patrick, et al.
Published: (2024)
by: Diehl, Patrick, et al.
Published: (2024)
Accelerating HDC-CNN Hybrid Models Using Custom Instructions on RISC-V GPUs
by: Matsumi, Wakuto, et al.
Published: (2025)
by: Matsumi, Wakuto, et al.
Published: (2025)
Efficient Architecture for RISC-V Vector Memory Access
by: Guan, Hongyi, et al.
Published: (2025)
by: Guan, Hongyi, et al.
Published: (2025)
Similar Items
-
GCV-Turbo: End-to-end Acceleration of GNN-based Computer Vision Tasks on FPGA
by: Zhang, Bingyi, et al.
Published: (2024) -
Divide and Save: Splitting Workload Among Containers in an Edge Device to Save Energy and Time
by: Khoshsirat, Aria, et al.
Published: (2023) -
IPMN Risk Assessment under Federated Learning Paradigm
by: Pan, Hongyi, et al.
Published: (2024) -
Decentralized Gossip Mutual Learning (GML) for automatic head and neck tumor segmentation
by: Chen, Jingyun, et al.
Published: (2024) -
Cryo-RALib -- a modular library for accelerating alignment in cryo-EM
by: Chung, Szu-Chi, et al.
Published: (2020)