TensorTEE: Unifying Heterogeneous TEE Granularity for Efficient Secure Collaborative Tensor Computing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Han, Husheng, Zheng, Xinyao, Wen, Yuanbo, Hao, Yifan, Feng, Erhu, Liang, Ling, Mu, Jianan, Li, Xiaqing, Ma, Tianyun, Jin, Pengwei, Song, Xinkai, Du, Zidong, Guo, Qi, Hu, Xing |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AGON: Automated Design Framework for Customizing Processors from ISA Documents
von: Li, Chongxiao, et al.
Veröffentlicht: (2024)
von: Li, Chongxiao, et al.
Veröffentlicht: (2024)
FlexMem: High-Parallel Near-Memory Architecture for Flexible Dataflow in Fully Homomorphic Encryption
von: Shi, Shangyi, et al.
Veröffentlicht: (2025)
von: Shi, Shangyi, et al.
Veröffentlicht: (2025)
HE^2: A Communication-Light Heterogeneous Architecture for Efficient Fully Homomorphic Encryption
von: Shi, Shangyi, et al.
Veröffentlicht: (2026)
von: Shi, Shangyi, et al.
Veröffentlicht: (2026)
AutoPPA: Automated Circuit PPA Optimization via Contrastive Code-based Rule Library Learning
von: Li, Chongxiao, et al.
Veröffentlicht: (2026)
von: Li, Chongxiao, et al.
Veröffentlicht: (2026)
Vulnerabilities in Partial TEE-Shielded LLM Inference with Precomputed Noise
von: Saini, Abhishek, et al.
Veröffentlicht: (2026)
von: Saini, Abhishek, et al.
Veröffentlicht: (2026)
QiMeng-CPU-v2: Automated Superscalar Processor Design by Learning Data Dependencies
von: Cheng, Shuyao, et al.
Veröffentlicht: (2025)
von: Cheng, Shuyao, et al.
Veröffentlicht: (2025)
Cambricon-LLM: A Chiplet-Based Hybrid Architecture for On-Device Inference of 70B LLM
von: Yu, Zhongkai, et al.
Veröffentlicht: (2024)
von: Yu, Zhongkai, et al.
Veröffentlicht: (2024)
Pecker: Bug Localization Framework for Sequential Designs via Causal Chain Reconstruction
von: Tang, Jiaping, et al.
Veröffentlicht: (2026)
von: Tang, Jiaping, et al.
Veröffentlicht: (2026)
StreamTensor: Make Tensors Stream in Dataflow Accelerators for LLMs
von: Ye, Hanchen, et al.
Veröffentlicht: (2025)
von: Ye, Hanchen, et al.
Veröffentlicht: (2025)
FADiff: Fusion-Aware Differentiable Optimization for DNN Scheduling on Tensor Accelerators
von: Jia, Shuao, et al.
Veröffentlicht: (2025)
von: Jia, Shuao, et al.
Veröffentlicht: (2025)
The Immutable Tensor Architecture: A Pure Dataflow Approach for Secure, Energy-Efficient AI Inference
von: Li, Fang
Veröffentlicht: (2025)
von: Li, Fang
Veröffentlicht: (2025)
RealBench: Benchmarking Verilog Generation Models with Real-World IP Designs
von: Jin, Pengwei, et al.
Veröffentlicht: (2025)
von: Jin, Pengwei, et al.
Veröffentlicht: (2025)
Tensor Manipulation Unit (TMU): Reconfigurable, Near-Memory Tensor Manipulation for High-Throughput AI SoC
von: Zhou, Weiyu, et al.
Veröffentlicht: (2025)
von: Zhou, Weiyu, et al.
Veröffentlicht: (2025)
Error Checking for Sparse Systolic Tensor Arrays
von: Peltekis, Christodoulos, et al.
Veröffentlicht: (2024)
von: Peltekis, Christodoulos, et al.
Veröffentlicht: (2024)
Think with Self-Decoupling and Self-Verification: Automated RTL Design with Backtrack-ToT
von: Chao, Zhiteng, et al.
Veröffentlicht: (2025)
von: Chao, Zhiteng, et al.
Veröffentlicht: (2025)
Open-source Stand-Alone Versatile Tensor Accelerator
von: Faure-Gignoux, Anthony, et al.
Veröffentlicht: (2025)
von: Faure-Gignoux, Anthony, et al.
Veröffentlicht: (2025)
ATiM: Autotuning Tensor Programs for Processing-in-DRAM
von: Shin, Yongwon, et al.
Veröffentlicht: (2024)
von: Shin, Yongwon, et al.
Veröffentlicht: (2024)
ATLAAS: Automatic Tensor-Level Abstraction of Accelerator Semantics
von: Gao, Ruijie, et al.
Veröffentlicht: (2026)
von: Gao, Ruijie, et al.
Veröffentlicht: (2026)
Tailors: Accelerating Sparse Tensor Algebra by Overbooking Buffer Capacity
von: Xue, Zi Yu, et al.
Veröffentlicht: (2023)
von: Xue, Zi Yu, et al.
Veröffentlicht: (2023)
N-TORC: Native Tensor Optimizer for Real-time Constraints
von: Singh, Suyash Vardhan, et al.
Veröffentlicht: (2025)
von: Singh, Suyash Vardhan, et al.
Veröffentlicht: (2025)
Tensor Memory Engine: On-the-fly Data Reorganization for Ideal Locality
von: Hoornaert, Denis, et al.
Veröffentlicht: (2026)
von: Hoornaert, Denis, et al.
Veröffentlicht: (2026)
AME-PIM: Can Memory be Your Next Tensor Accelerator?
von: Venieri, Emanuele, et al.
Veröffentlicht: (2026)
von: Venieri, Emanuele, et al.
Veröffentlicht: (2026)
FETTA: Flexible and Efficient Hardware Accelerator for Tensorized Neural Network Training
von: Lu, Jinming, et al.
Veröffentlicht: (2025)
von: Lu, Jinming, et al.
Veröffentlicht: (2025)
TeAAL: A Declarative Framework for Modeling Sparse Tensor Accelerators
von: Nayak, Nandeeka, et al.
Veröffentlicht: (2023)
von: Nayak, Nandeeka, et al.
Veröffentlicht: (2023)
RIROS: A Parallel RTL Fault SImulation FRamework with TwO-Dimensional Parallelism and Unified Schedule
von: Tang, Jiaping, et al.
Veröffentlicht: (2025)
von: Tang, Jiaping, et al.
Veröffentlicht: (2025)
QiMeng: Fully Automated Hardware and Software Design for Processor Chip
von: Zhang, Rui, et al.
Veröffentlicht: (2025)
von: Zhang, Rui, et al.
Veröffentlicht: (2025)
EN-T: Optimizing Tensor Computing Engines Performance via Encoder-Based Methodology
von: Wu, Qizhe, et al.
Veröffentlicht: (2024)
von: Wu, Qizhe, et al.
Veröffentlicht: (2024)
Scaling Photonic Tensor Cores with Unary and Homodyne Designs
von: Alo, Oluwaseun, et al.
Veröffentlicht: (2026)
von: Alo, Oluwaseun, et al.
Veröffentlicht: (2026)
Systolic Sparse Tensor Slices: FPGA Building Blocks for Sparse and Dense AI Acceleration
von: Taka, Endri, et al.
Veröffentlicht: (2025)
von: Taka, Endri, et al.
Veröffentlicht: (2025)
GTA: a new General Tensor Accelerator with Better Area Efficiency and Data Reuse
von: Ai, Chenyang, et al.
Veröffentlicht: (2024)
von: Ai, Chenyang, et al.
Veröffentlicht: (2024)
A Tensor-Train Decomposition based Compression of LLMs on Group Vector Systolic Accelerator
von: Huang, Sixiao, et al.
Veröffentlicht: (2025)
von: Huang, Sixiao, et al.
Veröffentlicht: (2025)
PyPIM: Integrating Digital Processing-in-Memory from Microarchitectural Design to Python Tensors
von: Leitersdorf, Orian, et al.
Veröffentlicht: (2023)
von: Leitersdorf, Orian, et al.
Veröffentlicht: (2023)
RTeAAL Sim: Using Tensor Algebra to Represent and Accelerate RTL Simulation (Extended Version)
von: Zhu, Yan, et al.
Veröffentlicht: (2026)
von: Zhu, Yan, et al.
Veröffentlicht: (2026)
TEMP: A Memory Efficient Physical-aware Tensor Partition-Mapping Framework on Wafer-scale Chips
von: Wang, Huizheng, et al.
Veröffentlicht: (2025)
von: Wang, Huizheng, et al.
Veröffentlicht: (2025)
RTLSeek: Boosting the LLM-Based RTL Generation with Multi-Stage Diversity-Oriented Reinforcement Learning
von: Zhang, Xinyu, et al.
Veröffentlicht: (2026)
von: Zhang, Xinyu, et al.
Veröffentlicht: (2026)
ITERA-LLM: Boosting Sub-8-Bit Large Language Model Inference via Iterative Tensor Decomposition
von: Zheng, Keran, et al.
Veröffentlicht: (2025)
von: Zheng, Keran, et al.
Veröffentlicht: (2025)
Extend IVerilog to Support Batch RTL Fault Simulation
von: Tang, Jiaping, et al.
Veröffentlicht: (2025)
von: Tang, Jiaping, et al.
Veröffentlicht: (2025)
Periodic Online Testing for Sparse Systolic Tensor Arrays
von: Peltekis, Christodoulos, et al.
Veröffentlicht: (2025)
von: Peltekis, Christodoulos, et al.
Veröffentlicht: (2025)
From Principles to Practice: A Systematic Study of LLM Serving on Multi-core NPUs
von: Zhu, Tianhao, et al.
Veröffentlicht: (2025)
von: Zhu, Tianhao, et al.
Veröffentlicht: (2025)
Accelerating Sparse Graph Neural Networks with Tensor Core Optimization
von: Wu, Ka Wai
Veröffentlicht: (2024)
von: Wu, Ka Wai
Veröffentlicht: (2024)
Ähnliche Einträge
-
AGON: Automated Design Framework for Customizing Processors from ISA Documents
von: Li, Chongxiao, et al.
Veröffentlicht: (2024) -
FlexMem: High-Parallel Near-Memory Architecture for Flexible Dataflow in Fully Homomorphic Encryption
von: Shi, Shangyi, et al.
Veröffentlicht: (2025) -
HE^2: A Communication-Light Heterogeneous Architecture for Efficient Fully Homomorphic Encryption
von: Shi, Shangyi, et al.
Veröffentlicht: (2026) -
AutoPPA: Automated Circuit PPA Optimization via Contrastive Code-based Rule Library Learning
von: Li, Chongxiao, et al.
Veröffentlicht: (2026) -
Vulnerabilities in Partial TEE-Shielded LLM Inference with Precomputed Noise
von: Saini, Abhishek, et al.
Veröffentlicht: (2026)