Federated Knowledge Distillation for Multi-Model Architectures Lithography Hotspot Detection
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Yuqi, Lin, Xingyou, Li, Yanli, Zhang, Kai, Yang, Chuanguang, Guo, Zhongliang, Gou, Jianping, Huang, Tingwen, Tian, Yingli |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Accelerating Multi-Scale Deformable Attention Using Near-Memory-Processing Architecture
von: Li, Huize, et al.
Veröffentlicht: (2026)
von: Li, Huize, et al.
Veröffentlicht: (2026)
ReGate: Enabling Power Gating in Neural Processing Units
von: Xue, Yuqi, et al.
Veröffentlicht: (2025)
von: Xue, Yuqi, et al.
Veröffentlicht: (2025)
FirePower: Towards a Foundation with Generalizable Knowledge for Architecture-Level Power Modeling
von: Zhang, Qijun, et al.
Veröffentlicht: (2024)
von: Zhang, Qijun, et al.
Veröffentlicht: (2024)
Optimized Spatial Architecture Mapping Flow for Transformer Accelerators
von: Xu, Haocheng, et al.
Veröffentlicht: (2024)
von: Xu, Haocheng, et al.
Veröffentlicht: (2024)
SkyByte: Architecting an Efficient Memory-Semantic CXL-based SSD with OS and Hardware Co-design
von: Zhang, Haoyang, et al.
Veröffentlicht: (2025)
von: Zhang, Haoyang, et al.
Veröffentlicht: (2025)
Near-Memory Architecture for Threshold-Ordinal Surface-Based Corner Detection of Event Cameras
von: Shang, Hongyang, et al.
Veröffentlicht: (2025)
von: Shang, Hongyang, et al.
Veröffentlicht: (2025)
Lifecycle Cost-Effectiveness Modeling for Redundancy-Enhanced Multi-Chiplet Architectures
von: Liu, Zizhen, et al.
Veröffentlicht: (2026)
von: Liu, Zizhen, et al.
Veröffentlicht: (2026)
APACHE: A Processing-Near-Memory Architecture for Multi-Scheme Fully Homomorphic Encryption
von: Ding, Lin, et al.
Veröffentlicht: (2024)
von: Ding, Lin, et al.
Veröffentlicht: (2024)
UniCAIM: A Unified CAM/CIM Architecture with Static-Dynamic KV Cache Pruning for Efficient Long-Context LLM Inference
von: Xu, Weikai, et al.
Veröffentlicht: (2025)
von: Xu, Weikai, et al.
Veröffentlicht: (2025)
The Immutable Tensor Architecture: A Pure Dataflow Approach for Secure, Energy-Efficient AI Inference
von: Li, Fang
Veröffentlicht: (2025)
von: Li, Fang
Veröffentlicht: (2025)
Cambricon-LLM: A Chiplet-Based Hybrid Architecture for On-Device Inference of 70B LLM
von: Yu, Zhongkai, et al.
Veröffentlicht: (2024)
von: Yu, Zhongkai, et al.
Veröffentlicht: (2024)
Energy-Oriented Computing Architecture Simulator for SNN Training
von: Ma, Yunhao, et al.
Veröffentlicht: (2025)
von: Ma, Yunhao, et al.
Veröffentlicht: (2025)
A System Architecture for Low Latency Multiprogramming Quantum Computing
von: Zhao, Yilun, et al.
Veröffentlicht: (2026)
von: Zhao, Yilun, et al.
Veröffentlicht: (2026)
LlamaF: An Efficient Llama2 Architecture Accelerator on Embedded FPGAs
von: Xu, Han, et al.
Veröffentlicht: (2024)
von: Xu, Han, et al.
Veröffentlicht: (2024)
Focus: A Streaming Concentration Architecture for Efficient Vision-Language Models
von: Wei, Chiyue, et al.
Veröffentlicht: (2025)
von: Wei, Chiyue, et al.
Veröffentlicht: (2025)
Configurable Multi-Port Memory Architecture for High-Speed Data Communication
von: Dhakad, Narendra Singh, et al.
Veröffentlicht: (2024)
von: Dhakad, Narendra Singh, et al.
Veröffentlicht: (2024)
KLiNQ: Knowledge Distillation-Assisted Lightweight Neural Network for Qubit Readout on FPGA
von: Guo, Xiaorang, et al.
Veröffentlicht: (2025)
von: Guo, Xiaorang, et al.
Veröffentlicht: (2025)
THERMOS: Thermally-Aware Multi-Objective Scheduling of AI Workloads on Heterogeneous Multi-Chiplet PIM Architectures
von: Kanani, Alish, et al.
Veröffentlicht: (2025)
von: Kanani, Alish, et al.
Veröffentlicht: (2025)
MFIT: Multi-Fidelity Thermal Modeling for 2.5D and 3D Multi-Chiplet Architectures
von: Pfromm, Lukas, et al.
Veröffentlicht: (2024)
von: Pfromm, Lukas, et al.
Veröffentlicht: (2024)
Benchmarking and Dissecting the Nvidia Hopper GPU Architecture
von: Luo, Weile, et al.
Veröffentlicht: (2024)
von: Luo, Weile, et al.
Veröffentlicht: (2024)
DMSA: A Decentralized Microservice Architecture for Edge Networks
von: Chen, Yuang, et al.
Veröffentlicht: (2025)
von: Chen, Yuang, et al.
Veröffentlicht: (2025)
A Low-Power Sparse Deep Learning Accelerator with Optimized Data Reuse
von: Hsu, Kai-Chieh, et al.
Veröffentlicht: (2025)
von: Hsu, Kai-Chieh, et al.
Veröffentlicht: (2025)
FlexMem: High-Parallel Near-Memory Architecture for Flexible Dataflow in Fully Homomorphic Encryption
von: Shi, Shangyi, et al.
Veröffentlicht: (2025)
von: Shi, Shangyi, et al.
Veröffentlicht: (2025)
Hermes: A Unified High-Performance NTT Architecture with Hybrid Dataflow
von: Gu, Hang, et al.
Veröffentlicht: (2026)
von: Gu, Hang, et al.
Veröffentlicht: (2026)
Chiplet Actuary: A Quantitative Cost Model and Multi-Chiplet Architecture Exploration
von: Feng, Yinxiao, et al.
Veröffentlicht: (2022)
von: Feng, Yinxiao, et al.
Veröffentlicht: (2022)
Optimizing Scalable Multi-Cluster Architectures for Next-Generation Wireless Sensing and Communication
von: Riedel, Samuel, et al.
Veröffentlicht: (2025)
von: Riedel, Samuel, et al.
Veröffentlicht: (2025)
A Tensor-Train Decomposition based Compression of LLMs on Group Vector Systolic Accelerator
von: Huang, Sixiao, et al.
Veröffentlicht: (2025)
von: Huang, Sixiao, et al.
Veröffentlicht: (2025)
Hemlet: A Heterogeneous Compute-in-Memory Chiplet Architecture for Vision Transformers with Group-Level Parallelism
von: Wang, Cong, et al.
Veröffentlicht: (2025)
von: Wang, Cong, et al.
Veröffentlicht: (2025)
NASiC: 3D NAND-based CAM-Selected Multibit CIM Architecture for Efficient On-Device Mixture-of-Experts LLM Inference
von: Xu, Weikai, et al.
Veröffentlicht: (2026)
von: Xu, Weikai, et al.
Veröffentlicht: (2026)
FireFly-S: Exploiting Dual-Side Sparsity for Spiking Neural Networks Acceleration with Reconfigurable Spatial Architecture
von: Li, Tenglong, et al.
Veröffentlicht: (2024)
von: Li, Tenglong, et al.
Veröffentlicht: (2024)
FireFly-T: High-Throughput Sparsity Exploitation for Spiking Transformer Acceleration with Dual-Engine Overlay Architecture
von: Li, Tenglong, et al.
Veröffentlicht: (2025)
von: Li, Tenglong, et al.
Veröffentlicht: (2025)
A Data-Driven Dynamic Execution Orchestration Architecture
von: Bai, Zhenyu, et al.
Veröffentlicht: (2026)
von: Bai, Zhenyu, et al.
Veröffentlicht: (2026)
Splatonic: Architecture Support for 3D Gaussian Splatting SLAM via Sparse Processing
von: Huang, Xiaotong, et al.
Veröffentlicht: (2025)
von: Huang, Xiaotong, et al.
Veröffentlicht: (2025)
Optimized Memory System Architecture for VESA VDC-M Decoder with Multi-Slice Support
von: Yang, Hannah, et al.
Veröffentlicht: (2025)
von: Yang, Hannah, et al.
Veröffentlicht: (2025)
EdgeLLM: A Highly Efficient CPU-FPGA Heterogeneous Edge Accelerator for Large Language Models
von: Huang, Mingqiang, et al.
Veröffentlicht: (2024)
von: Huang, Mingqiang, et al.
Veröffentlicht: (2024)
Survey on Characterizing and Understanding GNNs from a Computer Architecture Perspective
von: Wu, Meng, et al.
Veröffentlicht: (2024)
von: Wu, Meng, et al.
Veröffentlicht: (2024)
Nexus Machine: An Active Message Inspired Reconfigurable Architecture for Irregular Workloads
von: Juneja, Rohan, et al.
Veröffentlicht: (2025)
von: Juneja, Rohan, et al.
Veröffentlicht: (2025)
Overmind NSA: A Unified Neuro-Symbolic Computing Architecture with Approximate Nonlinear Activations and Preemptive Memory Bypass
von: Wang, Weilun, et al.
Veröffentlicht: (2026)
von: Wang, Weilun, et al.
Veröffentlicht: (2026)
AutoPower: Automated Few-Shot Architecture-Level Power Modeling by Power Group Decoupling
von: Zhang, Qijun, et al.
Veröffentlicht: (2025)
von: Zhang, Qijun, et al.
Veröffentlicht: (2025)
Microbenchmarking NVIDIA's Blackwell Architecture: An in-depth Architectural Analysis
von: Jarmusch, Aaron, et al.
Veröffentlicht: (2025)
von: Jarmusch, Aaron, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Accelerating Multi-Scale Deformable Attention Using Near-Memory-Processing Architecture
von: Li, Huize, et al.
Veröffentlicht: (2026) -
ReGate: Enabling Power Gating in Neural Processing Units
von: Xue, Yuqi, et al.
Veröffentlicht: (2025) -
FirePower: Towards a Foundation with Generalizable Knowledge for Architecture-Level Power Modeling
von: Zhang, Qijun, et al.
Veröffentlicht: (2024) -
Optimized Spatial Architecture Mapping Flow for Transformer Accelerators
von: Xu, Haocheng, et al.
Veröffentlicht: (2024) -
SkyByte: Architecting an Efficient Memory-Semantic CXL-based SSD with OS and Hardware Co-design
von: Zhang, Haoyang, et al.
Veröffentlicht: (2025)