Cocoon: A System Architecture for Differentially Private Training with Correlated Noises
Fuente:
arXiv
Salvato in:
| Autori principali: | Kim, Donghwan, Gu, Xin, Baek, Jinho, Lo, Timothy, Min, Younghoon, Shin, Kwangsik, Kim, Jongryool, Park, Jongse, Maeng, Kiwan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Characterization of Multi-Model Agentic AI Systems on General Tasks via Trace-Driven Simulation
di: Kim, Donghwan, et al.
Pubblicazione: (2026)
di: Kim, Donghwan, et al.
Pubblicazione: (2026)
Towards Understanding Systems Trade-offs in Retrieval-Augmented Generation Model Inference
di: Shen, Michael, et al.
Pubblicazione: (2024)
di: Shen, Michael, et al.
Pubblicazione: (2024)
MixDiT: Accelerating Image Diffusion Transformer Inference with Mixed-Precision MX Quantization
di: Kim, Daeun, et al.
Pubblicazione: (2025)
di: Kim, Daeun, et al.
Pubblicazione: (2025)
IBEX: Internal Bandwidth-Efficient Compression Architecture for Scalable CXL Memory Expansion
di: Ko, Younghoon, et al.
Pubblicazione: (2026)
di: Ko, Younghoon, et al.
Pubblicazione: (2026)
ONNXim: A Fast, Cycle-level Multi-core NPU Simulator
di: Ham, Hyungkyu, et al.
Pubblicazione: (2024)
di: Ham, Hyungkyu, et al.
Pubblicazione: (2024)
Accelerating String-Key Learned Index Structures via Memoization-based Incremental Training
di: Kim, Minsu, et al.
Pubblicazione: (2024)
di: Kim, Minsu, et al.
Pubblicazione: (2024)
LOCALUT: Harnessing Capacity-Computation Tradeoffs for LUT-Based Inference in DRAM-PIM
di: Hong, Junguk, et al.
Pubblicazione: (2026)
di: Hong, Junguk, et al.
Pubblicazione: (2026)
NeuPIMs: NPU-PIM Heterogeneous Acceleration for Batched LLM Inferencing
di: Heo, Guseul, et al.
Pubblicazione: (2024)
di: Heo, Guseul, et al.
Pubblicazione: (2024)
Oaken: Fast and Efficient LLM Serving with Online-Offline Hybrid KV Cache Quantization
di: Kim, Minsu, et al.
Pubblicazione: (2025)
di: Kim, Minsu, et al.
Pubblicazione: (2025)
LPU: A Latency-Optimized and Highly Scalable Processor for Large Language Model Inference
di: Moon, Seungjae, et al.
Pubblicazione: (2024)
di: Moon, Seungjae, et al.
Pubblicazione: (2024)
Smart-Infinity: Fast Large Language Model Training using Near-Storage Processing on a Real System
di: Jang, Hongsun, et al.
Pubblicazione: (2024)
di: Jang, Hongsun, et al.
Pubblicazione: (2024)
cMPI: Using CXL Memory Sharing for MPI One-Sided and Two-Sided Inter-Node Communications
di: Wang, Xi, et al.
Pubblicazione: (2025)
di: Wang, Xi, et al.
Pubblicazione: (2025)
Pimba: A Processing-in-Memory Acceleration for Post-Transformer Large Language Model Serving
di: Kim, Wonung, et al.
Pubblicazione: (2025)
di: Kim, Wonung, et al.
Pubblicazione: (2025)
A Cost-Effective Near-Storage Processing Solution for Offline Inference of Long-Context LLMs
di: Jang, Hongsun, et al.
Pubblicazione: (2025)
di: Jang, Hongsun, et al.
Pubblicazione: (2025)
SCRec: A Scalable Computational Storage System with Statistical Sharding and Tensor-train Decomposition for Recommendation Models
di: Yang, Jinho, et al.
Pubblicazione: (2025)
di: Yang, Jinho, et al.
Pubblicazione: (2025)
SoK: Systematizing a Decade of Architectural RowHammer Defenses Through the Lens of Streaming Algorithms
di: Kim, Michael Jaemin, et al.
Pubblicazione: (2025)
di: Kim, Michael Jaemin, et al.
Pubblicazione: (2025)
Neo: Real-Time On-Device 3D Gaussian Splatting with Reuse-and-Update Sorting Acceleration
di: Oh, Changhun, et al.
Pubblicazione: (2025)
di: Oh, Changhun, et al.
Pubblicazione: (2025)
RoMe: Row Granularity Access Memory System for Large Language Models
di: Nam, Hwayong, et al.
Pubblicazione: (2025)
di: Nam, Hwayong, et al.
Pubblicazione: (2025)
HURRY: Highly Utilized, Reconfigurable ReRAM-based In-situ Accelerator with Multifunctionality
di: Shin, Hery, et al.
Pubblicazione: (2024)
di: Shin, Hery, et al.
Pubblicazione: (2024)
PVAC: A RowHammer Mitigation Architecture Exploiting Per-victim-row Counting
di: Kim, Jumin, et al.
Pubblicazione: (2026)
di: Kim, Jumin, et al.
Pubblicazione: (2026)
Optimized Memory System Architecture for VESA VDC-M Decoder with Multi-Slice Support
di: Yang, Hannah, et al.
Pubblicazione: (2025)
di: Yang, Hannah, et al.
Pubblicazione: (2025)
DaCapo: Accelerating Continuous Learning in Autonomous Systems for Video Analytics
di: Kim, Yoonsung, et al.
Pubblicazione: (2024)
di: Kim, Yoonsung, et al.
Pubblicazione: (2024)
IANUS: Integrated Accelerator based on NPU-PIM Unified Memory System
di: Seo, Minseok, et al.
Pubblicazione: (2024)
di: Seo, Minseok, et al.
Pubblicazione: (2024)
IVE: An Accelerator for Single-Server Private Information Retrieval Using Versatile Processing Elements
di: Kim, Sangpyo, et al.
Pubblicazione: (2025)
di: Kim, Sangpyo, et al.
Pubblicazione: (2025)
Piccolo: Large-Scale Graph Processing with Fine-Grained In-Memory Scatter-Gather
di: Shin, Changmin, et al.
Pubblicazione: (2025)
di: Shin, Changmin, et al.
Pubblicazione: (2025)
SAL-PIM: A Subarray-level Processing-in-Memory Architecture with LUT-based Linear Interpolation for Transformer-based Text Generation
di: Han, Wontak, et al.
Pubblicazione: (2024)
di: Han, Wontak, et al.
Pubblicazione: (2024)
Pathfinding Future PIM Architectures by Demystifying a Commercial PIM Technology
di: Hyun, Bongjoon, et al.
Pubblicazione: (2023)
di: Hyun, Bongjoon, et al.
Pubblicazione: (2023)
Energy-Oriented Computing Architecture Simulator for SNN Training
di: Ma, Yunhao, et al.
Pubblicazione: (2025)
di: Ma, Yunhao, et al.
Pubblicazione: (2025)
STRAW: A Stress-Aware WL-Based Read Reclaim Technique for High-Density NAND Flash-Based SSDs
di: Chun, Myoungjun, et al.
Pubblicazione: (2025)
di: Chun, Myoungjun, et al.
Pubblicazione: (2025)
SliceMoE: Bit-Sliced Expert Caching under Miss-Rate Constraints for Efficient MoE Inference
di: Choi, Yuseon, et al.
Pubblicazione: (2025)
di: Choi, Yuseon, et al.
Pubblicazione: (2025)
Low-overhead General-purpose Near-Data Processing in CXL Memory Expanders
di: Ham, Hyungkyu, et al.
Pubblicazione: (2024)
di: Ham, Hyungkyu, et al.
Pubblicazione: (2024)
Token-Picker: Accelerating Attention in Text Generation with Minimized Memory Transfer via Probability Estimation
di: Park, Junyoung, et al.
Pubblicazione: (2024)
di: Park, Junyoung, et al.
Pubblicazione: (2024)
Energy-Efficient p-Bit-Based Fully-Connected Quantum-Inspired Simulated Annealer with Dual BRAM Architecture
di: Onizawa, Naoya, et al.
Pubblicazione: (2026)
di: Onizawa, Naoya, et al.
Pubblicazione: (2026)
AERO: Adaptive Erase Operation for Improving Lifetime and Performance of Modern NAND Flash-Based SSDs
di: Cho, Sungjun, et al.
Pubblicazione: (2024)
di: Cho, Sungjun, et al.
Pubblicazione: (2024)
Full System Architecture Modeling for Wearable Egocentric Contextual AI
di: Lee, Vincent T., et al.
Pubblicazione: (2025)
di: Lee, Vincent T., et al.
Pubblicazione: (2025)
Containerized In-Storage Processing and Computing-Enabled SSD Disaggregation
di: Kwon, Miryeong, et al.
Pubblicazione: (2025)
di: Kwon, Miryeong, et al.
Pubblicazione: (2025)
Jack Unit: An Area- and Energy-Efficient Multiply-Accumulate (MAC) Unit Supporting Diverse Data Formats
di: Noh, Seock-Hwan, et al.
Pubblicazione: (2025)
di: Noh, Seock-Hwan, et al.
Pubblicazione: (2025)
Gaze into the Pattern: Characterizing Spatial Patterns with Internal Temporal Correlations for Hardware Prefetching
di: Chen, Zixiao, et al.
Pubblicazione: (2024)
di: Chen, Zixiao, et al.
Pubblicazione: (2024)
Hybrid SLC-MLC RRAM Mixed-Signal Processing-in-Memory Architecture for Transformer Acceleration via Gradient Redistribution
di: Song, Chang Eun, et al.
Pubblicazione: (2025)
di: Song, Chang Eun, et al.
Pubblicazione: (2025)
FIGLUT: An Energy-Efficient Accelerator Design for FP-INT GEMM Using Look-Up Tables
di: Park, Gunho, et al.
Pubblicazione: (2025)
di: Park, Gunho, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Characterization of Multi-Model Agentic AI Systems on General Tasks via Trace-Driven Simulation
di: Kim, Donghwan, et al.
Pubblicazione: (2026) -
Towards Understanding Systems Trade-offs in Retrieval-Augmented Generation Model Inference
di: Shen, Michael, et al.
Pubblicazione: (2024) -
MixDiT: Accelerating Image Diffusion Transformer Inference with Mixed-Precision MX Quantization
di: Kim, Daeun, et al.
Pubblicazione: (2025) -
IBEX: Internal Bandwidth-Efficient Compression Architecture for Scalable CXL Memory Expansion
di: Ko, Younghoon, et al.
Pubblicazione: (2026) -
ONNXim: A Fast, Cycle-level Multi-core NPU Simulator
di: Ham, Hyungkyu, et al.
Pubblicazione: (2024)