zhongkaiyu/moe_exp_placement: Case Study2 for ISCA 2026 AE
Fuente:
Zenodo
Saved in:
| Main Authors: | Yu, Zhongkai, Guan, Yue, Yu, Zihao, Zhou, Chenyang, Hu, Zhengding, Pei, Shuyi, Kang, Yangwook, Ding, Yufei, Tsai, Po-An |
|---|---|
| Format: | Recurso digital |
| Published: |
Zenodo
2026
|
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Patterns behind Chaos: Forecasting Data Movement for Efficient Large-Scale MoE LLM Inference
by: Yu, Zhongkai, et al.
Published: (2025)
by: Yu, Zhongkai, et al.
Published: (2025)
AMMA: A Multi-Chiplet Memory-Centric Architecture for Low-Latency 1M Context Attention Serving
by: Yu, Zhongkai, et al.
Published: (2026)
by: Yu, Zhongkai, et al.
Published: (2026)
ScaleSim: Serving Large-Scale Multi-Agent Simulation with Invocation Distance-Based Memory Management
by: Pan, Zaifeng, et al.
Published: (2026)
by: Pan, Zaifeng, et al.
Published: (2026)
Pancake: Hierarchical Memory System for Multi-Agent LLM Serving
by: Hu, Zhengding, et al.
Published: (2026)
by: Hu, Zhengding, et al.
Published: (2026)
JigsawRL: Assembling RL Pipelines for Efficient LLM Post-Training
by: Hu, Zhengding, et al.
Published: (2026)
by: Hu, Zhengding, et al.
Published: (2026)
Syncopate: Efficient Multi-GPU AI Kernels via Automatic Chunk-Centric Compute-Communication Overlap
by: Qiang, Xinwei, et al.
Published: (2026)
by: Qiang, Xinwei, et al.
Published: (2026)
FlashEvolve: Accelerating Agent Self-Evolution with Asynchronous Stage Orchestration
by: Hu, Zhengding, et al.
Published: (2026)
by: Hu, Zhengding, et al.
Published: (2026)
KPerfIR: Towards an Open and Compiler-centric Ecosystem for GPU Kernel Performance Tooling on Modern AI Workloads
by: Guan, Yue, et al.
Published: (2025)
by: Guan, Yue, et al.
Published: (2025)
ISCA: A Framework for Interview-Style Conversational Agents
by: Welch, Charles, et al.
Published: (2025)
by: Welch, Charles, et al.
Published: (2025)
GNN-exps
by: Anonymous
Published: (2025)
by: Anonymous
Published: (2025)
ChipMATE: Multi-Agent Training via Reinforcement Learning for Enhanced RTL Generation
by: Yu, Zhongkai, et al.
Published: (2026)
by: Yu, Zhongkai, et al.
Published: (2026)
KVFlow: Efficient Prefix Caching for Accelerating LLM-Based Multi-Agent Workflows
by: Pan, Zaifeng, et al.
Published: (2025)
by: Pan, Zaifeng, et al.
Published: (2025)
Proceedings of the ISCA/ITG Workshop on Diversity in Large Speech and Language Models
by: Möller, Sebastian, et al.
Published: (2025)
by: Möller, Sebastian, et al.
Published: (2025)
ChipBench: A Next-Step Benchmark for Evaluating LLM Performance in AI-Aided Chip Design
by: Yu, Zhongkai, et al.
Published: (2026)
by: Yu, Zhongkai, et al.
Published: (2026)
(TTTCA)exp Drives the Genotype–Phenotype Correlation and Genetic Anticipation in FCMTE1
by: Xinhui Chen, et al.
Published: (2024)
by: Xinhui Chen, et al.
Published: (2024)
HedraRAG: Coordinating LLM Generation and Database Retrieval in Heterogeneous RAG Serving
by: Hu, Zhengding, et al.
Published: (2025)
by: Hu, Zhengding, et al.
Published: (2025)
Strategies to Enhance CAR ‐T Cell Persistence in Hematologic Malignancies: From Molecular Design to Clinical Optimization
by: Mengqi Yu, et al.
Published: (2025)
by: Mengqi Yu, et al.
Published: (2025)
TLX: Hardware-Native, Evolvable MIMW GPU Compiler for Large-scale Production Environments
by: Guan, Yue, et al.
Published: (2026)
by: Guan, Yue, et al.
Published: (2026)
SpadeMomo/SPAR-AE: SPAR Artifact for USENIX Security 2026
by: Ye, Hengdi
Published: (2026)
by: Ye, Hengdi
Published: (2026)
VMD opens applications for 2026 EMS placement
Published: (2026)
Published: (2026)
threewater-dot/MvAE: MvAE
by: threewater-dot
Published: (2026)
by: threewater-dot
Published: (2026)
APLICAÇÃO SISTEMÁTICA MECANIZADA DE ISCA FORMICIDA GRANULADA EM EUCALIPTAIS EM FASE DE MANUTENÇÃO
by: Marcelo de Almeida Reis
Published: (2015)
by: Marcelo de Almeida Reis
Published: (2015)
On Preparation Theorems for $\mathbb{R}_{an,exp}$-definable functions
by: Opris, Andre
Published: (2021)
by: Opris, Andre
Published: (2021)
Some non-algebraic forms of $\exp(A+B)$
by: Tapia-Valerdi, M. A., et al.
Published: (2024)
by: Tapia-Valerdi, M. A., et al.
Published: (2024)
Full capacity–volumetry of sharp exp‐integrability law
by: David R. Adams, et al.
Published: (2025)
by: David R. Adams, et al.
Published: (2025)
AntigenLM: Structure-Aware DNA Language Modeling for Influenza
by: Pei, Yue, et al.
Published: (2026)
by: Pei, Yue, et al.
Published: (2026)
Æ codes
by: Jain, Shubham P., et al.
Published: (2023)
by: Jain, Shubham P., et al.
Published: (2023)
The ERIC/AE Test Locator Service. ERIC/AE Digest.
by: Doolittle, Peter, et al.
Published: (1994)
by: Doolittle, Peter, et al.
Published: (1994)
Compositional Square Roots of $\exp(x)$ and $1+x^2$
by: Finch, Steven
Published: (2025)
by: Finch, Steven
Published: (2025)
Analytic holonomicity of real C$^{\mathrm{exp}}$-class distributions
by: Aizenbud, Avraham, et al.
Published: (2024)
by: Aizenbud, Avraham, et al.
Published: (2024)
eunhanka/mirage-eebl-detector: VehicleSec 2026 AE Final Submission (v2)
by: Eunhan
Published: (2026)
by: Eunhan
Published: (2026)
Non‐overlapping placement of macro cells based on reinforcement learning in chip design
by: Tao Yu, et al.
Published: (2024)
by: Tao Yu, et al.
Published: (2024)
OPAL: Omnidirectional Path-efficient Aerial 3D expLoration
by: Chappidi, Yoga Satwik, et al.
Published: (2026)
by: Chappidi, Yoga Satwik, et al.
Published: (2026)
Symbolic-numeric algorithm for parameter estimation in discrete-time models with $\exp$
by: Berman, Yosef, et al.
Published: (2024)
by: Berman, Yosef, et al.
Published: (2024)
Enhancing Unsupervised Anomaly Detection and Early Warning With Dual‐Attention LSTM ‐ AdvAE
by: Zhiyi Zhang, et al.
Published: (2026)
by: Zhiyi Zhang, et al.
Published: (2026)
C2S-AE: CSI to Sensing enabled by an Auto-Encoder-based Framework
by: Jiang, Jun, et al.
Published: (2025)
by: Jiang, Jun, et al.
Published: (2025)
Profit Maximization for Electric Vehicle Charging Stations Using Multiagent Reinforcement Learning
by: Jiang, Kun-Yan, et al.
Published: (2026)
by: Jiang, Kun-Yan, et al.
Published: (2026)
DeCode: Decoupling Content and Delivery for Medical QA
by: Ko, Po-Jen, et al.
Published: (2026)
by: Ko, Po-Jen, et al.
Published: (2026)
Dynamic Channel Charting: An LSTM-AE-based Approach
by: Gao, Yuan, et al.
Published: (2026)
by: Gao, Yuan, et al.
Published: (2026)
Bilateral Unsymmetrical Graph Contrastive Learning for Recommendation
by: Yu, Jiaheng, et al.
Published: (2024)
by: Yu, Jiaheng, et al.
Published: (2024)
Similar Items
-
Patterns behind Chaos: Forecasting Data Movement for Efficient Large-Scale MoE LLM Inference
by: Yu, Zhongkai, et al.
Published: (2025) -
AMMA: A Multi-Chiplet Memory-Centric Architecture for Low-Latency 1M Context Attention Serving
by: Yu, Zhongkai, et al.
Published: (2026) -
ScaleSim: Serving Large-Scale Multi-Agent Simulation with Invocation Distance-Based Memory Management
by: Pan, Zaifeng, et al.
Published: (2026) -
Pancake: Hierarchical Memory System for Multi-Agent LLM Serving
by: Hu, Zhengding, et al.
Published: (2026) -
JigsawRL: Assembling RL Pipelines for Efficient LLM Post-Training
by: Hu, Zhengding, et al.
Published: (2026)