PANDA: Noise-Resilient Antagonist Identification in Production Datacenters
Fuente:
arXiv
Saved in:
| Main Authors: | Zhou, Sixiang, Deng, Nan, Rzadca, Krzysiek, Lin, Xiaojun, Hu, Y. Charlie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Automated PMC-based Power Modeling Methodology for Modern Mobile GPUs
by: Dash, Pranab, et al.
Published: (2024)
by: Dash, Pranab, et al.
Published: (2024)
Long-term Monitoring of Kernel and Hardware Events to Understand Latency Variance
by: Zhou, Fang, et al.
Published: (2026)
by: Zhou, Fang, et al.
Published: (2026)
An Inquiry into Datacenter TCO for LLM Inference with FP8
by: Kim, Jiwoo, et al.
Published: (2025)
by: Kim, Jiwoo, et al.
Published: (2025)
Denoising Application Performance Models with Noise-Resilient Priors
by: de Morais, Gustavo, et al.
Published: (2025)
by: de Morais, Gustavo, et al.
Published: (2025)
Modeling the Impact of Fiber Latency on Compute-Communication Overlap in Geo-Distributed Multi-Datacenter AI Training
by: Papavasileiou, Ioannis, et al.
Published: (2026)
by: Papavasileiou, Ioannis, et al.
Published: (2026)
Noise Injection for__Performance Bottleneck Analysis
by: Delval, Aurélien, et al.
Published: (2025)
by: Delval, Aurélien, et al.
Published: (2025)
Fine-Grained Clustering-Based Power Identification for Multicores
by: Elshamy, Mohamed R., et al.
Published: (2024)
by: Elshamy, Mohamed R., et al.
Published: (2024)
Efficient Data-Driven Production Scheduling in Pharmaceutical Manufacturing
by: Balatsos, Ioannis, et al.
Published: (2026)
by: Balatsos, Ioannis, et al.
Published: (2026)
Graph-Based Product Form
by: Comte, Céline, et al.
Published: (2025)
by: Comte, Céline, et al.
Published: (2025)
SysOM-AI: Continuous Cross-Layer Performance Diagnosis for Production AI Training
by: Zheng, Yusheng, et al.
Published: (2026)
by: Zheng, Yusheng, et al.
Published: (2026)
Computational Algorithms for the Product Form Solution of Closed Queuing Networks with Finite Buffers and Skip-Over Policy
by: Balbo, Gianfranco, et al.
Published: (2024)
by: Balbo, Gianfranco, et al.
Published: (2024)
SONIQ: System-Optimized Noise-Injected Ultra-Low-Precision Quantization with Full-Precision Parity
by: Zhou, Cyrus, et al.
Published: (2023)
by: Zhou, Cyrus, et al.
Published: (2023)
Profiling Large Language Model Inference on Apple Silicon: A Quantization Perspective
by: Benazir, Afsara, et al.
Published: (2025)
by: Benazir, Afsara, et al.
Published: (2025)
Two Criteria for Performance Analysis of Optimization Algorithms
by: Jing, Yunpeng, et al.
Published: (2024)
by: Jing, Yunpeng, et al.
Published: (2024)
The Theoretical Limit of Radar Target Detection
by: Xu, Dazhuan, et al.
Published: (2021)
by: Xu, Dazhuan, et al.
Published: (2021)
Profiling Apple Silicon Performance for ML Training
by: Feng, Dahua, et al.
Published: (2025)
by: Feng, Dahua, et al.
Published: (2025)
A Microbenchmark Framework for Performance Evaluation of OpenMP Target Offloading
by: Atif, Mohammad, et al.
Published: (2025)
by: Atif, Mohammad, et al.
Published: (2025)
The Price of Interoperability: Exploring Cross-Chain Bridges and Their Economic Consequences
by: Cao, Yiyue, et al.
Published: (2026)
by: Cao, Yiyue, et al.
Published: (2026)
How Much Parallelism Is "Free"? A Principle of Near-Free Parallelism for Parallel Decoding
by: He, Minghua, et al.
Published: (2026)
by: He, Minghua, et al.
Published: (2026)
Mosaic: Cross-Modal Clustering for Efficient Video Understanding
by: Wang, Tuowei, et al.
Published: (2026)
by: Wang, Tuowei, et al.
Published: (2026)
CXL-Interference: Analysis and Characterization in Modern Computer Systems
by: Mao, Shunyu, et al.
Published: (2024)
by: Mao, Shunyu, et al.
Published: (2024)
SparseX: Efficient Segment-Level KV Cache Sharing for Interleaved LLM Serving
by: Zhang, Quqing, et al.
Published: (2026)
by: Zhang, Quqing, et al.
Published: (2026)
RAPID-LLM: Resilience-Aware Performance analysis of Infrastructure for Distributed LLM Training and Inference
by: Karfakis, George, et al.
Published: (2025)
by: Karfakis, George, et al.
Published: (2025)
Beamforming-based Achievable Rate Maximization in ISAC System for Multi-UAV Networking
by: Zhou, Shengcai, et al.
Published: (2025)
by: Zhou, Shengcai, et al.
Published: (2025)
Achieving Consistent and Comparable CPU Evaluation
by: Wang, Chenxi, et al.
Published: (2024)
by: Wang, Chenxi, et al.
Published: (2024)
Large-Scale Data Parallelization of Product Quantization and Inverted Indexing Using Dask
by: Abraham, Ashley N., et al.
Published: (2026)
by: Abraham, Ashley N., et al.
Published: (2026)
A Modular Graph-Native Query Optimization Framework
by: Lyu, Bingqing, et al.
Published: (2024)
by: Lyu, Bingqing, et al.
Published: (2024)
PASTA: A Modular Program Analysis Tool Framework for Accelerators
by: Lin, Mao, et al.
Published: (2026)
by: Lin, Mao, et al.
Published: (2026)
A relação entre a «performance» social e a «performance» económico-financeira
by: Daniel Taborda
Published: (2007)
by: Daniel Taborda
Published: (2007)
Overhead Measurement Noise in Different Runtime Environments
by: Reichelt, David Georg, et al.
Published: (2024)
by: Reichelt, David Georg, et al.
Published: (2024)
StiffGIPC: Advancing GPU IPC for stiff affine-deformable simulation
by: Huang, Kemeng, et al.
Published: (2024)
by: Huang, Kemeng, et al.
Published: (2024)
TurboSpec: Closed-loop Speculation Control System for Optimizing LLM Serving Goodput
by: Liu, Xiaoxuan, et al.
Published: (2024)
by: Liu, Xiaoxuan, et al.
Published: (2024)
Efficient Graph Knowledge Distillation from GNNs to Kolmogorov--Arnold Networks via Self-Attention Dynamic Sampling
by: Cui, Can, et al.
Published: (2025)
by: Cui, Can, et al.
Published: (2025)
LOS MÁRGENES TOMAN LA ESCENA. EL USO DE LA PERFORMANCE EN LA LUCHA SUBALTERNA. UNA VISIÓN ANTROPOLÓGICA.
by: Iván Alvarado
Published: (2013)
by: Iván Alvarado
Published: (2013)
Da artificação do sagrado nos museus: entre o teatro e a sacralidade
by: Bruno Brulon
Published: (2013)
by: Bruno Brulon
Published: (2013)
Adaptive Orchestration for Large-Scale Inference on Heterogeneous Accelerator Systems Balancing Cost, Performance, and Resilience
by: Biran, Yahav, et al.
Published: (2025)
by: Biran, Yahav, et al.
Published: (2025)
TrainMover: An Interruption-Resilient Runtime for ML Training
by: Lao, ChonLam, et al.
Published: (2024)
by: Lao, ChonLam, et al.
Published: (2024)
Discourse Analysis of Stress in Postmodernity: The Precariousness of Labor, the Cult of Performance, and Production and Consumption
by: Nogueira da Silva, João Paulo, et al.
Published: (2025)
by: Nogueira da Silva, João Paulo, et al.
Published: (2025)
Improving LLM Performance Through Black-Box Online Tuning: A Case for Adding System Specs to Factsheets for Trusted AI
by: Atinafu, Yonas, et al.
Published: (2026)
by: Atinafu, Yonas, et al.
Published: (2026)
RWKV-edge: Deeply Compressed RWKV for Resource-Constrained Devices
by: Choe, Wonkyo, et al.
Published: (2024)
by: Choe, Wonkyo, et al.
Published: (2024)
Similar Items
-
Automated PMC-based Power Modeling Methodology for Modern Mobile GPUs
by: Dash, Pranab, et al.
Published: (2024) -
Long-term Monitoring of Kernel and Hardware Events to Understand Latency Variance
by: Zhou, Fang, et al.
Published: (2026) -
An Inquiry into Datacenter TCO for LLM Inference with FP8
by: Kim, Jiwoo, et al.
Published: (2025) -
Denoising Application Performance Models with Noise-Resilient Priors
by: de Morais, Gustavo, et al.
Published: (2025) -
Modeling the Impact of Fiber Latency on Compute-Communication Overlap in Geo-Distributed Multi-Datacenter AI Training
by: Papavasileiou, Ioannis, et al.
Published: (2026)