Scaler: Efficient and Effective Cross Flow Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Steven, Tang, Xiang, Mingcan, Wang, Yang, Wu, Bo, Chen, Jianjun, Liu, Tongping |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Long-term Monitoring of Kernel and Hardware Events to Understand Latency Variance
by: Zhou, Fang, et al.
Published: (2026)
by: Zhou, Fang, et al.
Published: (2026)
ProTrain: Efficient LLM Training via Memory-Aware Techniques
by: Yang, Hanmei, et al.
Published: (2024)
by: Yang, Hanmei, et al.
Published: (2024)
SparseX: Efficient Segment-Level KV Cache Sharing for Interleaved LLM Serving
by: Zhang, Quqing, et al.
Published: (2026)
by: Zhang, Quqing, et al.
Published: (2026)
AdapMTL: Adaptive Pruning Framework for Multitask Learning Model
by: Xiang, Mingcan, et al.
Published: (2024)
by: Xiang, Mingcan, et al.
Published: (2024)
Mosaic: Cross-Modal Clustering for Efficient Video Understanding
by: Wang, Tuowei, et al.
Published: (2026)
by: Wang, Tuowei, et al.
Published: (2026)
Understanding and Alleviating Memory Consumption in RLHF for LLMs
by: Zhou, Jin, et al.
Published: (2024)
by: Zhou, Jin, et al.
Published: (2024)
Obfuscation as an Effective Signal for Prioritizing Cross-Chain Smart Contract Audits: Large-Scale Measurement and Risk Profiling
by: Zhao, Yao, et al.
Published: (2026)
by: Zhao, Yao, et al.
Published: (2026)
DeepContext: A Context-aware, Cross-platform, and Cross-framework Tool for Performance Profiling and Analysis of Deep Learning Workloads
by: Zhao, Qidong, et al.
Published: (2024)
by: Zhao, Qidong, et al.
Published: (2024)
From Profiling to Optimization: Unveiling the Profile Guided Optimization
by: Liu, Bingxin, et al.
Published: (2025)
by: Liu, Bingxin, et al.
Published: (2025)
Accurate Performance Modeling And Uncertainty Analysis of Lossy Compression in Scientific Applications
by: Liu, Youyuan, et al.
Published: (2024)
by: Liu, Youyuan, et al.
Published: (2024)
ExpertFlow: Adaptive Expert Scheduling and Memory Coordination for Efficient MoE Inference
by: Shen, Zixu, et al.
Published: (2025)
by: Shen, Zixu, et al.
Published: (2025)
Spatiotemporal Non-Uniformity-Aware Online Task Scheduling in Collaborative Edge Computing for Industrial Internet of Things
by: Li, Yang, et al.
Published: (2025)
by: Li, Yang, et al.
Published: (2025)
The Price of Interoperability: Exploring Cross-Chain Bridges and Their Economic Consequences
by: Cao, Yiyue, et al.
Published: (2026)
by: Cao, Yiyue, et al.
Published: (2026)
Beamforming-based Achievable Rate Maximization in ISAC System for Multi-UAV Networking
by: Zhou, Shengcai, et al.
Published: (2025)
by: Zhou, Shengcai, et al.
Published: (2025)
Application Research On Real-Time Perception Of Device Performance Status
by: Wang, Zhe, et al.
Published: (2024)
by: Wang, Zhe, et al.
Published: (2024)
Two Criteria for Performance Analysis of Optimization Algorithms
by: Jing, Yunpeng, et al.
Published: (2024)
by: Jing, Yunpeng, et al.
Published: (2024)
Pinching-Antenna Systems For Indoor Immersive Communications: A 3D-Modeling Based Performance Analysis
by: Wang, Yulei, et al.
Published: (2025)
by: Wang, Yulei, et al.
Published: (2025)
PerfSeer: An Efficient and Accurate Deep Learning Models Performance Predictor
by: Zhao, Xinlong, et al.
Published: (2025)
by: Zhao, Xinlong, et al.
Published: (2025)
EcoSpa: Efficient Transformer Training with Coupled Sparsity
by: Xiao, Jinqi, et al.
Published: (2025)
by: Xiao, Jinqi, et al.
Published: (2025)
Non-Asymptotic Performance Analysis of DOA Estimation Based on Real-Valued Root-MUSIC
by: Liu, Junyang, et al.
Published: (2025)
by: Liu, Junyang, et al.
Published: (2025)
HistogramTools for Efficient Data Analysis and Distribution Representation in Large Data Sets
by: Malhotra, Shubham
Published: (2025)
by: Malhotra, Shubham
Published: (2025)
Atys: An Efficient Profiling Framework for Identifying Hotspot Functions in Large-scale Cloud Microservices
by: Sun, Jiaqi, et al.
Published: (2025)
by: Sun, Jiaqi, et al.
Published: (2025)
Closed-Loop Integrated Sensing, Communication, and Control for Efficient Drone Flight
by: Li, Jingli, et al.
Published: (2026)
by: Li, Jingli, et al.
Published: (2026)
H2EAL: Hybrid-Bonding Architecture with Hybrid Sparse Attention for Efficient Long-Context LLM Inference
by: Fu, Zizhuo, et al.
Published: (2025)
by: Fu, Zizhuo, et al.
Published: (2025)
GreenLLM: SLO-Aware Dynamic Frequency Scaling for Energy-Efficient LLM Serving
by: Liu, Qunyou, et al.
Published: (2025)
by: Liu, Qunyou, et al.
Published: (2025)
An Efficient Wireless iBCI Headstage with Adaptive ADC Sample Rate
by: Liu, Hongyao, et al.
Published: (2026)
by: Liu, Hongyao, et al.
Published: (2026)
Efficient Performance Analysis of Modular Rewritable Petri Nets
by: Capra, Lorenzo, et al.
Published: (2024)
by: Capra, Lorenzo, et al.
Published: (2024)
FLEXIS: FLEXible Frequent Subgraph Mining using Maximal Independent Sets
by: Sharma, Akshit, et al.
Published: (2024)
by: Sharma, Akshit, et al.
Published: (2024)
Blending Learning to Rank and Dense Representations for Efficient and Effective Cascades
by: Nardini, Franco Maria, et al.
Published: (2025)
by: Nardini, Franco Maria, et al.
Published: (2025)
CEBench: A Benchmarking Toolkit for the Cost-Effectiveness of LLM Pipelines
by: Sun, Wenbo, et al.
Published: (2024)
by: Sun, Wenbo, et al.
Published: (2024)
Efficient Top-k Vulnerable Nodes Detection in Uncertain Graphs
by: Cheng, Dawei, et al.
Published: (2019)
by: Cheng, Dawei, et al.
Published: (2019)
gDist: Efficient Distance Computation between 3D Meshes on GPU
by: Fang, Peng, et al.
Published: (2024)
by: Fang, Peng, et al.
Published: (2024)
Benchmark-based Study of CPU/GPU Power-Related Features through JAX and TensorFlow
by: Tchakoute, Roblex Nana, et al.
Published: (2025)
by: Tchakoute, Roblex Nana, et al.
Published: (2025)
SysOM-AI: Continuous Cross-Layer Performance Diagnosis for Production AI Training
by: Zheng, Yusheng, et al.
Published: (2026)
by: Zheng, Yusheng, et al.
Published: (2026)
Fault-Tolerant Hybrid-Parallel Training at Scale with Reliable and Efficient In-memory Checkpointing
by: Wang, Yuxin, et al.
Published: (2023)
by: Wang, Yuxin, et al.
Published: (2023)
Efficient Data-Driven Production Scheduling in Pharmaceutical Manufacturing
by: Balatsos, Ioannis, et al.
Published: (2026)
by: Balatsos, Ioannis, et al.
Published: (2026)
TINA: Acceleration of Non-NN Signal Processing Algorithms Using NN Accelerators
by: Boerkamp, Christiaan, et al.
Published: (2024)
by: Boerkamp, Christiaan, et al.
Published: (2024)
Tail Bounds for Queues with Abandonment: Constant, Moderate, Large Deviations, and Efficient Concentration
by: Wang, Zedong, et al.
Published: (2026)
by: Wang, Zedong, et al.
Published: (2026)
SAfEPaTh: A System-Level Approach for Efficient Power and Thermal Estimation of Convolutional Neural Network Accelerator
by: Chen, Yukai, et al.
Published: (2024)
by: Chen, Yukai, et al.
Published: (2024)
Energy Efficiency Analysis of Active RIS-enhanced Wireless Network under Power-Sum Constraint
by: Xin, Jingdie, et al.
Published: (2025)
by: Xin, Jingdie, et al.
Published: (2025)
Similar Items
-
Long-term Monitoring of Kernel and Hardware Events to Understand Latency Variance
by: Zhou, Fang, et al.
Published: (2026) -
ProTrain: Efficient LLM Training via Memory-Aware Techniques
by: Yang, Hanmei, et al.
Published: (2024) -
SparseX: Efficient Segment-Level KV Cache Sharing for Interleaved LLM Serving
by: Zhang, Quqing, et al.
Published: (2026) -
AdapMTL: Adaptive Pruning Framework for Multitask Learning Model
by: Xiang, Mingcan, et al.
Published: (2024) -
Mosaic: Cross-Modal Clustering for Efficient Video Understanding
by: Wang, Tuowei, et al.
Published: (2026)