1.5 Million Messages Per Second on 3 Machines: Benchmarking and Latency Optimization of Apache Pulsar at Enterprise Scale
Fuente:
arXiv
Saved in:
| Main Author: | Mukkolakkal, Muhamed Ramees Cheriya |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Deploy, Calibrate, Monitor, Heal -- No Human Required: An Autonomous AI SRE Agent for Elasticsearch
by: Mukkolakkal, Muhamed Ramees Cheriya
Published: (2026)
by: Mukkolakkal, Muhamed Ramees Cheriya
Published: (2026)
On Efficiently Partitioning a Topic in Apache Kafka
by: Raptis, Theofanis P., et al.
Published: (2022)
by: Raptis, Theofanis P., et al.
Published: (2022)
Performance Evaluation of Brokerless Messaging Libraries
by: La Corte, Lorenzo, et al.
Published: (2025)
by: La Corte, Lorenzo, et al.
Published: (2025)
Short-circuiting Rings for Low-Latency AllReduce
by: Hammer, Sarah-Michelle, et al.
Published: (2025)
by: Hammer, Sarah-Michelle, et al.
Published: (2025)
Solving AI Foundational Model Latency with Telco Infrastructure
by: Barros, Sebastian
Published: (2025)
by: Barros, Sebastian
Published: (2025)
Throughput-Optimized Networks at Scale
by: Green, Conor James, et al.
Published: (2026)
by: Green, Conor James, et al.
Published: (2026)
Trivance: Latency-Optimal AllReduce by Shortcutting Multiport Networks
by: Juerss, Anton, et al.
Published: (2026)
by: Juerss, Anton, et al.
Published: (2026)
PerLLM: Personalized Inference Scheduling with Edge-Cloud Collaboration for Diverse LLM Services
by: Yang, Zheming, et al.
Published: (2024)
by: Yang, Zheming, et al.
Published: (2024)
Risk-Aware and Stable Edge Server Selection Under Network Latency SLOs
by: Liyanage, Mohan, et al.
Published: (2026)
by: Liyanage, Mohan, et al.
Published: (2026)
Lightweight Latency Prediction Scheme for Edge Applications: A Rational Modelling Approach
by: Liyanage, Mohan, et al.
Published: (2025)
by: Liyanage, Mohan, et al.
Published: (2025)
COREC: Concurrent Non-Blocking Single-Queue Receive Driver for Low Latency Networking
by: Faltelli, Marco, et al.
Published: (2024)
by: Faltelli, Marco, et al.
Published: (2024)
RouterWise: Joint Resource Allocation and Routing for Latency-Aware Multi-Model LLM Serving
by: Kasnavieh, Hossein Hosseini, et al.
Published: (2026)
by: Kasnavieh, Hossein Hosseini, et al.
Published: (2026)
GORGO: Maximizing KV-Cache Reuse While Minimizing Network Latency in Cross-Region LLM Load Balancing
by: Toniolo, Alessio Ricci, et al.
Published: (2026)
by: Toniolo, Alessio Ricci, et al.
Published: (2026)
CRAFT: Latency and Cost-Aware Genetic-Based Framework for Node Placement in Edge-Fog Environments
by: Mahdizadeh, Soheil, et al.
Published: (2025)
by: Mahdizadeh, Soheil, et al.
Published: (2025)
Optimizing Split Learning Latency in TinyML-Based IoT Systems
by: Jenhani, Zied, et al.
Published: (2025)
by: Jenhani, Zied, et al.
Published: (2025)
FogROS2-PLR: Probabilistic Latency-Reliability For Cloud Robotics
by: Chen, Kaiyuan, et al.
Published: (2024)
by: Chen, Kaiyuan, et al.
Published: (2024)
Extreme-Scale Interconnection Networks
by: Cano, Alejandro, et al.
Published: (2026)
by: Cano, Alejandro, et al.
Published: (2026)
Design and Operation of Shared Machine Learning Clusters on Campus
by: Xu, Kaiqiang, et al.
Published: (2021)
by: Xu, Kaiqiang, et al.
Published: (2021)
CCL-Bench 1.0: A Trace-Based Benchmark for LLM Infrastructure
by: Ding, Eric, et al.
Published: (2026)
by: Ding, Eric, et al.
Published: (2026)
Toward Co-adapting Machine Learning Job Shape and Cluster Topology
by: Chen, Shawn Shuoshuo, et al.
Published: (2025)
by: Chen, Shawn Shuoshuo, et al.
Published: (2025)
SplitSim: Large-Scale Simulations for Evaluating Network Systems Research
by: Li, Hejing, et al.
Published: (2024)
by: Li, Hejing, et al.
Published: (2024)
Decentralized Stratified Sampling for Low-Latency Approximate Geospatial Data Stream Processing in Edge-Cloud Architectures
by: Jawarneh, Isam Mashhour Al, et al.
Published: (2026)
by: Jawarneh, Isam Mashhour Al, et al.
Published: (2026)
Diffusion Models on the Edge: Challenges, Optimizations, and Applications
by: Zheng, Dongqi
Published: (2025)
by: Zheng, Dongqi
Published: (2025)
Optimal Oblivious Load-Balancing for Sparse Traffic in Large-Scale Satellite Networks
by: Ramakanth, Rudrapatna Vallabh, et al.
Published: (2026)
by: Ramakanth, Rudrapatna Vallabh, et al.
Published: (2026)
Ethereal: Divide and Conquer Network Load Balancing in Large-Scale Distributed Training
by: Addanki, Vamsi, et al.
Published: (2024)
by: Addanki, Vamsi, et al.
Published: (2024)
Harvest: Adaptive Photonic Switching Schedules for Collective Communication in Scale-up Domains
by: Rahman, Mahir, et al.
Published: (2026)
by: Rahman, Mahir, et al.
Published: (2026)
When Light Bends to the Collective Will: A Theory and Vision for Adaptive Photonic Scale-up Domains
by: Addanki, Vamsi
Published: (2025)
by: Addanki, Vamsi
Published: (2025)
Distributed Simulation for Digital Twins of Large-Scale Real-World DiffServ-Based Networks
by: Huang, Zhuoyao, et al.
Published: (2024)
by: Huang, Zhuoyao, et al.
Published: (2024)
Topology-aware Microservice Architecture in Edge Networks: Deployment Optimization and Implementation
by: Chen, Yuang, et al.
Published: (2025)
by: Chen, Yuang, et al.
Published: (2025)
Real-Time In-Network Machine Learning on P4-Programmable FPGA SmartNICs with Fixed-Point Arithmetic and Taylor
by: Sada, Mohammad Firas, et al.
Published: (2025)
by: Sada, Mohammad Firas, et al.
Published: (2025)
Low-Latency Video Conferencing via Optimized Packet Routing and Reordering
by: Xiao, Yao, et al.
Published: (2023)
by: Xiao, Yao, et al.
Published: (2023)
PSMOA: Policy Support Multi-Objective Optimization Algorithm for Decentralized Data Replication
by: Wang, Xi, et al.
Published: (2025)
by: Wang, Xi, et al.
Published: (2025)
LIMO: Load-balanced Offloading with MAPE and Particle Swarm Optimization in Mobile Fog Networks
by: Seraj, Yasaman, et al.
Published: (2024)
by: Seraj, Yasaman, et al.
Published: (2024)
A Study on 5G Network Slice Isolation Based on Native Cloud and Edge Computing Tools
by: Andrade, Maiko, et al.
Published: (2025)
by: Andrade, Maiko, et al.
Published: (2025)
Towards Integrated Energy-Communication-Transportation Hub: A Base-Station-Centric Design in 5G and Beyond
by: Shen, Linfeng, et al.
Published: (2025)
by: Shen, Linfeng, et al.
Published: (2025)
Diving into 3D Parallelism with Heterogeneous Spot Instance GPUs: Design and Implications
by: Wang, Yuxiao, et al.
Published: (2025)
by: Wang, Yuxiao, et al.
Published: (2025)
RailX: A Flexible, Scalable, and Low-Cost Network Architecture for Hyper-Scale LLM Training Systems
by: Feng, Yinxiao, et al.
Published: (2025)
by: Feng, Yinxiao, et al.
Published: (2025)
Time Complexity of Broadcast and Consensus for Randomized Oblivious Message Adversaries
by: El-Hayek, Antoine, et al.
Published: (2023)
by: El-Hayek, Antoine, et al.
Published: (2023)
ScaleAcross Explorer: Exploring Communication Optimization for Scale-Across AI Model Training
by: Li, Minghao, et al.
Published: (2026)
by: Li, Minghao, et al.
Published: (2026)
EvalNet: A Practical Toolchain for Generation and Analysis of Extreme-Scale Interconnects
by: Besta, Maciej, et al.
Published: (2021)
by: Besta, Maciej, et al.
Published: (2021)
Similar Items
-
Deploy, Calibrate, Monitor, Heal -- No Human Required: An Autonomous AI SRE Agent for Elasticsearch
by: Mukkolakkal, Muhamed Ramees Cheriya
Published: (2026) -
On Efficiently Partitioning a Topic in Apache Kafka
by: Raptis, Theofanis P., et al.
Published: (2022) -
Performance Evaluation of Brokerless Messaging Libraries
by: La Corte, Lorenzo, et al.
Published: (2025) -
Short-circuiting Rings for Low-Latency AllReduce
by: Hammer, Sarah-Michelle, et al.
Published: (2025) -
Solving AI Foundational Model Latency with Telco Infrastructure
by: Barros, Sebastian
Published: (2025)