Reducing End-to-End Latency of Cause-Effect Chains with Shared Cache Analysis
Fuente:
arXiv
Saved in:
| Main Authors: | Zhu, Yixuan, Gao, Yinkang, Zhang, Bo, Gong, Xiaohang, Jiang, Binze, Gong, Lei, Lou, Wenqi, Wang, Teng, Wang, Chao, Li, Xi, Zhou, Xuehai |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Scheduling Cause-Effect Chains without Timing Anomalies in End-to-End Latency
by: Zhu, Yixuan, et al.
Published: (2026)
by: Zhu, Yixuan, et al.
Published: (2026)
A Timing-Anomaly Free Dynamic Scheduling on Heterogeneous Systems
by: Zhu, Yixuan, et al.
Published: (2026)
by: Zhu, Yixuan, et al.
Published: (2026)
Reducing End-to-End Latencies of Multi-Rate Cause-Effect Chains for the LET Model
by: Maia, Luiz, et al.
Published: (2023)
by: Maia, Luiz, et al.
Published: (2023)
Crystal-KV: Efficient KV Cache Management for Chain-of-Thought LLMs via Answer-First Principle
by: Wang, Zihan, et al.
Published: (2026)
by: Wang, Zihan, et al.
Published: (2026)
Decreasing Utilization of Systems with Multi-Rate Cause-Effect Chains While Reducing End-to-End Latencies
by: Maia, Luiz, et al.
Published: (2025)
by: Maia, Luiz, et al.
Published: (2025)
Window-Diffusion: Accelerating Diffusion Language Model Inference with Windowed Token Pruning and Caching
by: Zuo, Fengrui, et al.
Published: (2026)
by: Zuo, Fengrui, et al.
Published: (2026)
UbiMoE: A Ubiquitous Mixture-of-Experts Vision Transformer Accelerator With Hybrid Computation Pattern on FPGA
by: Dong, Jiale, et al.
Published: (2025)
by: Dong, Jiale, et al.
Published: (2025)
Hermes: A Unified High-Performance NTT Architecture with Hybrid Dataflow
by: Gu, Hang, et al.
Published: (2026)
by: Gu, Hang, et al.
Published: (2026)
CoQMoE: Co-Designed Quantization and Computation Orchestration for Mixture-of-Experts Vision Transformer on FPGA
by: Dong, Jiale, et al.
Published: (2025)
by: Dong, Jiale, et al.
Published: (2025)
ActionFlow: A Pipelined Action Acceleration for Vision Language Models on Edge
by: Dai, Yuntao, et al.
Published: (2025)
by: Dai, Yuntao, et al.
Published: (2025)
Bits-to-Photon: End-to-End Learned Scalable Point Cloud Compression for Direct Rendering
by: Hu, Yueyu, et al.
Published: (2024)
by: Hu, Yueyu, et al.
Published: (2024)
SynthVC: Leveraging Synthetic Data for End-to-End Low Latency Streaming Voice Conversion
by: Guo, Zhao, et al.
Published: (2025)
by: Guo, Zhao, et al.
Published: (2025)
SQLens: An End-to-End Framework for Error Detection and Correction in Text-to-SQL
by: Gong, Yue, et al.
Published: (2025)
by: Gong, Yue, et al.
Published: (2025)
End-to-End Evaluation for Low-Latency Simultaneous Speech Translation
by: Huber, Christian, et al.
Published: (2023)
by: Huber, Christian, et al.
Published: (2023)
We Can Hear You with mmWave Radar! An End-to-End Eavesdropping System
by: Han, Dachao, et al.
Published: (2025)
by: Han, Dachao, et al.
Published: (2025)
Is the Molecular Weight Dependence of the Glass Transition Temperature Caused by a Chain End Effect?
by: Drayer, William F., et al.
Published: (2023)
by: Drayer, William F., et al.
Published: (2023)
ReaGeo: Reasoning-Enhanced End-to-End Geocoding with LLMs
by: Cui, Jian, et al.
Published: (2026)
by: Cui, Jian, et al.
Published: (2026)
LKV: End-to-End Learning of Head-wise Budgets and Token Selection for LLM KV Cache Eviction
by: Zhou, Enshuai, et al.
Published: (2026)
by: Zhou, Enshuai, et al.
Published: (2026)
SimpleVSF: VLM-Scoring Fusion for Trajectory Prediction of End-to-End Autonomous Driving
by: Zheng, Peiru, et al.
Published: (2025)
by: Zheng, Peiru, et al.
Published: (2025)
mmFHE: mmWave Sensing with End-to-End Fully Homomorphic Encryption
by: Ahmed, Tanvir, et al.
Published: (2026)
by: Ahmed, Tanvir, et al.
Published: (2026)
Accelerating Precise End-to-End Simulation: Latency-Sensitive Many-core System Modeling
by: Li, Yinrong, et al.
Published: (2026)
by: Li, Yinrong, et al.
Published: (2026)
Embodied Cognition Augmented End2End Autonomous Driving
by: Niu, Ling, et al.
Published: (2025)
by: Niu, Ling, et al.
Published: (2025)
End-to-End Long Document Summarization using Gradient Caching
by: Saxena, Rohit, et al.
Published: (2025)
by: Saxena, Rohit, et al.
Published: (2025)
How Well Do Large Language Models Serve as End-to-End Secure Code Agents for Python?
by: Gong, Jianian, et al.
Published: (2024)
by: Gong, Jianian, et al.
Published: (2024)
Safety-Critical Edge Robotics Architecture with Bounded End-to-End Latency
by: Gala, Gautam, et al.
Published: (2024)
by: Gala, Gautam, et al.
Published: (2024)
Minimizing End-to-End Latency for Joint Source-Channel Coding Systems
by: Chi, Kaiyi, et al.
Published: (2024)
by: Chi, Kaiyi, et al.
Published: (2024)
End-to-End Latency Measurement Methodology for Connected and Autonomous Vehicle Teleoperation
by: Provost, François, et al.
Published: (2026)
by: Provost, François, et al.
Published: (2026)
LP-LLM: End-to-End Real-World Degraded License Plate Text Recognition via Large Multimodal Models
by: Gong, Haoyan, et al.
Published: (2026)
by: Gong, Haoyan, et al.
Published: (2026)
VADv2: End-to-End Vectorized Autonomous Driving via Probabilistic Planning
by: Jiang, Bo, et al.
Published: (2024)
by: Jiang, Bo, et al.
Published: (2024)
EGA-V1: Unifying Online Advertising with End-to-End Learning
by: Qiu, Junyan, et al.
Published: (2025)
by: Qiu, Junyan, et al.
Published: (2025)
CommonKV: Compressing KV Cache with Cross-layer Parameter Sharing
by: Wang, Yixuan, et al.
Published: (2025)
by: Wang, Yixuan, et al.
Published: (2025)
An End-to-End Robust Point Cloud Semantic Segmentation Network with Single-Step Conditional Diffusion Models
by: Qu, Wentao, et al.
Published: (2024)
by: Qu, Wentao, et al.
Published: (2024)
A Study of Rule Omission in Raven's Progressive Matrices
by: Li, Binze
Published: (2025)
by: Li, Binze
Published: (2025)
DriveCoT: Integrating Chain-of-Thought Reasoning with End-to-End Driving
by: Wang, Tianqi, et al.
Published: (2024)
by: Wang, Tianqi, et al.
Published: (2024)
Senna-2: Aligning VLM and End-to-End Driving Policy for Consistent Decision Making and Planning
by: Song, Yuehao, et al.
Published: (2026)
by: Song, Yuehao, et al.
Published: (2026)
Chain-of-Agents: End-to-End Agent Foundation Models via Multi-Agent Distillation and Agentic RL
by: Li, Weizhen, et al.
Published: (2025)
by: Li, Weizhen, et al.
Published: (2025)
MaskFuser: Masked Fusion of Joint Multi-Modal Tokenization for End-to-End Autonomous Driving
by: Duan, Yiqun, et al.
Published: (2024)
by: Duan, Yiqun, et al.
Published: (2024)
Dynamic Interval Scheduling with Random Start and End Times
by: Gong, Rui, et al.
Published: (2026)
by: Gong, Rui, et al.
Published: (2026)
SimpleLLM4AD: An End-to-End Vision-Language Model with Graph Visual Question Answering for Autonomous Driving
by: Zheng, Peiru, et al.
Published: (2024)
by: Zheng, Peiru, et al.
Published: (2024)
FlowDrive: Energy Flow Field for End-to-End Autonomous Driving
by: Jiang, Hao, et al.
Published: (2025)
by: Jiang, Hao, et al.
Published: (2025)
Similar Items
-
Scheduling Cause-Effect Chains without Timing Anomalies in End-to-End Latency
by: Zhu, Yixuan, et al.
Published: (2026) -
A Timing-Anomaly Free Dynamic Scheduling on Heterogeneous Systems
by: Zhu, Yixuan, et al.
Published: (2026) -
Reducing End-to-End Latencies of Multi-Rate Cause-Effect Chains for the LET Model
by: Maia, Luiz, et al.
Published: (2023) -
Crystal-KV: Efficient KV Cache Management for Chain-of-Thought LLMs via Answer-First Principle
by: Wang, Zihan, et al.
Published: (2026) -
Decreasing Utilization of Systems with Multi-Rate Cause-Effect Chains While Reducing End-to-End Latencies
by: Maia, Luiz, et al.
Published: (2025)