MeanCache: From Instantaneous to Average Velocity for Accelerating Flow Matching Inference
Fuente:
arXiv
Saved in:
| Main Authors: | Gao, Huanlin, Chen, Ping, Shi, Fuyuan, Wu, Ruijia, YanTao, Li, Hui, Qiang, You, Yuren, Lu, Ting, Tan, Chao, Zhao, Shaoan, Liu, Zhaoxiang, Zhao, Fang, Wang, Kai, Lian, Shiguo |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LeMiCa: Lexicographic Minimax Path Caching for Efficient Diffusion-Based Video Generation
by: Gao, Huanlin, et al.
Published: (2025)
by: Gao, Huanlin, et al.
Published: (2025)
HiMo-CLIP: Modeling Semantic Hierarchy and Monotonicity in Vision-Language Alignment
by: Wu, Ruijia, et al.
Published: (2025)
by: Wu, Ruijia, et al.
Published: (2025)
MediaClaw: Multimodal Intelligent-Agent Platform Technical Report
by: Zhao, Shaoan, et al.
Published: (2026)
by: Zhao, Shaoan, et al.
Published: (2026)
MeanCache: User-Centric Semantic Caching for LLM Web Services
by: Gill, Waris, et al.
Published: (2024)
by: Gill, Waris, et al.
Published: (2024)
TP3M: Transformer-based Pseudo 3D Image Matching with Reference Image
by: Han, Liming, et al.
Published: (2024)
by: Han, Liming, et al.
Published: (2024)
Patch-wise Auto-Encoder for Visual Anomaly Detection
by: Cui, Yajie, et al.
Published: (2023)
by: Cui, Yajie, et al.
Published: (2023)
Mean Flow Policy with Instantaneous Velocity Constraint for One-step Action Generation
by: Zhan, Guojian, et al.
Published: (2026)
by: Zhan, Guojian, et al.
Published: (2026)
LayerCache: Exploiting Layer-wise Velocity Heterogeneity for Efficient Flow Matching Inference
by: Li, Guandong
Published: (2026)
by: Li, Guandong
Published: (2026)
KAConvNet: Kolmogorov-Arnold Convolutional Networks for Vision Recognition
by: Liu, Zhaoxiang, et al.
Published: (2026)
by: Liu, Zhaoxiang, et al.
Published: (2026)
Compose Yourself: Average-Velocity Flow Matching for One-Step Speech Enhancement
by: Yang, Gang, et al.
Published: (2025)
by: Yang, Gang, et al.
Published: (2025)
Pave-GRPO: Beyond Instantaneous Guidance through Principled Average Velocity Decomposition
by: Ling, Pengyang, et al.
Published: (2026)
by: Ling, Pengyang, et al.
Published: (2026)
GlitchMiner: Mining Glitch Tokens in Large Language Models via Gradient-based Discrete Optimization
by: Wu, Zihui, et al.
Published: (2024)
by: Wu, Zihui, et al.
Published: (2024)
A Large Vision-Language Model based Environment Perception System for Visually Impaired People
by: Chen, Zezhou, et al.
Published: (2025)
by: Chen, Zezhou, et al.
Published: (2025)
Piculet: Specialized Models-Guided Hallucination Decrease for MultiModal Large Language Models
by: Wang, Kohou, et al.
Published: (2024)
by: Wang, Kohou, et al.
Published: (2024)
Hierarchical Deep Fusion Framework for Multi-dimensional Facial Forgery Detection -- The 2024 Global Deepfake Image Detection Challenge
by: Wang, Kohou, et al.
Published: (2025)
by: Wang, Kohou, et al.
Published: (2025)
Beyond Fixed Inference: Quantitative Flow Matching for Adaptive Image Denoising
by: Duan, Jigang, et al.
Published: (2026)
by: Duan, Jigang, et al.
Published: (2026)
Flow Matching for Averaged Systems
by: Adu, Daniel Owusu, et al.
Published: (2025)
by: Adu, Daniel Owusu, et al.
Published: (2025)
What is the best model? Application-driven Evaluation for Large Language Models
by: Lian, Shiguo, et al.
Published: (2024)
by: Lian, Shiguo, et al.
Published: (2024)
CHiSafetyBench: A Chinese Hierarchical Safety Benchmark for Large Language Models
by: Zhang, Wenjing, et al.
Published: (2024)
by: Zhang, Wenjing, et al.
Published: (2024)
FastFlow: Accelerating The Generative Flow Matching Models with Bandit Inference
by: Bajpai, Divya Jyoti, et al.
Published: (2026)
by: Bajpai, Divya Jyoti, et al.
Published: (2026)
ALISA: Accelerating Large Language Model Inference via Sparsity-Aware KV Caching
by: Zhao, Youpeng, et al.
Published: (2024)
by: Zhao, Youpeng, et al.
Published: (2024)
PSTF-AttControl: Per-Subject-Tuning-Free Personalized Image Generation with Controllable Face Attributes
by: liu, Xiang, et al.
Published: (2025)
by: liu, Xiang, et al.
Published: (2025)
Fatigue Life Analysis of Cranes Based on Load Spectrum Prediction and Fracture Mechanics
by: Mantang Hu, et al.
Published: (2025)
by: Mantang Hu, et al.
Published: (2025)
FAVE: Flow-based Average Velocity Establishment for Sequential Recommendation
by: Shi, Ke, et al.
Published: (2026)
by: Shi, Ke, et al.
Published: (2026)
Accelerating LLM Inference Throughput via Asynchronous KV Cache Prefetching
by: Dong, Yanhao, et al.
Published: (2025)
by: Dong, Yanhao, et al.
Published: (2025)
Coreset-Induced Conditional Velocity Flow Matching
by: Wang, Xiao, et al.
Published: (2026)
by: Wang, Xiao, et al.
Published: (2026)
Consistency Flow Matching: Defining Straight Flows with Velocity Consistency
by: Yang, Ling, et al.
Published: (2024)
by: Yang, Ling, et al.
Published: (2024)
A Systematic Security Evaluation of OpenClaw and Its Variants
by: Wang, Yuhang, et al.
Published: (2026)
by: Wang, Yuhang, et al.
Published: (2026)
A Multimodal Benchmark Dataset and Model for Crop Disease Diagnosis
by: Liu, Xiang, et al.
Published: (2025)
by: Liu, Xiang, et al.
Published: (2025)
SLearnLLM: A Self-Learning Framework for Efficient Domain-Specific Adaptation of Large Language Models
by: Liu, Xiang, et al.
Published: (2025)
by: Liu, Xiang, et al.
Published: (2025)
Optimizing for the Shortest Path in Denoising Diffusion Model
by: Chen, Ping, et al.
Published: (2025)
by: Chen, Ping, et al.
Published: (2025)
SenCache: Accelerating Diffusion Model Inference via Sensitivity-Aware Caching
by: Haghighi, Yasaman, et al.
Published: (2026)
by: Haghighi, Yasaman, et al.
Published: (2026)
Firm theories in neoclassical institutional economics
by: Shaoan Huang
Published: (2024)
by: Shaoan Huang
Published: (2024)
Geometric Erasure by Contrastive Velocity Matching in Rectified Flows
by: Grebe, Jonas Henry, et al.
Published: (2026)
by: Grebe, Jonas Henry, et al.
Published: (2026)
Stable Velocity: A Variance Perspective on Flow Matching
by: Yang, Donglin, et al.
Published: (2026)
by: Yang, Donglin, et al.
Published: (2026)
The Velocity Deficit: Initial Energy Injection for Flow Matching
by: Li, Linze, et al.
Published: (2026)
by: Li, Linze, et al.
Published: (2026)
Window-Diffusion: Accelerating Diffusion Language Model Inference with Windowed Token Pruning and Caching
by: Zuo, Fengrui, et al.
Published: (2026)
by: Zuo, Fengrui, et al.
Published: (2026)
Agentic AI for Particle-Based Simulation: Automating SPH Workflows for Debris Flow Modeling
by: Zhang, Danrong, et al.
Published: (2026)
by: Zhang, Danrong, et al.
Published: (2026)
ProCache: Constraint-Aware Feature Caching with Selective Computation for Diffusion Transformer Acceleration
by: Cao, Fanpu, et al.
Published: (2025)
by: Cao, Fanpu, et al.
Published: (2025)
Parallelized Instantaneous Velocity and Heading Estimation of Objects using Single Imaging Radar
by: Singh, Nihal, et al.
Published: (2020)
by: Singh, Nihal, et al.
Published: (2020)
Similar Items
-
LeMiCa: Lexicographic Minimax Path Caching for Efficient Diffusion-Based Video Generation
by: Gao, Huanlin, et al.
Published: (2025) -
HiMo-CLIP: Modeling Semantic Hierarchy and Monotonicity in Vision-Language Alignment
by: Wu, Ruijia, et al.
Published: (2025) -
MediaClaw: Multimodal Intelligent-Agent Platform Technical Report
by: Zhao, Shaoan, et al.
Published: (2026) -
MeanCache: User-Centric Semantic Caching for LLM Web Services
by: Gill, Waris, et al.
Published: (2024) -
TP3M: Transformer-based Pseudo 3D Image Matching with Reference Image
by: Han, Liming, et al.
Published: (2024)