MeanCache: From Instantaneous to Average Velocity for Accelerating Flow Matching Inference
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gao, Huanlin, Chen, Ping, Shi, Fuyuan, Wu, Ruijia, YanTao, Li, Hui, Qiang, You, Yuren, Lu, Ting, Tan, Chao, Zhao, Shaoan, Liu, Zhaoxiang, Zhao, Fang, Wang, Kai, Lian, Shiguo |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
LeMiCa: Lexicographic Minimax Path Caching for Efficient Diffusion-Based Video Generation
von: Gao, Huanlin, et al.
Veröffentlicht: (2025)
von: Gao, Huanlin, et al.
Veröffentlicht: (2025)
HiMo-CLIP: Modeling Semantic Hierarchy and Monotonicity in Vision-Language Alignment
von: Wu, Ruijia, et al.
Veröffentlicht: (2025)
von: Wu, Ruijia, et al.
Veröffentlicht: (2025)
MediaClaw: Multimodal Intelligent-Agent Platform Technical Report
von: Zhao, Shaoan, et al.
Veröffentlicht: (2026)
von: Zhao, Shaoan, et al.
Veröffentlicht: (2026)
MeanCache: User-Centric Semantic Caching for LLM Web Services
von: Gill, Waris, et al.
Veröffentlicht: (2024)
von: Gill, Waris, et al.
Veröffentlicht: (2024)
TP3M: Transformer-based Pseudo 3D Image Matching with Reference Image
von: Han, Liming, et al.
Veröffentlicht: (2024)
von: Han, Liming, et al.
Veröffentlicht: (2024)
Patch-wise Auto-Encoder for Visual Anomaly Detection
von: Cui, Yajie, et al.
Veröffentlicht: (2023)
von: Cui, Yajie, et al.
Veröffentlicht: (2023)
Mean Flow Policy with Instantaneous Velocity Constraint for One-step Action Generation
von: Zhan, Guojian, et al.
Veröffentlicht: (2026)
von: Zhan, Guojian, et al.
Veröffentlicht: (2026)
LayerCache: Exploiting Layer-wise Velocity Heterogeneity for Efficient Flow Matching Inference
von: Li, Guandong
Veröffentlicht: (2026)
von: Li, Guandong
Veröffentlicht: (2026)
KAConvNet: Kolmogorov-Arnold Convolutional Networks for Vision Recognition
von: Liu, Zhaoxiang, et al.
Veröffentlicht: (2026)
von: Liu, Zhaoxiang, et al.
Veröffentlicht: (2026)
Compose Yourself: Average-Velocity Flow Matching for One-Step Speech Enhancement
von: Yang, Gang, et al.
Veröffentlicht: (2025)
von: Yang, Gang, et al.
Veröffentlicht: (2025)
Pave-GRPO: Beyond Instantaneous Guidance through Principled Average Velocity Decomposition
von: Ling, Pengyang, et al.
Veröffentlicht: (2026)
von: Ling, Pengyang, et al.
Veröffentlicht: (2026)
GlitchMiner: Mining Glitch Tokens in Large Language Models via Gradient-based Discrete Optimization
von: Wu, Zihui, et al.
Veröffentlicht: (2024)
von: Wu, Zihui, et al.
Veröffentlicht: (2024)
A Large Vision-Language Model based Environment Perception System for Visually Impaired People
von: Chen, Zezhou, et al.
Veröffentlicht: (2025)
von: Chen, Zezhou, et al.
Veröffentlicht: (2025)
Piculet: Specialized Models-Guided Hallucination Decrease for MultiModal Large Language Models
von: Wang, Kohou, et al.
Veröffentlicht: (2024)
von: Wang, Kohou, et al.
Veröffentlicht: (2024)
Hierarchical Deep Fusion Framework for Multi-dimensional Facial Forgery Detection -- The 2024 Global Deepfake Image Detection Challenge
von: Wang, Kohou, et al.
Veröffentlicht: (2025)
von: Wang, Kohou, et al.
Veröffentlicht: (2025)
Beyond Fixed Inference: Quantitative Flow Matching for Adaptive Image Denoising
von: Duan, Jigang, et al.
Veröffentlicht: (2026)
von: Duan, Jigang, et al.
Veröffentlicht: (2026)
Flow Matching for Averaged Systems
von: Adu, Daniel Owusu, et al.
Veröffentlicht: (2025)
von: Adu, Daniel Owusu, et al.
Veröffentlicht: (2025)
What is the best model? Application-driven Evaluation for Large Language Models
von: Lian, Shiguo, et al.
Veröffentlicht: (2024)
von: Lian, Shiguo, et al.
Veröffentlicht: (2024)
CHiSafetyBench: A Chinese Hierarchical Safety Benchmark for Large Language Models
von: Zhang, Wenjing, et al.
Veröffentlicht: (2024)
von: Zhang, Wenjing, et al.
Veröffentlicht: (2024)
FastFlow: Accelerating The Generative Flow Matching Models with Bandit Inference
von: Bajpai, Divya Jyoti, et al.
Veröffentlicht: (2026)
von: Bajpai, Divya Jyoti, et al.
Veröffentlicht: (2026)
ALISA: Accelerating Large Language Model Inference via Sparsity-Aware KV Caching
von: Zhao, Youpeng, et al.
Veröffentlicht: (2024)
von: Zhao, Youpeng, et al.
Veröffentlicht: (2024)
PSTF-AttControl: Per-Subject-Tuning-Free Personalized Image Generation with Controllable Face Attributes
von: liu, Xiang, et al.
Veröffentlicht: (2025)
von: liu, Xiang, et al.
Veröffentlicht: (2025)
Fatigue Life Analysis of Cranes Based on Load Spectrum Prediction and Fracture Mechanics
von: Mantang Hu, et al.
Veröffentlicht: (2025)
von: Mantang Hu, et al.
Veröffentlicht: (2025)
FAVE: Flow-based Average Velocity Establishment for Sequential Recommendation
von: Shi, Ke, et al.
Veröffentlicht: (2026)
von: Shi, Ke, et al.
Veröffentlicht: (2026)
Accelerating LLM Inference Throughput via Asynchronous KV Cache Prefetching
von: Dong, Yanhao, et al.
Veröffentlicht: (2025)
von: Dong, Yanhao, et al.
Veröffentlicht: (2025)
Coreset-Induced Conditional Velocity Flow Matching
von: Wang, Xiao, et al.
Veröffentlicht: (2026)
von: Wang, Xiao, et al.
Veröffentlicht: (2026)
Consistency Flow Matching: Defining Straight Flows with Velocity Consistency
von: Yang, Ling, et al.
Veröffentlicht: (2024)
von: Yang, Ling, et al.
Veröffentlicht: (2024)
A Systematic Security Evaluation of OpenClaw and Its Variants
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
von: Wang, Yuhang, et al.
Veröffentlicht: (2026)
A Multimodal Benchmark Dataset and Model for Crop Disease Diagnosis
von: Liu, Xiang, et al.
Veröffentlicht: (2025)
von: Liu, Xiang, et al.
Veröffentlicht: (2025)
SLearnLLM: A Self-Learning Framework for Efficient Domain-Specific Adaptation of Large Language Models
von: Liu, Xiang, et al.
Veröffentlicht: (2025)
von: Liu, Xiang, et al.
Veröffentlicht: (2025)
Optimizing for the Shortest Path in Denoising Diffusion Model
von: Chen, Ping, et al.
Veröffentlicht: (2025)
von: Chen, Ping, et al.
Veröffentlicht: (2025)
SenCache: Accelerating Diffusion Model Inference via Sensitivity-Aware Caching
von: Haghighi, Yasaman, et al.
Veröffentlicht: (2026)
von: Haghighi, Yasaman, et al.
Veröffentlicht: (2026)
Firm theories in neoclassical institutional economics
von: Shaoan Huang
Veröffentlicht: (2024)
von: Shaoan Huang
Veröffentlicht: (2024)
Geometric Erasure by Contrastive Velocity Matching in Rectified Flows
von: Grebe, Jonas Henry, et al.
Veröffentlicht: (2026)
von: Grebe, Jonas Henry, et al.
Veröffentlicht: (2026)
Stable Velocity: A Variance Perspective on Flow Matching
von: Yang, Donglin, et al.
Veröffentlicht: (2026)
von: Yang, Donglin, et al.
Veröffentlicht: (2026)
The Velocity Deficit: Initial Energy Injection for Flow Matching
von: Li, Linze, et al.
Veröffentlicht: (2026)
von: Li, Linze, et al.
Veröffentlicht: (2026)
Window-Diffusion: Accelerating Diffusion Language Model Inference with Windowed Token Pruning and Caching
von: Zuo, Fengrui, et al.
Veröffentlicht: (2026)
von: Zuo, Fengrui, et al.
Veröffentlicht: (2026)
Agentic AI for Particle-Based Simulation: Automating SPH Workflows for Debris Flow Modeling
von: Zhang, Danrong, et al.
Veröffentlicht: (2026)
von: Zhang, Danrong, et al.
Veröffentlicht: (2026)
ProCache: Constraint-Aware Feature Caching with Selective Computation for Diffusion Transformer Acceleration
von: Cao, Fanpu, et al.
Veröffentlicht: (2025)
von: Cao, Fanpu, et al.
Veröffentlicht: (2025)
Parallelized Instantaneous Velocity and Heading Estimation of Objects using Single Imaging Radar
von: Singh, Nihal, et al.
Veröffentlicht: (2020)
von: Singh, Nihal, et al.
Veröffentlicht: (2020)
Ähnliche Einträge
-
LeMiCa: Lexicographic Minimax Path Caching for Efficient Diffusion-Based Video Generation
von: Gao, Huanlin, et al.
Veröffentlicht: (2025) -
HiMo-CLIP: Modeling Semantic Hierarchy and Monotonicity in Vision-Language Alignment
von: Wu, Ruijia, et al.
Veröffentlicht: (2025) -
MediaClaw: Multimodal Intelligent-Agent Platform Technical Report
von: Zhao, Shaoan, et al.
Veröffentlicht: (2026) -
MeanCache: User-Centric Semantic Caching for LLM Web Services
von: Gill, Waris, et al.
Veröffentlicht: (2024) -
TP3M: Transformer-based Pseudo 3D Image Matching with Reference Image
von: Han, Liming, et al.
Veröffentlicht: (2024)