Déjà Vu: Efficient Video-Language Query Engine with Learning-based Inter-Frame Computation Reuse
Fuente:
arXiv
Saved in:
| Main Authors: | Hwang, Jinwoo, Kim, Daeun, Lee, Sangyeop, Kim, Yoonsung, Heo, Guseul, Kim, Hojoon, Jeong, Yunseok, Meaza, Tadiwos, Park, Eunhyeok, Ahn, Jeongseob, Park, Jongse |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Understanding the Performance Behaviors of End-to-End Protein Design Pipelines on GPUs
by: Hwang, Jinwoo, et al.
Published: (2026)
by: Hwang, Jinwoo, et al.
Published: (2026)
Neo: Real-Time On-Device 3D Gaussian Splatting with Reuse-and-Update Sorting Acceleration
by: Oh, Changhun, et al.
Published: (2025)
by: Oh, Changhun, et al.
Published: (2025)
Accelerating String-Key Learned Index Structures via Memoization-based Incremental Training
by: Kim, Minsu, et al.
Published: (2024)
by: Kim, Minsu, et al.
Published: (2024)
MixDiT: Accelerating Image Diffusion Transformer Inference with Mixed-Precision MX Quantization
by: Kim, Daeun, et al.
Published: (2025)
by: Kim, Daeun, et al.
Published: (2025)
LLMServingSim: A HW/SW Co-Simulation Infrastructure for LLM Inference Serving at Scale
by: Cho, Jaehong, et al.
Published: (2024)
by: Cho, Jaehong, et al.
Published: (2024)
ONNXim: A Fast, Cycle-level Multi-core NPU Simulator
by: Ham, Hyungkyu, et al.
Published: (2024)
by: Ham, Hyungkyu, et al.
Published: (2024)
LLMServingSim 2.0: A Unified Simulator for Heterogeneous and Disaggregated LLM Serving Infrastructure
by: Cho, Jaehong, et al.
Published: (2026)
by: Cho, Jaehong, et al.
Published: (2026)
NeuPIMs: NPU-PIM Heterogeneous Acceleration for Batched LLM Inferencing
by: Heo, Guseul, et al.
Published: (2024)
by: Heo, Guseul, et al.
Published: (2024)
Detecting Time Series Anomalies Like an Expert: A Multi-Agent LLM Framework with Specialized Analyzers
by: Kang, Hyeongwon, et al.
Published: (2026)
by: Kang, Hyeongwon, et al.
Published: (2026)
Déjà Vu
by: Houppermans, Sjef, et al.
Published: (2016)
by: Houppermans, Sjef, et al.
Published: (2016)
DaCapo: Accelerating Continuous Learning in Autonomous Systems for Video Analytics
by: Kim, Yoonsung, et al.
Published: (2024)
by: Kim, Yoonsung, et al.
Published: (2024)
The Oversmoothing Fallacy: A Misguided Narrative in GNN Research
by: Park, MoonJeong, et al.
Published: (2025)
by: Park, MoonJeong, et al.
Published: (2025)
Let Multimodal Embedders Learn When to Augment Query via Adaptive Query Augmentation
by: Kim, Wongyu, et al.
Published: (2025)
by: Kim, Wongyu, et al.
Published: (2025)
Pimba: A Processing-in-Memory Acceleration for Post-Transformer Large Language Model Serving
by: Kim, Wonung, et al.
Published: (2025)
by: Kim, Wonung, et al.
Published: (2025)
Deja Vu from the Bridge.
by: Sullivan, Peggy
Published: (1979)
by: Sullivan, Peggy
Published: (1979)
HLQ: Fast and Efficient Backpropagation via Hadamard Low-rank Quantization
by: Kim, Seonggon, et al.
Published: (2024)
by: Kim, Seonggon, et al.
Published: (2024)
GraLoRA: Granular Low-Rank Adaptation for Parameter-Efficient Fine-Tuning
by: Jung, Yeonjoon, et al.
Published: (2025)
by: Jung, Yeonjoon, et al.
Published: (2025)
Déjà Vu Memorization in Vision-Language Models
by: Jayaraman, Bargav, et al.
Published: (2024)
by: Jayaraman, Bargav, et al.
Published: (2024)
The Electronic Revolution in Libraries: Microfilm Deja Vu?
by: Cady, Susan A.
Published: (1990)
by: Cady, Susan A.
Published: (1990)
Info Lit 2.0 or Déjà Vu?
by: Ianuzzi, Patricia Anne
Published: (2013)
by: Ianuzzi, Patricia Anne
Published: (2013)
Upcycling Candidate Tokens of Large Language Models for Query Expansion
by: Kim, Jinseok, et al.
Published: (2025)
by: Kim, Jinseok, et al.
Published: (2025)
FRDiff : Feature Reuse for Universal Training-free Acceleration of Diffusion Models
by: So, Junhyuk, et al.
Published: (2023)
by: So, Junhyuk, et al.
Published: (2023)
Diffusion Model Compression for Image-to-Image Translation
by: Kim, Geonung, et al.
Published: (2024)
by: Kim, Geonung, et al.
Published: (2024)
Size‐Dependent Characteristics of InGaN‐Based Blue and Green Micro‐Light‐Emitting Diodes
by: Shyam Mohan, et al.
Published: (2024)
by: Shyam Mohan, et al.
Published: (2024)
DejaVu: A Minimalistic Mechanism for Distributed Plurality Consensus
by: d'Amore, Francesco, et al.
Published: (2026)
by: d'Amore, Francesco, et al.
Published: (2026)
Déjà Vu? Decoding Repeated Reading from Eye Movements
by: Meiri, Yoav, et al.
Published: (2025)
by: Meiri, Yoav, et al.
Published: (2025)
News Deja Vu: Connecting Past and Present with Semantic Search
by: Franklin, Brevin, et al.
Published: (2024)
by: Franklin, Brevin, et al.
Published: (2024)
School and Public Library Relationships: Deja Vu or New Beginnings.
by: Fitzgibbons, Shirley A.
Published: (2001)
by: Fitzgibbons, Shirley A.
Published: (2001)
Efficient Latent Semantic Clustering for Scaling Test-Time Computation of LLMs
by: Lee, Sungjae, et al.
Published: (2025)
by: Lee, Sungjae, et al.
Published: (2025)
Efficient LLM Inference with Activation Checkpointing and Hybrid Caching
by: Lee, Sanghyeon, et al.
Published: (2025)
by: Lee, Sanghyeon, et al.
Published: (2025)
Deep Filter Estimation from Inter-Frame Correlations for Monaural Speech Dereverberation
by: Shin, Ui-Hyeop, et al.
Published: (2026)
by: Shin, Ui-Hyeop, et al.
Published: (2026)
Framing Matters: Addressing Framing Sensitivity in Decision-Making through Behaviorally-Grounded Value Alignment
by: Hwang, Seojin, et al.
Published: (2026)
by: Hwang, Seojin, et al.
Published: (2026)
EMO100DB: An Open Dataset of Improvised Songs with Emotion Data
by: Hwang, Daeun, et al.
Published: (2025)
by: Hwang, Daeun, et al.
Published: (2025)
In‐Memory Euclidean Distance Computation in a Stacked Memristor Crossbar for Hardware Self‐Organizing Maps
by: Jinwoo Park, et al.
Published: (2026)
by: Jinwoo Park, et al.
Published: (2026)
HD Maps are Lane Detection Generalizers: A Novel Generative Framework for Single-Source Domain Generalization
by: Lee, Daeun, et al.
Published: (2023)
by: Lee, Daeun, et al.
Published: (2023)
OWQ: Outlier-Aware Weight Quantization for Efficient Fine-Tuning and Inference of Large Language Models
by: Lee, Changhun, et al.
Published: (2023)
by: Lee, Changhun, et al.
Published: (2023)
Noisy Neighbor: Exploiting RDMA for Resource Exhaustion Attacks in Containerized Clouds
by: Kim, Gunwoo, et al.
Published: (2025)
by: Kim, Gunwoo, et al.
Published: (2025)
Merge-Friendly Post-Training Quantization for Multi-Target Domain Adaptation
by: Shin, Juncheol, et al.
Published: (2025)
by: Shin, Juncheol, et al.
Published: (2025)
PTQ4VM: Post-Training Quantization for Visual Mamba
by: Cho, Younghyun, et al.
Published: (2024)
by: Cho, Younghyun, et al.
Published: (2024)
LLM-guided Plan and Retrieval: A Strategic Alignment for Interpretable User Satisfaction Estimation in Dialogue
by: Kim, Sangyeop, et al.
Published: (2025)
by: Kim, Sangyeop, et al.
Published: (2025)
Similar Items
-
Understanding the Performance Behaviors of End-to-End Protein Design Pipelines on GPUs
by: Hwang, Jinwoo, et al.
Published: (2026) -
Neo: Real-Time On-Device 3D Gaussian Splatting with Reuse-and-Update Sorting Acceleration
by: Oh, Changhun, et al.
Published: (2025) -
Accelerating String-Key Learned Index Structures via Memoization-based Incremental Training
by: Kim, Minsu, et al.
Published: (2024) -
MixDiT: Accelerating Image Diffusion Transformer Inference with Mixed-Precision MX Quantization
by: Kim, Daeun, et al.
Published: (2025) -
LLMServingSim: A HW/SW Co-Simulation Infrastructure for LLM Inference Serving at Scale
by: Cho, Jaehong, et al.
Published: (2024)