EchoPrune: Interpreting Redundancy as Temporal Echoes for Efficient VideoLLMs
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Li, Jiameng, Wu, Minye, Cao, Jiezhang, Tiulpin, Aleksei, Blaschko, Matthew B. |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
MI-Pruner: Crossmodal Mutual Information-guided Token Pruner for Efficient MLLMs
par: Li, Jiameng, et autres
Publié: (2026)
par: Li, Jiameng, et autres
Publié: (2026)
SoftCFG: Uncertainty-guided Stable Guidance for Visual Autoregressive Model
par: Xu, Dongli, et autres
Publié: (2025)
par: Xu, Dongli, et autres
Publié: (2025)
CARE: Confidence-aware Ratio Estimation for Medical Biomarkers
par: Li, Jiameng, et autres
Publié: (2025)
par: Li, Jiameng, et autres
Publié: (2025)
Language-Guided Temporal Token Pruning for Efficient VideoLLM Processing
par: Kumar, Yogesh
Publié: (2025)
par: Kumar, Yogesh
Publié: (2025)
Sharp Eyes and Memory for VideoLLMs: Information-Aware Visual Token Pruning for Efficient and Reliable VideoLLM Reasoning
par: Qin, Jialong, et autres
Publié: (2025)
par: Qin, Jialong, et autres
Publié: (2025)
Lost in Time: A New Temporal Benchmark for VideoLLMs
par: Cores, Daniel, et autres
Publié: (2024)
par: Cores, Daniel, et autres
Publié: (2024)
Mitigating Hallucination in VideoLLMs via Temporal-Aware Activation Engineering
par: Cai, Jianfeng, et autres
Publié: (2025)
par: Cai, Jianfeng, et autres
Publié: (2025)
CLASH: A Benchmark for Cross-Modal Contradiction Detection
par: Popordanoska, Teodora, et autres
Publié: (2025)
par: Popordanoska, Teodora, et autres
Publié: (2025)
Diversity-Driven View Subset Selection for Indoor Novel View Synthesis
par: Wang, Zehao, et autres
Publié: (2024)
par: Wang, Zehao, et autres
Publié: (2024)
Efficient Video Sampling: Pruning Temporally Redundant Tokens for Faster VLM Inference
par: Bagrov, Natan, et autres
Publié: (2025)
par: Bagrov, Natan, et autres
Publié: (2025)
Video Streaming Thinking: VideoLLMs Can Watch and Think Simultaneously
par: Guan, Yiran, et autres
Publié: (2026)
par: Guan, Yiran, et autres
Publié: (2026)
Map the Flow: Revealing Hidden Pathways of Information in VideoLLMs
par: Kim, Minji, et autres
Publié: (2025)
par: Kim, Minji, et autres
Publié: (2025)
Reinforcement Learning Tuning for VideoLLMs: Reward Design and Data Efficiency
par: Li, Hongyu, et autres
Publié: (2025)
par: Li, Hongyu, et autres
Publié: (2025)
SiNGR: Brain Tumor Segmentation via Signed Normalized Geodesic Transform Regression
par: Dang, Trung, et autres
Publié: (2024)
par: Dang, Trung, et autres
Publié: (2024)
Image-level Regression for Uncertainty-aware Retinal Image Segmentation
par: Dang, Trung, et autres
Publié: (2024)
par: Dang, Trung, et autres
Publié: (2024)
Memory-efficient Streaming VideoLLMs for Real-time Procedural Video Understanding
par: Chatterjee, Dibyadip, et autres
Publié: (2025)
par: Chatterjee, Dibyadip, et autres
Publié: (2025)
Implicit Gaussian Splatting with Efficient Multi-Level Tri-Plane Representation
par: Wu, Minye, et autres
Publié: (2024)
par: Wu, Minye, et autres
Publié: (2024)
Geometry-Guided Camera Motion Understanding in VideoLLMs
par: Feng, Haoan, et autres
Publié: (2026)
par: Feng, Haoan, et autres
Publié: (2026)
VideoThinker: Building Agentic VideoLLMs with LLM-Guided Tool Reasoning
par: Li, Chenglin, et autres
Publié: (2026)
par: Li, Chenglin, et autres
Publié: (2026)
Video Parallel Scaling: Aggregating Diverse Frame Subsets for VideoLLMs
par: Chung, Hyungjin, et autres
Publié: (2025)
par: Chung, Hyungjin, et autres
Publié: (2025)
Grounded-VideoLLM: Sharpening Fine-grained Temporal Grounding in Video Large Language Models
par: Wang, Haibo, et autres
Publié: (2024)
par: Wang, Haibo, et autres
Publié: (2024)
WeaveTime: Stream from Earlier Frames into Emergent Memory in VideoLLMs
par: Zhang, Yulin, et autres
Publié: (2026)
par: Zhang, Yulin, et autres
Publié: (2026)
VideoLLM-MoD: Efficient Video-Language Streaming with Mixture-of-Depths Vision Computation
par: Wu, Shiwei, et autres
Publié: (2024)
par: Wu, Shiwei, et autres
Publié: (2024)
VideoLLM-online: Online Video Large Language Model for Streaming Video
par: Chen, Joya, et autres
Publié: (2024)
par: Chen, Joya, et autres
Publié: (2024)
Mipmap-GS: Let Gaussians Deform with Scale-specific Mipmap for Anti-aliasing Rendering
par: Li, Jiameng, et autres
Publié: (2024)
par: Li, Jiameng, et autres
Publié: (2024)
Dynamic-VLM: Simple Dynamic Visual Token Compression for VideoLLM
par: Wang, Han, et autres
Publié: (2024)
par: Wang, Han, et autres
Publié: (2024)
Predicting Knee Osteoarthritis Progression from Structural MRI using Deep Learning
par: Panfilov, Egor, et autres
Publié: (2022)
par: Panfilov, Egor, et autres
Publié: (2022)
VideoLLM Benchmarks and Evaluation: A Survey
par: Kumar, Yogesh
Publié: (2025)
par: Kumar, Yogesh
Publié: (2025)
LoG-VMamba: Local-Global Vision Mamba for Medical Image Segmentation
par: Dang, Trung Dinh Quoc, et autres
Publié: (2024)
par: Dang, Trung Dinh Quoc, et autres
Publié: (2024)
Proact-VL: A Proactive VideoLLM for Real-Time AI Companions
par: Yan, Weicai, et autres
Publié: (2026)
par: Yan, Weicai, et autres
Publié: (2026)
ReTaKe: Reducing Temporal and Knowledge Redundancy for Long Video Understanding
par: Wang, Xiao, et autres
Publié: (2024)
par: Wang, Xiao, et autres
Publié: (2024)
NAB: Neural Adaptive Binning for Sparse-View CT reconstruction
par: Xie, Wangduo, et autres
Publié: (2026)
par: Xie, Wangduo, et autres
Publié: (2026)
When and What: Diffusion-Grounded VideoLLM with Entity Aware Segmentation for Long Video Understanding
par: Fang, Pengcheng, et autres
Publié: (2025)
par: Fang, Pengcheng, et autres
Publié: (2025)
Structural Pruning via Spatial-aware Information Redundancy for Semantic Segmentation
par: Wu, Dongyue, et autres
Publié: (2024)
par: Wu, Dongyue, et autres
Publié: (2024)
Temporal Aware Pruning for Efficient Diffusion-based Video Generation
par: Li, Sheng, et autres
Publié: (2026)
par: Li, Sheng, et autres
Publié: (2026)
RGS-DR: Deferred Reflections and Residual Shading in 2D Gaussian Splatting
par: Kouros, Georgios, et autres
Publié: (2025)
par: Kouros, Georgios, et autres
Publié: (2025)
EchoGen: Generating Visual Echoes in Any Scene via Feed-Forward Subject-Driven Auto-Regressive Model
par: Dong, Ruixiao, et autres
Publié: (2025)
par: Dong, Ruixiao, et autres
Publié: (2025)
Improving Robustness of Deep Learning Based Knee MRI Segmentation: Mixup and Adversarial Domain Adaptation
par: Panfilov, Egor, et autres
Publié: (2019)
par: Panfilov, Egor, et autres
Publié: (2019)
IPFormer-VideoLLM: Enhancing Multi-modal Video Understanding for Multi-shot Scenes
par: Liang, Yujia, et autres
Publié: (2025)
par: Liang, Yujia, et autres
Publié: (2025)
Spec-Gloss Surfels and Normal-Diffuse Priors for Relightable Glossy Objects
par: Kouros, Georgios, et autres
Publié: (2025)
par: Kouros, Georgios, et autres
Publié: (2025)
Documents similaires
-
MI-Pruner: Crossmodal Mutual Information-guided Token Pruner for Efficient MLLMs
par: Li, Jiameng, et autres
Publié: (2026) -
SoftCFG: Uncertainty-guided Stable Guidance for Visual Autoregressive Model
par: Xu, Dongli, et autres
Publié: (2025) -
CARE: Confidence-aware Ratio Estimation for Medical Biomarkers
par: Li, Jiameng, et autres
Publié: (2025) -
Language-Guided Temporal Token Pruning for Efficient VideoLLM Processing
par: Kumar, Yogesh
Publié: (2025) -
Sharp Eyes and Memory for VideoLLMs: Information-Aware Visual Token Pruning for Efficient and Reliable VideoLLM Reasoning
par: Qin, Jialong, et autres
Publié: (2025)