CHAI: CacHe Attention Inference for text2video
Fuente:
arXiv
Salvato in:
| Autori principali: | Cherian, Joel Mathew, Bharadwaj, Ashutosh Muralidhara, Gupta, Vima, Iyer, Anand Padmanabha |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Lynx: Enabling Efficient MoE Inference through Dynamic Batch-Aware Expert Selection
di: Gupta, Vima, et al.
Pubblicazione: (2024)
di: Gupta, Vima, et al.
Pubblicazione: (2024)
CLoRA: Parameter-Efficient Continual Learning with Low-Rank Adaptation
di: Muralidhara, Shishir, et al.
Pubblicazione: (2025)
di: Muralidhara, Shishir, et al.
Pubblicazione: (2025)
High-resolution Multi-spectral Image Guided DEM Super-resolution using Sinkhorn Regularized Adversarial Network
di: Paul, Subhajit, et al.
Pubblicazione: (2023)
di: Paul, Subhajit, et al.
Pubblicazione: (2023)
Fake It To Make It: Virtual Multiviews to Enhance Monocular Indoor Semantic Scene Completion
di: Selvakumar, Anith, et al.
Pubblicazione: (2025)
di: Selvakumar, Anith, et al.
Pubblicazione: (2025)
Neuron Incidence Redistribution for Fairness in Medical Image Classification
di: Shoby, Abin, et al.
Pubblicazione: (2026)
di: Shoby, Abin, et al.
Pubblicazione: (2026)
MI CAM: Mutual Information Weighted Activation Mapping for Causal Visual Explanations of Convolutional Neural Networks
di: Iyer, Ram S, et al.
Pubblicazione: (2025)
di: Iyer, Ram S, et al.
Pubblicazione: (2025)
Edge Attention Module for Object Classification
di: Roy, Santanu, et al.
Pubblicazione: (2025)
di: Roy, Santanu, et al.
Pubblicazione: (2025)
ASAP: Attention-Shift-Aware Pruning for Efficient LVLM Inference
di: Pathak, Surendra, et al.
Pubblicazione: (2026)
di: Pathak, Surendra, et al.
Pubblicazione: (2026)
EgoCHARM: Resource-Efficient Hierarchical Activity Recognition using an Egocentric IMU Sensor
di: Padmanabha, Akhil, et al.
Pubblicazione: (2025)
di: Padmanabha, Akhil, et al.
Pubblicazione: (2025)
EfficientSign: An Attention-Enhanced Lightweight Architecture for Indian Sign Language Recognition
di: Gupta, Rishabh, et al.
Pubblicazione: (2026)
di: Gupta, Rishabh, et al.
Pubblicazione: (2026)
Resolving Spatio-Temporal Entanglement in Video Prediction via Multi-Modal Attention
di: Gupta, Shreyam, et al.
Pubblicazione: (2025)
di: Gupta, Shreyam, et al.
Pubblicazione: (2025)
Improving text-conditioned latent diffusion for cancer pathology
di: Rao, Aakash Madhav, et al.
Pubblicazione: (2024)
di: Rao, Aakash Madhav, et al.
Pubblicazione: (2024)
SHaSaM: Submodular Hard Sample Mining for Fair Facial Attribute Recognition
di: Majee, Anay, et al.
Pubblicazione: (2026)
di: Majee, Anay, et al.
Pubblicazione: (2026)
Improving Image Clustering with Artifacts Attenuation via Inference-Time Attention Engineering
di: Nakamura, Kazumoto, et al.
Pubblicazione: (2024)
di: Nakamura, Kazumoto, et al.
Pubblicazione: (2024)
Recurrent Attention-based Token Selection for Efficient Streaming Video-LLMs
di: Dorovatas, Vaggelis, et al.
Pubblicazione: (2025)
di: Dorovatas, Vaggelis, et al.
Pubblicazione: (2025)
Attention-aware Inference Optimizations for Large Vision-Language Models with Memory-efficient Decoding
di: Ilhan, Fatih, et al.
Pubblicazione: (2026)
di: Ilhan, Fatih, et al.
Pubblicazione: (2026)
VROOM - Visual Reconstruction over Onboard Multiview
di: Yadav, Yajat, et al.
Pubblicazione: (2025)
di: Yadav, Yajat, et al.
Pubblicazione: (2025)
Robust Mitigation of Age-Dependent Confounding Effects via Sample-Difficulty Decorrelation
di: Kurian, Nikhil Cherian, et al.
Pubblicazione: (2026)
di: Kurian, Nikhil Cherian, et al.
Pubblicazione: (2026)
STAR: Stage-Wise Attention-Guided Token Reduction for Efficient Large Vision-Language Models Inference
di: Guo, Yichen, et al.
Pubblicazione: (2025)
di: Guo, Yichen, et al.
Pubblicazione: (2025)
Looking Beyond the Known: Towards a Data Discovery Guided Open-World Object Detection
di: Majee, Anay, et al.
Pubblicazione: (2025)
di: Majee, Anay, et al.
Pubblicazione: (2025)
LEDA: Log-Euclidean Diffeomorphism Autoencoder for Efficient Statistical Analysis of Diffeomorphisms
di: Iyer, Krithika, et al.
Pubblicazione: (2024)
di: Iyer, Krithika, et al.
Pubblicazione: (2024)
TurboEdit: Instant text-based image editing
di: Wu, Zongze, et al.
Pubblicazione: (2024)
di: Wu, Zongze, et al.
Pubblicazione: (2024)
TechING: Towards Real World Technical Image Understanding via VLMs
di: Nadeem, Tafazzul, et al.
Pubblicazione: (2026)
di: Nadeem, Tafazzul, et al.
Pubblicazione: (2026)
The Silent Brush: Evaluating Artistic Style Leakage in AI Art Generation
di: Joshi, Ninad, et al.
Pubblicazione: (2026)
di: Joshi, Ninad, et al.
Pubblicazione: (2026)
Towards a text-based quantitative and explainable histopathology image analysis
di: Nguyen, Anh Tien, et al.
Pubblicazione: (2024)
di: Nguyen, Anh Tien, et al.
Pubblicazione: (2024)
Hier-COS: Making Deep Features Hierarchy-aware via Composition of Orthogonal Subspaces
di: Sani, Depanshu, et al.
Pubblicazione: (2025)
di: Sani, Depanshu, et al.
Pubblicazione: (2025)
Oh That Looks Familiar: A Novel Similarity Measure for Spreadsheet Template Discovery
di: Krishnakumar, Anand, et al.
Pubblicazione: (2025)
di: Krishnakumar, Anand, et al.
Pubblicazione: (2025)
LASERS: LAtent Space Encoding for Representations with Sparsity for Generative Modeling
di: Li, Xin, et al.
Pubblicazione: (2024)
di: Li, Xin, et al.
Pubblicazione: (2024)
Synthesizing Proton-Density Fat Fraction and $R_2^*$ from 2-point Dixon MRI with Generative Machine Learning
di: Anand, Suma, et al.
Pubblicazione: (2024)
di: Anand, Suma, et al.
Pubblicazione: (2024)
QMViT: A Mushroom is worth 16x16 Words
di: Dutta, Siddhant, et al.
Pubblicazione: (2024)
di: Dutta, Siddhant, et al.
Pubblicazione: (2024)
SCoRe: Submodular Combinatorial Representation Learning
di: Majee, Anay, et al.
Pubblicazione: (2023)
di: Majee, Anay, et al.
Pubblicazione: (2023)
FAR: Function-preserving Attention Replacement for IMC-friendly Inference
di: Ren, Yuxin, et al.
Pubblicazione: (2025)
di: Ren, Yuxin, et al.
Pubblicazione: (2025)
Universal representations:The missing link between faces, text, planktons, and cat breeds
di: Bilen, Hakan, et al.
Pubblicazione: (2017)
di: Bilen, Hakan, et al.
Pubblicazione: (2017)
DiffuSAM: Diffusion Guided Zero-Shot Object Grounding for Remote Sensing Imagery
di: Sethi, Geet, et al.
Pubblicazione: (2026)
di: Sethi, Geet, et al.
Pubblicazione: (2026)
No Training Wheels: Steering Vectors for Bias Correction at Inference Time
di: Gupta, Aviral, et al.
Pubblicazione: (2025)
di: Gupta, Aviral, et al.
Pubblicazione: (2025)
SegHeD+: Segmentation of Heterogeneous Data for Multiple Sclerosis Lesions with Anatomical Constraints and Lesion-aware Augmentation
di: Basaran, Berke Doga, et al.
Pubblicazione: (2024)
di: Basaran, Berke Doga, et al.
Pubblicazione: (2024)
$R_\text{dm}$: Re-conceptualizing Distribution Matching as a Reward for Diffusion Distillation
di: Fan, Linqian, et al.
Pubblicazione: (2026)
di: Fan, Linqian, et al.
Pubblicazione: (2026)
AI-Based Copyright Detection Of An Image In a Video Using Degree Of Similarity And Image Hashing
di: Ashutosh, et al.
Pubblicazione: (2024)
di: Ashutosh, et al.
Pubblicazione: (2024)
LLMPhy: Parameter-Identifiable Physical Reasoning Combining Large Language Models and Physics Engines
di: Cherian, Anoop, et al.
Pubblicazione: (2024)
di: Cherian, Anoop, et al.
Pubblicazione: (2024)
Foul prediction with estimated poses from soccer broadcast video
di: Fang, Jiale, et al.
Pubblicazione: (2024)
di: Fang, Jiale, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Lynx: Enabling Efficient MoE Inference through Dynamic Batch-Aware Expert Selection
di: Gupta, Vima, et al.
Pubblicazione: (2024) -
CLoRA: Parameter-Efficient Continual Learning with Low-Rank Adaptation
di: Muralidhara, Shishir, et al.
Pubblicazione: (2025) -
High-resolution Multi-spectral Image Guided DEM Super-resolution using Sinkhorn Regularized Adversarial Network
di: Paul, Subhajit, et al.
Pubblicazione: (2023) -
Fake It To Make It: Virtual Multiviews to Enhance Monocular Indoor Semantic Scene Completion
di: Selvakumar, Anith, et al.
Pubblicazione: (2025) -
Neuron Incidence Redistribution for Fairness in Medical Image Classification
di: Shoby, Abin, et al.
Pubblicazione: (2026)