Jarvis: Towards Personalized AI Assistant via Personal KV-Cache Retrieval
Fuente:
arXiv
Salvato in:
| Autori principali: | Xu, Binxiao, Feng, Junyu, Lu, Shaolin, Luo, Yulin, Yan, Shilin, Liang, Hao, Lu, Ming, Zhang, Wentao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
M2A: Multimodal Memory Agent with Dual-Layer Hybrid Memory for Long-Term Personalized Interactions
di: Feng, Junyu, et al.
Pubblicazione: (2026)
di: Feng, Junyu, et al.
Pubblicazione: (2026)
OpenJarvis: Personal AI, On Personal Devices
di: Saad-Falcon, Jon, et al.
Pubblicazione: (2026)
di: Saad-Falcon, Jon, et al.
Pubblicazione: (2026)
AD-MIR: Bridging the Gap from Perception to Persuasion in Advertising Video Understanding via Structured Reasoning
di: Xu, Binxiao, et al.
Pubblicazione: (2026)
di: Xu, Binxiao, et al.
Pubblicazione: (2026)
ProphetKV: User-Query-Driven Selective Recomputation for Efficient KV Cache Reuse in Retrieval-Augmented Generation
di: Wang, Shihao, et al.
Pubblicazione: (2026)
di: Wang, Shihao, et al.
Pubblicazione: (2026)
MC-LLaVA: Multi-Concept Personalized Vision-Language Model
di: An, Ruichuan, et al.
Pubblicazione: (2025)
di: An, Ruichuan, et al.
Pubblicazione: (2025)
MC-LLaVA: Multi-Concept Personalized Vision-Language Model
di: An, Ruichuan, et al.
Pubblicazione: (2024)
di: An, Ruichuan, et al.
Pubblicazione: (2024)
EgoSelf: From Memory to Personalized Egocentric Assistant
di: Wang, Yanshuo, et al.
Pubblicazione: (2026)
di: Wang, Yanshuo, et al.
Pubblicazione: (2026)
Towards Threshold-Free KV Cache Pruning
di: Ni, Xuanfan, et al.
Pubblicazione: (2025)
di: Ni, Xuanfan, et al.
Pubblicazione: (2025)
Small-Large Collaboration: Training-efficient Concept Personalization for Large VLM using a Meta Personalized Small VLM
di: Yang, Sihan, et al.
Pubblicazione: (2025)
di: Yang, Sihan, et al.
Pubblicazione: (2025)
LouisKV: Efficient KV Cache Retrieval for Long Input-Output Sequences
di: Wu, Wenbo, et al.
Pubblicazione: (2025)
di: Wu, Wenbo, et al.
Pubblicazione: (2025)
Aligning VLM Assistants with Personalized Situated Cognition
di: Li, Yongqi, et al.
Pubblicazione: (2025)
di: Li, Yongqi, et al.
Pubblicazione: (2025)
TravelAgent: An AI Assistant for Personalized Travel Planning
di: Chen, Aili, et al.
Pubblicazione: (2024)
di: Chen, Aili, et al.
Pubblicazione: (2024)
FreeKV: Boosting KV Cache Retrieval for Efficient LLM Inference
di: Liu, Guangda, et al.
Pubblicazione: (2025)
di: Liu, Guangda, et al.
Pubblicazione: (2025)
DeltaKV: Residual-Based KV Cache Compression via Long-Range Similarity
di: Hao, Jitai, et al.
Pubblicazione: (2026)
di: Hao, Jitai, et al.
Pubblicazione: (2026)
Concept-as-Tree: A Controllable Synthetic Data Framework Makes Stronger Personalized VLMs
di: An, Ruichuan, et al.
Pubblicazione: (2025)
di: An, Ruichuan, et al.
Pubblicazione: (2025)
HiMeS: Hippocampus-inspired Memory System for Personalized AI Assistants
di: Li, Hailong, et al.
Pubblicazione: (2026)
di: Li, Hailong, et al.
Pubblicazione: (2026)
Personality Modeling for Persuasion of Misinformation using AI Agent
di: Lou, Qianmin, et al.
Pubblicazione: (2025)
di: Lou, Qianmin, et al.
Pubblicazione: (2025)
G-KV: Decoding-Time KV Cache Eviction with Global Attention
di: Liao, Mengqi, et al.
Pubblicazione: (2025)
di: Liao, Mengqi, et al.
Pubblicazione: (2025)
HeteroCache: A Dynamic Retrieval Approach to Heterogeneous KV Cache Compression for Long-Context LLM Inference
di: Shi, Zhiyuan, et al.
Pubblicazione: (2026)
di: Shi, Zhiyuan, et al.
Pubblicazione: (2026)
Towards Next-Generation Recommender Systems: A Benchmark for Personalized Recommendation Assistant with LLMs
di: Huang, Jiani, et al.
Pubblicazione: (2025)
di: Huang, Jiani, et al.
Pubblicazione: (2025)
LoopGuard: Breaking Self-Reinforcing Attention Loops via Dynamic KV Cache Intervention
di: Xu, Dongjie, et al.
Pubblicazione: (2026)
di: Xu, Dongjie, et al.
Pubblicazione: (2026)
Towards Ethical Personal AI Applications: Practical Considerations for AI Assistants with Long-Term Memory
di: Lee, Eunhae
Pubblicazione: (2024)
di: Lee, Eunhae
Pubblicazione: (2024)
PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling
di: Cai, Zefan, et al.
Pubblicazione: (2024)
di: Cai, Zefan, et al.
Pubblicazione: (2024)
CoKV: Optimizing KV Cache Allocation via Cooperative Game
di: Sun, Qiheng, et al.
Pubblicazione: (2025)
di: Sun, Qiheng, et al.
Pubblicazione: (2025)
Lookahead Q-Cache: Achieving More Consistent KV Cache Eviction via Pseudo Query
di: Wang, Yixuan, et al.
Pubblicazione: (2025)
di: Wang, Yixuan, et al.
Pubblicazione: (2025)
Hierarchical Adaptive Eviction for KV Cache Management in Multimodal Language Models
di: Ma, Xindian, et al.
Pubblicazione: (2026)
di: Ma, Xindian, et al.
Pubblicazione: (2026)
PaRT: Enhancing Proactive Social Chatbots with Personalized Real-Time Retrieval
di: Niu, Zihan, et al.
Pubblicazione: (2025)
di: Niu, Zihan, et al.
Pubblicazione: (2025)
AI-driven Personalized Privacy Assistants: a Systematic Literature Review
di: Morel, Victor, et al.
Pubblicazione: (2025)
di: Morel, Victor, et al.
Pubblicazione: (2025)
SQuat: Subspace-orthogonal KV Cache Quantization
di: Wang, Hao, et al.
Pubblicazione: (2025)
di: Wang, Hao, et al.
Pubblicazione: (2025)
ReCalKV: Low-Rank KV Cache Compression via Head Reordering and Offline Calibration
di: Yan, Xianglong, et al.
Pubblicazione: (2025)
di: Yan, Xianglong, et al.
Pubblicazione: (2025)
LagKV: Lag-Relative Information of the KV Cache Tells Which Tokens Are Important
di: Liang, Manlai, et al.
Pubblicazione: (2025)
di: Liang, Manlai, et al.
Pubblicazione: (2025)
From Assistant to Double Agent: Formalizing and Benchmarking Attacks on OpenClaw for Personalized Local AI Agent
di: Wang, Yuhang, et al.
Pubblicazione: (2026)
di: Wang, Yuhang, et al.
Pubblicazione: (2026)
GOD model: Privacy Preserved AI School for Personal Assistant
di: PIN AI Team, et al.
Pubblicazione: (2025)
di: PIN AI Team, et al.
Pubblicazione: (2025)
MPFormer: Adaptive Framework for Industrial Multi-Task Personalized Sequential Retriever
di: Sun, Yijia, et al.
Pubblicazione: (2025)
di: Sun, Yijia, et al.
Pubblicazione: (2025)
R-KV: Redundancy-aware KV Cache Compression for Reasoning Models
di: Cai, Zefan, et al.
Pubblicazione: (2025)
di: Cai, Zefan, et al.
Pubblicazione: (2025)
Accelerating LLM Inference Throughput via Asynchronous KV Cache Prefetching
di: Dong, Yanhao, et al.
Pubblicazione: (2025)
di: Dong, Yanhao, et al.
Pubblicazione: (2025)
KnowU-Bench: Towards Interactive, Proactive, and Personalized Mobile Agent Evaluation
di: Chen, Tongbo, et al.
Pubblicazione: (2026)
di: Chen, Tongbo, et al.
Pubblicazione: (2026)
Meta-Soft: Leveraging Composable Meta-Tokens for Context-Preserving KV Cache Compression
di: Luo, Wei, et al.
Pubblicazione: (2026)
di: Luo, Wei, et al.
Pubblicazione: (2026)
RelayCaching: Accelerating LLM Collaboration via Decoding KV Cache Reuse
di: Geng, Yingsheng, et al.
Pubblicazione: (2026)
di: Geng, Yingsheng, et al.
Pubblicazione: (2026)
CachePrune: Teaching LLMs What Not to Follow via KV-Cache Editing
di: Wang, Rui, et al.
Pubblicazione: (2025)
di: Wang, Rui, et al.
Pubblicazione: (2025)
Documenti analoghi
-
M2A: Multimodal Memory Agent with Dual-Layer Hybrid Memory for Long-Term Personalized Interactions
di: Feng, Junyu, et al.
Pubblicazione: (2026) -
OpenJarvis: Personal AI, On Personal Devices
di: Saad-Falcon, Jon, et al.
Pubblicazione: (2026) -
AD-MIR: Bridging the Gap from Perception to Persuasion in Advertising Video Understanding via Structured Reasoning
di: Xu, Binxiao, et al.
Pubblicazione: (2026) -
ProphetKV: User-Query-Driven Selective Recomputation for Efficient KV Cache Reuse in Retrieval-Augmented Generation
di: Wang, Shihao, et al.
Pubblicazione: (2026) -
MC-LLaVA: Multi-Concept Personalized Vision-Language Model
di: An, Ruichuan, et al.
Pubblicazione: (2025)