Towards General Continuous Memory for Vision-Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wu, Wenyi, Song, Zixuan, Zhou, Kun, Shao, Yifei, Hu, Zhiting, Huang, Biwei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Auto-scaling Continuous Memory for GUI Agent
by: Wu, Wenyi, et al.
Published: (2025)
by: Wu, Wenyi, et al.
Published: (2025)
Hybrid Self-evolving Structured Memory for GUI Agents
by: Zhu, Sibo, et al.
Published: (2026)
by: Zhu, Sibo, et al.
Published: (2026)
Planner Matters! An Efficient and Unbalanced Multi-agent Collaboration Framework for Long-horizon Planning
by: Wu, Wenyi, et al.
Published: (2026)
by: Wu, Wenyi, et al.
Published: (2026)
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models
by: Zhao, Zekai, et al.
Published: (2025)
by: Zhao, Zekai, et al.
Published: (2025)
PolarMem: A Training-Free Polarized Latent Graph Memory for Verifiable Vision-Language Models
by: Chen, Zhisheng, et al.
Published: (2026)
by: Chen, Zhisheng, et al.
Published: (2026)
A Fast Kernel-based Conditional Independence test with Application to Causal Discovery
by: Schacht, Oliver, et al.
Published: (2025)
by: Schacht, Oliver, et al.
Published: (2025)
Emerging Synergies in Causality and Deep Generative Models: A Survey
by: Zhou, Guanglin, et al.
Published: (2023)
by: Zhou, Guanglin, et al.
Published: (2023)
Learning Discrete Concepts in Latent Hierarchical Models
by: Kong, Lingjing, et al.
Published: (2024)
by: Kong, Lingjing, et al.
Published: (2024)
Robust Bidirectional Associative Memory via Regularization Inspired by the Subspace Rotation Algorithm
by: Lin, Ci, et al.
Published: (2025)
by: Lin, Ci, et al.
Published: (2025)
Ability Transfer and Recovery via Modularized Parameters Localization
by: Jin, Songyao, et al.
Published: (2026)
by: Jin, Songyao, et al.
Published: (2026)
Continuous Autoregressive Language Models
by: Shao, Chenze, et al.
Published: (2025)
by: Shao, Chenze, et al.
Published: (2025)
General Agentic Planning Through Simulative Reasoning with World Models
by: Deng, Mingkai, et al.
Published: (2025)
by: Deng, Mingkai, et al.
Published: (2025)
Generalized Independent Noise Condition for Estimating Causal Structure with Latent Variables
by: Xie, Feng, et al.
Published: (2023)
by: Xie, Feng, et al.
Published: (2023)
Continuous Reasoning for Vision-Language-Action
by: Wu, Yueh-Hua, et al.
Published: (2026)
by: Wu, Yueh-Hua, et al.
Published: (2026)
Information-Theoretic Dual Memory System for Continual Learning
by: Wu, RunQing, et al.
Published: (2025)
by: Wu, RunQing, et al.
Published: (2025)
Transformer Is Inherently a Causal Learner
by: Wang, Xinyue, et al.
Published: (2026)
by: Wang, Xinyue, et al.
Published: (2026)
Learning Modal-Mixed Chain-of-Thought Reasoning with Latent Embeddings
by: Shao, Yifei, et al.
Published: (2026)
by: Shao, Yifei, et al.
Published: (2026)
TAME: A Trustworthy Test-Time Evolution of Agent Memory with Systematic Benchmarking
by: Cheng, Yu, et al.
Published: (2026)
by: Cheng, Yu, et al.
Published: (2026)
Towards Graph Foundation Models: Training on Knowledge Graphs Enables Transferability to General Graphs
by: Wang, Kai, et al.
Published: (2024)
by: Wang, Kai, et al.
Published: (2024)
KALIE: Fine-Tuning Vision-Language Models for Open-World Manipulation without Robot Data
by: Tang, Grace, et al.
Published: (2024)
by: Tang, Grace, et al.
Published: (2024)
Multimodal Health Risk Prediction System for Chronic Diseases via Vision-Language Fusion and Large Language Models
by: Lu, Dingxin, et al.
Published: (2025)
by: Lu, Dingxin, et al.
Published: (2025)
CPGD: Toward Stable Rule-based Reinforcement Learning for Language Models
by: Liu, Zongkai, et al.
Published: (2025)
by: Liu, Zongkai, et al.
Published: (2025)
Cross-Modality Controlled Molecule Generation with Diffusion Language Model
by: Zhang, Yunzhe, et al.
Published: (2025)
by: Zhang, Yunzhe, et al.
Published: (2025)
Towards Long-window Anchoring in Vision-Language Model Distillation
by: Zhou, Haoyi, et al.
Published: (2025)
by: Zhou, Haoyi, et al.
Published: (2025)
Task-Core Memory Management and Consolidation for Long-term Continual Learning
by: Huai, Tianyu, et al.
Published: (2025)
by: Huai, Tianyu, et al.
Published: (2025)
Ada-Diffuser: Latent-Aware Adaptive Diffusion for Decision-Making
by: Feng, Fan, et al.
Published: (2026)
by: Feng, Fan, et al.
Published: (2026)
Revis: Sparse Latent Steering to Mitigate Object Hallucination in Large Vision-Language Models
by: Wu, Jialin, et al.
Published: (2026)
by: Wu, Jialin, et al.
Published: (2026)
Boosting Efficiency in Task-Agnostic Exploration through Causal Knowledge
by: Yang, Yupei, et al.
Published: (2024)
by: Yang, Yupei, et al.
Published: (2024)
Understanding Retrieval-Augmented Task Adaptation for Vision-Language Models
by: Ming, Yifei, et al.
Published: (2024)
by: Ming, Yifei, et al.
Published: (2024)
Adaptive Defense against Harmful Fine-Tuning for Large Language Models via Bayesian Data Scheduler
by: Hu, Zixuan, et al.
Published: (2025)
by: Hu, Zixuan, et al.
Published: (2025)
CRL-VLA: Continual Vision-Language-Action Learning
by: Zeng, Qixin, et al.
Published: (2026)
by: Zeng, Qixin, et al.
Published: (2026)
ArcMemo: Abstract Reasoning Composition with Lifelong LLM Memory
by: Ho, Matthew, et al.
Published: (2025)
by: Ho, Matthew, et al.
Published: (2025)
ReasoningWeekly: A General Knowledge and Verbal Reasoning Challenge for Large Language Models
by: Wu, Zixuan, et al.
Published: (2025)
by: Wu, Zixuan, et al.
Published: (2025)
CoMemNet: Contrastive Sampling with Memory Replay Network for Continual Traffic Prediction
by: Wu, Mei, et al.
Published: (2026)
by: Wu, Mei, et al.
Published: (2026)
FOREVER: Forgetting Curve-Inspired Memory Replay for Language Model Continual Learning
by: Feng, Yujie, et al.
Published: (2026)
by: Feng, Yujie, et al.
Published: (2026)
Memory-Statistics Tradeoff in Continual Learning with Structural Regularization
by: Li, Haoran, et al.
Published: (2025)
by: Li, Haoran, et al.
Published: (2025)
Towards Automated Semantic Interpretability in Reinforcement Learning via Vision-Language Models
by: Li, Zhaoxin, et al.
Published: (2025)
by: Li, Zhaoxin, et al.
Published: (2025)
VisMem: Latent Vision Memory Unlocks Potential of Vision-Language Models
by: Yu, Xinlei, et al.
Published: (2025)
by: Yu, Xinlei, et al.
Published: (2025)
Toward Universal and Transferable Jailbreak Attacks on Vision-Language Models
by: Cui, Kaiyuan, et al.
Published: (2026)
by: Cui, Kaiyuan, et al.
Published: (2026)
RPO: Fine-Tuning Visual Generative Models via Rich Vision-Language Preferences
by: Zhao, Hanyang, et al.
Published: (2025)
by: Zhao, Hanyang, et al.
Published: (2025)
Similar Items
-
Auto-scaling Continuous Memory for GUI Agent
by: Wu, Wenyi, et al.
Published: (2025) -
Hybrid Self-evolving Structured Memory for GUI Agents
by: Zhu, Sibo, et al.
Published: (2026) -
Planner Matters! An Efficient and Unbalanced Multi-agent Collaboration Framework for Long-horizon Planning
by: Wu, Wenyi, et al.
Published: (2026) -
Activation Control for Efficiently Eliciting Long Chain-of-thought Ability of Language Models
by: Zhao, Zekai, et al.
Published: (2025) -
PolarMem: A Training-Free Polarized Latent Graph Memory for Verifiable Vision-Language Models
by: Chen, Zhisheng, et al.
Published: (2026)