VLM Q-Learning: Aligning Vision-Language Models for Interactive Decision-Making
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Grigsby, Jake, Zhu, Yuke, Ryoo, Michael, Niebles, Juan Carlos |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
AMAGO: Scalable In-Context Reinforcement Learning for Adaptive Agents
von: Grigsby, Jake, et al.
Veröffentlicht: (2023)
von: Grigsby, Jake, et al.
Veröffentlicht: (2023)
Human-Level Competitive Pokémon via Scalable Offline Reinforcement Learning with Transformers
von: Grigsby, Jake, et al.
Veröffentlicht: (2025)
von: Grigsby, Jake, et al.
Veröffentlicht: (2025)
AMAGO-2: Breaking the Multi-Task Barrier in Meta-Reinforcement Learning with Transformers
von: Grigsby, Jake, et al.
Veröffentlicht: (2024)
von: Grigsby, Jake, et al.
Veröffentlicht: (2024)
Aligning Learning and Endogenous Decision-Making
von: Cristian, Rares, et al.
Veröffentlicht: (2025)
von: Cristian, Rares, et al.
Veröffentlicht: (2025)
LAM SIMULATOR: Advancing Data Generation for Large Action Model Training via Online Exploration and Trajectory Feedback
von: Hoang, Thai, et al.
Veröffentlicht: (2025)
von: Hoang, Thai, et al.
Veröffentlicht: (2025)
HandsOnVLM: Vision-Language Models for Hand-Object Interaction Prediction
von: Bao, Chen, et al.
Veröffentlicht: (2024)
von: Bao, Chen, et al.
Veröffentlicht: (2024)
Pretrained Vision-Language-Action Models are Surprisingly Resistant to Forgetting in Continual Learning
von: Liu, Huihan, et al.
Veröffentlicht: (2026)
von: Liu, Huihan, et al.
Veröffentlicht: (2026)
SDM-Q: Cost-Aware Staged Decision-Making for Multi-Omics Classification with Deep Q-Learning
von: Mu, Nan, et al.
Veröffentlicht: (2026)
von: Mu, Nan, et al.
Veröffentlicht: (2026)
Efficient Sequential Decision Making with Large Language Models
von: Chen, Dingyang, et al.
Veröffentlicht: (2024)
von: Chen, Dingyang, et al.
Veröffentlicht: (2024)
Model-Based Runtime Monitoring with Interactive Imitation Learning
von: Liu, Huihan, et al.
Veröffentlicht: (2023)
von: Liu, Huihan, et al.
Veröffentlicht: (2023)
DriVLM: Domain Adaptation of Vision-Language Models in Autonomous Driving
von: Zheng, Xuran, et al.
Veröffentlicht: (2025)
von: Zheng, Xuran, et al.
Veröffentlicht: (2025)
GraphVLM: Benchmarking Vision Language Models for Multimodal Graph Learning
von: Liu, Jiajin, et al.
Veröffentlicht: (2026)
von: Liu, Jiajin, et al.
Veröffentlicht: (2026)
AsymVLM: Asymmetric Token Pruning for Efficient Vision-Language Model Inference
von: Feng, Yilin, et al.
Veröffentlicht: (2026)
von: Feng, Yilin, et al.
Veröffentlicht: (2026)
Strefer: Empowering Video LLMs with Space-Time Referring and Reasoning via Synthetic Instruction Data
von: Zhou, Honglu, et al.
Veröffentlicht: (2025)
von: Zhou, Honglu, et al.
Veröffentlicht: (2025)
Gen-DFL: Decision-Focused Generative Learning for Robust Decision Making
von: Wang, Prince Zizhuang, et al.
Veröffentlicht: (2025)
von: Wang, Prince Zizhuang, et al.
Veröffentlicht: (2025)
RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
von: Wang, Yufei, et al.
Veröffentlicht: (2024)
von: Wang, Yufei, et al.
Veröffentlicht: (2024)
On the Importance of Uncertainty in Decision-Making with Large Language Models
von: Felicioni, Nicolò, et al.
Veröffentlicht: (2024)
von: Felicioni, Nicolò, et al.
Veröffentlicht: (2024)
FastVLM: Efficient Vision Encoding for Vision Language Models
von: Vasu, Pavan Kumar Anasosalu, et al.
Veröffentlicht: (2024)
von: Vasu, Pavan Kumar Anasosalu, et al.
Veröffentlicht: (2024)
DRESS: Instructing Large Vision-Language Models to Align and Interact with Humans via Natural Language Feedback
von: Chen, Yangyi, et al.
Veröffentlicht: (2023)
von: Chen, Yangyi, et al.
Veröffentlicht: (2023)
DocVLM: Make Your VLM an Efficient Reader
von: Nacson, Mor Shpigel, et al.
Veröffentlicht: (2024)
von: Nacson, Mor Shpigel, et al.
Veröffentlicht: (2024)
InvestAlign: Overcoming Data Scarcity in Aligning Large Language Models with Investor Decision-Making Processes under Herd Behavior
von: Wang, Huisheng, et al.
Veröffentlicht: (2025)
von: Wang, Huisheng, et al.
Veröffentlicht: (2025)
The Relative Value of Prediction in Algorithmic Decision Making
von: Perdomo, Juan Carlos
Veröffentlicht: (2023)
von: Perdomo, Juan Carlos
Veröffentlicht: (2023)
FastVLM: Self-Speculative Decoding for Fast Vision-Language Model Inference
von: Bajpai, Divya Jyoti, et al.
Veröffentlicht: (2025)
von: Bajpai, Divya Jyoti, et al.
Veröffentlicht: (2025)
Geometry of Decision Making in Language Models
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
von: Joshi, Abhinav, et al.
Veröffentlicht: (2025)
Human-Aligned Calibration for AI-Assisted Decision Making
von: Benz, Nina L. Corvelo, et al.
Veröffentlicht: (2023)
von: Benz, Nina L. Corvelo, et al.
Veröffentlicht: (2023)
Aligning Diffusion Behaviors with Q-functions for Efficient Continuous Control
von: Chen, Huayu, et al.
Veröffentlicht: (2024)
von: Chen, Huayu, et al.
Veröffentlicht: (2024)
Unifying Specialized Visual Encoders for Video Language Models
von: Chung, Jihoon, et al.
Veröffentlicht: (2025)
von: Chung, Jihoon, et al.
Veröffentlicht: (2025)
CF-VLM:CounterFactual Vision-Language Fine-tuning
von: Zhang, Jusheng, et al.
Veröffentlicht: (2025)
von: Zhang, Jusheng, et al.
Veröffentlicht: (2025)
GROW: Aligning GRPO with State-Action Modeling for Open-World VLM Agents
von: Wu, Xiongbin, et al.
Veröffentlicht: (2026)
von: Wu, Xiongbin, et al.
Veröffentlicht: (2026)
Memory-Driven Self-Improvement for Decision Making with Large Language Models
von: Yan, Xue, et al.
Veröffentlicht: (2025)
von: Yan, Xue, et al.
Veröffentlicht: (2025)
Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
von: Zhai, Yuexiang, et al.
Veröffentlicht: (2024)
von: Zhai, Yuexiang, et al.
Veröffentlicht: (2024)
Can Vision Language Models Learn Intuitive Physics from Interaction?
von: Buschoff, Luca M. Schulze, et al.
Veröffentlicht: (2026)
von: Buschoff, Luca M. Schulze, et al.
Veröffentlicht: (2026)
Vision-Language-Action Models for Robotics: A Review Towards Real-World Applications
von: Kawaharazuka, Kento, et al.
Veröffentlicht: (2025)
von: Kawaharazuka, Kento, et al.
Veröffentlicht: (2025)
Aligning Large Vision-Language Models by Deep Reinforcement Learning and Direct Preference Optimization
von: Nguyen, Thanh Thi, et al.
Veröffentlicht: (2025)
von: Nguyen, Thanh Thi, et al.
Veröffentlicht: (2025)
Causal-aware Large Language Models: Enhancing Decision-Making Through Learning, Adapting and Acting
von: Chen, Wei, et al.
Veröffentlicht: (2025)
von: Chen, Wei, et al.
Veröffentlicht: (2025)
Reinforcement Learning for Aligning Large Language Models Agents with Interactive Environments: Quantifying and Mitigating Prompt Overfitting
von: Aissi, Mohamed Salim, et al.
Veröffentlicht: (2024)
von: Aissi, Mohamed Salim, et al.
Veröffentlicht: (2024)
Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making
von: Chen, Ruoyu, et al.
Veröffentlicht: (2026)
von: Chen, Ruoyu, et al.
Veröffentlicht: (2026)
Continual Learning in Vision-Language Models via Aligned Model Merging
von: Sokar, Ghada, et al.
Veröffentlicht: (2025)
von: Sokar, Ghada, et al.
Veröffentlicht: (2025)
Aligning Language Models from User Interactions
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2026)
von: Buening, Thomas Kleine, et al.
Veröffentlicht: (2026)
GeoFlowVLM: Geometry-Aware Joint Uncertainty for Frozen Vision-Language Embedding
von: Nautiyal, Mayank, et al.
Veröffentlicht: (2026)
von: Nautiyal, Mayank, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
AMAGO: Scalable In-Context Reinforcement Learning for Adaptive Agents
von: Grigsby, Jake, et al.
Veröffentlicht: (2023) -
Human-Level Competitive Pokémon via Scalable Offline Reinforcement Learning with Transformers
von: Grigsby, Jake, et al.
Veröffentlicht: (2025) -
AMAGO-2: Breaking the Multi-Task Barrier in Meta-Reinforcement Learning with Transformers
von: Grigsby, Jake, et al.
Veröffentlicht: (2024) -
Aligning Learning and Endogenous Decision-Making
von: Cristian, Rares, et al.
Veröffentlicht: (2025) -
LAM SIMULATOR: Advancing Data Generation for Large Action Model Training via Online Exploration and Trajectory Feedback
von: Hoang, Thai, et al.
Veröffentlicht: (2025)