Generative Models in Decision Making: A Survey
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shao, Xinyu, Zhang, Jianping, Wang, Haozhi, Brunswic, Leo Maxime, Zhou, Kaiwen, Dong, Jiqian, Guo, Kaiyang, Chen, Zhitang, Wang, Jun, Hao, Jianye, Li, Xiu, Li, Yinchuan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Theory of Multi-Agent Generative Flow Networks
von: Brunswic, Leo Maxime, et al.
Veröffentlicht: (2025)
von: Brunswic, Leo Maxime, et al.
Veröffentlicht: (2025)
Proximalized Preference Optimization for Diverse Feedback Types: A Decomposed Perspective on DPO
von: Guo, Kaiyang, et al.
Veröffentlicht: (2025)
von: Guo, Kaiyang, et al.
Veröffentlicht: (2025)
A Theory of Non-Acyclic Generative Flow Networks
von: Brunswic, Leo Maxime, et al.
Veröffentlicht: (2023)
von: Brunswic, Leo Maxime, et al.
Veröffentlicht: (2023)
Boosting Cross-problem Generalization in Diffusion-Based Neural Combinatorial Solver via Inference Time Adaptation
von: Lei, Haoyu, et al.
Veröffentlicht: (2025)
von: Lei, Haoyu, et al.
Veröffentlicht: (2025)
Ergodic Generative Flows
von: Brunswic, Leo Maxime, et al.
Veröffentlicht: (2025)
von: Brunswic, Leo Maxime, et al.
Veröffentlicht: (2025)
Alexandrov Theorem for 2+1 flat radiant spacetimes
von: Brunswic, Léo
Veröffentlicht: (2020)
von: Brunswic, Léo
Veröffentlicht: (2020)
On branched coverings of singular $(G,X)$-manifolds
von: Brunswic, Léo
Veröffentlicht: (2020)
von: Brunswic, Léo
Veröffentlicht: (2020)
EmbodiedMAE: A Unified 3D Multi-Modal Representation for Robot Manipulation
von: Dong, Zibin, et al.
Veröffentlicht: (2025)
von: Dong, Zibin, et al.
Veröffentlicht: (2025)
Conditioning Matters: Training Diffusion Policies is Faster Than You Think
von: Dong, Zibin, et al.
Veröffentlicht: (2025)
von: Dong, Zibin, et al.
Veröffentlicht: (2025)
Omni-Thinker: Scaling Multi-Task RL in LLMs with Hybrid Reward and Task Scheduling
von: Li, Derek, et al.
Veröffentlicht: (2025)
von: Li, Derek, et al.
Veröffentlicht: (2025)
More than A Point: Capturing Uncertainty with Adaptive Affordance Heatmaps for Spatial Grounding in Robotic Tasks
von: Shao, Xinyu, et al.
Veröffentlicht: (2025)
von: Shao, Xinyu, et al.
Veröffentlicht: (2025)
JPmHC Dynamical Isometry via Orthogonal Hyper-Connections
von: Sengupta, Biswa, et al.
Veröffentlicht: (2026)
von: Sengupta, Biswa, et al.
Veröffentlicht: (2026)
STAR: Learning Diverse Robot Skill Abstractions through Rotation-Augmented Vector Quantization
von: Li, Hao, et al.
Veröffentlicht: (2025)
von: Li, Hao, et al.
Veröffentlicht: (2025)
Federated Learning via Variational Bayesian Inference: Personalization, Sparsity and Clustering
von: Zhang, Xu, et al.
Veröffentlicht: (2023)
von: Zhang, Xu, et al.
Veröffentlicht: (2023)
Two-Steps Diffusion Policy for Robotic Manipulation via Genetic Denoising
von: Clemente, Mateo, et al.
Veröffentlicht: (2025)
von: Clemente, Mateo, et al.
Veröffentlicht: (2025)
CleanDiffuser: An Easy-to-use Modularized Library for Diffusion Models in Decision Making
von: Dong, Zibin, et al.
Veröffentlicht: (2024)
von: Dong, Zibin, et al.
Veröffentlicht: (2024)
Spatial-Temporal Graph Diffusion Policy with Kinematic Modeling for Bimanual Robotic Manipulation
von: Lv, Qi, et al.
Veröffentlicht: (2025)
von: Lv, Qi, et al.
Veröffentlicht: (2025)
Hista and Numca: Estimate State Value Effectively for LLM Reinforcement Learning
von: Chen, Zizhe, et al.
Veröffentlicht: (2026)
von: Chen, Zizhe, et al.
Veröffentlicht: (2026)
Scaling DRL for Decision Making: A Survey on Data, Network, and Training Budget Strategies
von: Ma, Yi, et al.
Veröffentlicht: (2025)
von: Ma, Yi, et al.
Veröffentlicht: (2025)
ActionCodec: What Makes for Good Action Tokenizers
von: Dong, Zibin, et al.
Veröffentlicht: (2026)
von: Dong, Zibin, et al.
Veröffentlicht: (2026)
Generalized Out-of-Distribution Detection: A Survey
von: Yang, Jingkang, et al.
Veröffentlicht: (2021)
von: Yang, Jingkang, et al.
Veröffentlicht: (2021)
VideoAgent2: Enhancing the LLM-Based Agent System for Long-Form Video Understanding by Uncertainty-Aware CoT
von: Zhi, Zhuo, et al.
Veröffentlicht: (2025)
von: Zhi, Zhuo, et al.
Veröffentlicht: (2025)
Tensor Generalized Approximate Message Passing
von: Li, Yinchuan, et al.
Veröffentlicht: (2025)
von: Li, Yinchuan, et al.
Veröffentlicht: (2025)
$A^2Flow:$ Automating Agentic Workflow Generation via Self-Adaptive Abstraction Operators
von: Zhao, Mingming, et al.
Veröffentlicht: (2025)
von: Zhao, Mingming, et al.
Veröffentlicht: (2025)
Parametric Feature Transfer: One-shot Federated Learning with Foundation Models
von: Beitollahi, Mahdi, et al.
Veröffentlicht: (2024)
von: Beitollahi, Mahdi, et al.
Veröffentlicht: (2024)
The Ride‐Hailing and Pricing Strategy Decision With the Application of Autonomous Vehicles
von: Bin Li, et al.
Veröffentlicht: (2025)
von: Bin Li, et al.
Veröffentlicht: (2025)
Mirage-1: Augmenting and Updating GUI Agent with Hierarchical Multimodal Skills
von: Xie, Yuquan, et al.
Veröffentlicht: (2025)
von: Xie, Yuquan, et al.
Veröffentlicht: (2025)
3D-AffordanceLLM: Harnessing Large Language Models for Open-Vocabulary Affordance Detection in 3D Worlds
von: Chu, Hengshuo, et al.
Veröffentlicht: (2025)
von: Chu, Hengshuo, et al.
Veröffentlicht: (2025)
Less is More: Empowering GUI Agent with Context-Aware Simplification
von: Chen, Gongwei, et al.
Veröffentlicht: (2025)
von: Chen, Gongwei, et al.
Veröffentlicht: (2025)
GUI-explorer: Autonomous Exploration and Mining of Transition-aware Knowledge for GUI Agent
von: Xie, Bin, et al.
Veröffentlicht: (2025)
von: Xie, Bin, et al.
Veröffentlicht: (2025)
Research on Short-Video Platform User Decision-Making via Multimodal Temporal Modeling and Reinforcement Learning
von: Wang, Jinmeiyang, et al.
Veröffentlicht: (2025)
von: Wang, Jinmeiyang, et al.
Veröffentlicht: (2025)
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning
von: Wu, Qingyuan, et al.
Veröffentlicht: (2025)
von: Wu, Qingyuan, et al.
Veröffentlicht: (2025)
Succeed or Learn Slowly: Sample Efficient Off-Policy Reinforcement Learning for Mobile App Control
von: Papoudakis, Georgios, et al.
Veröffentlicht: (2025)
von: Papoudakis, Georgios, et al.
Veröffentlicht: (2025)
Integrating Generative AI into Financial Market Prediction for Improved Decision Making
von: Che, Chang, et al.
Veröffentlicht: (2024)
von: Che, Chang, et al.
Veröffentlicht: (2024)
CAPE: Context-Aware Diffusion Policy Via Proximal Mode Expansion for Collision Avoidance
von: Yang, Rui Heng, et al.
Veröffentlicht: (2025)
von: Yang, Rui Heng, et al.
Veröffentlicht: (2025)
Digital Economy, Industrial Structure Optimization, and Agricultural Green Development Efficiency: A Double Machine Learning Causal Analysis
von: Mi Zhou, et al.
Veröffentlicht: (2025)
von: Mi Zhou, et al.
Veröffentlicht: (2025)
Few-Shot Vision-Language Action-Incremental Policy Learning
von: Song, Mingchen, et al.
Veröffentlicht: (2025)
von: Song, Mingchen, et al.
Veröffentlicht: (2025)
Dynamic Decision-Making under Model Misspecification
von: Dai, Xinyu
Veröffentlicht: (2025)
von: Dai, Xinyu
Veröffentlicht: (2025)
DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents
von: Wang, Taiyi, et al.
Veröffentlicht: (2024)
von: Wang, Taiyi, et al.
Veröffentlicht: (2024)
Diffusion Models for Smarter UAVs: Decision-Making and Modeling
von: Emami, Yousef, et al.
Veröffentlicht: (2025)
von: Emami, Yousef, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
A Theory of Multi-Agent Generative Flow Networks
von: Brunswic, Leo Maxime, et al.
Veröffentlicht: (2025) -
Proximalized Preference Optimization for Diverse Feedback Types: A Decomposed Perspective on DPO
von: Guo, Kaiyang, et al.
Veröffentlicht: (2025) -
A Theory of Non-Acyclic Generative Flow Networks
von: Brunswic, Leo Maxime, et al.
Veröffentlicht: (2023) -
Boosting Cross-problem Generalization in Diffusion-Based Neural Combinatorial Solver via Inference Time Adaptation
von: Lei, Haoyu, et al.
Veröffentlicht: (2025) -
Ergodic Generative Flows
von: Brunswic, Leo Maxime, et al.
Veröffentlicht: (2025)