Dr. Strategy: Model-Based Generalist Agents with Strategic Dreaming
Fuente:
arXiv
Saved in:
| Main Authors: | Hamed, Hany, Kim, Subin, Kim, Dongyeong, Yoon, Jaesik, Ahn, Sungjin |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Spatially-Aware Transformer for Embodied Agents
by: Cho, Junmo, et al.
Published: (2024)
by: Cho, Junmo, et al.
Published: (2024)
Inference-Time Scaling in Diffusion Models through Iterative Partial Refinement
by: Kang, Taegu, et al.
Published: (2026)
by: Kang, Taegu, et al.
Published: (2026)
Compositional Monte Carlo Tree Diffusion for Extendable Planning
by: Yoon, Jaesik, et al.
Published: (2025)
by: Yoon, Jaesik, et al.
Published: (2025)
TransDreamer: Reinforcement Learning with Transformer World Models
by: Chen, Chang, et al.
Published: (2022)
by: Chen, Chang, et al.
Published: (2022)
Loopholing Discrete Diffusion: Deterministic Bypass of the Sampling Wall
by: Jo, Mingyu, et al.
Published: (2025)
by: Jo, Mingyu, et al.
Published: (2025)
Monte Carlo Tree Diffusion for System 2 Planning
by: Yoon, Jaesik, et al.
Published: (2025)
by: Yoon, Jaesik, et al.
Published: (2025)
An Investigation into Pre-Training Object-Centric Representations for Reinforcement Learning
by: Yoon, Jaesik, et al.
Published: (2023)
by: Yoon, Jaesik, et al.
Published: (2023)
Adaptive Inference-Time Scaling via Cyclic Diffusion Search
by: Lee, Gyubin, et al.
Published: (2025)
by: Lee, Gyubin, et al.
Published: (2025)
Extendable Planning via Multiscale Diffusion
by: Chen, Chang, et al.
Published: (2025)
by: Chen, Chang, et al.
Published: (2025)
MrSteve: Instruction-Following Agents in Minecraft with What-Where-When Memory
by: Park, Junyeong, et al.
Published: (2024)
by: Park, Junyeong, et al.
Published: (2024)
Budget-Aware Sequential Brick Assembly with Efficient Constraint Satisfaction
by: Ahn, Seokjun, et al.
Published: (2022)
by: Ahn, Seokjun, et al.
Published: (2022)
FlowerFormer: Empowering Neural Architecture Encoding using a Flow-aware Graph Transformer
by: Hwang, Dongyeong, et al.
Published: (2024)
by: Hwang, Dongyeong, et al.
Published: (2024)
Neural Language of Thought Models
by: Wu, Yi-Fu, et al.
Published: (2024)
by: Wu, Yi-Fu, et al.
Published: (2024)
Probing the Impact of Scale on Data-Efficient, Generalist Transformer World Models for Atari
by: Kim, Jooyeon
Published: (2026)
by: Kim, Jooyeon
Published: (2026)
Stabilizing Consistency Training: A Flow Map Analysis and Self-Distillation
by: Kim, Youngjoong, et al.
Published: (2026)
by: Kim, Youngjoong, et al.
Published: (2026)
Towards Diverse Perspective Learning with Selection over Multiple Temporal Poolings
by: Seong, Jihyeon, et al.
Published: (2024)
by: Seong, Jihyeon, et al.
Published: (2024)
Capsule Neural Networks as Noise Stabilizer for Time Series Data
by: Kim, Soyeon, et al.
Published: (2024)
by: Kim, Soyeon, et al.
Published: (2024)
Fast Monte Carlo Tree Diffusion: 100x Speedup via Parallel Sparse Planning
by: Yoon, Jaesik, et al.
Published: (2025)
by: Yoon, Jaesik, et al.
Published: (2025)
Latent Veracity Inference for Identifying Errors in Stepwise Reasoning
by: Kim, Minsu, et al.
Published: (2025)
by: Kim, Minsu, et al.
Published: (2025)
PnPXAI: A Universal XAI Framework Providing Automatic Explanations Across Diverse Modalities and Models
by: Kim, Seongun, et al.
Published: (2025)
by: Kim, Seongun, et al.
Published: (2025)
NitroGen: An Open Foundation Model for Generalist Gaming Agents
by: Magne, Loïc, et al.
Published: (2026)
by: Magne, Loïc, et al.
Published: (2026)
Identifying the Source of Generation for Large Language Models
by: Park, Bumjin, et al.
Published: (2024)
by: Park, Bumjin, et al.
Published: (2024)
GOAL: A Generalist Combinatorial Optimization Agent Learner
by: Drakulic, Darko, et al.
Published: (2024)
by: Drakulic, Darko, et al.
Published: (2024)
DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos
by: Gao, Shenyuan, et al.
Published: (2026)
by: Gao, Shenyuan, et al.
Published: (2026)
Dreamweaver: Learning Compositional World Models from Pixels
by: Baek, Junyeob, et al.
Published: (2025)
by: Baek, Junyeob, et al.
Published: (2025)
Learning to Compose: Improving Object Centric Learning by Injecting Compositionality
by: Jung, Whie, et al.
Published: (2024)
by: Jung, Whie, et al.
Published: (2024)
Effective Tuning Strategies for Generalist Robot Manipulation Policies
by: Zhang, Wenbo, et al.
Published: (2024)
by: Zhang, Wenbo, et al.
Published: (2024)
Federated Active Learning (F-AL): an Efficient Annotation Strategy for Federated Learning
by: Ahn, Jin-Hyun, et al.
Published: (2022)
by: Ahn, Jin-Hyun, et al.
Published: (2022)
Contrastive Conceptor Activation Steering (COAST): Unlocking Vision-Language-Action Models through Hidden States
by: Miao, Miranda Muqing, et al.
Published: (2026)
by: Miao, Miranda Muqing, et al.
Published: (2026)
Probabilistic Multi-Agent Aircraft Landing Time Prediction
by: Kim, Kyungmin, et al.
Published: (2025)
by: Kim, Kyungmin, et al.
Published: (2025)
From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons
by: Szot, Andrew, et al.
Published: (2024)
by: Szot, Andrew, et al.
Published: (2024)
Simple Hierarchical Planning with Diffusion
by: Chen, Chang, et al.
Published: (2024)
by: Chen, Chang, et al.
Published: (2024)
Learning to Theorize the World from Observation
by: Baek, Doojin, et al.
Published: (2026)
by: Baek, Doojin, et al.
Published: (2026)
Flexible MOF Generation with Torsion-Aware Flow Matching
by: Kim, Nayoung, et al.
Published: (2025)
by: Kim, Nayoung, et al.
Published: (2025)
Rethinking Caching for LLM Serving Systems: Beyond Traditional Heuristics
by: Kim, Jungwoo, et al.
Published: (2025)
by: Kim, Jungwoo, et al.
Published: (2025)
Multimodal Deep Generative Model for Semi-Supervised Learning under Class Imbalance
by: Yoon, Heegeon, et al.
Published: (2026)
by: Yoon, Heegeon, et al.
Published: (2026)
Turning Video Models into Generalist Robot Policies
by: Li, Sizhe Lester, et al.
Published: (2026)
by: Li, Sizhe Lester, et al.
Published: (2026)
Generating Multi-Table Time Series EHR from Latent Space with Minimal Preprocessing
by: Cho, Eunbyeol, et al.
Published: (2025)
by: Cho, Eunbyeol, et al.
Published: (2025)
SheetAgent: Towards A Generalist Agent for Spreadsheet Reasoning and Manipulation via Large Language Models
by: Chen, Yibin, et al.
Published: (2024)
by: Chen, Yibin, et al.
Published: (2024)
Symmetric Replay Training: Enhancing Sample Efficiency in Deep Reinforcement Learning for Combinatorial Optimization
by: Kim, Hyeonah, et al.
Published: (2023)
by: Kim, Hyeonah, et al.
Published: (2023)
Similar Items
-
Spatially-Aware Transformer for Embodied Agents
by: Cho, Junmo, et al.
Published: (2024) -
Inference-Time Scaling in Diffusion Models through Iterative Partial Refinement
by: Kang, Taegu, et al.
Published: (2026) -
Compositional Monte Carlo Tree Diffusion for Extendable Planning
by: Yoon, Jaesik, et al.
Published: (2025) -
TransDreamer: Reinforcement Learning with Transformer World Models
by: Chen, Chang, et al.
Published: (2022) -
Loopholing Discrete Diffusion: Deterministic Bypass of the Sampling Wall
by: Jo, Mingyu, et al.
Published: (2025)