HAZARD Challenge: Embodied Decision Making in Dynamically Changing Environments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Qinhong, Chen, Sunli, Wang, Yisong, Xu, Haozhe, Du, Weihua, Zhang, Hongxin, Du, Yilun, Tenenbaum, Joshua B., Gan, Chuang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Building Cooperative Embodied Agents Modularly with Large Language Models
von: Zhang, Hongxin, et al.
Veröffentlicht: (2023)
von: Zhang, Hongxin, et al.
Veröffentlicht: (2023)
Ella: Embodied Social Agents with Lifelong Memory
von: Zhang, Hongxin, et al.
Veröffentlicht: (2025)
von: Zhang, Hongxin, et al.
Veröffentlicht: (2025)
Constrained Human-AI Cooperation: An Inclusive Embodied Social Intelligence Challenge
von: Du, Weihua, et al.
Veröffentlicht: (2024)
von: Du, Weihua, et al.
Veröffentlicht: (2024)
Sentinel: Embodied Cooperative Spatial Reasoning and Planning
von: Lin, Xiangye, et al.
Veröffentlicht: (2026)
von: Lin, Xiangye, et al.
Veröffentlicht: (2026)
COMBO: Compositional World Models for Embodied Multi-Agent Cooperation
von: Zhang, Hongxin, et al.
Veröffentlicht: (2024)
von: Zhang, Hongxin, et al.
Veröffentlicht: (2024)
TesserAct: Learning 4D Embodied World Models
von: Zhen, Haoyu, et al.
Veröffentlicht: (2025)
von: Zhen, Haoyu, et al.
Veröffentlicht: (2025)
3D-Mem: 3D Scene Memory for Embodied Exploration and Reasoning
von: Yang, Yuncong, et al.
Veröffentlicht: (2024)
von: Yang, Yuncong, et al.
Veröffentlicht: (2024)
STAR: A Benchmark for Situated Reasoning in Real-World Videos
von: Wu, Bo, et al.
Veröffentlicht: (2024)
von: Wu, Bo, et al.
Veröffentlicht: (2024)
SALMON: Self-Alignment with Instructable Reward Models
von: Sun, Zhiqing, et al.
Veröffentlicht: (2023)
von: Sun, Zhiqing, et al.
Veröffentlicht: (2023)
Multiagent Finetuning: Self Improvement with Diverse Reasoning Chains
von: Subramaniam, Vighnesh, et al.
Veröffentlicht: (2025)
von: Subramaniam, Vighnesh, et al.
Veröffentlicht: (2025)
Improving Reinforcement Learning from Human Feedback with Efficient Reward Model Ensemble
von: Zhang, Shun, et al.
Veröffentlicht: (2024)
von: Zhang, Shun, et al.
Veröffentlicht: (2024)
Embodied Agent Interface: Benchmarking LLMs for Embodied Decision Making
von: Li, Manling, et al.
Veröffentlicht: (2024)
von: Li, Manling, et al.
Veröffentlicht: (2024)
Virtual Community: An Open World for Humans, Robots, and Society
von: Zhou, Qinhong, et al.
Veröffentlicht: (2025)
von: Zhou, Qinhong, et al.
Veröffentlicht: (2025)
SOK-Bench: A Situated Video Reasoning Benchmark with Aligned Open-World Knowledge
von: Wang, Andong, et al.
Veröffentlicht: (2024)
von: Wang, Andong, et al.
Veröffentlicht: (2024)
Learning 3D Persistent Embodied World Models
von: Zhou, Siyuan, et al.
Veröffentlicht: (2025)
von: Zhou, Siyuan, et al.
Veröffentlicht: (2025)
MSI-Agent: Incorporating Multi-Scale Insight into Embodied Agents for Superior Planning and Decision-Making
von: Fu, Dayuan, et al.
Veröffentlicht: (2024)
von: Fu, Dayuan, et al.
Veröffentlicht: (2024)
What Makes a Maze Look Like a Maze?
von: Hsu, Joy, et al.
Veröffentlicht: (2024)
von: Hsu, Joy, et al.
Veröffentlicht: (2024)
3D-VLA: A 3D Vision-Language-Action Generative World Model
von: Zhen, Haoyu, et al.
Veröffentlicht: (2024)
von: Zhen, Haoyu, et al.
Veröffentlicht: (2024)
Potential Based Diffusion Motion Planning
von: Luo, Yunhao, et al.
Veröffentlicht: (2024)
von: Luo, Yunhao, et al.
Veröffentlicht: (2024)
Reasoning with Sampling: Your Base Model is Smarter Than You Think
von: Karan, Aayush, et al.
Veröffentlicht: (2025)
von: Karan, Aayush, et al.
Veröffentlicht: (2025)
Strategist: Self-improvement of LLM Decision Making via Bi-Level Tree Search
von: Light, Jonathan, et al.
Veröffentlicht: (2024)
von: Light, Jonathan, et al.
Veröffentlicht: (2024)
DSGBench: A Diverse Strategic Game Benchmark for Evaluating LLM-based Agents in Complex Decision-Making Environments
von: Tang, Wenjie, et al.
Veröffentlicht: (2025)
von: Tang, Wenjie, et al.
Veröffentlicht: (2025)
Building Decision Making Models Through Language Model Regime
von: Zhang, Yu, et al.
Veröffentlicht: (2024)
von: Zhang, Yu, et al.
Veröffentlicht: (2024)
Minding the Politeness Gap in Cross-cultural Communication
von: Machino, Yuka, et al.
Veröffentlicht: (2025)
von: Machino, Yuka, et al.
Veröffentlicht: (2025)
Neuro-Symbolic Concepts
von: Mao, Jiayuan, et al.
Veröffentlicht: (2025)
von: Mao, Jiayuan, et al.
Veröffentlicht: (2025)
Shoot First, Ask Questions Later? Building Rational Agents that Explore and Act Like People
von: Grand, Gabriel, et al.
Veröffentlicht: (2025)
von: Grand, Gabriel, et al.
Veröffentlicht: (2025)
Loose LIPS Sink Ships: Asking Questions in Battleship with Language-Informed Program Sampling
von: Grand, Gabriel, et al.
Veröffentlicht: (2024)
von: Grand, Gabriel, et al.
Veröffentlicht: (2024)
Label-Confidence-Aware Uncertainty Estimation in Natural Language Generation
von: Lin, Qinhong, et al.
Veröffentlicht: (2024)
von: Lin, Qinhong, et al.
Veröffentlicht: (2024)
Compositional Image Decomposition with Diffusion Models
von: Su, Jocelin, et al.
Veröffentlicht: (2024)
von: Su, Jocelin, et al.
Veröffentlicht: (2024)
Large Language Models Are Bad Dice Players: LLMs Struggle to Generate Random Numbers from Statistical Distributions
von: Zhao, Minda, et al.
Veröffentlicht: (2026)
von: Zhao, Minda, et al.
Veröffentlicht: (2026)
On the Same Wavelength? Evaluating Pragmatic Reasoning in Language Models across Broad Concepts
von: Qiu, Linlu, et al.
Veröffentlicht: (2025)
von: Qiu, Linlu, et al.
Veröffentlicht: (2025)
Optimizing Temperature for Language Models with Multi-Sample Inference
von: Du, Weihua, et al.
Veröffentlicht: (2025)
von: Du, Weihua, et al.
Veröffentlicht: (2025)
MultiPLY: A Multisensory Object-Centric Embodied Large Language Model in 3D World
von: Hong, Yining, et al.
Veröffentlicht: (2024)
von: Hong, Yining, et al.
Veröffentlicht: (2024)
Do Large Language Models Discriminate in Hiring Decisions on the Basis of Race, Ethnicity, and Gender?
von: An, Haozhe, et al.
Veröffentlicht: (2024)
von: An, Haozhe, et al.
Veröffentlicht: (2024)
The Consensus Trap: Rescuing Multi-Agent LLMs from Adversarial Majorities via Token-Level Collaboration
von: Liu, Jiayuan, et al.
Veröffentlicht: (2026)
von: Liu, Jiayuan, et al.
Veröffentlicht: (2026)
Finding structure in logographic writing with library learning
von: Jiang, Guangyuan, et al.
Veröffentlicht: (2024)
von: Jiang, Guangyuan, et al.
Veröffentlicht: (2024)
LLMs and Agentic AI in Insurance Decision-Making: Opportunities and Challenges For Africa
von: Hill, Graham, et al.
Veröffentlicht: (2025)
von: Hill, Graham, et al.
Veröffentlicht: (2025)
Challenging Multilingual LLMs: A New Taxonomy and Benchmark for Unraveling Hallucination in Translation
von: Wu, Xinwei, et al.
Veröffentlicht: (2025)
von: Wu, Xinwei, et al.
Veröffentlicht: (2025)
Self-Improving Language Models with Bidirectional Evolutionary Search
von: Xu, Guowei, et al.
Veröffentlicht: (2026)
von: Xu, Guowei, et al.
Veröffentlicht: (2026)
DynaThink: Fast or Slow? A Dynamic Decision-Making Framework for Large Language Models
von: Pan, Jiabao, et al.
Veröffentlicht: (2024)
von: Pan, Jiabao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Building Cooperative Embodied Agents Modularly with Large Language Models
von: Zhang, Hongxin, et al.
Veröffentlicht: (2023) -
Ella: Embodied Social Agents with Lifelong Memory
von: Zhang, Hongxin, et al.
Veröffentlicht: (2025) -
Constrained Human-AI Cooperation: An Inclusive Embodied Social Intelligence Challenge
von: Du, Weihua, et al.
Veröffentlicht: (2024) -
Sentinel: Embodied Cooperative Spatial Reasoning and Planning
von: Lin, Xiangye, et al.
Veröffentlicht: (2026) -
COMBO: Compositional World Models for Embodied Multi-Agent Cooperation
von: Zhang, Hongxin, et al.
Veröffentlicht: (2024)