LangSuitE: Planning, Controlling and Interacting with Large Language Models in Embodied Text Environments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jia, Zixia, Wang, Mengmeng, Tong, Baichen, Zhu, Song-Chun, Zheng, Zilong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Combining Supervised Learning and Reinforcement Learning for Multi-Label Classification Tasks with Partial Labels
von: Jia, Zixia, et al.
Veröffentlicht: (2024)
von: Jia, Zixia, et al.
Veröffentlicht: (2024)
LooGLE: Can Long-Context Language Models Understand Long Contexts?
von: Li, Jiaqi, et al.
Veröffentlicht: (2023)
von: Li, Jiaqi, et al.
Veröffentlicht: (2023)
TALES: Text Adventure Learning Environment Suite
von: Cui, Christopher Zhang, et al.
Veröffentlicht: (2025)
von: Cui, Christopher Zhang, et al.
Veröffentlicht: (2025)
Seek in the Dark: Reasoning via Test-Time Instance-Level Policy Gradient in Latent Space
von: Li, Hengli, et al.
Veröffentlicht: (2025)
von: Li, Hengli, et al.
Veröffentlicht: (2025)
RAM: Towards an Ever-Improving Memory System by Learning from Communications
von: Li, Jiaqi, et al.
Veröffentlicht: (2024)
von: Li, Jiaqi, et al.
Veröffentlicht: (2024)
TokenSwift: Lossless Acceleration of Ultra Long Sequence Generation
von: Wu, Tong, et al.
Veröffentlicht: (2025)
von: Wu, Tong, et al.
Veröffentlicht: (2025)
Understanding and Leveraging the Expert Specialization of Context Faithfulness in Mixture-of-Experts LLMs
von: Bai, Jun, et al.
Veröffentlicht: (2025)
von: Bai, Jun, et al.
Veröffentlicht: (2025)
SCOPE: Language Models as One-Time Teacher for Hierarchical Planning in Text Environments
von: Lu, Haoye, et al.
Veröffentlicht: (2025)
von: Lu, Haoye, et al.
Veröffentlicht: (2025)
LangVAE and LangSpace: Building and Probing for Language Model VAEs
von: Carvalho, Danilo S., et al.
Veröffentlicht: (2025)
von: Carvalho, Danilo S., et al.
Veröffentlicht: (2025)
Multi-Turn Interactions for Text-to-SQL with Large Language Models
von: Xiong, Guanming, et al.
Veröffentlicht: (2024)
von: Xiong, Guanming, et al.
Veröffentlicht: (2024)
Runaway is Ashamed, But Helpful: On the Early-Exit Behavior of Large Language Model-based Agents in Embodied Environments
von: Lu, Qingyu, et al.
Veröffentlicht: (2025)
von: Lu, Qingyu, et al.
Veröffentlicht: (2025)
Sparser is Faster and Less is More: Efficient Sparse Attention for Long-Range Transformers
von: Lou, Chao, et al.
Veröffentlicht: (2024)
von: Lou, Chao, et al.
Veröffentlicht: (2024)
A Lightweight Multi Aspect Controlled Text Generation Solution For Large Language Models
von: Zhang, Chenyang, et al.
Veröffentlicht: (2024)
von: Zhang, Chenyang, et al.
Veröffentlicht: (2024)
Brain in a Vat: On Missing Pieces Towards Artificial General Intelligence in Large Language Models
von: Ma, Yuxi, et al.
Veröffentlicht: (2023)
von: Ma, Yuxi, et al.
Veröffentlicht: (2023)
Native Parallel Reasoner: Reasoning in Parallelism via Self-Distilled Reinforcement Learning
von: Wu, Tong, et al.
Veröffentlicht: (2025)
von: Wu, Tong, et al.
Veröffentlicht: (2025)
Probing and Inducing Combinational Creativity in Vision-Language Models
von: Peng, Yongqian, et al.
Veröffentlicht: (2025)
von: Peng, Yongqian, et al.
Veröffentlicht: (2025)
Foundations of Large Language Models
von: Xiao, Tong, et al.
Veröffentlicht: (2025)
von: Xiao, Tong, et al.
Veröffentlicht: (2025)
Mars: Situated Inductive Reasoning in an Open-World Environment
von: Tang, Xiaojuan, et al.
Veröffentlicht: (2024)
von: Tang, Xiaojuan, et al.
Veröffentlicht: (2024)
EduEval: A Hierarchical Cognitive Benchmark for Evaluating Large Language Models in Chinese Education
von: Ma, Guoqing, et al.
Veröffentlicht: (2025)
von: Ma, Guoqing, et al.
Veröffentlicht: (2025)
PSALM-V: Automating Symbolic Planning in Interactive Visual Environments with Large Language Models
von: Zhu, Wang Bill, et al.
Veröffentlicht: (2025)
von: Zhu, Wang Bill, et al.
Veröffentlicht: (2025)
XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models
von: Röttger, Paul, et al.
Veröffentlicht: (2023)
von: Röttger, Paul, et al.
Veröffentlicht: (2023)
CFDLLMBench: A Benchmark Suite for Evaluating Large Language Models in Computational Fluid Dynamics
von: Somasekharan, Nithin, et al.
Veröffentlicht: (2025)
von: Somasekharan, Nithin, et al.
Veröffentlicht: (2025)
ViLLM-Eval: A Comprehensive Evaluation Suite for Vietnamese Large Language Models
von: Nguyen, Trong-Hieu, et al.
Veröffentlicht: (2024)
von: Nguyen, Trong-Hieu, et al.
Veröffentlicht: (2024)
FVEL: Interactive Formal Verification Environment with Large Language Models via Theorem Proving
von: Lin, Xiaohan, et al.
Veröffentlicht: (2024)
von: Lin, Xiaohan, et al.
Veröffentlicht: (2024)
Safety of Embodied Navigation: A Survey
von: Wang, Zixia, et al.
Veröffentlicht: (2025)
von: Wang, Zixia, et al.
Veröffentlicht: (2025)
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
von: Yang, Rui, et al.
Veröffentlicht: (2025)
von: Yang, Rui, et al.
Veröffentlicht: (2025)
Automatically Planning Optimal Parallel Strategy for Large Language Models
von: Li, Zongbiao, et al.
Veröffentlicht: (2024)
von: Li, Zongbiao, et al.
Veröffentlicht: (2024)
Large Language Models as Generalizable Policies for Embodied Tasks
von: Szot, Andrew, et al.
Veröffentlicht: (2023)
von: Szot, Andrew, et al.
Veröffentlicht: (2023)
When Language Overrules: Revealing Text Dominance in Multimodal Large Language Models
von: Wu, Huyu, et al.
Veröffentlicht: (2025)
von: Wu, Huyu, et al.
Veröffentlicht: (2025)
Discrete Markov Bridge
von: Li, Hengli, et al.
Veröffentlicht: (2025)
von: Li, Hengli, et al.
Veröffentlicht: (2025)
A Comprehensive Evaluation of Large Language Models on Benchmark Biomedical Text Processing Tasks
von: Jahan, Israt, et al.
Veröffentlicht: (2023)
von: Jahan, Israt, et al.
Veröffentlicht: (2023)
Knowledge Editing for Large Language Models: A Survey
von: Wang, Song, et al.
Veröffentlicht: (2023)
von: Wang, Song, et al.
Veröffentlicht: (2023)
Can ChatGPT replace StackOverflow? A Study on Robustness and Reliability of Large Language Model Code Generation
von: Zhong, Li, et al.
Veröffentlicht: (2023)
von: Zhong, Li, et al.
Veröffentlicht: (2023)
Multi-Agent Simulator Drives Language Models for Legal Intensive Interaction
von: Yue, Shengbin, et al.
Veröffentlicht: (2025)
von: Yue, Shengbin, et al.
Veröffentlicht: (2025)
MLLMGuard: A Multi-dimensional Safety Evaluation Suite for Multimodal Large Language Models
von: Gu, Tianle, et al.
Veröffentlicht: (2024)
von: Gu, Tianle, et al.
Veröffentlicht: (2024)
Adaptive Preference Optimization with Uncertainty-aware Utility Anchor
von: Wang, Xiaobo, et al.
Veröffentlicht: (2025)
von: Wang, Xiaobo, et al.
Veröffentlicht: (2025)
Parameter-Efficient Fine-Tuning of Large Language Models via Deconvolution in Subspace
von: Zhang, Jia-Chen, et al.
Veröffentlicht: (2025)
von: Zhang, Jia-Chen, et al.
Veröffentlicht: (2025)
Agent AI with LangGraph: A Modular Framework for Enhancing Machine Translation Using Large Language Models
von: Wang, Jialin, et al.
Veröffentlicht: (2024)
von: Wang, Jialin, et al.
Veröffentlicht: (2024)
UrbanPlanBench: A Comprehensive Urban Planning Benchmark for Evaluating Large Language Models
von: Zheng, Yu, et al.
Veröffentlicht: (2025)
von: Zheng, Yu, et al.
Veröffentlicht: (2025)
Unifying Text Semantics and Graph Structures for Temporal Text-attributed Graphs with Large Language Models
von: Zhang, Siwei, et al.
Veröffentlicht: (2025)
von: Zhang, Siwei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Combining Supervised Learning and Reinforcement Learning for Multi-Label Classification Tasks with Partial Labels
von: Jia, Zixia, et al.
Veröffentlicht: (2024) -
LooGLE: Can Long-Context Language Models Understand Long Contexts?
von: Li, Jiaqi, et al.
Veröffentlicht: (2023) -
TALES: Text Adventure Learning Environment Suite
von: Cui, Christopher Zhang, et al.
Veröffentlicht: (2025) -
Seek in the Dark: Reasoning via Test-Time Instance-Level Policy Gradient in Latent Space
von: Li, Hengli, et al.
Veröffentlicht: (2025) -
RAM: Towards an Ever-Improving Memory System by Learning from Communications
von: Li, Jiaqi, et al.
Veröffentlicht: (2024)