Gespeichert in:
| Hauptverfasser: | Goel, Shivam, Wei, Yichen, Lymperopoulos, Panagiotis, Chura, Klara, Scheutz, Matthias, Sinapov, Jivko |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2401.03546 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
FLEX: A Framework for Learning Robot-Agnostic Force-based Skills Involving Sustained Contact Object Manipulation
von: Fang, Shijie, et al.
Veröffentlicht: (2025)
von: Fang, Shijie, et al.
Veröffentlicht: (2025)
Novelty Adaptation Through Hybrid Large Language Model (LLM)-Symbolic Planning and LLM-guided Reinforcement Learning
von: Lu, Hong, et al.
Veröffentlicht: (2026)
von: Lu, Hong, et al.
Veröffentlicht: (2026)
Creative Problem Solving in Large Language and Vision Models -- What Would it Take?
von: Nair, Lakshmi, et al.
Veröffentlicht: (2024)
von: Nair, Lakshmi, et al.
Veröffentlicht: (2024)
Curiosity-Driven Imagination: Discovering Plan Operators and Learning Associated Policies for Open-World Adaptation
von: Lorang, Pierrick, et al.
Veröffentlicht: (2025)
von: Lorang, Pierrick, et al.
Veröffentlicht: (2025)
Logical Specifications-guided Dynamic Task Sampling for Reinforcement Learning Agents
von: Shukla, Yash, et al.
Veröffentlicht: (2024)
von: Shukla, Yash, et al.
Veröffentlicht: (2024)
Towards Reinforcement Learning from Neural Feedback: Mapping fNIRS Signals to Agent Performance
von: Santaniello, Julia, et al.
Veröffentlicht: (2025)
von: Santaniello, Julia, et al.
Veröffentlicht: (2025)
Graph Pruning for Enumeration of Minimal Unsatisfiable Subsets
von: Lymperopoulos, Panagiotis, et al.
Veröffentlicht: (2024)
von: Lymperopoulos, Panagiotis, et al.
Veröffentlicht: (2024)
Are AI Machines Making Humans Obsolete?
von: Scheutz, Matthias
Veröffentlicht: (2025)
von: Scheutz, Matthias
Veröffentlicht: (2025)
Mapping Neural Signals to Agent Performance, A Step Towards Reinforcement Learning from Neural Feedback
von: Santaniello, Julia, et al.
Veröffentlicht: (2025)
von: Santaniello, Julia, et al.
Veröffentlicht: (2025)
The BrowserGym Ecosystem for Web Agent Research
von: De Chezelles, Thibault Le Sellier, et al.
Veröffentlicht: (2024)
von: De Chezelles, Thibault Le Sellier, et al.
Veröffentlicht: (2024)
Are you with me? A Framework for Detecting Mental Model Discrepancies in Task-Based Team Dialogues
von: Kowalyshyn, Katharine, et al.
Veröffentlicht: (2026)
von: Kowalyshyn, Katharine, et al.
Veröffentlicht: (2026)
MOSAIC: Learning Unified Multi-Sensory Object Property Representations for Robot Learning via Interactive Perception
von: Tatiya, Gyan, et al.
Veröffentlicht: (2023)
von: Tatiya, Gyan, et al.
Veröffentlicht: (2023)
Tools in the Loop: Quantifying Uncertainty of LLM Question Answering Systems That Use Tools
von: Lymperopoulos, Panagiotis, et al.
Veröffentlicht: (2025)
von: Lymperopoulos, Panagiotis, et al.
Veröffentlicht: (2025)
Build on Priors: Vision--Language--Guided Neuro-Symbolic Imitation Learning for Data-Efficient Real-World Robot Manipulation
von: Lorang, Pierrick, et al.
Veröffentlicht: (2026)
von: Lorang, Pierrick, et al.
Veröffentlicht: (2026)
ResearchGym: Evaluating Language Model Agents on Real-World AI Research
von: Garikaparthi, Aniketh, et al.
Veröffentlicht: (2026)
von: Garikaparthi, Aniketh, et al.
Veröffentlicht: (2026)
WorldGym: World Model as An Environment for Policy Evaluation
von: Quevedo, Julian, et al.
Veröffentlicht: (2025)
von: Quevedo, Julian, et al.
Veröffentlicht: (2025)
Noise Injection Systemically Degrades Large Language Model Safety Guardrails
von: Shahani, Prithviraj Singh, et al.
Veröffentlicht: (2025)
von: Shahani, Prithviraj Singh, et al.
Veröffentlicht: (2025)
FormGym: Doing Paperwork with Agents
von: Toles, Matthew, et al.
Veröffentlicht: (2025)
von: Toles, Matthew, et al.
Veröffentlicht: (2025)
CASSANDRA: Programmatic and Probabilistic Learning and Inference for Stochastic World Modeling
von: Lymperopoulos, Panagiotis, et al.
Veröffentlicht: (2026)
von: Lymperopoulos, Panagiotis, et al.
Veröffentlicht: (2026)
CyberGym: Evaluating AI Agents' Real-World Cybersecurity Capabilities at Scale
von: Wang, Zhun, et al.
Veröffentlicht: (2025)
von: Wang, Zhun, et al.
Veröffentlicht: (2025)
ViroGym: Realistic Large-Scale Benchmarks for Evaluating Viral Proteins
von: Zhou, Yichen, et al.
Veröffentlicht: (2026)
von: Zhou, Yichen, et al.
Veröffentlicht: (2026)
Smart Language Agents in Real-World Planning
von: Miin, Annabelle, et al.
Veröffentlicht: (2024)
von: Miin, Annabelle, et al.
Veröffentlicht: (2024)
Probing a Vision-Language-Action Model for Symbolic States and Integration into a Cognitive Architecture
von: Lu, Hong, et al.
Veröffentlicht: (2025)
von: Lu, Hong, et al.
Veröffentlicht: (2025)
PersonaGym: Evaluating Persona Agents and LLMs
von: Samuel, Vinay, et al.
Veröffentlicht: (2024)
von: Samuel, Vinay, et al.
Veröffentlicht: (2024)
Mini Amusement Parks (MAPs): A Testbed for Modelling Business Decisions
von: Aroca-Ouellette, Stéphane, et al.
Veröffentlicht: (2025)
von: Aroca-Ouellette, Stéphane, et al.
Veröffentlicht: (2025)
Gym-Anything: Turn any Software into an Agent Environment
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2026)
von: Aggarwal, Pranjal, et al.
Veröffentlicht: (2026)
Gym4ReaL: A Suite for Benchmarking Real-World Reinforcement Learning
von: Salaorni, Davide, et al.
Veröffentlicht: (2025)
von: Salaorni, Davide, et al.
Veröffentlicht: (2025)
HDDLGym: A Tool for Studying Multi-Agent Hierarchical Problems Defined in HDDL with OpenAI Gym
von: La, Ngoc, et al.
Veröffentlicht: (2025)
von: La, Ngoc, et al.
Veröffentlicht: (2025)
EcoGym: Evaluating LLMs for Long-Horizon Plan-and-Execute in Interactive Economies
von: Hu, Xavier, et al.
Veröffentlicht: (2026)
von: Hu, Xavier, et al.
Veröffentlicht: (2026)
Where Norms and References Collide: Evaluating LLMs on Normative Reasoning
von: Abrams, Mitchell, et al.
Veröffentlicht: (2026)
von: Abrams, Mitchell, et al.
Veröffentlicht: (2026)
AgentGym: Evolving Large Language Model-based Agents across Diverse Environments
von: Xi, Zhiheng, et al.
Veröffentlicht: (2024)
von: Xi, Zhiheng, et al.
Veröffentlicht: (2024)
SearchGym: Bootstrapping Real-World Search Agents via Cost-Effective and High-Fidelity Environment Simulation
von: Zhang, Xichen, et al.
Veröffentlicht: (2026)
von: Zhang, Xichen, et al.
Veröffentlicht: (2026)
InnoGym: Benchmarking the Innovation Potential of AI Agents
von: Zhang, Jintian, et al.
Veröffentlicht: (2025)
von: Zhang, Jintian, et al.
Veröffentlicht: (2025)
Novelty Accommodating Multi-Agent Planning in High Fidelity Simulated Open World
von: Chao, James, et al.
Veröffentlicht: (2023)
von: Chao, James, et al.
Veröffentlicht: (2023)
CybORG++: An Enhanced Gym for the Development of Autonomous Cyber Agents
von: Emerson, Harry, et al.
Veröffentlicht: (2024)
von: Emerson, Harry, et al.
Veröffentlicht: (2024)
ClawGym: A Scalable Framework for Building Effective Claw Agents
von: Bai, Fei, et al.
Veröffentlicht: (2026)
von: Bai, Fei, et al.
Veröffentlicht: (2026)
EduGym: An Environment and Notebook Suite for Reinforcement Learning Education
von: Moerland, Thomas M., et al.
Veröffentlicht: (2023)
von: Moerland, Thomas M., et al.
Veröffentlicht: (2023)
EO-Gym: A Multimodal, Interactive Environment for Earth Observation Agents
von: Ma, Sai, et al.
Veröffentlicht: (2026)
von: Ma, Sai, et al.
Veröffentlicht: (2026)
HyPlan: Hybrid Learning-Assisted Planning Under Uncertainty for Safe Autonomous Driving
von: Pfaffmann, Donald, et al.
Veröffentlicht: (2025)
von: Pfaffmann, Donald, et al.
Veröffentlicht: (2025)
LLMs and their Limited Theory of Mind: Evaluating Mental State Annotations in Situated Dialogue
von: Kowalyshyn, Katharine, et al.
Veröffentlicht: (2025)
von: Kowalyshyn, Katharine, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
FLEX: A Framework for Learning Robot-Agnostic Force-based Skills Involving Sustained Contact Object Manipulation
von: Fang, Shijie, et al.
Veröffentlicht: (2025) -
Novelty Adaptation Through Hybrid Large Language Model (LLM)-Symbolic Planning and LLM-guided Reinforcement Learning
von: Lu, Hong, et al.
Veröffentlicht: (2026) -
Creative Problem Solving in Large Language and Vision Models -- What Would it Take?
von: Nair, Lakshmi, et al.
Veröffentlicht: (2024) -
Curiosity-Driven Imagination: Discovering Plan Operators and Learning Associated Policies for Open-World Adaptation
von: Lorang, Pierrick, et al.
Veröffentlicht: (2025) -
Logical Specifications-guided Dynamic Task Sampling for Reinforcement Learning Agents
von: Shukla, Yash, et al.
Veröffentlicht: (2024)