The Sandbox Environment for Generalizable Agent Research (SEGAR)
Fuente:
arXiv
Guardado en:
| Autores principales: | Hjelm, R Devon, Mazoure, Bogdan, Golemo, Florian, Kahou, Samira Ebrahimi, Braga, Pedro, Frujeri, Felipe, Jalobeanu, Mihai, Kolobov, Andrey |
|---|---|
| Formato: | Preprint |
| Publicado: |
2022
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
On the Modeling Capabilities of Large Language Models for Sequential Decision Making
por: Klissarov, Martin, et al.
Publicado: (2024)
por: Klissarov, Martin, et al.
Publicado: (2024)
GRACE: A Language Model Framework for Explainable Inverse Reinforcement Learning
por: Sapora, Silvia, et al.
Publicado: (2025)
por: Sapora, Silvia, et al.
Publicado: (2025)
Grounding Multimodal Large Language Models in Actions
por: Szot, Andrew, et al.
Publicado: (2024)
por: Szot, Andrew, et al.
Publicado: (2024)
Large Language Models as Generalizable Policies for Embodied Tasks
por: Szot, Andrew, et al.
Publicado: (2023)
por: Szot, Andrew, et al.
Publicado: (2023)
From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons
por: Szot, Andrew, et al.
Publicado: (2024)
por: Szot, Andrew, et al.
Publicado: (2024)
CAMMARL: Conformal Action Modeling in Multi Agent Reinforcement Learning
por: Gupta, Nikunj, et al.
Publicado: (2023)
por: Gupta, Nikunj, et al.
Publicado: (2023)
Locally Constrained Representations in Reinforcement Learning
por: Nath, Somjit, et al.
Publicado: (2022)
por: Nath, Somjit, et al.
Publicado: (2022)
Estimation of Head Motion in Structural MRI and its Impact on Cortical Thickness Measurements in Retrospective Data
por: Bricout, Charles, et al.
Publicado: (2025)
por: Bricout, Charles, et al.
Publicado: (2025)
Zero-Shot Anomaly Detection with Dual-Branch Prompt Selection
por: Wang, Zihan, et al.
Publicado: (2025)
por: Wang, Zihan, et al.
Publicado: (2025)
Adaptive Group Robust Ensemble Knowledge Distillation
por: Kenfack, Patrik, et al.
Publicado: (2024)
por: Kenfack, Patrik, et al.
Publicado: (2024)
Towards Fair In-Context Learning with Tabular Foundation Models
por: Kenfack, Patrik, et al.
Publicado: (2025)
por: Kenfack, Patrik, et al.
Publicado: (2025)
Learning to Play Atari in a World of Tokens
por: Agarwal, Pranav, et al.
Publicado: (2024)
por: Agarwal, Pranav, et al.
Publicado: (2024)
Fairness Under Demographic Scarce Regime
por: Kenfack, Patrik Joslin, et al.
Publicado: (2023)
por: Kenfack, Patrik Joslin, et al.
Publicado: (2023)
Chapter Making Della Porta a Baconian Philosopher: <i>Magia naturalis </i>in the Context of the English Experimental Philosophy
por: Dana, Jalobeanu
Publicado: (2026)
por: Dana, Jalobeanu
Publicado: (2026)
Vector Quantized Latent Concepts: A Scalable Alternative to Clustering-Based Concept Discovery
por: Yu, Xuemin, et al.
Publicado: (2026)
por: Yu, Xuemin, et al.
Publicado: (2026)
Cross-Layer Discrete Concept Discovery for Interpreting Language Models
por: Garg, Ankur, et al.
Publicado: (2025)
por: Garg, Ankur, et al.
Publicado: (2025)
Behaviour Discovery and Attribution for Explainable Reinforcement Learning
por: Rishav, Rishav, et al.
Publicado: (2025)
por: Rishav, Rishav, et al.
Publicado: (2025)
KD-LoRA: A Hybrid Approach to Efficient Fine-Tuning with LoRA and Knowledge Distillation
por: Azimi, Rambod, et al.
Publicado: (2024)
por: Azimi, Rambod, et al.
Publicado: (2024)
Learning Multi-agent Multi-machine Tending by Mobile Robots
por: Abdalwhab, Abdalwhab, et al.
Publicado: (2024)
por: Abdalwhab, Abdalwhab, et al.
Publicado: (2024)
SEGAR: Selective Enhancement for Generative Augmented Reality
por: Bu, Fanjun, et al.
Publicado: (2026)
por: Bu, Fanjun, et al.
Publicado: (2026)
Indignados en España e Indecisos en Polonia. La inspiración española en el contexto polaco y el fracaso de la protesta en el país de “Solidarnosć”
por: Karolina Golemo
Publicado: (2014)
por: Karolina Golemo
Publicado: (2014)
Goal Representations for Instruction Following: A Semi-Supervised Language Interface to Control
por: Myers, Vivek, et al.
Publicado: (2023)
por: Myers, Vivek, et al.
Publicado: (2023)
Handling Delay in Real-Time Reinforcement Learning
por: Anokhin, Ivan, et al.
Publicado: (2025)
por: Anokhin, Ivan, et al.
Publicado: (2025)
Survey on AI Ethics: A Socio-technical Perspective
por: Mbiazi, Dave, et al.
Publicado: (2023)
por: Mbiazi, Dave, et al.
Publicado: (2023)
Solving Models of Economic Dynamics with Ridgeless Kernel Regressions
por: Kahou, Mahdi Ebrahimi, et al.
Publicado: (2024)
por: Kahou, Mahdi Ebrahimi, et al.
Publicado: (2024)
On the benefits of pixel-based hierarchical policies for task generalization
por: Cristea-Platon, Tudor, et al.
Publicado: (2024)
por: Cristea-Platon, Tudor, et al.
Publicado: (2024)
GHIssuemarket: A Sandbox Environment for SWE-Agents Economic Experimentation
por: Fouad, Mohamed A., et al.
Publicado: (2024)
por: Fouad, Mohamed A., et al.
Publicado: (2024)
Prediction of Final Phosphorus Content of Steel in a Scrap-Based Electric Arc Furnace Using Artificial Neural Networks
por: Azzaz, Riadh, et al.
Publicado: (2024)
por: Azzaz, Riadh, et al.
Publicado: (2024)
Empowering Clinicians with Medical Decision Transformers: A Framework for Sepsis Treatment
por: Rahman, Aamer Abdul, et al.
Publicado: (2024)
por: Rahman, Aamer Abdul, et al.
Publicado: (2024)
Peter L. Berger on Religion
por: Hjelm, Titus
Publicado: (2026)
por: Hjelm, Titus
Publicado: (2026)
Uskonto, kieli ja yhteiskunta
por: Hjelm, Titus
Publicado: (2021)
por: Hjelm, Titus
Publicado: (2021)
Learning From the Past with Cascading Eligibility Traces
por: Ralambomihanta, Tokiniaina Raharison, et al.
Publicado: (2025)
por: Ralambomihanta, Tokiniaina Raharison, et al.
Publicado: (2025)
Reinforcement Learning for Sequence Design Leveraging Protein Language Models
por: Subramanian, Jithendaraa, et al.
Publicado: (2024)
por: Subramanian, Jithendaraa, et al.
Publicado: (2024)
Comparative Analysis of Diffusion Generative Models in Computational Pathology
por: Thakkar, Denisha, et al.
Publicado: (2024)
por: Thakkar, Denisha, et al.
Publicado: (2024)
WindSeer: Real-time volumetric wind prediction over complex terrain aboard a small UAV
por: Achermann, Florian, et al.
Publicado: (2024)
por: Achermann, Florian, et al.
Publicado: (2024)
Sari Sandbox: A Virtual Retail Store Environment for Embodied AI Agents
por: Gajo, Janika Deborah, et al.
Publicado: (2025)
por: Gajo, Janika Deborah, et al.
Publicado: (2025)
On the Limits of Multi-modal Meta-Learning with Auxiliary Task Modulation Using Conditional Batch Normalization
por: Armengol-Estapé, Jordi, et al.
Publicado: (2024)
por: Armengol-Estapé, Jordi, et al.
Publicado: (2024)
Scaling Synthetic Task Generation for Agents via Exploration
por: Ramrakhya, Ram, et al.
Publicado: (2025)
por: Ramrakhya, Ram, et al.
Publicado: (2025)
ClustRecNet: A Novel End-to-End Deep Learning Framework for Clustering Algorithm Recommendation
por: Bakhtyari, Mohammadreza, et al.
Publicado: (2025)
por: Bakhtyari, Mohammadreza, et al.
Publicado: (2025)
Okapi: Efficiently Safeguarding Speculative Data Accesses in Sandboxed Environments
por: Schmitz, Philipp, et al.
Publicado: (2023)
por: Schmitz, Philipp, et al.
Publicado: (2023)
Ejemplares similares
-
On the Modeling Capabilities of Large Language Models for Sequential Decision Making
por: Klissarov, Martin, et al.
Publicado: (2024) -
GRACE: A Language Model Framework for Explainable Inverse Reinforcement Learning
por: Sapora, Silvia, et al.
Publicado: (2025) -
Grounding Multimodal Large Language Models in Actions
por: Szot, Andrew, et al.
Publicado: (2024) -
Large Language Models as Generalizable Policies for Embodied Tasks
por: Szot, Andrew, et al.
Publicado: (2023) -
From Multimodal LLMs to Generalist Embodied Agents: Methods and Lessons
por: Szot, Andrew, et al.
Publicado: (2024)