NiceWebRL: a Python library for human subject experiments with reinforcement learning environments
Fuente:
arXiv
Salvato in:
| Autori principali: | Carvalho, Wilka, Goddla, Vikram, Sinha, Ishaan, Shin, Hoon, Jha, Kunal |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Optimal Policy Sparsification and Low Rank Decomposition for Deep Reinforcement Learning
di: Goddla, Vikram
Pubblicazione: (2024)
di: Goddla, Vikram
Pubblicazione: (2024)
Task diversity produces systematic transfer but inhibits continual reinforcement learning
di: Seth, Purab, et al.
Pubblicazione: (2026)
di: Seth, Purab, et al.
Pubblicazione: (2026)
Cross-environment Cooperation Enables Zero-shot Multi-agent Coordination
di: Jha, Kunal, et al.
Pubblicazione: (2025)
di: Jha, Kunal, et al.
Pubblicazione: (2025)
Naturalistic Computational Cognitive Science: Towards generalizable models and theories that capture the full range of natural behavior
di: Carvalho, Wilka, et al.
Pubblicazione: (2025)
di: Carvalho, Wilka, et al.
Pubblicazione: (2025)
Online library learning in human visual puzzle solving
di: Zhao, Pinzhe, et al.
Pubblicazione: (2026)
di: Zhao, Pinzhe, et al.
Pubblicazione: (2026)
Decoding Latent Attack Surfaces in LLMs: Prompt Injection via HTML in Web Summarization
di: Verma, Ishaan, et al.
Pubblicazione: (2025)
di: Verma, Ishaan, et al.
Pubblicazione: (2025)
Delayed homomorphic reinforcement learning for environments with delayed feedback
di: Lee, Jongsoo, et al.
Pubblicazione: (2026)
di: Lee, Jongsoo, et al.
Pubblicazione: (2026)
Found-RL: foundation model-enhanced reinforcement learning for autonomous driving
di: Qu, Yansong, et al.
Pubblicazione: (2026)
di: Qu, Yansong, et al.
Pubblicazione: (2026)
seqme: a Python library for evaluating biological sequence design
di: Møller-Larsen, Rasmus, et al.
Pubblicazione: (2025)
di: Møller-Larsen, Rasmus, et al.
Pubblicazione: (2025)
xai-cola: A Python library for sparsifying counterfactual explanations
di: Zhu, Lin, et al.
Pubblicazione: (2026)
di: Zhu, Lin, et al.
Pubblicazione: (2026)
Counterfactual experience augmented off-policy reinforcement learning
di: Lee, Sunbowen, et al.
Pubblicazione: (2025)
di: Lee, Sunbowen, et al.
Pubblicazione: (2025)
Detecting the Disturbance: A Nuanced View of Introspective Abilities in LLMs
di: Hahami, Ely, et al.
Pubblicazione: (2025)
di: Hahami, Ely, et al.
Pubblicazione: (2025)
Predictive representations: building blocks of intelligence
di: Carvalho, Wilka, et al.
Pubblicazione: (2024)
di: Carvalho, Wilka, et al.
Pubblicazione: (2024)
An efficient deep reinforcement learning environment for flexible job-shop scheduling
di: Wu, Xinquan, et al.
Pubblicazione: (2025)
di: Wu, Xinquan, et al.
Pubblicazione: (2025)
Scilab-RL: A software framework for efficient reinforcement learning and cognitive modeling research
di: Dohmen, Jan, et al.
Pubblicazione: (2024)
di: Dohmen, Jan, et al.
Pubblicazione: (2024)
MaxInfoRL: Boosting exploration in reinforcement learning through information gain maximization
di: Sukhija, Bhavya, et al.
Pubblicazione: (2024)
di: Sukhija, Bhavya, et al.
Pubblicazione: (2024)
SuperCoder2.0: Technical Report on Exploring the feasibility of LLMs as Autonomous Programmer
di: Gautam, Anmol, et al.
Pubblicazione: (2024)
di: Gautam, Anmol, et al.
Pubblicazione: (2024)
Multi-Task Learning with Additive U-Net for Image Denoising and Classification
di: Lakkavalli, Vikram, et al.
Pubblicazione: (2026)
di: Lakkavalli, Vikram, et al.
Pubblicazione: (2026)
Learning to summarize user information for personalized reinforcement learning from human feedback
di: Nam, Hyunji, et al.
Pubblicazione: (2025)
di: Nam, Hyunji, et al.
Pubblicazione: (2025)
BenchRL-QAS: Benchmarking reinforcement learning algorithms for quantum architecture search
di: Ikhtiarudin, Azhar, et al.
Pubblicazione: (2025)
di: Ikhtiarudin, Azhar, et al.
Pubblicazione: (2025)
1000 Layer Networks for Self-Supervised RL: Scaling Depth Can Enable New Goal-Reaching Capabilities
di: Wang, Kevin, et al.
Pubblicazione: (2025)
di: Wang, Kevin, et al.
Pubblicazione: (2025)
Safe reinforcement learning with online filtering for fatigue-predictive human-robot task planning and allocation in production
di: Xue, Jintao, et al.
Pubblicazione: (2026)
di: Xue, Jintao, et al.
Pubblicazione: (2026)
On Minimizing Adversarial Counterfactual Error in Adversarial RL
di: Belaire, Roman, et al.
Pubblicazione: (2024)
di: Belaire, Roman, et al.
Pubblicazione: (2024)
GUIDE: Graphical User Interface Data for Execution
di: Chawla, Rajat, et al.
Pubblicazione: (2024)
di: Chawla, Rajat, et al.
Pubblicazione: (2024)
Collocational bootstrapping: A hypothesis about the learning of subject-verb agreement in humans and neural networks
di: Hobbs, Claire, et al.
Pubblicazione: (2026)
di: Hobbs, Claire, et al.
Pubblicazione: (2026)
This human study did not involve human subjects: Validating LLM simulations as behavioral evidence
di: Hullman, Jessica, et al.
Pubblicazione: (2026)
di: Hullman, Jessica, et al.
Pubblicazione: (2026)
Unified Smart Factory Model: A model-based Approach for Integrating Industry 4.0 and Sustainability for Manufacturing Systems
di: Kaushal, Ishaan, et al.
Pubblicazione: (2025)
di: Kaushal, Ishaan, et al.
Pubblicazione: (2025)
MicroProbe: Efficient Reliability Assessment for Foundation Models with Minimal Data
di: Bansal, Aayam, et al.
Pubblicazione: (2025)
di: Bansal, Aayam, et al.
Pubblicazione: (2025)
AgentComm-Bench: Stress-Testing Cooperative Embodied AI Under Latency, Packet Loss, and Bandwidth Collapse
di: Bansal, Aayam, et al.
Pubblicazione: (2026)
di: Bansal, Aayam, et al.
Pubblicazione: (2026)
Beyond human subjectivity and error: a novel AI grading system
di: Gobrecht, Alexandra, et al.
Pubblicazione: (2024)
di: Gobrecht, Alexandra, et al.
Pubblicazione: (2024)
A hierarchical spatial-aware algorithm with efficient reinforcement learning for human-robot task planning and allocation in production
di: Xue, Jintao, et al.
Pubblicazione: (2026)
di: Xue, Jintao, et al.
Pubblicazione: (2026)
Nice Fold or Hero Call: Learning Budget-Efficient Thinking for Adaptive Reasoning
di: Zhou, Zhaomeng, et al.
Pubblicazione: (2026)
di: Zhou, Zhaomeng, et al.
Pubblicazione: (2026)
Safeguarding AI Agents: Developing and Analyzing Safety Architectures
di: Domkundwar, Ishaan, et al.
Pubblicazione: (2024)
di: Domkundwar, Ishaan, et al.
Pubblicazione: (2024)
Traffic expertise meets residual RL: Knowledge-informed model-based residual reinforcement learning for CAV trajectory control
di: Sheng, Zihao, et al.
Pubblicazione: (2024)
di: Sheng, Zihao, et al.
Pubblicazione: (2024)
RL-SPH: Learning to Achieve Feasible Solutions for Integer Linear Programs
di: Lee, Tae-Hoon, et al.
Pubblicazione: (2024)
di: Lee, Tae-Hoon, et al.
Pubblicazione: (2024)
PyRelationAL: a python library for active learning research and development
di: Scherer, Paul, et al.
Pubblicazione: (2022)
di: Scherer, Paul, et al.
Pubblicazione: (2022)
SCOPE-RL: A Python Library for Offline Reinforcement Learning and Off-Policy Evaluation
di: Kiyohara, Haruka, et al.
Pubblicazione: (2023)
di: Kiyohara, Haruka, et al.
Pubblicazione: (2023)
Steward: Natural Language Web Automation
di: Tang, Brian, et al.
Pubblicazione: (2024)
di: Tang, Brian, et al.
Pubblicazione: (2024)
Normalization and effective learning rates in reinforcement learning
di: Lyle, Clare, et al.
Pubblicazione: (2024)
di: Lyle, Clare, et al.
Pubblicazione: (2024)
Reason-to-Transmit: Deliberative Adaptive Communication for Cooperative Perception
di: Bansal, Aayam, et al.
Pubblicazione: (2026)
di: Bansal, Aayam, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Optimal Policy Sparsification and Low Rank Decomposition for Deep Reinforcement Learning
di: Goddla, Vikram
Pubblicazione: (2024) -
Task diversity produces systematic transfer but inhibits continual reinforcement learning
di: Seth, Purab, et al.
Pubblicazione: (2026) -
Cross-environment Cooperation Enables Zero-shot Multi-agent Coordination
di: Jha, Kunal, et al.
Pubblicazione: (2025) -
Naturalistic Computational Cognitive Science: Towards generalizable models and theories that capture the full range of natural behavior
di: Carvalho, Wilka, et al.
Pubblicazione: (2025) -
Online library learning in human visual puzzle solving
di: Zhao, Pinzhe, et al.
Pubblicazione: (2026)