LuckyMera: a Modular AI Framework for Building Hybrid NetHack Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Quarantiello, Luigi, Marzeddu, Simone, Guzzi, Antonio, Lomonaco, Vincenzo |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Task-Agnostic Experts Composition for Continual Learning
by: Quarantiello, Luigi, et al.
Published: (2025)
by: Quarantiello, Luigi, et al.
Published: (2025)
Parameter-Efficient Continual Fine-Tuning: A Survey
by: Coleman, Eric Nuertey, et al.
Published: (2025)
by: Coleman, Eric Nuertey, et al.
Published: (2025)
Playing NetHack with LLMs: Potential & Limitations as Zero-Shot Agents
by: Jeurissen, Dominik, et al.
Published: (2024)
by: Jeurissen, Dominik, et al.
Published: (2024)
The Future of Continual Learning in the Era of Foundation Models: Three Key Directions
by: Bell, Jack, et al.
Published: (2025)
by: Bell, Jack, et al.
Published: (2025)
HAM: Hierarchical Adapter Merging for Scalable Continual Learning
by: Coleman, Eric Nuertey, et al.
Published: (2025)
by: Coleman, Eric Nuertey, et al.
Published: (2025)
Modular Memory is the Key to Continual Learning Agents
by: Dorovatas, Vaggelis, et al.
Published: (2026)
by: Dorovatas, Vaggelis, et al.
Published: (2026)
Continual Policy Distillation of Reinforcement Learning-based Controllers for Soft Robotic In-Hand Manipulation
by: Li, Lanpei, et al.
Published: (2024)
by: Li, Lanpei, et al.
Published: (2024)
Combining Pre-Trained Models for Enhanced Feature Representation in Reinforcement Learning
by: Piccoli, Elia, et al.
Published: (2025)
by: Piccoli, Elia, et al.
Published: (2025)
Calibration of Continual Learning Models
by: Li, Lanpei, et al.
Published: (2024)
by: Li, Lanpei, et al.
Published: (2024)
Hack-Verifiable Environments: Towards Evaluating Reward Hacking at Scale
by: Roth, Amit, et al.
Published: (2026)
by: Roth, Amit, et al.
Published: (2026)
Reward Hacking Benchmark: Measuring Exploits in LLM Agents with Tool Use
by: Thaman, Kunvar
Published: (2026)
by: Thaman, Kunvar
Published: (2026)
Mitigating Preference Hacking in Policy Optimization with Pessimism
by: Gupta, Dhawal, et al.
Published: (2025)
by: Gupta, Dhawal, et al.
Published: (2025)
Book your room in the Turing Hotel! A symmetric and distributed Turing Test with multiple AIs and humans
by: Di Maio, Christian, et al.
Published: (2026)
by: Di Maio, Christian, et al.
Published: (2026)
Reward Hacking Mitigation using Verifiable Composite Rewards
by: Tarek, Mirza Farhan Bin, et al.
Published: (2025)
by: Tarek, Mirza Farhan Bin, et al.
Published: (2025)
Repairing Reward Functions with Feedback to Mitigate Reward Hacking
by: Hatgis-Kessell, Stephane, et al.
Published: (2025)
by: Hatgis-Kessell, Stephane, et al.
Published: (2025)
HyCARD-Net: A Synergistic Hybrid Intelligence Framework for Cardiovascular Disease Diagnosis
by: Gupta, Rajan Das, et al.
Published: (2026)
by: Gupta, Rajan Das, et al.
Published: (2026)
LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking
by: Helff, Lukas, et al.
Published: (2026)
by: Helff, Lukas, et al.
Published: (2026)
Adversarial Reward Auditing for Active Detection and Mitigation of Reward Hacking
by: Beigi, Mohammad, et al.
Published: (2026)
by: Beigi, Mohammad, et al.
Published: (2026)
HyBattNet: Hybrid Framework for Predicting the Remaining Useful Life of Lithium-Ion Batteries
by: Tran, Khoa, et al.
Published: (2025)
by: Tran, Khoa, et al.
Published: (2025)
xAI-Drop: Don't Use What You Cannot Explain
by: De Luca, Vincenzo Marco, et al.
Published: (2024)
by: De Luca, Vincenzo Marco, et al.
Published: (2024)
Architectures for Building Agentic AI
by: Nowaczyk, Sławomir
Published: (2025)
by: Nowaczyk, Sławomir
Published: (2025)
Correlated Proxies: A New Definition and Improved Mitigation for Reward Hacking
by: Laidlaw, Cassidy, et al.
Published: (2024)
by: Laidlaw, Cassidy, et al.
Published: (2024)
HackAtari: Atari Learning Environments for Robust and Continual Reinforcement Learning
by: Delfosse, Quentin, et al.
Published: (2024)
by: Delfosse, Quentin, et al.
Published: (2024)
Beyond Reward Hacking: Causal Rewards for Large Language Model Alignment
by: Wang, Chaoqi, et al.
Published: (2025)
by: Wang, Chaoqi, et al.
Published: (2025)
On Teacher Hacking in Language Model Distillation
by: Tiapkin, Daniil, et al.
Published: (2025)
by: Tiapkin, Daniil, et al.
Published: (2025)
FedModule: A Modular Federated Learning Framework
by: Chen, Chuyi, et al.
Published: (2024)
by: Chen, Chuyi, et al.
Published: (2024)
BuilderBench: The Building Blocks of Intelligent Agents
by: Ghugare, Raj, et al.
Published: (2025)
by: Ghugare, Raj, et al.
Published: (2025)
ODIN: Disentangled Reward Mitigates Hacking in RLHF
by: Chen, Lichang, et al.
Published: (2024)
by: Chen, Lichang, et al.
Published: (2024)
Reward Shaping to Mitigate Reward Hacking in RLHF
by: Fu, Jiayi, et al.
Published: (2025)
by: Fu, Jiayi, et al.
Published: (2025)
IR$^3$: Contrastive Inverse Reinforcement Learning for Interpretable Detection and Mitigation of Reward Hacking
by: Beigi, Mohammad, et al.
Published: (2026)
by: Beigi, Mohammad, et al.
Published: (2026)
Honesty to Subterfuge: In-Context Reinforcement Learning Can Make Honest Models Reward Hack
by: McKee-Reid, Leo, et al.
Published: (2024)
by: McKee-Reid, Leo, et al.
Published: (2024)
InfoRM: Mitigating Reward Hacking in RLHF via Information-Theoretic Reward Modeling
by: Miao, Yuchun, et al.
Published: (2024)
by: Miao, Yuchun, et al.
Published: (2024)
ClawGym: A Scalable Framework for Building Effective Claw Agents
by: Bai, Fei, et al.
Published: (2026)
by: Bai, Fei, et al.
Published: (2026)
A Compositional Paradigm for Foundation Models: Towards Smarter Robotic Agents
by: Quarantiello, Luigi, et al.
Published: (2025)
by: Quarantiello, Luigi, et al.
Published: (2025)
Harness as an Asset: Enforcing Determinism via the Convergent AI Agent Framework (CAAF)
by: Zhang, Tianbao
Published: (2026)
by: Zhang, Tianbao
Published: (2026)
MONA: Myopic Optimization with Non-myopic Approval Can Mitigate Multi-step Reward Hacking
by: Farquhar, Sebastian, et al.
Published: (2025)
by: Farquhar, Sebastian, et al.
Published: (2025)
Configurable Foundation Models: Building LLMs from a Modular Perspective
by: Xiao, Chaojun, et al.
Published: (2024)
by: Xiao, Chaojun, et al.
Published: (2024)
Fairness Hacking: The Malicious Practice of Shrouding Unfairness in Algorithms
by: Meding, Kristof, et al.
Published: (2023)
by: Meding, Kristof, et al.
Published: (2023)
NetArena: Dynamic Benchmarks for AI Agents in Network Automation
by: Zhou, Yajie, et al.
Published: (2025)
by: Zhou, Yajie, et al.
Published: (2025)
AttentionSmithy: A Modular Framework for Rapid Transformer Development and Customization
by: Cranney, Caleb, et al.
Published: (2025)
by: Cranney, Caleb, et al.
Published: (2025)
Similar Items
-
Task-Agnostic Experts Composition for Continual Learning
by: Quarantiello, Luigi, et al.
Published: (2025) -
Parameter-Efficient Continual Fine-Tuning: A Survey
by: Coleman, Eric Nuertey, et al.
Published: (2025) -
Playing NetHack with LLMs: Potential & Limitations as Zero-Shot Agents
by: Jeurissen, Dominik, et al.
Published: (2024) -
The Future of Continual Learning in the Era of Foundation Models: Three Key Directions
by: Bell, Jack, et al.
Published: (2025) -
HAM: Hierarchical Adapter Merging for Scalable Continual Learning
by: Coleman, Eric Nuertey, et al.
Published: (2025)