Hyperagents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Jenny, Zhao, Bingchen, Yang, Wannan, Foerster, Jakob, Clune, Jeff, Jiang, Minqi, Devlin, Sam, Shavrina, Tatiana |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
APRES: An Agentic Paper Revision and Evaluation System
von: Zhao, Bingchen, et al.
Veröffentlicht: (2026)
von: Zhao, Bingchen, et al.
Veröffentlicht: (2026)
OMNI-EPIC: Open-endedness via Models of human Notions of Interestingness with Environments Programmed in Code
von: Faldor, Maxence, et al.
Veröffentlicht: (2024)
von: Faldor, Maxence, et al.
Veröffentlicht: (2024)
OMNI: Open-endedness via Models of human Notions of Interestingness
von: Zhang, Jenny, et al.
Veröffentlicht: (2023)
von: Zhang, Jenny, et al.
Veröffentlicht: (2023)
The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
von: Lu, Chris, et al.
Veröffentlicht: (2024)
von: Lu, Chris, et al.
Veröffentlicht: (2024)
Darwin Godel Machine: Open-Ended Evolution of Self-Improving Agents
von: Zhang, Jenny, et al.
Veröffentlicht: (2025)
von: Zhang, Jenny, et al.
Veröffentlicht: (2025)
Thought Cloning: Learning to Think while Acting by Imitating Human Thinking
von: Hu, Shengran, et al.
Veröffentlicht: (2023)
von: Hu, Shengran, et al.
Veröffentlicht: (2023)
First-Explore, then Exploit: Meta-Learning to Solve Hard Exploration-Exploitation Trade-Offs
von: Norman, Ben, et al.
Veröffentlicht: (2023)
von: Norman, Ben, et al.
Veröffentlicht: (2023)
Asking the Right Questions: Improving Reasoning with Generated Stepping Stones
von: Hu, Hengyuan, et al.
Veröffentlicht: (2026)
von: Hu, Hengyuan, et al.
Veröffentlicht: (2026)
Quality Diversity through Human Feedback: Towards Open-Ended Diversity-Driven Optimization
von: Ding, Li, et al.
Veröffentlicht: (2023)
von: Ding, Li, et al.
Veröffentlicht: (2023)
Learning to Continually Learn via Meta-learning Agentic Memory Designs
von: Xiong, Yiming, et al.
Veröffentlicht: (2026)
von: Xiong, Yiming, et al.
Veröffentlicht: (2026)
Automated Design of Agentic Systems
von: Hu, Shengran, et al.
Veröffentlicht: (2024)
von: Hu, Shengran, et al.
Veröffentlicht: (2024)
Refining Minimax Regret for Unsupervised Environment Design
von: Beukman, Michael, et al.
Veröffentlicht: (2024)
von: Beukman, Michael, et al.
Veröffentlicht: (2024)
The AI Scientist-v2: Workshop-Level Automated Scientific Discovery via Agentic Tree Search
von: Yamada, Yutaro, et al.
Veröffentlicht: (2025)
von: Yamada, Yutaro, et al.
Veröffentlicht: (2025)
Foundation Model Self-Play: Open-Ended Strategy Innovation via Foundation Models
von: Dharna, Aaron, et al.
Veröffentlicht: (2025)
von: Dharna, Aaron, et al.
Veröffentlicht: (2025)
AI & Human Co-Improvement for Safer Co-Superintelligence
von: Weston, Jason, et al.
Veröffentlicht: (2025)
von: Weston, Jason, et al.
Veröffentlicht: (2025)
Automated Capability Discovery via Foundation Model Self-Exploration
von: Lu, Cong, et al.
Veröffentlicht: (2025)
von: Lu, Cong, et al.
Veröffentlicht: (2025)
Intelligent Go-Explore: Standing on the Shoulders of Giant Foundation Models
von: Lu, Cong, et al.
Veröffentlicht: (2024)
von: Lu, Cong, et al.
Veröffentlicht: (2024)
Learning to Act without Actions
von: Schmidt, Dominik, et al.
Veröffentlicht: (2023)
von: Schmidt, Dominik, et al.
Veröffentlicht: (2023)
The Automated LLM Speedrunning Benchmark: Reproducing NanoGPT Improvements
von: Zhao, Bingchen, et al.
Veröffentlicht: (2025)
von: Zhao, Bingchen, et al.
Veröffentlicht: (2025)
Continual learning under domain transfer with sparse synaptic bursting
von: Beaulieu, Shawn L., et al.
Veröffentlicht: (2021)
von: Beaulieu, Shawn L., et al.
Veröffentlicht: (2021)
PARDEN, Can You Repeat That? Defending against Jailbreaks via Repetition
von: Zhang, Ziyang, et al.
Veröffentlicht: (2024)
von: Zhang, Ziyang, et al.
Veröffentlicht: (2024)
JaxUED: A simple and useable UED library in Jax
von: Coward, Samuel, et al.
Veröffentlicht: (2024)
von: Coward, Samuel, et al.
Veröffentlicht: (2024)
QuanForge: A Mutation Testing Framework for Quantum Neural Networks
von: Shao, Minqi, et al.
Veröffentlicht: (2026)
von: Shao, Minqi, et al.
Veröffentlicht: (2026)
A Model-Based Solution to the Offline Multi-Agent Reinforcement Learning Coordination Problem
von: Barde, Paul, et al.
Veröffentlicht: (2023)
von: Barde, Paul, et al.
Veröffentlicht: (2023)
Kinetix: Investigating the Training of General Agents through Open-Ended Physics-Based Control Tasks
von: Matthews, Michael, et al.
Veröffentlicht: (2024)
von: Matthews, Michael, et al.
Veröffentlicht: (2024)
AQA-Bench: An Interactive Benchmark for Evaluating LLMs' Sequential Reasoning Ability
von: Yang, Siwei, et al.
Veröffentlicht: (2024)
von: Yang, Siwei, et al.
Veröffentlicht: (2024)
AgentBreeder: Mitigating the AI Safety Risks of Multi-Agent Scaffolds via Self-Improvement
von: Rosser, J, et al.
Veröffentlicht: (2025)
von: Rosser, J, et al.
Veröffentlicht: (2025)
JaxLife: An Open-Ended Agentic Simulator
von: Lu, Chris, et al.
Veröffentlicht: (2024)
von: Lu, Chris, et al.
Veröffentlicht: (2024)
Oasis: One Image is All You Need for Multimodal Instruction Data Synthesis
von: Zhang, Letian, et al.
Veröffentlicht: (2025)
von: Zhang, Letian, et al.
Veröffentlicht: (2025)
The Generalization Gap in Offline Reinforcement Learning
von: Mediratta, Ishita, et al.
Veröffentlicht: (2023)
von: Mediratta, Ishita, et al.
Veröffentlicht: (2023)
minimax: Efficient Baselines for Autocurricula in JAX
von: Jiang, Minqi, et al.
Veröffentlicht: (2023)
von: Jiang, Minqi, et al.
Veröffentlicht: (2023)
Artificial Generational Intelligence: Cultural Accumulation in Reinforcement Learning
von: Cook, Jonathan, et al.
Veröffentlicht: (2024)
von: Cook, Jonathan, et al.
Veröffentlicht: (2024)
The Edge-of-Reach Problem in Offline Model-Based Reinforcement Learning
von: Sims, Anya, et al.
Veröffentlicht: (2024)
von: Sims, Anya, et al.
Veröffentlicht: (2024)
Mirror Learning: A Unifying Framework of Policy Optimisation
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2022)
von: Kuba, Jakub Grudzien, et al.
Veröffentlicht: (2022)
Learning Multi-Agent Communication with Contrastive Learning
von: Lo, Yat Long, et al.
Veröffentlicht: (2023)
von: Lo, Yat Long, et al.
Veröffentlicht: (2023)
The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind
von: Lupu, Andrei, et al.
Veröffentlicht: (2025)
von: Lupu, Andrei, et al.
Veröffentlicht: (2025)
Local Descriptors Weighted Adaptive Threshold Filtering For Few-Shot Learning
von: Yan, Bingchen
Veröffentlicht: (2024)
von: Yan, Bingchen
Veröffentlicht: (2024)
AI Research Agents for Machine Learning: Search, Exploration, and Generalization in MLE-bench
von: Toledo, Edan, et al.
Veröffentlicht: (2025)
von: Toledo, Edan, et al.
Veröffentlicht: (2025)
Hallucination reduction with CASAL: Contrastive Activation Steering For Amortized Learning
von: Wannan, et al.
Veröffentlicht: (2025)
von: Wannan, et al.
Veröffentlicht: (2025)
SpecBench: Measuring Reward Hacking in Long-Horizon Coding Agents
von: Zhao, Bingchen, et al.
Veröffentlicht: (2026)
von: Zhao, Bingchen, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
APRES: An Agentic Paper Revision and Evaluation System
von: Zhao, Bingchen, et al.
Veröffentlicht: (2026) -
OMNI-EPIC: Open-endedness via Models of human Notions of Interestingness with Environments Programmed in Code
von: Faldor, Maxence, et al.
Veröffentlicht: (2024) -
OMNI: Open-endedness via Models of human Notions of Interestingness
von: Zhang, Jenny, et al.
Veröffentlicht: (2023) -
The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
von: Lu, Chris, et al.
Veröffentlicht: (2024) -
Darwin Godel Machine: Open-Ended Evolution of Self-Improving Agents
von: Zhang, Jenny, et al.
Veröffentlicht: (2025)