Rainbow Teaming: Open-Ended Generation of Diverse Adversarial Prompts
Fuente:
arXiv
Salvato in:
| Autori principali: | Samvelyan, Mikayel, Raparthy, Sharath Chandra, Lupu, Andrei, Hambro, Eric, Markosyan, Aram H., Bhatt, Manish, Mao, Yuning, Jiang, Minqi, Parker-Holder, Jack, Foerster, Jakob, Rocktäschel, Tim, Raileanu, Roberta |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Multi-Agent Diagnostics for Robustness via Illuminated Diversity
di: Samvelyan, Mikayel, et al.
Pubblicazione: (2024)
di: Samvelyan, Mikayel, et al.
Pubblicazione: (2024)
Robust Agents in Open-Ended Worlds
di: Samvelyan, Mikayel
Pubblicazione: (2025)
di: Samvelyan, Mikayel
Pubblicazione: (2025)
DéjàQ: Open-Ended Evolution of Diverse, Learnable and Verifiable Problems
di: Röpke, Willem, et al.
Pubblicazione: (2026)
di: Röpke, Willem, et al.
Pubblicazione: (2026)
GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements
di: Havrilla, Alex, et al.
Pubblicazione: (2024)
di: Havrilla, Alex, et al.
Pubblicazione: (2024)
Craftax: A Lightning-Fast Benchmark for Open-Ended Reinforcement Learning
di: Matthews, Michael, et al.
Pubblicazione: (2024)
di: Matthews, Michael, et al.
Pubblicazione: (2024)
Teaching Large Language Models to Reason with Reinforcement Learning
di: Havrilla, Alex, et al.
Pubblicazione: (2024)
di: Havrilla, Alex, et al.
Pubblicazione: (2024)
LLM-First Search: Self-Guided Exploration of the Solution Space
di: Herr, Nathan, et al.
Pubblicazione: (2025)
di: Herr, Nathan, et al.
Pubblicazione: (2025)
The Decrypto Benchmark for Multi-Agent Reasoning and Theory of Mind
di: Lupu, Andrei, et al.
Pubblicazione: (2025)
di: Lupu, Andrei, et al.
Pubblicazione: (2025)
Open-Endedness is Essential for Artificial Superhuman Intelligence
di: Hughes, Edward, et al.
Pubblicazione: (2024)
di: Hughes, Edward, et al.
Pubblicazione: (2024)
Outliers and Calibration Sets have Diminishing Effect on Quantization of Modern LLMs
di: Paglieri, Davide, et al.
Pubblicazione: (2024)
di: Paglieri, Davide, et al.
Pubblicazione: (2024)
SPARQ: Synthetic Problem Generation for Reasoning via Quality-Diversity Algorithms
di: Havrilla, Alex, et al.
Pubblicazione: (2025)
di: Havrilla, Alex, et al.
Pubblicazione: (2025)
Bootstrapping Task Spaces for Self-Improvement
di: Jiang, Minqi, et al.
Pubblicazione: (2025)
di: Jiang, Minqi, et al.
Pubblicazione: (2025)
Imagined Autocurricula
di: Güzel, Ahmet H., et al.
Pubblicazione: (2025)
di: Güzel, Ahmet H., et al.
Pubblicazione: (2025)
CURATe: Benchmarking Personalised Alignment of Conversational AI Assistants
di: Alberts, Lize, et al.
Pubblicazione: (2024)
di: Alberts, Lize, et al.
Pubblicazione: (2024)
The Generalization Gap in Offline Reinforcement Learning
di: Mediratta, Ishita, et al.
Pubblicazione: (2023)
di: Mediratta, Ishita, et al.
Pubblicazione: (2023)
Learning When to Plan: Efficiently Allocating Test-Time Compute for LLM Agents
di: Paglieri, Davide, et al.
Pubblicazione: (2025)
di: Paglieri, Davide, et al.
Pubblicazione: (2025)
JaxMARL: Multi-Agent RL Environments and Algorithms in JAX
di: Rutherford, Alexander, et al.
Pubblicazione: (2023)
di: Rutherford, Alexander, et al.
Pubblicazione: (2023)
Understanding the Effects of RLHF on LLM Generalisation and Diversity
di: Kirk, Robert, et al.
Pubblicazione: (2023)
di: Kirk, Robert, et al.
Pubblicazione: (2023)
Behaviour Distillation
di: Lupu, Andrei, et al.
Pubblicazione: (2024)
di: Lupu, Andrei, et al.
Pubblicazione: (2024)
JaxLife: An Open-Ended Agentic Simulator
di: Lu, Chris, et al.
Pubblicazione: (2024)
di: Lu, Chris, et al.
Pubblicazione: (2024)
TICKing All the Boxes: Generated Checklists Improve LLM Evaluation and Generation
di: Cook, Jonathan, et al.
Pubblicazione: (2024)
di: Cook, Jonathan, et al.
Pubblicazione: (2024)
minimax: Efficient Baselines for Autocurricula in JAX
di: Jiang, Minqi, et al.
Pubblicazione: (2023)
di: Jiang, Minqi, et al.
Pubblicazione: (2023)
OvercookedV2: Rethinking Overcooked for Zero-Shot Coordination
di: Gessler, Tobias, et al.
Pubblicazione: (2025)
di: Gessler, Tobias, et al.
Pubblicazione: (2025)
Discovering Minimal Reinforcement Learning Environments
di: Liesen, Jarek, et al.
Pubblicazione: (2024)
di: Liesen, Jarek, et al.
Pubblicazione: (2024)
BALROG: Benchmarking Agentic LLM and VLM Reasoning On Games
di: Paglieri, Davide, et al.
Pubblicazione: (2024)
di: Paglieri, Davide, et al.
Pubblicazione: (2024)
Kinetix: Investigating the Training of General Agents through Open-Ended Physics-Based Control Tasks
di: Matthews, Michael, et al.
Pubblicazione: (2024)
di: Matthews, Michael, et al.
Pubblicazione: (2024)
Evolving Many Worlds: Towards Open-Ended Discovery in Petri Dish NCA via Population-Based Training
di: Berdica, Uljad, et al.
Pubblicazione: (2026)
di: Berdica, Uljad, et al.
Pubblicazione: (2026)
Source2Synth: Synthetic Data Generation and Curation Grounded in Real Data Sources
di: Lupidi, Alisia, et al.
Pubblicazione: (2024)
di: Lupidi, Alisia, et al.
Pubblicazione: (2024)
Adam on Local Time: Addressing Nonstationarity in RL with Relative Adam Timesteps
di: Ellis, Benjamin, et al.
Pubblicazione: (2024)
di: Ellis, Benjamin, et al.
Pubblicazione: (2024)
Multi-Agent Craftax: Benchmarking Open-Ended Multi-Agent Reinforcement Learning at the Hyperscale
di: Omari, Bassel Al, et al.
Pubblicazione: (2025)
di: Omari, Bassel Al, et al.
Pubblicazione: (2025)
Bailouts and Redistribution
di: Sukiasyan, Mikayel
Pubblicazione: (2025)
di: Sukiasyan, Mikayel
Pubblicazione: (2025)
Programming by Backprop: An Instruction is Worth 100 Examples When Finetuning LLMs
di: Cook, Jonathan, et al.
Pubblicazione: (2025)
di: Cook, Jonathan, et al.
Pubblicazione: (2025)
The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery
di: Lu, Chris, et al.
Pubblicazione: (2024)
di: Lu, Chris, et al.
Pubblicazione: (2024)
Predictive Coding and Information Bottleneck for Hallucination Detection in Large Language Models
di: Bhatt, Manish
Pubblicazione: (2026)
di: Bhatt, Manish
Pubblicazione: (2026)
Bhatt Conjectures: On Necessary-But-Not-Sufficient Benchmark Tautology for Human Like Reasoning
di: Bhatt, Manish
Pubblicazione: (2025)
di: Bhatt, Manish
Pubblicazione: (2025)
Synthetic Data is Sufficient for Zero-Shot Visual Generalization from Offline Data
di: Güzel, Ahmet H., et al.
Pubblicazione: (2025)
di: Güzel, Ahmet H., et al.
Pubblicazione: (2025)
Epistemic Dissonance and Modal Boundaries
di: Raileanu, Dragos
Pubblicazione: (2025)
di: Raileanu, Dragos
Pubblicazione: (2025)
Scaling Opponent Shaping to High Dimensional Games
di: Khan, Akbir, et al.
Pubblicazione: (2023)
di: Khan, Akbir, et al.
Pubblicazione: (2023)
Large Empirical Case Study: Go-Explore adapted for AI Red Team Testing
di: Bhatt, Manish, et al.
Pubblicazione: (2025)
di: Bhatt, Manish, et al.
Pubblicazione: (2025)
Asking the Right Questions: Improving Reasoning with Generated Stepping Stones
di: Hu, Hengyuan, et al.
Pubblicazione: (2026)
di: Hu, Hengyuan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Multi-Agent Diagnostics for Robustness via Illuminated Diversity
di: Samvelyan, Mikayel, et al.
Pubblicazione: (2024) -
Robust Agents in Open-Ended Worlds
di: Samvelyan, Mikayel
Pubblicazione: (2025) -
DéjàQ: Open-Ended Evolution of Diverse, Learnable and Verifiable Problems
di: Röpke, Willem, et al.
Pubblicazione: (2026) -
GLoRe: When, Where, and How to Improve LLM Reasoning via Global and Local Refinements
di: Havrilla, Alex, et al.
Pubblicazione: (2024) -
Craftax: A Lightning-Fast Benchmark for Open-Ended Reinforcement Learning
di: Matthews, Michael, et al.
Pubblicazione: (2024)