Learning Multi-Agent Communication with Contrastive Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Lo, Yat Long, Sengupta, Biswa, Foerster, Jakob, Noukhovitch, Michael |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Representation Without Reward: A JEPA Audit for LLM Fine-Tuning
by: Sengupta, Biswa
Published: (2026)
by: Sengupta, Biswa
Published: (2026)
HARBOR: Automated Harness Optimization
by: Sengupta, Biswa, et al.
Published: (2026)
by: Sengupta, Biswa, et al.
Published: (2026)
JPmHC Dynamical Isometry via Orthogonal Hyper-Connections
by: Sengupta, Biswa, et al.
Published: (2026)
by: Sengupta, Biswa, et al.
Published: (2026)
A Model-Based Solution to the Offline Multi-Agent Reinforcement Learning Coordination Problem
by: Barde, Paul, et al.
Published: (2023)
by: Barde, Paul, et al.
Published: (2023)
Goal-Conditioned Agents that Learn Everything All at Once
by: Matthews, Michael, et al.
Published: (2026)
by: Matthews, Michael, et al.
Published: (2026)
Kinetix: Investigating the Training of General Agents through Open-Ended Physics-Based Control Tasks
by: Matthews, Michael, et al.
Published: (2024)
by: Matthews, Michael, et al.
Published: (2024)
Gradient Regularization Prevents Reward Hacking in Reinforcement Learning from Human Feedback and Verifiable Rewards
by: Ackermann, Johannes, et al.
Published: (2026)
by: Ackermann, Johannes, et al.
Published: (2026)
JaxUED: A simple and useable UED library in Jax
by: Coward, Samuel, et al.
Published: (2024)
by: Coward, Samuel, et al.
Published: (2024)
Mirror Learning: A Unifying Framework of Policy Optimisation
by: Kuba, Jakub Grudzien, et al.
Published: (2022)
by: Kuba, Jakub Grudzien, et al.
Published: (2022)
The Edge-of-Reach Problem in Offline Model-Based Reinforcement Learning
by: Sims, Anya, et al.
Published: (2024)
by: Sims, Anya, et al.
Published: (2024)
Abstraction for Offline Goal-Conditioned Reinforcement Learning
by: Wibault, Clarisse, et al.
Published: (2026)
by: Wibault, Clarisse, et al.
Published: (2026)
How Should We Meta-Learn Reinforcement Learning Algorithms?
by: Goldie, Alexander David, et al.
Published: (2025)
by: Goldie, Alexander David, et al.
Published: (2025)
DITTO: Offline Imitation Learning with World Models
by: DeMoss, Branton, et al.
Published: (2023)
by: DeMoss, Branton, et al.
Published: (2023)
Can Learned Optimization Make Reinforcement Learning Less Difficult?
by: Goldie, Alexander David, et al.
Published: (2024)
by: Goldie, Alexander David, et al.
Published: (2024)
Recurrent Reinforcement Learning with Memoroids
by: Morad, Steven, et al.
Published: (2024)
by: Morad, Steven, et al.
Published: (2024)
Meta-Learning Objectives for Preference Optimization
by: Alfano, Carlo, et al.
Published: (2024)
by: Alfano, Carlo, et al.
Published: (2024)
SOReL and TOReL: Two Methods for Fully Offline Reinforcement Learning
by: Fellows, Mattie, et al.
Published: (2025)
by: Fellows, Mattie, et al.
Published: (2025)
Learning to Reason at the Frontier of Learnability
by: Foster, Thomas, et al.
Published: (2025)
by: Foster, Thomas, et al.
Published: (2025)
Compositional Discrete Latent Code for High Fidelity, Productive Diffusion Models
by: Lavoie, Samuel, et al.
Published: (2025)
by: Lavoie, Samuel, et al.
Published: (2025)
Rethinking Out-of-Distribution Detection for Reinforcement Learning: Advancing Methods for Evaluation and Detection
by: Nasvytis, Linas, et al.
Published: (2024)
by: Nasvytis, Linas, et al.
Published: (2024)
A Clean Slate for Offline Reinforcement Learning
by: Jackson, Matthew Thomas, et al.
Published: (2025)
by: Jackson, Matthew Thomas, et al.
Published: (2025)
The Yokai Learning Environment: Tracking Beliefs Over Space and Time
by: Ruhdorfer, Constantin, et al.
Published: (2025)
by: Ruhdorfer, Constantin, et al.
Published: (2025)
Discovering Temporally-Aware Reinforcement Learning Algorithms
by: Jackson, Matthew Thomas, et al.
Published: (2024)
by: Jackson, Matthew Thomas, et al.
Published: (2024)
From Translation to Superset: Benchmark-Driven Evolution of a Production AI Agent from Rust to Python
by: Wang, Jinhua, et al.
Published: (2026)
by: Wang, Jinhua, et al.
Published: (2026)
Dual-Kernel Graph Community Contrastive Learning
by: Chen, Xiang, et al.
Published: (2025)
by: Chen, Xiang, et al.
Published: (2025)
Multi-Agent Deep Reinforcement Learning Under Constrained Communications
by: Shaik, Shahil, et al.
Published: (2026)
by: Shaik, Shahil, et al.
Published: (2026)
Generative Image as Action Models
by: Shridhar, Mohit, et al.
Published: (2024)
by: Shridhar, Mohit, et al.
Published: (2024)
Active Learning for Communication Structure Optimization in LLM-Based Multi-Agent Systems
by: Yang, Huchen, et al.
Published: (2026)
by: Yang, Huchen, et al.
Published: (2026)
Why the Agent Made that Decision: Contrastive Explanation Learning for Reinforcement Learning
by: Zuo, Rui, et al.
Published: (2024)
by: Zuo, Rui, et al.
Published: (2024)
Learning Representations in Video Game Agents with Supervised Contrastive Imitation Learning
by: Celemin, Carlos, et al.
Published: (2025)
by: Celemin, Carlos, et al.
Published: (2025)
Refining Minimax Regret for Unsupervised Environment Design
by: Beukman, Michael, et al.
Published: (2024)
by: Beukman, Michael, et al.
Published: (2024)
ProtoECGNet: Case-Based Interpretable Deep Learning for Multi-Label ECG Classification with Contrastive Learning
by: Sethi, Sahil, et al.
Published: (2025)
by: Sethi, Sahil, et al.
Published: (2025)
Tensor-Fused Multi-View Graph Contrastive Learning
by: Wu, Yujia, et al.
Published: (2024)
by: Wu, Yujia, et al.
Published: (2024)
Positive Unlabeled Contrastive Learning
by: Acharya, Anish, et al.
Published: (2022)
by: Acharya, Anish, et al.
Published: (2022)
Symmetry-Breaking Augmentations for Ad Hoc Teamwork
by: Hammond, Ravi, et al.
Published: (2024)
by: Hammond, Ravi, et al.
Published: (2024)
Advancing Wildfire Risk Prediction via Morphology-Aware Curriculum Contrastive Learning
by: Scudo, Fabrizio Lo, et al.
Published: (2025)
by: Scudo, Fabrizio Lo, et al.
Published: (2025)
Pun Intended: Multi-Agent Translation of Wordplay with Contrastive Learning and Phonetic-Semantic Embeddings
by: Taylor, Russell, et al.
Published: (2025)
by: Taylor, Russell, et al.
Published: (2025)
Asynchronous RLHF: Faster and More Efficient Off-Policy RL for Language Models
by: Noukhovitch, Michael, et al.
Published: (2024)
by: Noukhovitch, Michael, et al.
Published: (2024)
Behaviour Distillation
by: Lupu, Andrei, et al.
Published: (2024)
by: Lupu, Andrei, et al.
Published: (2024)
Learning to Communicate Locally for Large-Scale Multi-Agent Pathfinding
by: Vyaltsev, Valeriy, et al.
Published: (2026)
by: Vyaltsev, Valeriy, et al.
Published: (2026)
Similar Items
-
Representation Without Reward: A JEPA Audit for LLM Fine-Tuning
by: Sengupta, Biswa
Published: (2026) -
HARBOR: Automated Harness Optimization
by: Sengupta, Biswa, et al.
Published: (2026) -
JPmHC Dynamical Isometry via Orthogonal Hyper-Connections
by: Sengupta, Biswa, et al.
Published: (2026) -
A Model-Based Solution to the Offline Multi-Agent Reinforcement Learning Coordination Problem
by: Barde, Paul, et al.
Published: (2023) -
Goal-Conditioned Agents that Learn Everything All at Once
by: Matthews, Michael, et al.
Published: (2026)