Downloads · 30 days
0
chandar-lab/r3d2_agent
r3d2_agent is a machine learning model from chandar-lab. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
Recurrent Replay Relevance Distributed DQN (R3D2) is a generalist multi-agent reinforcement learning (MARL) agent designed to play Hanabi across all game settings while adapting to unfamiliar collaborators. Unlike tra…
Downloads · 30 days
0
Access
Public
Updated Mar 27, 2025
Repo size
107 MB
Likes
0
Public
Click a slice to open those files.
.pthw107 MB · 100%
From the Hugging Face model README
Recurrent Replay Relevance Distributed DQN (R3D2) is a generalist multi-agent reinforcement learning (MARL) agent designed to play Hanabi across all game settings while adapting to unfamiliar collaborators. Unlike traditional MARL agents that struggle with transferability and cooperation beyond their training setting, R3D2 utilizes language-based reformulation and a distributed learning approach to handle dynamic observation and action spaces. This allows it to generalize across different game configurations and effectively collaborate with diverse algorithmic agents.
Generalized MARL agent: Play Hanabi across different player settings (2-to-5 players) without changing architecture or retraining from scratch.
Adaptive cooperation: Capable of collaborating with unfamiliar partners, overcoming limitations of traditional MARL systems.
Language-based task reformulation: Utilizes text representations to enhance transfer learning and generalization.
Distributed Learning Framework: Employs a scalable MARL algorithm to handle dynamic observations and actions effectively.

Follow the steps here: