open_spiel vs stable-baselines3

open_spiel

OpenSpiel is a collection of environments and algorithms for research in general reinforcement learning and search/planning in games. (by google-deepmind)

Source Code

Suggest alternative

Edit details

stable-baselines3

PyTorch version of Stable Baselines, reliable implementations of reinforcement learning algorithms. (by DLR-RM)

reinforcement-learning reinforcement-learning-algorithms Machine Learning Gym openai baselines Toolbox stable-baselines Python Pytorch Robotics Sde gsde sb3

Source Code

stable-baselines3.readthedocs.io

Suggest alternative

Edit details

Our great sponsors

InfluxDB - Power Real-Time Data Analytics at Scale

WorkOS - The modern identity platform for B2B SaaS

SaaSHub - Software Alternatives and Reviews

Our great sponsors

open_spiel		stable-baselines3
	Project
44	Mentions	46
3,999	Stars	7,953
1.5%	Growth	5.2%
9.5	Activity	8.2
5 days ago	Latest Commit	2 days ago
C++	Language	Python
Apache License 2.0	License	MIT License

The number of mentions indicates the total number of mentions that we've tracked plus the number of user suggested alternatives.
Stars - the number of stars that a project has on GitHub. Growth - month over month growth in stars.
Activity is a relative number indicating how actively a project is being developed. Recent commits have higher weight than older ones.
For example, an activity of 9.0 indicates that a project is amongst the top 10% of the most actively developed projects that we are tracking.

open_spiel

Posts with mentions or reviews of open_spiel. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2023-05-26.

What projects or open-source contributions can impress Jane Street recruiters for a Quant SWE role ?
1 project | /r/csMajors | 3 Jul 2023

Deep mind actually has a repository where they applied this algorithm for incomplete-knowledge games. You could use it for reference: https://github.com/deepmind/open_spiel/tree/master/open_spiel/python/algorithms
I want to build a learning agent for a combinatorial game
1 project | /r/reinforcementlearning | 24 Jun 2023

+1. You can also find an implementation of Clobber and AlphaZero (and many other basic RL algorithms) in OpenSpiel: https://github.com/deepmind/open_spiel
minimax for imperfect-information turn-games?
1 project | /r/reinforcementlearning | 24 Jun 2023

You can find a lot of code online if you look, and many of these applied to Poker. There's a general implementation of both in Python and C++ in OpenSpiel, with some examples applied to small poker games. It's nice code to learn from because the algorithms operate over generic game descriptions, so there aren't game-specific design choices mixed up with the implementation of the algorithms, and you can create your own poker game and just run them on it.
OpenSpiel 1.3 Released!
1 project | /r/reinforcementlearning | 1 Jun 2023

And many other additions and improvements. See all the details here: https://github.com/deepmind/open_spiel/releases/tag/v1.3
What's a good OpenAI Gym Environment for applying centralized multi-agent learning using expected SARSA with tile coding?
1 project | /r/reinforcementlearning | 30 May 2023

I would checkout the openspiel package. It's main focus is RL in games (multi-agent environments). You'll find RL examples there and games that are small enough to solve without deep RL. There's also a wide range of environments from fully cooperative to adversarial zero-sum.
Competitive reinforcement learning for turn-based games
2 projects | /r/reinforcementlearning | 26 May 2023

Hi, you can check out OpenSpiel: https://github.com/deepmind/open_spiel/
Reinforcement learning and Game Theory a turn-based game
1 project | /r/reinforcementlearning | 8 May 2023

as for algorithms , openspiel repository has few implementations some of these are not related to imperfect information games , and others are not for multiagent environment and others are tabular algorithms .
Shimmy 1.0: Gymnasium & PettingZoo bindings for popular external RL environments
10 projects | /r/reinforcementlearning | 25 Apr 2023

This includes single-agent Gymnasium wrappers for DM Control, DM Lab, Behavior Suite, Arcade Learning Environment, OpenAI Gym V21 & V26. Multi-agent PettingZoo wrappers support DM Control Soccer, OpenSpiel and Melting Pot. For more information, read the release notes here:
How to deal with situations where the RL agent cannot act at every time step?
1 project | /r/reinforcementlearning | 21 Mar 2023

I've had some success using Action Masking - you can refer to here https://github.com/deepmind/open_spiel/blob/120420a74a69354d64c10b51cd129d4587f9f325/open_spiel/python/algorithms/dqn.py but for DQN you need to mask out q values for invalid actions (as well as masking them during prediction). In my case I'm able to place my mask in the observation so can fetch it quite easily during prediction but if that's not possible you could query it from the environment and store it in the replay buffer (like they do in the link I shared)
How to search the game tree with depth-first search?
1 project | /r/reinforcementlearning | 14 Mar 2023

Take a look at this simple implementation: https://github.com/deepmind/open_spiel/blob/master/open_spiel/algorithms/minimax.cc

stable-baselines3

Posts with mentions or reviews of stable-baselines3. We have used some of these posts to build our list of alternatives and similar projects. The last one was on 2023-12-09.

Sim-to-real RL pipeline for open-source wheeled bipeds
2 projects | /r/robotics | 9 Dec 2023

The latest release (v3.0.0) of Upkie's software brings a functional sim-to-real reinforcement learning pipeline based on Stable Baselines3, with standard sim-to-real tricks. The pipeline trains on the Gymnasium environments distributed in upkie.envs (setup: pip install upkie) and is implemented in the PPO balancer. Here is a policy running on an Upkie:
[P] PettingZoo 1.24.0 has been released (including Stable-Baselines3 tutorials)
4 projects | /r/reinforcementlearning | 24 Aug 2023

PettingZoo 1.24.0 is now live! This release includes Python 3.11 support, updated Chess and Hanabi environment versions, and many bugfixes, documentation updates and testing expansions. We are also very excited to announce 3 tutorials using Stable-Baselines3, and a full training script using CleanRL with TensorBoard and WandB.
[Question] Why there is so few algorithms implemented in SB3?
1 project | /r/reinforcementlearning | 22 Jul 2023

I am wondering why there is so few algorithms in Stable Baselines 3 (SB3, https://github.com/DLR-RM/stable-baselines3/tree/master)? I was expecting some algorithms like ICM, HIRO, DIAYN, ... Why there is no model-based, skill-chaining, hierarchical-RL, ... algorithms implemented there?
Stable baselines! Where my people at?
1 project | /r/reinforcementlearning | 5 Jul 2023

Discord is more focused, and they have a page for people who wants to contribute https://github.com/DLR-RM/stable-baselines3/blob/master/CONTRIBUTING.md
SB3 - NotImplementedError: Box([-1. -1. -8.], [1. 1. 8.], (3,), <class 'numpy.float32'>) observation space is not supported
2 projects | /r/reinforcementlearning | 19 Jun 2023

Therefore, I debugged this error to the ReplayBuffer that was imported from `SB3`. This is the problem function -
Exporting an A2C model created with stable-baselines3 to PyTorch
1 project | /r/reinforcementlearning | 5 Jun 2023
Shimmy 1.0: Gymnasium & PettingZoo bindings for popular external RL environments
10 projects | /r/reinforcementlearning | 25 Apr 2023

Have you ever wanted to use dm-control with stable-baselines3? Within Reinforcement learning (RL), a number of APIs are used to implement environments, with limited ability to convert between them. This makes training agents across different APIs highly difficult, and has resulted in a fractured ecosystem.
Stable-Baselines3 v1.8 Release
2 projects | /r/reinforcementlearning | 12 Apr 2023

Changelog: https://github.com/DLR-RM/stable-baselines3/releases/tag/v1.8.0
[P] Reinforcement learning evolutionary hyperparameter optimization - 10x speed up
3 projects | /r/MachineLearning | 24 Mar 2023

Great project! One question though, is there any reason why you are not using existing RL models instead of creating your own, such as stable baselines?
Is stable-baselines3 compatible with gymnasium/gymnasium-robotics?
1 project | /r/reinforcementlearning | 13 Feb 2023

What are some alternatives?

When comparing open_spiel and stable-baselines3 you can also consider the following projects:

muzero-general - MuZero

Ray - Ray is a unified framework for scaling AI and Python applications. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.

PettingZoo - An API standard for multi-agent reinforcement learning environments, with popular reference environments and related utilities

stable-baselines - A fork of OpenAI Baselines, implementations of reinforcement learning algorithms

gym - A toolkit for developing and comparing reinforcement learning algorithms.

Pytorch - Tensors and Dynamic neural networks in Python with strong GPU acceleration

rlcard - Reinforcement Learning / AI Bots in Card (Poker) Games - Blackjack, Leduc, Texas, DouDizhu, Mahjong, UNO.

cleanrl - High-quality single file implementation of Deep Reinforcement Learning algorithms with research-friendly features (PPO, DQN, C51, DDPG, TD3, SAC, PPG)

gym-battleship - Battleship environment for reinforcement learning tasks

tianshou - An elegant PyTorch deep reinforcement learning library.

TexasHoldemSolverJava - A Java implemented Texas holdem and short deck Solver

Super-mario-bros-PPO-pytorch - Proximal Policy Optimization (PPO) algorithm for Super Mario Bros

open_spiel vs muzero-general stable-baselines3 vs Ray open_spiel vs PettingZoo stable-baselines3 vs stable-baselines open_spiel vs gym stable-baselines3 vs Pytorch open_spiel vs rlcard stable-baselines3 vs cleanrl open_spiel vs gym-battleship stable-baselines3 vs tianshou open_spiel vs TexasHoldemSolverJava stable-baselines3 vs Super-mario-bros-PPO-pytorch

Compare open_spiel vs stable-baselines3 and see what are their differences.

open_spiel

stable-baselines3

open_spiel

stable-baselines3

What are some alternatives?