AlphaZero_Gomoku

Gomoku AI model

An implementation of the AlphaZero algorithm for playing Gomoku from pure self-play training

An implementation of the AlphaZero algorithm for Gomoku (also called Gobang or Five in a Row)

GitHub

3k stars
102 watching
972 forks
Language: Python
last commit: over 2 years ago
alphagoalphago-zeroalphazeroboard-gamegobanggomokumctsmonte-carlo-tree-searchpytorchreinforcement-learningrlself-learningtensorflow

Related projects:

RepositoryDescriptionStars
zeta36/chess-alpha-zeroAn implementation of AlphaGo Zero's reinforcement learning approach to master the game of chess2,137
gorgonia/gorgoniaA low-level machine learning and graph computation library for Go.5,582
leela-zero/leela-zeroA Go program implementing a neural network-based AI system designed to play the game of Go without human-provided knowledge.5,368
packtpublishing/hands-on-intelligent-agents-with-openai-gymTeaching software developers to build intelligent agents using deep reinforcement learning and OpenAI Gym374
ikostrikov/pytorch-a2c-ppo-acktr-gailAn open-source implementation of several reinforcement learning algorithms in PyTorch3,644
tju-drl-lab/ai-optimizerA next-generation deep reinforcement learning toolkit with libraries for multiagent, self-supervised, offline, and transfer/reinforcement learning4,848
farama-foundation/gymnasiumDevelops and compares reinforcement learning algorithms by providing a standard API to communicate between learning algorithms and environments7,613
inancgumus/learngoA repository of thousands of Go examples and exercises to help developers learn the language by fixing and solving problems.18,987
thu-ml/tianshouA high-performance reinforcement learning library with modular interfaces and user-friendly APIs for building deep learning agents.8,069
dvyukov/go-fuzzA tool for generating and testing random inputs to ensure software reliability4,790
hubfire/muti-branch-ddpg-carlaAn implementation of a reinforcement learning algorithm using multi-branch architecture and Deep Deterministic Policy Gradients (DDPG) to control autonomous vehicles in simulation environments.81
erikbern/deep-pinkAn AI system designed to play chess using deep learning techniques813
alexis-jacq/lola_dicePyTorch implementation of LOLA using DiCE for decision-making in game-playing AI91
joonspk-research/generative_agentsA research project simulating human behavior in interactive environments.17,888
golang/goA programming language designed to build simple, reliable, and efficient software124,564