AlphaZero_Gomoku
Gomoku AI model
An implementation of the AlphaZero algorithm for playing Gomoku from pure self-play training
An implementation of the AlphaZero algorithm for Gomoku (also called Gobang or Five in a Row)
3k stars
102 watching
972 forks
Language: Python
last commit: over 2 years agoalphagoalphago-zeroalphazeroboard-gamegobanggomokumctsmonte-carlo-tree-searchpytorchreinforcement-learningrlself-learningtensorflow
Related projects:
| Repository | Description | Stars |
|---|---|---|
| An implementation of AlphaGo Zero's reinforcement learning approach to master the game of chess | 2,137 | |
| A low-level machine learning and graph computation library for Go. | 5,582 | |
| A Go program implementing a neural network-based AI system designed to play the game of Go without human-provided knowledge. | 5,368 | |
| Teaching software developers to build intelligent agents using deep reinforcement learning and OpenAI Gym | 374 | |
| An open-source implementation of several reinforcement learning algorithms in PyTorch | 3,644 | |
| A next-generation deep reinforcement learning toolkit with libraries for multiagent, self-supervised, offline, and transfer/reinforcement learning | 4,848 | |
| Develops and compares reinforcement learning algorithms by providing a standard API to communicate between learning algorithms and environments | 7,613 | |
| A repository of thousands of Go examples and exercises to help developers learn the language by fixing and solving problems. | 18,987 | |
| A high-performance reinforcement learning library with modular interfaces and user-friendly APIs for building deep learning agents. | 8,069 | |
| A tool for generating and testing random inputs to ensure software reliability | 4,790 | |
| An implementation of a reinforcement learning algorithm using multi-branch architecture and Deep Deterministic Policy Gradients (DDPG) to control autonomous vehicles in simulation environments. | 81 | |
| An AI system designed to play chess using deep learning techniques | 813 | |
| PyTorch implementation of LOLA using DiCE for decision-making in game-playing AI | 91 | |
| A research project simulating human behavior in interactive environments. | 17,888 | |
| A programming language designed to build simple, reliable, and efficient software | 124,564 |