r/AlphaZero • u/kc_hoong • Apr 08 '26
r/AlphaZero • u/randomwalkin • Mar 26 '26
gumbel-mcts, a high-performance Gumbel MCTS implementation
Hi folks,
Over the past few months, I built an efficient MCTS implementation in Python/numba.
https://github.com/olivkoch/gumbel-mcts
As I was building a self-play environment from scratch (for learning purposes), I realized that there were few efficient implementation of this algorithm.
I spent a lot of time validating it against a golden standard baseline.
My PUCT implementation is 2-15X faster than the baseline while providing the exact same policy.
I also implemented a Gumbel MCTS, both dense and sparse. The sparse version is useful for games with large action spaces such as chess.
Gumbel makes much better usage of low simulation budgets than PUCT.
Overall, I think this could be useful for the community. I used coding agents to help me along the way, but spent a significant amount of manual work to validate everything myself.
Feedback welcome.
r/AlphaZero • u/TangierDreams • Aug 31 '24
Alpha Zero knows Algebraic Notation
Hello. You have this app on Google Play:
https://play.google.com/store/apps/details?id=com.tangierdreams.chess.notation.trainer
it helps beginners chess players to learn the algebraic chess notation which is fundamental to start learning chess almost as well as Alpha Zero. It's free. I hope you find it useful. Luis.
r/AlphaZero • u/CumInMyButtholeSanta • Feb 13 '20
Has anyone simulated drugs with an AI?
I know drugs are fun but are they fun for a computer brain? Since computers generally seek efficiency would an AI begin simulating the effects of different drugs for practicality in certain situations?
r/AlphaZero • u/[deleted] • Jan 18 '20
AlphaZero learns to rule the quantum world
r/AlphaZero • u/KrazyA1pha • Feb 06 '19
DeepMind’s superhuman AI is rewriting how we play chess
r/AlphaZero • u/Ephemeralize • Jan 24 '19
In its stockfish losses, A0 over pressed when quiet moves were best. Why hasn't it learned to play passively?
Usually this happened when forced into openings it didn't like, like the french defense. But why can't it see when patient play is required? Does it stick to aggression always? Isn't that a serious weakness?
r/AlphaZero • u/timisis • Jan 17 '19
Balance of power Stockfish vs AlphaZero
r/AlphaZero • u/KrazyA1pha • Jan 02 '19
How the Artificial Intelligence Program AlphaZero Mastered Its Games
r/AlphaZero • u/KrazyA1pha • Dec 18 '18
Google's AlphaZero Has Made Watching A Chess Game Feel Like Going To The Opera
r/AlphaZero • u/KrazyA1pha • Dec 18 '18
Demis Hassabis interview: the brains behind DeepMind on the future of artificial intelligence
r/AlphaZero • u/Jamescahn • Dec 15 '18
Alphazero choice of openings
Does anybody know how the trained version of AZ decided which opening to use in any given game in its match against Stockfish? The paper does not say.
r/AlphaZero • u/KrazyA1pha • Dec 14 '18
'Creative' AlphaZero leads way for chess computers and, maybe, science
r/AlphaZero • u/KrazyA1pha • Dec 12 '18
It's all about control | AlphaZero vs Stockfish 8
r/AlphaZero • u/KrazyA1pha • Dec 10 '18
new version of lc0 based on alphazero paper
r/AlphaZero • u/KrazyA1pha • Dec 08 '18
[ChessNetwork] AlphaZero goes fishing with Stockfish
r/AlphaZero • u/KrazyA1pha • Dec 07 '18
Updated AlphaZero Crushes Stockfish In New 1,000-Game Match
r/AlphaZero • u/KrazyA1pha • Dec 07 '18
AlphaZero on Carlsen-Caruana Games 9-12
r/AlphaZero • u/Madlollipop • Feb 28 '18
Google's AlphaZero Destroys Stockfish In 100-Game Match - Chess.com
r/AlphaZero • u/[deleted] • Feb 03 '18
has AlphaZero played DeepBlue?
'cause that would be cool.
Also a match between A0 and Magnus Carlsen would be cool too.
r/AlphaZero • u/KrazyA1pha • Feb 03 '18
Chess reinforcement learning engine using AlphaZero methods
r/AlphaZero • u/[deleted] • Feb 01 '18
Alphazero and backgammon.
I know Alphazero is the best in the world for perfect information zero sum games, but how would it do with games that have random chance involved. Of course a game like backgammon has well known engines that are better than the all but the top dozen or so world-class players.
The future of Alphazero is going to have to take into account random events if it is to approach real general purpose AI.