halfrost / threes-ai Sponsor Star 163 Code Issues Pull requests 🏆 Threes! AI: deck-aware expectimax, N-tuple TD learning, and AlphaZerogolangreinforcement-learninggame-enginedeep-reinforcement-learningartificial-intelligencepuzzle-gamebitboardmonte-carlo-tree-search2048game-aithreesexpectimaxvalue-functionalphazeroself-playtemporal-difference-learningboard-game-ain-tuple-network Updated Jul 28, 2026Python
dellalibera / td-gammon Star 53 Code Issues Pull requests TD-Gammon implementationgamereinforcement-learningneural-networkpytorchartificial-intelligenceconvolutional-neural-networksbackgammonvalue-functiontemporal-differencing-learningself-play Updated Sep 25, 2023Python
YyzHarry / SV-RL Star 33 Code Issues Pull requests [ICLR 2020, Oral] Harnessing Structures for Value-Based Planning and Reinforcement Learningreinforcement-learningdeep-reinforcement-learningplanningcontrolsmatrix-completionvalue-iterationvalue-functioniclrlow-rankiclr2020 Updated Feb 1, 2020Python
Multi-Shot-Approximation-of-MDPs / Self-Guided-ALPs-Discounted-Cost Star 12 Code Issues Pull requests Multi-Shot Approximation of Discounted Cost MDPsreinforcement-learninglinear-programmingreinforcement-learning-algorithmsinventory-managementvalue-functionkernel-trickapproximate-dynamic-programmingrandom-features Updated Jul 23, 2024Python
Jerry1314520 / DLGAN-Zoo Star 6 Code Issues Pull requests GAN zoo include GAN, ACGAN, EBGAN, BEGAN, LSGAN, SAGAN, CVAE.ganvalue-functiongan-list Updated Feb 6, 2023Python