raklokesh / ReinforcementLearning_Sutton-Barto_Solutions Star 20 Code Issues Pull requests Solutions and figures for problems from Reinforcement Learning: An Introduction Sutton&Bartoreinforcement-learningqlearningmountain-carsarsagradient-descentfeature-engineeringbandit-algorithmsutton-gamblersutton-bookdynaqsutton-gridworldblackjack-montecarlobatch-updatemaximization-biasinfinite-variancerl-suttonsemi-gradient-sarsashort-corridoroptimal-policy Updated Jul 16, 2019Python
adik993 / reinforcement-learning-sutton Star 16 Code Issues Pull requests reinforcement-learningq-learningsarsagridworldmulti-armed-banditsrandom-walkracecarbandit-algorithmsutton-booktd-lambdadyna-qcliffwalking Updated Mar 4, 2020Python
vinaychetnani / Q-Learning-for-Non-Competitive-Bridge-Bidding Star 6 Code Issues Pull requests reinforcement-learningdeep-learningbandit-algorithm Updated Jan 23, 2018Python
mgpopinjay / bandit-algorithms Star 4 Code Issues Pull requests A small collection of Bandit Algorithms (ETC, E-Greedy, Elimination, UCB, Exp3, LinearUCB, and Thompson Sampling)online-learningbandit-algorithm Updated May 25, 2022Python
NickKaparinos / Stanford-CS-234-RL-2022 Star 3 Code Issues Pull requests Solutions to the Stanford CS:234 Reinforcement Learning 2022 course assignments.deep-reinforcement-learningstanford-universitypytorchdqnbandit-algorithmpolicy-gradients Updated Jun 27, 2022Python
riccardopoiani / unimodal-pure-exp Star 1 Code Issues Pull requests Code for the paper "Best-Arm Identification in Unimodal Bandits" (AISTATS 2025)multi-armed-banditsbandit-algorithmbest-arm-identification Updated Mar 6, 2025Python