Problem 14 Dynamic Programming Solutions
Rational Decision-Making Under Uncertainty: Observed Betting Patterns on a Biased Coin
Openai/gym: A Toolkit for Developing and Comparing Reinforcement Learning Algorithms.
memoise
Faster Hash Maps in R
Array-Memoize: Memoization Combinators Using Arrays for Finite Sub-Domains of Functions
Vector: Efficient Arrays
https://github.com/FeepingCreature
KC Exact Solution
https://x.com/ArthurB/status/823241996244422656
Repoze.lru: Tiny LRU Cache
Apache MXNet
Meta-learning of Sequential Strategies
Deep Reinforcement Learning for Keras
Keras: Deep Learning for Humans
TensorFlow
Maxpumperla/hyperas: Keras + Hyperopt: A Very Simple Wrapper for Convenient Hyperparameter Optimization
Hyperopt Documentation
Hyperopt: A Python Library for Optimizing the Hyperparameters of Machine Learning Algorithms
Playing Atari with Deep Reinforcement Learning
Deep DPG (DDPG): Continuous control with deep reinforcement learning
Keras-Rl/examples/dqn_cartpole.py at Master
Deep Reinforcement Learning: Pong from Pixels
Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift
Keras-Rl/examples/ddpg_pendulum.py at Master
The Gambler’s Problem and Beyond