Skip to content

Folders and files

NameName
Last commit message
Last commit date

Latest commit

 

History

1 Commit
 
 

Repository files navigation

Reinforcement_Learning

For the second project of CS3600, we are building a MDP process through the understanding of Policy Iteration, Value Iteration, and Policy Extraction. These implementations are useful to show why companies like Nvidia have strong parallel concurrencies being simultaneous and mathematically efficient.

About

For the second project of CS3600, we are building a MDP process through the understanding of Policy Iteration, Value Iteration, and Policy Extraction. These implementations are useful to show why companies like Nvidia have strong parallel concurrencies being simultaneous and mathematically efficient.

Resources

Stars

0 stars

Watchers

1 watching

Forks

Releases

Packages

Contributors