Top AI Repos — open-source AI, indexed and scored
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
Top AI Repos tracks AI repositories on GitHub and answers two different questions about each one: is it moving right now, and would you bet a product on it.
[Coursera] Reinforcement Learning Specialization by "University of Alberta" & "Alberta Machine Intelligence Institute"
| Date | Stars |
|---|---|
| 2026-07-31 | 267 |
| 2026-08-01 | 267 |
| 2026-08-02 | 267 |
| 2026-08-06 | 267 |
Today
— stars today
This week
— stars this week
This month
— stars this month
Momentum
0.0
growth rate 0.00%/day
# Reinforcement Learning
**[Reinforcement Learning Specialization](https://www.coursera.org/specializations/reinforcement-learning)**
+ **[Fundamentals of Reinforcement Learning](https://www.coursera.org/learn/fundamentals-of-reinforcement-learning)**
+ Week 1
+ [Practice Quiz: Exploration-Exploitation](https://github.com/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Fundamentals%20of%20Reinforcement%20Learning/Week%201/Practice%20Quiz:%20Exploration-Exploitation.pdf)
+ [Notebook: Bandits and Exploration/Exploitation](https://nbviewer.jupyter.org/github/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Fundamentals%20of%20Reinforcement%20Learning/Week%201/Notebook%3A%20Bandits%20and%20Exploration-Exploitation/C1M1-Assignment1-v8.ipynb)
+ Week 2
+ [Practice Quiz: MDPs](https://github.com/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Fundamentals%20of%20Reinforcement%20Learning/Week%202/Practice%20Quiz:%20MDPs.pdf)
+ Week 3
+ [Practice Quiz: Value Functions and Bellman Equations](https://github.com/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Fundamentals%20of%20Reinforcement%20Learning/Week%203/Practice%20Quiz:%20Value%20Functions%20and%20Bellman%20Equations.pdf)
+ [Quiz: Value Functions and Bellman Equations](https://github.com/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Fundamentals%20of%20Reinforcement%20Learning/Week%203/Quiz:%20Value%20Functions%20and%20Bellman%20Equations.pdf)
+ Week 4
+ [Practice Quiz: Dynamic Programming](https://github.com/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Fundamentals%20of%20Reinforcement%20Learning/Week%204/Practice%20Quiz:%20Dynamic%20Programming.pdf)
+ [Notebook: Optimal Policies with Dynamic Programming](https://nbviewer.jupyter.org/github/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Fundamentals%20of%20Reinforcement%20Learning/Week%204/Notebook%3A%20Optimal%20Policies%20with%20Dynamic%20Programming/C1M4_Assignment2-v2.ipynb)
+ **[Sample-based Learning Methods](https://www.coursera.org/learn/sample-based-learning-methods)**
+ Week 2
+ [Quiz: Graded Quiz](https://github.com/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Sample-based%20Learning%20Methods/Week%202/Quiz:%20Graded%20Quiz.pdf)
+ [Notebook: Blackjack](https://nbviewer.jupyter.org/github/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Sample-based%20Learning%20Methods/Week%202/Notebook%3A%20Blackjack/Blackjack.ipynb)
+ Week 3
+ [Quiz: Practice Quiz](https://github.com/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Sample-based%20Learning%20Methods/Week%203/Quiz:%20Practice%20Quiz.pdf)
+ [Notebook: Policy Evaluation with Temporal Difference Learning](https://github.com/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Sample-based%20Learning%20Methods/Week%203/Notebook:%20Policy%20Evaluation%20with%20Temporal%20Difference%20Learning/C2M2-Assignment-v4.ipynb)
+ Week 4
+ [Quiz: Practice Quiz](https://github.com/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Sample-based%20Learning%20Methods/Week%204/Quiz:%20Practice%20Quiz.pdf)
+ [Notebook: Q-Learning and Expected Sarsa](https://nbviewer.jupyter.org/github/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Sample-based%20Learning%20Methods/Week%204/Notebook%3A%20Q-Learning%20and%20Expected%20Sarsa/C2M3_Assignment2_v6.ipynb)
+ Week 5
+ [Quiz: Practice Assessment](https://github.com/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Sample-based%20Learning%20Methods/Week%205/Quiz:%20Practice%20Assessment.png)
+ [Notebook: Dyna-Q and Dyna-Q+](https://github.com/ChanchalKumarMaji/Reinforcement-Learning-Specialization/blob/master/Sample-based%20Learning%20Methods/Week%205/Notebook:%20Dyna-Q%20and%20Dyna-Q%2B/Planning_AssignmeExcerpt of 9,524 characters
Read on GitHubWould you bet a product on this? Bounded 0–100 and slow moving.
matched fp:5100ec4b21ad49ec, topic:reinforcement-learning, name:reinforcement learning, desc:reinforcement learning