"The structure, the practice exams, the instructor — all top tier. Passed first try."
Course Outline
What the programme covers, module by module.
Module 1: Introduction to Reinforcement Learning
- Fundamentals of Reinforcement Learning
- AI Learning Paradigms
- RL Workflow
- Agent and Environment
- Industry Applications
Module 2: Mathematical Foundations
- Probability Concepts
- State Spaces
- Rewards
- Discount Factors
- Utility Functions
Module 3: Markov Decision Processes
- States and Actions
- Transition Models
- Reward Functions
- Policy Definition
- Value Functions
Module 4: Dynamic Programming
- Bellman Equations
- Policy Evaluation
- Policy Improvement
- Value Iteration
- Policy Iteration
Module 5: Multi-Armed Bandits
- Exploration vs Exploitation
- Epsilon-Greedy Strategy
- Upper Confidence Bound
- Thompson Sampling
- Performance Comparison
Module 6: Monte Carlo Methods
- Monte Carlo Prediction
- Monte Carlo Control
- Episode Sampling
- Policy Evaluation
- Practical Examples
Module 7: Temporal Difference Learning
- TD Prediction
- TD Control
- Online Learning
- Bootstrapping
- Learning Updates
Module 8: Q-Learning
- Q-Value Functions
- Bellman Updates
- Off-Policy Learning
- Convergence
- Practical Implementation
Module 9: SARSA Algorithm
- On-Policy Learning
- SARSA Workflow
- Exploration Strategies
- Performance Analysis
- Algorithm Comparison
Module 10: Function Approximation
- Linear Approximation
- Neural Networks
- Generalization
- Feature Representation
- Model Scaling
Module 11: Deep Reinforcement Learning
- Deep Q Networks
- Experience Replay
- Target Networks
- Stability Techniques
- Deep RL Workflow
Module 12: Policy Gradient Methods
- Policy Optimization
- Gradient Estimation
- REINFORCE Algorithm
- Reward Maximization
- Variance Reduction
Module 13: Actor-Critic Algorithms
- Actor-Critic Architecture
- Advantage Functions
- Policy Updates
- Critic Learning
- Model Optimization
Module 14: Advanced RL Algorithms
- PPO Concepts
- TRPO Overview
- DDPG Concepts
- Soft Actor-Critic
- Algorithm Comparison
Module 15: Model-Based Reinforcement Learning
- Environment Models
- Planning
- Simulation
- Decision Optimization
- Hybrid Learning
Module 16: Multi-Agent Reinforcement Learning
- Cooperative Agents
- Competitive Agents
- Communication Strategies
- Coordination
- Team Learning
Module 17: Reinforcement Learning Frameworks
- OpenAI Gym
- Gymnasium
- Stable-Baselines
- RLlib Overview
- Environment Development
Module 18: Robotics Applications
- Autonomous Navigation
- Robot Control
- Motion Planning
- Industrial Robotics
- Simulation Platforms
Module 19: Reinforcement Learning in Games
- Game Environments
- Strategy Learning
- Decision Making
- Performance Evaluation
- Simulation-Based Training
Module 20: Reinforcement Learning for Optimization
- Resource Allocation
- Scheduling
- Recommendation Systems
- Financial Applications
- Industrial Optimization
Module 21: Model Evaluation
- Performance Metrics
- Reward Analysis
- Policy Evaluation
- Hyperparameter Tuning
- Benchmarking
Module 22: Responsible AI in Reinforcement Learning
- Safe Reinforcement Learning
- Ethical Decision Making
- Bias Awareness
- Risk Management
- Governance Practices
Module 23: Deployment of RL Models
- Production Deployment
- Model Serving
- Monitoring
- Scaling Strategies
- Maintenance
Module 24: End-to-End Reinforcement Learning Project
- Problem Definition
- Environment Design
- Agent Development
- Training and Evaluation
- Deployment Strategy
Who it's for & what's included
Pick a delivery method to see exactly who it suits and everything you receive.
Classroom
Best for learners who want face-to-face tuition and to network with peers in person.
Everything you get
- ✓ Live instructor on-site
- ✓ Printed workbook & materials
- ✓ Group exercises & case studies
Online Instructor-Led
Best for learners who want a live instructor and a fixed schedule, without the travel.
Everything you get
- ✓ Live instructor via video call
- ✓ Digital workbook & resources
- ✓ Session recordings
Self-Paced
Best for self-motivated learners who need maximum flexibility around work and life.
Everything you get
- ✓ On-demand video lessons
- ✓ Interactive quizzes
- ✓ 24/7 access on any device