Foundations of Reinforcement Learning

99 students enrolled
Reinforcement learning studies how agents learn to make better decisions through interaction with an environment. Agents act, observe consequences, receive feedback, and adapt future behavior. This specialization develops reinforcement learning as a framework for sequential decision-making under uncertainty, progressing from classical foundations to scalable deep learning methods and reward design. The first course, Classical Reinforcement Learning, introduces finite-state decision problems, Markov chains, Markov decision processes, discounted rewards, Bellman equations, planning with known models, and learning from sampled experience. Learners study value iteration, policy iteration, Monte Carlo methods, temporal-difference learning, SARSA, and Q-learning. The second course, Deep Reinforcement Learning, shows how reinforcement learning scales beyond tabular settings using neural-network-based function approximation. Learners study Deep Q-Networks, replay buffers, target networks, policy-gradient methods, actor–critic algorithms, and modern methods such as PPO, DDPG, and SAC, with attention to stability, diagnosis, evaluation, and reproducibility. The third course, Reward Programming, addresses how to design, infer, monitor, and revise objectives so agents learn intended behavior. Learners study temporal logic, automata, reward machines, reward shaping, inverse reinforcement learning, preference feedback, safety, shielding, auditing, and stress testing.
CERTIFICATEKatılım Sertifikası
FORMAT100% Online
DURATION1 ay

What you'll learn

  • Reinforcement Learning
  • Markov Model
  • Probability Distribution
  • Model Optimization
  • Artificial Intelligence and Machine Learning (AIu002FML)
  • Sampling (Statistics)
  • Statistical Machine Learning
  • Machine Learning
  • Theoretical Computer Science
  • Probability & Statistics
  • Decision Intelligence
  • Algorithms

Course Content

3 topics
  1. Kurs 1 Mastering Classic Reinforcement Learning Algorithms
  2. Kurs 2 Deep Reinforcement Learning: From Theory to Practice
  3. Kurs 3 Reward Programming: Optimizing RL Efficiency and Safety

Details

  • ProviderUniversity of Colorado Boulder
  • TypeCourse
  • CategorySoftware & Programming
  • LevelIntermediate
  • Duration1 ay
  • LanguageEnglish

Öğrenenlerimiz ne diyor?

Türkiye'nin yüz akı üniversitelerince hazırlanan; akademik doyuruculuğa sahip eğitim içeriklerinin, etkileşimli videolarla bir araya getirildiği bir üniversiteden eğitim almak istiyorsanız doğru yerdesiniz.
Ramazan Bölükbaşı
Çok yoğun programı olan öğrenciler için büyük bir fırsat. Bir şeylerin gelişmesi değişmesi için çabalamalıyız.
Melis Gülsar
Hızlı desteği ve üst düzey hizmeti ile Campus Online ve Sosyal Medya Sertifika Programı hizmeti sağlayan Adnan Menderes Üniversitesine sonsuz teşekkürlerimi sunuyorum.
Nazif Bayram
$49
CampusOnline Assistant
courses saved

Course Comparison

Institution
Rating
Turkish Subtitles
Level
Duration
Price
Certificate