University of Colorado Boulder

Spécialisation "Foundations of Reinforcement Learning"

La Fête du Travail commence avec plus de 70 $ d'économies sur Coursera Plus. Bénéficiez de 40 % de réduction pendant 3 mois.

Ce spécialisation n'est pas disponible en Français (France)

Nous sommes actuellement en train de le traduire dans plus de langues.
University of Colorado Boulder

Spécialisation "Foundations of Reinforcement Learning"

Master Reinforcement Learning.

Build foundations in classical RL, deep RL, and reward design.

Ashutosh Trivedi

Instructeur : Ashutosh Trivedi

Inclus avec Coursera PlusEn savoir plus

Demander à Coursera

Approfondissez votre connaissance d’un sujet
niveau Intermédiaire

Expérience recommandée

4 semaines à compléter
à 10 heures par semaine
Planning flexible
Apprenez à votre propre rythme
Approfondissez votre connaissance d’un sujet
niveau Intermédiaire

Expérience recommandée

4 semaines à compléter
à 10 heures par semaine
Planning flexible
Apprenez à votre propre rythme

Ce que vous apprendrez

  • Explain the mathematical foundations of reinforcement learning.

  • Analyze and compare tabular, approximate, and deep reinforcement learning algorithms .

  • Explain how function approximation and neural networks extend reinforcement learning beyond finite tabular settings

  • Design, infer, and assess reward structures and specification-based objectives that align learned behavior with intended task goals.

Compétences que vous acquerrez

  • Catégorie : Artificial Intelligence
  • Catégorie : Theoretical Computer Science
  • Catégorie : Decision Intelligence
  • Catégorie : Responsible AI
  • Catégorie : Machine Learning Algorithms
  • Catégorie : Artificial Intelligence and Machine Learning (AI/ML)
  • Catégorie : Computational Logic
  • Catégorie : Deep Learning
  • Catégorie : Reinforcement Learning
  • Catégorie : Statistical Machine Learning
  • Catégorie : Machine Learning
  • Catégorie : Model Optimization
  • Catégorie : Artificial Neural Networks
  • Catégorie : Machine Learning Methods
  • Catégorie : Applied Mathematics
  • Catégorie : Markov Model
  • Catégorie : Agentic systems
  • Catégorie : Model Evaluation
  • Catégorie : Algorithms

Outils que vous découvrirez

  • Catégorie : AI Workflows

Détails à connaître

Certificat partageable

Ajouter à votre profil LinkedIn

Enseigné en Anglais
Récemment mis à jour !

juillet 2026

Découvrez comment les employés des entreprises prestigieuses maîtrisent des compétences recherchées

 logos de Petrobras, TATA, Danone, Capgemini, P&G et L'Oreal

Améliorez votre expertise en la matière

  • Acquérez des compétences recherchées auprès d’universités et d’experts du secteur
  • Maîtrisez un sujet ou un outil avec des projets pratiques
  • Développez une compréhension approfondie de concepts clés
  • Obtenez un certificat professionnel auprès de University of Colorado Boulder

Spécialisation - série de 3 cours

Mastering Classic Reinforcement Learning Algorithms

Mastering Classic Reinforcement Learning Algorithms

COURS 1, 14 heures

Ce que vous apprendrez

  • Formulate sequential decision-making problems as deterministic decision processes, Markov chains, and finite Markov decision processes.

  • Explain and apply core reinforcement-learning concepts, including discounting, value functions, policies, Bellman equations, and optimality.

  • Implement planning algorithms for finite Markov decision processes, including value iteration, policy iteration, and linear programming formulations.

  • Compare tabular reinforcement-learning algorithms, including bandits, Monte Carlo methods, temporal-difference learning, SARSA, and Q-learning.

Compétences que vous acquerrez

Catégorie : Reinforcement Learning
Catégorie : Probability Distribution
Catégorie : Model Optimization
Catégorie : Markov Model
Catégorie : Statistical Machine Learning
Catégorie : Probability & Statistics
Catégorie : Machine Learning Algorithms
Catégorie : Machine Learning
Catégorie : Decision Intelligence
Catégorie : Artificial Intelligence and Machine Learning (AI/ML)
Catégorie : Algorithms
Catégorie : Sampling (Statistics)
Catégorie : Applied Mathematics
Deep Reinforcement Learning: From Theory to Practice

Deep Reinforcement Learning: From Theory to Practice

COURS 2, 14 heures

Ce que vous apprendrez

  • Explain how neural-network-based function approximation extends reinforcement learning beyond finite tabular settings.

  • Implement and evaluate value-based deep reinforcement learning algorithms, including Deep Q-Networks and stabilizing techniques.

  • Derive and implement policy-gradient methods, including REINFORCE, baselines, and advantage-based updates.

  • Explain and analyze actor–critic methods that combine policy optimization with value estimation.

Compétences que vous acquerrez

Catégorie : Reinforcement Learning
Catégorie : Deep Learning
Catégorie : System Design and Implementation
Catégorie : Artificial Neural Networks
Catégorie : Machine Learning Algorithms
Catégorie : Algorithms
Catégorie : Machine Learning
Catégorie : Model Training
Catégorie : Model Evaluation
Catégorie : Artificial Intelligence
Catégorie : Applied Machine Learning
Catégorie : Machine Learning Methods
Catégorie : Model Optimization
Catégorie : Agentic systems

Ce que vous apprendrez

  • Identify limitations of standard scalar reward formulations, including reward hacking, specification gaming, and brittle proxies.

  • Express structured learning objectives using formal tools such as temporal logic, automata, and reward machines.

  • Construct and analyze reward mechanisms based on temporal logic, automata, product MDPs, reward machines, and reward shaping.

  • Model reward-programming problems under hidden state, memory, hierarchy, multiagent interaction, and continuous-time dynamics

Compétences que vous acquerrez

Catégorie : Reinforcement Learning
Catégorie : Markov Model
Catégorie : Machine Learning Methods
Catégorie : Functional Specification
Catégorie : Continuous Monitoring
Catégorie : AI Workflows
Catégorie : Machine Learning
Catégorie : Safety and Security
Catégorie : Theoretical Computer Science
Catégorie : Computational Logic
Catégorie : Model Optimization
Catégorie : Verification And Validation
Catégorie : Model Evaluation
Catégorie : Responsible AI
Catégorie : Agentic systems

Obtenez un certificat professionnel

Ajoutez ce titre à votre profil LinkedIn, à votre curriculum vitae ou à votre CV. Partagez-le sur les médias sociaux et dans votre évaluation des performances.

Instructeur

Ashutosh Trivedi
University of Colorado Boulder
3 Cours582 apprenants

Offert par

Pour quelles raisons les étudiants sur Coursera nous choisissent-ils pour leur carrière ?

Felipe M.

Étudiant(e) depuis 2018
’Pouvoir suivre des cours à mon rythme à été une expérience extraordinaire. Je peux apprendre chaque fois que mon emploi du temps me le permet et en fonction de mon humeur.’

Jennifer J.

Étudiant(e) depuis 2020
’J'ai directement appliqué les concepts et les compétences que j'ai appris de mes cours à un nouveau projet passionnant au travail.’

Larry W.

Étudiant(e) depuis 2021
’Lorsque j'ai besoin de cours sur des sujets que mon université ne propose pas, Coursera est l'un des meilleurs endroits où se rendre.’

Chaitanya A.

’Apprendre, ce n'est pas seulement s'améliorer dans son travail : c'est bien plus que cela. Coursera me permet d'apprendre sans limites.’
  • cplus logo

    Profitez de plus de 10 000 programmes grâce à notre offre spéciale pour la Fête du Travail

  • Coursera for teams logo

    Commencez par réaliser des économies simples pour vos équipes qui travaillent dur

    30% off team training

Foire Aux Questions