Get in Touch

Course Outline

Introduction to Reinforcement Learning and Agentic AI

  • Decision-making under uncertainty and the principles of sequential planning
  • Core components of RL: agents, environments, states, and rewards
  • The role of RL in fostering adaptive and agentic AI systems

Markov Decision Processes (MDPs)

  • Formal definitions and properties of MDPs
  • Value functions, Bellman equations, and dynamic programming techniques
  • Processes of policy evaluation, improvement, and iteration

Model-Free Reinforcement Learning

  • Monte Carlo methods and Temporal-Difference (TD) learning
  • Q-learning and SARSA algorithms
  • Practical application: implementing tabular RL methods in Python

Deep Reinforcement Learning

  • Integrating neural networks with RL for function approximation
  • Deep Q-Networks (DQN) and the use of experience replay
  • Actor-Critic architectures and policy gradient methods
  • Practical application: training agents using DQN and PPO with Stable-Baselines3

Exploration Strategies and Reward Shaping

  • Balancing exploration against exploitation (including ε-greedy, UCB, and entropy methods)
  • Crafting reward functions and mitigating unintended behaviors
  • Techniques for reward shaping and curriculum learning

Advanced Topics in RL and Decision-Making

  • Multi-agent reinforcement learning and cooperative strategies
  • Hierarchical reinforcement learning and the options framework
  • Offline RL and imitation learning for safer deployment scenarios

Simulation Environments and Evaluation

  • Leveraging OpenAI Gym and custom-built environments
  • Distinguishing between continuous and discrete action spaces
  • Metrics for assessing agent performance, stability, and sample efficiency

Integrating RL into Agentic AI Systems

  • Blending reasoning and RL within hybrid agent architectures
  • Combining reinforcement learning with tool-using agents
  • Operational considerations for scaling and deployment

Capstone Project

  • Designing and implementing a reinforcement learning agent for a simulated task
  • Analyzing training performance and optimizing hyperparameters
  • Demonstrating adaptive behavior and decision-making within an agentic context

Summary and Next Steps

Requirements

  • Advanced proficiency in Python programming
  • A robust understanding of machine learning and deep learning concepts
  • Familiarity with linear algebra, probability theory, and fundamental optimization methods

Target Audience

  • Reinforcement learning engineers and applied AI researchers
  • Developers specializing in robotics and automation
  • Engineering teams developing adaptive and agentic AI systems
 28 Hours

Custom Corporate Training

Training solutions designed exclusively for businesses.

  • Customized Content: We adapt the syllabus and practical exercises to the real goals and needs of your project.
  • Flexible Schedule: Dates and times adapted to your team's agenda.
  • Format: Online (live), In-company (at your offices), or Hybrid.
Investment

Price per private group, online live training, starting from 6400 € + VAT*

Contact us for an exact quote and to hear our latest promotions

Testimonials (3)

Provisional Upcoming Courses (Contact Us For More Information)

Related Categories