Book Overview & Buying
Table Of Contents
Feedback & Rating

Hands-On Intelligent Agents with OpenAI Gym

By : Palanisamy

2 (3)

Buy this Book

Hands-On Intelligent Agents with OpenAI Gym

2 (3)

By: Palanisamy

Buy this Book

Overview of this book

Many real-world problems can be broken down into tasks that require a series of decisions to be made or actions to be taken. The ability to solve such tasks without a machine being programmed requires a machine to be artificially intelligent and capable of learning to adapt. This book is an easy-to-follow guide to implementing learning algorithms for machine software agents in order to solve discrete or continuous sequential decision making and control tasks. Hands-On Intelligent Agents with OpenAI Gym takes you through the process of building intelligent agent algorithms using deep reinforcement learning starting from the implementation of the building blocks for configuring, training, logging, visualizing, testing, and monitoring the agent. You will walk through the process of building intelligent agents from scratch to perform a variety of tasks. In the closing chapters, the book provides an overview of the latest learning environments and learning algorithms, along with pointers to more resources that will help you take your deep reinforcement learning skills to the next level.

Preface

Who this book is for

What this book covers

To get the most out of this book

Get in touch

Free Chapter

Introduction to Intelligent Agents and Learning Environments

What is an intelligent agent?

Learning environments

What is OpenAI Gym?

Understanding the features of OpenAI Gym

What can you do with the OpenAI Gym toolkit?

Creating your first OpenAI Gym environment

Summary

Reinforcement Learning and Deep Reinforcement Learning

What is reinforcement learning?

Understanding what AI means and what's in it in an intuitive way

Practical reinforcement learning

Markov Decision Process

Planning with dynamic programming

Monte Carlo learning and temporal difference learning

SARSA and Q-learning

Deep reinforcement learning

Practical applications of reinforcement and deep reinforcement learning algorithms

Summary

Getting Started with OpenAI Gym and Deep Reinforcement Learning

Code repository, setup, and configuration

Installing tools and libraries needed for deep reinforcement learning

Summary

Exploring the Gym and its Features

Exploring the list of environments and nomenclature

Understanding the Gym interface

Spaces in the Gym

Summary

Implementing your First Learning Agent - Solving the Mountain Car problem

Understanding the Mountain Car problem

Implementing a Q-learning agent from scratch

Training the reinforcement learning agent at the Gym

Testing and recording the performance of the agent

A simple and complete Q-Learner implementation for solving the Mountain Car problem

Summary

Implementing an Intelligent Agent for Optimal Control using Deep Q-Learning

Improving the Q-learning agent

Implementing a deep Q-learning agent

The Atari Gym environment

Training the deep Q-learner to play Atari games

Summary

Creating Custom OpenAI Gym Environments - CARLA Driving Simulator

Understanding the anatomy of Gym environments

Creating an OpenAI Gym-compatible CARLA driving simulator environment

Summary

Implementing an Intelligent - Autonomous Car Driving Agent using Deep Actor-Critic Algorithm

The deep n-step advantage actor-critic algorithm

Implementing a deep n-step advantage actor critic agent

Training an intelligent and autonomous driving agent

Summary

Exploring the Learning Environment Landscape - Roboschool, Gym-Retro, StarCraft-II, DeepMindLab

Gym interface-compatible environments

Other open source Python-based learning environments

Summary

Exploring the Learning Algorithm Landscape - DDPG (Actor-Critic), PPO (Policy-Gradient), Rainbow (Value-Based)

Deep Deterministic Policy Gradients

Proximal Policy Optimization

Rainbow

Summary

Other Books You May Enjoy

Leave a review - let other readers know what you think

Customer Reviews

2 (3)

5 star

4 star

33.3%

3 star

2 star

1 star

66.7%

The deep n-step advantage actor-critic algorithm

In our deep Q-learner-based intelligent agent implementation, we used a deep neural network as the function approximator to represent the action-value function. The agent then used the action-value function to come up with a policy based on the value function. In particular, we used the -greedy algorithm in our implementation. So, we understand that ultimately the agent has to know what actions are good to take given an observation/state. Instead of parametrizing or approximating a state/action action function and then deriving a policy based on that function, can we not parametrize the policy directly? Yes we can! That is the exact idea behind policy gradient methods.

In the following subsections, we will briefly look at policy gradient-based learning methods and then transition to actor-critic methods that combine and make use...