Packt+ | Advance your knowledge in tech

You're reading from Python Reinforcement Learning Solve complex real-world problems by mastering reinforcement learning algorithms using OpenAI Gym and TensorFlow

Product type Course

Published in Apr 2019

Publisher Packt

ISBN-13 9781838649777

Length 496 pages

Edition 1st Edition

Languages

Python

Tools

OpenAI Gym

Concepts

Reinforcement Learning

Authors (4):

Yang Wenzhuo

Sean Saito

Sudharsan Ravichandiran

Rajalingappaa Shanmugamani

View More author details

Table of Contents (27) Chapters

Title Page

About Packt

Contributors

Preface

1. Introduction to Reinforcement Learning FREE CHAPTER

2. Getting Started with OpenAI and TensorFlow

3. The Markov Decision Process and Dynamic Programming

4. Gaming with Monte Carlo Methods

5. Temporal Difference Learning

6. Multi-Armed Bandit Problem

7. Playing Atari Games

8. Atari Games with Deep Q Network

9. Playing Doom with a Deep Recurrent Q Network

10. The Asynchronous Advantage Actor Critic Network

11. Policy Gradients and Optimization

12. Balancing CartPole

13. Simulating Control Tasks

14. Building Virtual Worlds in Minecraft

15. Learning to Play Go

16. Creating a Chatbot

17. Generating a Deep Learning Image Classifier

18. Predicting Future Stock Prices

19. Capstone Project - Car Racing Using DQN

20. Looking Ahead

1. Assessments

2. Other Books You May Enjoy

Leave a review - let other readers know what you think

Index

Implementation of DQN

This chapter will show you how to implement all the components, for example, Q-network, replay memory, trainer, and Q-learning optimizer, of the deep Q-learning algorithm with Python and TensorFlow.

We will implement the QNetwork class for the Q-network that we discussed in the previous chapter, which is defined as follows:

class QNetwork:

    def __init__(self, input_shape=(84, 84, 4), n_outputs=4, 
                 network_type='cnn', scope='q_network'):

        self.width = input_shape[0]
        self.height = input_shape[1]
        self.channel = input_shape[2]
        self.n_outputs = n_outputs
        self.network_type = network_type
        self.scope = scope

        # Frame images
        self.x = tf.placeholder(dtype=tf.float32, 
                                shape=(None, self.channel, 
                                       self.width, self.height))
        # Estimates of Q-value
        self.y = tf.placeholder(dtype=tf.float32, shape=(None,))
       ...

The rest of the chapter is locked

You're reading from Python Reinforcement Learning Solve complex real-world problems by mastering reinforcement learning algorithms using OpenAI Gym and TensorFlow

Table of Contents (27) Chapters

Implementation of DQN

Unlock this book and the full library FREE for 7 days

Authors (4)

Personalised recommendations for you