You're reading from Mastering PyTorch Build powerful neural network architectures using advanced PyTorch 1.x features

Product type Paperback

Published in Feb 2021

Publisher Packt

ISBN-13 9781789614381

Length 450 pages

Edition 1st Edition

Languages

Python

Tools

PyTorch

Concepts

Deep Learning

Author (1):

Ashish Ranjan Jha

View More author details

Table of Contents (20) Chapters

Preface

1. Section 1: PyTorch Overview

2. Chapter 1: Overview of Deep Learning using PyTorch FREE CHAPTER

3. Chapter 2: Combining CNNs and LSTMs

4. Section 2: Working with Advanced Neural Network Architectures

5. Chapter 3: Deep CNN Architectures

6. Chapter 4: Deep Recurrent Model Architectures

7. Chapter 5: Hybrid Advanced Models

8. Section 3: Generative Models and Deep Reinforcement Learning

9. Chapter 6: Music and Text Generation with PyTorch

10. Chapter 7: Neural Style Transfer

11. Chapter 8: Deep Convolutional GANs

12. Chapter 9: Deep Reinforcement Learning

13. Section 4: PyTorch in Production Systems

14. Chapter 10: Operationalizing PyTorch Models into Production

15. Chapter 11: Distributed Training

16. Chapter 12: PyTorch and AutoML

17. Chapter 13: PyTorch and Explainable AI

18. Chapter 14: Rapid Prototyping with PyTorch

19. Other Books You May Enjoy

Leave a review - let other readers know what you think

Building a transformer-based text generator with PyTorch

We built a transformer-based language model using PyTorch in the previous chapter. Because a language model models the probability of a certain word following a given sequence of words, we are more than half-way through in building our own text generator. In this section, we will learn how to extend this language model as a deep generative model that can generate arbitrary yet meaningful sentences, given an initial textual cue in the form of a sequence of words.

Training the transformer-based language model

In the previous chapter, we trained a language model for 5 epochs. In this section, we will follow those exact same steps but will train the model for longer; that is, 50 epochs. The goal here is to obtain a better performing language model that can then generate realistic sentences. Please note that model training can take several hours. Hence, it is recommended to train it in the background; for example, overnight...