Subscription

Explore Products

Best Sellers

New Releases

Books

Videos

Audiobooks

Learning Hub

Conferences

Free Learning

You're reading from Hands-On Computer Vision with TensorFlow 2 Leverage deep learning to create powerful image processing apps with TensorFlow 2.0 and Keras

Product type Paperback

Published in May 2019

Publisher Packt

ISBN-13 9781788830645

Length 372 pages

Edition 1st Edition

Languages

Python

Tools

Keras

Concepts

Computer Vision

Authors (2):

Eliot Andres

Benjamin Planche

View More author details

Table of Contents (16) Chapters

Preface

1. Section 1: TensorFlow 2 and Deep Learning Applied to Computer Vision FREE CHAPTER

2. Computer Vision and Neural Networks

3. TensorFlow Basics and Training a Model

4. Modern Neural Networks

5. Section 2: State-of-the-Art Solutions for Classic Recognition Problems

6. Influential Classification Tools

7. Object Detection Models

8. Enhancing and Segmenting Images

9. Section 3: Advanced Concepts and New Frontiers of Computer Vision

10. Training on Complex and Scarce Datasets

11. Video and Recurrent Neural Networks

12. Optimizing Models and Deploying on Mobile Devices

13. Migrating from TensorFlow 1 to TensorFlow 2

14. Assessments

15. Other Books You May Enjoy

Leave a review - let other readers know what you think

Summary

We expanded our knowledge of neural networks by describing the general principles of RNNs. After covering the inner workings of the basic RNN, we extended backpropagation to apply it to recurrent networks. As presented in this chapter, BPTT suffers from gradient vanishing when applied to RNNs. This can be worked around by using truncated backpropagation, or by using a different type of architecture—LSTM networks.

We applied those theoretical principles to a practical problem—action recognition in videos. By combining CNNs and LSTMs, we successfully trained a network to classify videos in 101 categories, introducing video-specific techniques such as frame sampling and padding.

In the next chapter, we will broaden our knowledge of neural network applications by covering new platforms—mobile devices and web browsers.

The rest of the chapter is locked

A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.

Unlock this book and the full library FREE for 7 days

Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of

Start free trial

Renews at $19.99/month. Cancel anytime

Authors (2)

Benjamin Planche

Dr. Benjamin Planche is a passionate research scientist in computer vision and machine learning. His main research efforts focus on data scarcity problems and industrial vision systems, leading to numerous patents and publications at international conferences. He worked in various research labs around the world (including in France, Japan, Germany, and the USA). Benjamin obtained his Ph.D. summa cum laude from the Faculty of Computer Science and Mathematics at the University of Passau, under the supervision of Prof. Dr. Harald Kosch. He also has a double master's degree from INSA-Lyon (France) and the University of Passau (Germany), with first-class honors and a multinational excellence award. He also likes sharing his knowledge and experience on various platforms or applying them to the creation of aesthetic demos.

See other products by Benjamin Planche

Eliot Andres

Eliot Andres is a freelance deep learning and computer vision engineer. He has more than 3 years' experience in the field, applying his skills to a variety of industries, such as banking, health, social media, and video streaming. Eliot has a double master's degree from cole des Ponts and Tlcom, Paris. His focus is industrialization: delivering value by applying new technologies to business problems. Eliot keeps his knowledge up to date by publishing articles on his blog and by building prototypes using the latest technologies.

See other products by Eliot Andres