Explore Products

Best Sellers

New Releases

Books

Videos

Audiobooks

Learning Hub

Newsletter Hub

Free Learning

You're reading from Hands-On Computer Vision with TensorFlow 2 Leverage deep learning to create powerful image processing apps with TensorFlow 2.0 and Keras

Product type Paperback

Published in May 2019

Publisher Packt

ISBN-13 9781788830645

Length 372 pages

Edition 1st Edition

Languages

Python

Tools

Keras

Concepts

Computer Vision

Authors (2):

Eliot Andres

Benjamin Planche

View More author details

Table of Contents (16) Chapters

Preface

1. Section 1: TensorFlow 2 and Deep Learning Applied to Computer Vision FREE CHAPTER

2. Computer Vision and Neural Networks

3. TensorFlow Basics and Training a Model

4. Modern Neural Networks

5. Section 2: State-of-the-Art Solutions for Classic Recognition Problems

6. Influential Classification Tools

7. Object Detection Models

8. Enhancing and Segmenting Images

9. Section 3: Advanced Concepts and New Frontiers of Computer Vision

10. Training on Complex and Scarce Datasets

11. Video and Recurrent Neural Networks

12. Optimizing Models and Deploying on Mobile Devices

13. Migrating from TensorFlow 1 to TensorFlow 2

14. Assessments

15. Other Books You May Enjoy

Leave a review - let other readers know what you think

Setting up the task

Classifying images of handwritten digits (that is, recognizing whether an image contains a 0 or a 1 and so on) is a historical problem in computer vision. The Modified National Institute of Standards and Technology (MNIST) dataset (http://yann.lecun.com/exdb/mnist/), which contains 70,000 grayscale images (28 × 28 pixels) of such digits, has been used as a reference over the years so that people can test their methods for this recognition task (Yann LeCun and Corinna Cortes hold all copyrights for this dataset, which is shown in the following diagram):

Figure 1.14: Ten samples of each digit from the MNIST dataset

For digit classification, what we want is a network that takes one of these images as input and returns an output vector expressing how strongly the network believes the image corresponds to each class. The input vector has 28 × 28 = 784 values, while the output has 10 values (for the 10 different digits, from 0 to 9). In-between...

The rest of the chapter is locked

A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.

Unlock this book and the full library FREE for 7 days

Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of

Start free trial

Renews at €18.99/month. Cancel anytime

Authors (2)

Benjamin Planche

Dr. Benjamin Planche is a passionate research scientist in computer vision and machine learning. His main research efforts focus on data scarcity problems and industrial vision systems, leading to numerous patents and publications at international conferences. He worked in various research labs around the world (including in France, Japan, Germany, and the USA). Benjamin obtained his Ph.D. summa cum laude from the Faculty of Computer Science and Mathematics at the University of Passau, under the supervision of Prof. Dr. Harald Kosch. He also has a double master's degree from INSA-Lyon (France) and the University of Passau (Germany), with first-class honors and a multinational excellence award. He also likes sharing his knowledge and experience on various platforms or applying them to the creation of aesthetic demos.

See other products by Benjamin Planche

Eliot Andres

Eliot Andres is a freelance deep learning and computer vision engineer. He has more than 3 years' experience in the field, applying his skills to a variety of industries, such as banking, health, social media, and video streaming. Eliot has a double master's degree from cole des Ponts and Tlcom, Paris. His focus is industrialization: delivering value by applying new technologies to business problems. Eliot keeps his knowledge up to date by publishing articles on his blog and by building prototypes using the latest technologies.

See other products by Eliot Andres