You're reading from Computer Vision Projects with OpenCV and Python 3

Product typeBook

Published inDec 2018

Reading LevelIntermediate

PublisherPackt

ISBN-139781789954555

Edition1st Edition

Languages

Python

Tools

OpenCV

Concepts

Computer Vision

Author (1)

Matthew Rever

Deep Learning Image Classification with TensorFlow

In this chapter, we will learn how to classify images using TensorFlow. First, we will use a pre-trained model, and then we'll proceed with training our own model using custom images.

Toward the end of the chapter, we will make use of the GPU to help us speed up our computations.

In this chapter, we will cover the following:

A deep introduction to TensorFlow
Using a pre-trained model (Inception) for image classification
Retraining with our own images
Speeding up computation with the GPU

Technical requirements

Along with knowledge of Python and the basics of image processing and computer vision, you will need the following libraries:

TensorFlow
NVIDIA CUDA® Deep Neural Network

The code used in the chapter has been added to the following GitHub repository:
https://github.com/PacktPublishing/Computer-Vision-Projects-with-OpenCV-and-Python-3

An introduction to TensorFlow

In this chapter, we'll go deeper into TensorFlow and see how we can build a general-purpose image classifier using its deep learning method.

This will be an extension of what we learned in Chapter 2, Handwritten Digit Recognition with scikit-learn and TensorFlow, where we learned how to classify handwritten digits. However, this method is quite a bit more powerful, as it will work on general images of people, animals, food, everyday objects, and so on.

To start, let's talk a little bit about what TensorFlow does, and the general workflow of TensorFlow.

To begin, what is a tensor? Wikipedia states this:

"In mathematics, tensors are geometric objects that describe linear relations between geometric vectors, scalars, and other tensors... Given a reference basis of vectors, a tensor can be represented as an organized multi-dimensional array...

Using Inception for image classification

In this section, we're going to use a pre-trained model, Inception, from Google to perform image classification. We'll then move on and build our own model—or, at least do some retraining on the model in order to train on our own images and classify our own objects.

For now, we want to see what we can do with a model that's already trained, which would take a lot of time to reproduce from scratch. Let's get started with the code.

Let's go back to Jupyter Notebook. The Notebook file can be found at https://github.com/PacktPublishing/Computer-Vision-Projects-with-OpenCV-and-Python-3/Chapter04.

In order to run the code, we're going to need to download a file from TensorFlow's website, from the following link: http://download.tensorflow.org/models/image/imagenet/inception-2015-12-05.tgz. This is the...

Retraining with our own images

In this section, we're going to go beyond what we did with the pre-built classifier and use our own images with our own labels.

The first thing I should mention is that this isn't really training from scratch with deep learning—there are multiple layers and algorithms for training the whole thing, which are very time-consuming—but we can take advantage of something called transfer learning, where we take the first few layers that were trained with a very large number of images, as illustrated in the following diagram:

It's one of the caveats of deep learning that having a few hundred or a few thousand images isn't enough. You need hundreds of thousands or even millions of samples in order to get good results, and gathering that much data is very time-consuming. Also, running it on a personal computer, which I expect...

Speeding up computation with your GPU

In this section, we'll talk briefly about speeding up computations with your GPU. The good news is that TensorFlow is actually very smart about using the GPU, so if you have everything set up, then it's pretty simple.

Let's see what things look like if we have the GPU properly set up. First, import TensorFlow as follows:

import tensorflow

Next, we print tensorflow.Session(). This just gives us information about our CPU and GPU (if it is properly set up):

print(tensorflow.Session())

The output is as follows:

As we can see from the output, we're using a laptop with a GeForce GTX 970M, which is CUDA-compatible. This is needed in order to run TensorFlow with the GPU. If everything is set up properly, you will see a message very similar to the preceding output for your GPU, your card model, and details about it such as its...

Summary

In this chapter, we learned how to classify images using a pre-trained model based on TensorFlow. We then retrained our model to work with custom images.

Finally, we had a brief overview of how to speed up the classification process by carrying out the computation on a GPU.

Using the examples covered in this book, you will be able to carry your our custom projects using Python, OpenCV, and TensorFlow.

The rest of the chapter is locked

You have been reading a chapter from

Computer Vision Projects with OpenCV and Python 3

Published in: Dec 2018Publisher: PacktISBN-13: 9781789954555

A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.

undefined

Unlock this book and the full library FREE for 7 days

Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of

Start free trial

Renews at $15.99/month. Cancel anytime

Author (1)

Matthew Rever

Matthew Rever received his PhD. in electrical engineering from the University of Michigan, Ann Arbor. His career revolves around image processing, computer vision, and machine learning for scientific research applications. He started programming in C++, a language he still uses today, over 20 years ago, and has also used Matlab and most heavily Python in the past few years, using OpenCV, SciPy, scikit-learn, TensorFlow, and PyTorch. He believes it is important to stay up to date on the latest tools to be as productive as possible. Dr. Rever is the author of Packt's Computer Vision Projects with Python 3 and Advanced Computer Vision Projects.
Read more about Matthew Rever

Other recommended products

Related to this chapter

Python Image Processing Cookbook

Advancements in wireless devices and mobile technology have enabled the acquisition of a tremendous amount of graphics, pictures, and videos. Through cutting edge recipes, this book provides coverage on tools, algorithms, and analysis for image processing. This book provides solutions addressing the challenges and complex tasks of image processing.

BookApr 2020438 pages

Deep Learning Essentials

Deep Learning is one of the trending topics in the field of Artificial Intelligence today and can be considered to be an advanced form of machine learning. This book will help you take your first steps when it comes to training efficient deep learning models, and apply them in various practical scenarios. You will model, train and deploy different kinds of neural networks such as Convolutional Neural Network, Recurrent Neural Network, and see their applications in real-world domains such as computer vision, natural language processing, and speech recognition. This book also covers solutions to tackle different problems you might come across while training your models and ensure their high performance. This book does not assume any prior knowledge of deep learning. By the end of this book, you will have a firm understanding of the basics of deep learning and neural network modeling, along with their practical applications.

BookJan 2018284 pages3

The Computer Vision Workshop

With The Computer Vision Workshop, you’ll explore the basic and advanced techniques in video and image processing using OpenCV and Python. It is filled with real-world exercises and activities that will make the learning process easy and enjoyable.

BookJul 2020568 pages

Intelligent Mobile Projects with TensorFlow

Google TensorFlow is used to train all the models deployed and running on mobile devices. This book covers 10 projects on the implementation of all major AI areas of iOS, Android, and Raspberry Pi: computer vision, speech and language processing, and machine learning.

BookMay 2018404 pages

Raspberry Pi Computer Vision Programming

You will learn the basics of hardware and software required for image processing and computer vision with Raspberry Pi and Python 3. You will have a look at all the major image processing, manipulation, and computer vision techniques and algorithms in detail using engaging examples. You will build a lot of real-life computer vision applications.

BookJun 2020306 pages5

Mastering Computer Vision with TensorFlow 2.x

You will learn the principles of computer vision and deep learning, and understand various models and architectures with their pros and cons. You will learn how to use TensorFlow 2.x to build your own neural network model and apply it to various computer vision tasks such as image acquiring, processing, and analyzing.

BookMay 2020430 pages

Deep Learning with TensorFlow

Machine learning is concerned with algorithms for transforming data into actionable intelligence and predictive analytics. Deep learning is a branch of machine learning based on multiple levels of representations. This book introduces the core concepts of deep learning using the latest version of TensorFlow to get implementation and research details on cutting-edge architectures. You will learn deep learning with the hands-on model building, data collection and transformation and even more!

BookApr 2017320 pages

Deep Learning for Computer Vision

Deep learning has shown its power in several application areas of Artificial Intelligence, especially in Computer Vision, the science of manipulating and processing images. In this book, you will learn different techniques in deep learning to accomplish tasks related to object classification, object detection, image segmentation, captioning, image generation, and more. You will also explore their application using the popular Python libraries such as TensorFlow and Keras. With practical examples, you will learn to develop Computer Vision applications by leveraging the power of deep learning.

BookJan 2018310 pages

Mastering OpenCV 4 with Python

Mastering OpenCV 4 with Python is a comprehensive guide to help you to get acquainted with various computer vision algorithms running in real-time. This book will help you to build complete projects on image processing, motion detection, and image segmentation where you can gain advanced computer vision techniques.

BookMar 2019532 pages

Machine Learning with TensorFlow 1.x

TensorFlow 1.x is an open source software library for numerical computation using data flow graphs. This book approaches common commercial machine learning problems using Google’s TensorFlow 1.x library. It covers unique features of the library such as Data Flow Graphs, training, visualization of performance with TensorBoard—all within a context rich with examples, using problems from multiple industries.

BookNov 2017304 pages

Personalised recommendations for you

Based on your interests and search pattern

Et al.

Ever wonder why speech recognition systems don't understand the Scottish accent, or what would happen if an astronaut only ate mac 'n' cheese, or other spurious reflections you'd have at a bar? We did, then collated those deliberations into absurd research articles with fake figures and methodologies inspired by even more fictionally absurd studies.

BookAug 2023230 pages5

Generative AI with LangChain

This book is a comprehensive introduction to LLMs and LangChain, demystifying the basic mechanics of LangChain, its functionalities, and the myriad of applications it can be integrated into.

BookDec 2023360 pages4

Generative AI with LangChain

This book is a comprehensive introduction to LLMs and LangChain, demystifying the basic mechanics of LangChain, its functionalities, and the myriad of applications it can be integrated into.

BookDec 2023360 pages5

Generative AI with LangChain

This book is a comprehensive introduction to LLMs and LangChain, demystifying the basic mechanics of LangChain, its functionalities, and the myriad of applications it can be integrated into.

BookDec 2023360 pages1

Generative AI with LangChain

This book is a comprehensive introduction to LLMs and LangChain, demystifying the basic mechanics of LangChain, its functionalities, and the myriad of applications it can be integrated into.

BookDec 2023360 pages5

Mastering Tableau 2023

This book is a comprehensive resource to mastering your Tableau skills and becoming a BI expert. As you progress, you will learn how to build advanced dashboards and improve your storytelling to derive key business insight, as well as make you well-versed with advanced functionalities of Tableau in the business intelligence domain.

BookAug 2023684 pages

Building AI Applications with ChatGPT APIs

This guide covers all ChatGPT API features for effortless creation of robust AI powered apps. With its help, you’ll be able to leverage ChatGPT’s cutting-edge NLP models to take your app development skills to the next level. You’ll also work on ten exciting projects that will give you the practical know-how that you can apply to your existing applications.

BookSep 2023258 pages5

Building AI Applications with ChatGPT APIs

This guide covers all ChatGPT API features for effortless creation of robust AI powered apps. With its help, you’ll be able to leverage ChatGPT’s cutting-edge NLP models to take your app development skills to the next level. You’ll also work on ten exciting projects that will give you the practical know-how that you can apply to your existing applications.

BookSep 2023258 pages2

Data Engineering with AWS

Embark on a journey to master data engineering pipelines on AWS! Our book offers a hands-on experience of AWS services for ingesting, transforming, and consuming data. Whether you're an absolute beginner or someone with basic data engineering experience, this guide is an indispensable resource.

BookOct 2023636 pages5

Modern Data Architecture on AWS

Every organization wants an agile, performant, and cost-effective data platform that meets all their current and future business needs. Purpose-built AWS analytics services and their features play a big part in building such a modern data platform. This book brings to you all the design and architectural patterns that’ll help you achieve this goal.

BookAug 2023420 pages5

Practical Guide to Applied Conformal Prediction in Python

Discover the power of Conformal Prediction with the "Practical Guide to Applied Conformal Prediction in Python." Master the latest techniques to quantify uncertainty in machine learning and computer vision models, and seamlessly apply them to your industry applications.

BookDec 2023240 pages

TinyML Cookbook

With over 70 project-based recipes, the TinyML Cookbook is a practical guide that will help you to get the most out of your microcontrollers. It provides a comprehensive understanding of the theoretical foundations while giving you hands-on experience training ML models for deployment on Arduino Nano 33 BLE Sense, Raspberry Pi Pico, and SparkFun RedBoard Artemis Nano microcontrollers.

BookNov 2023664 pages