Explore Products

Best Sellers

New Releases

Books

Videos

Audiobooks

Learning Hub

Conferences

Free Learning

You're reading from Hands-On GPU-Accelerated Computer Vision with OpenCV and CUDA Effective techniques for processing complex image data in real time using GPUs

Product type Paperback

Published in Sep 2018

Publisher Packt

ISBN-13 9781789348293

Length 380 pages

Edition 1st Edition

Languages

C++

Tools

CUDA

Concepts

Computer Vision

Author (1):

Bhaumik Vaidya

View More author details

Table of Contents (15) Chapters

Preface

1. Introducing CUDA and Getting Started with CUDA FREE CHAPTER

2. Parallel Programming using CUDA C

3. Threads, Synchronization, and Memory

4. Advanced Concepts in CUDA

5. Getting Started with OpenCV with CUDA Support

6. Basic Computer Vision Operations Using OpenCV and CUDA

7. Object Detection and Tracking Using OpenCV and CUDA

8. Introduction to the Jetson TX1 Development Board and Installing OpenCV on Jetson TX1

9. Deploying Computer Vision Applications on Jetson TX1

10. Getting Started with PyCUDA

11. Working with PyCUDA

12. Basic Computer Vision Applications Using PyCUDA

13. Assessments

14. Other Books You May Enjoy

Leave a review - let other readers know what you think

Chapter 11

C/C++ programming language is used to write kernel function inside SourceModule class, and this kernel function is compiled by nvcc (Nvidia C ) Compiler.
The kernel call function is as follows:

myfirst_kernel(block=(512,512,1),grid=(1024,1014,1))

False. The order of block execution is random in PyCUDA program, and it can't be determined by PyCUDA programmer.

The directives from driver class remove the need of separate allocation of memory for the Array, uploading it to the device and downloading the result back to host. All operations are performed simultaneously during a kernel call. This makes the code simpler and easy to read.
The PyCUDA code for adding two to every element in an array is shown below:

import pycuda.gpuarray as gpuarray
import numpy
import pycuda.driver as drv

start = drv.Event()
end=drv.Event()
start.record()
start.synchronize()
n=10
h_b = numpy...

The rest of the chapter is locked

A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.

Unlock this book and the full library FREE for 7 days

Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of

Start free trial

Renews at $19.99/month. Cancel anytime

Authors (1)

Vaidya

Bhaumik Vaidya is an experienced computer vision engineer and mentor. He has worked extensively on OpenCV Library in solving computer vision problems. He is a University gold medalist in masters and is now doing a PhD in the acceleration of computer vision algorithms built using OpenCV and deep learning libraries on GPUs. He has a background in teaching and has guided many projects in computer vision and VLSI(Very-large-scale integration). He has worked in the VLSI domain previously as an ASIC verification engineer, so he has very good knowledge of hardware architectures also. He has published many research papers in reputable journals to his credit. He, along with his PhD mentor, has also received an NVIDIA Jetson TX1 embedded development platform as a research grant from NVIDIA.

See other products by Vaidya