Subscription

Explore Products

Best Sellers

New Releases

Books

Videos

Audiobooks

Learning Hub

Newsletter Hub

Free Learning

You're reading from Hands-On GPU-Accelerated Computer Vision with OpenCV and CUDA Effective techniques for processing complex image data in real time using GPUs

Product type Paperback

Published in Sep 2018

Publisher Packt

ISBN-13 9781789348293

Length 380 pages

Edition 1st Edition

Languages

C++

Tools

CUDA

Concepts

Computer Vision

Author (1):

Bhaumik Vaidya

View More author details

Table of Contents (15) Chapters

Preface

1. Introducing CUDA and Getting Started with CUDA FREE CHAPTER

2. Parallel Programming using CUDA C

3. Threads, Synchronization, and Memory

4. Advanced Concepts in CUDA

5. Getting Started with OpenCV with CUDA Support

6. Basic Computer Vision Operations Using OpenCV and CUDA

7. Object Detection and Tracking Using OpenCV and CUDA

8. Introduction to the Jetson TX1 Development Board and Installing OpenCV on Jetson TX1

9. Deploying Computer Vision Applications on Jetson TX1

10. Getting Started with PyCUDA

11. Working with PyCUDA

12. Basic Computer Vision Applications Using PyCUDA

13. Assessments

14. Other Books You May Enjoy

Leave a review - let other readers know what you think

Threads

The CUDA has a hierarchical architecture in terms of parallel execution. The kernel execution can be done in parallel with multiple blocks. Each block is further divided into multiple threads. In the last chapter, we saw that CUDA runtime can carry out parallel operations by launching the same copies of the kernel multiple times. We saw that it can be done in two ways: either by launching multiple blocks in parallel, with one thread per block, or by launching a single block, with many threads in parallel. So, two questions you might ask are, which method should I use in my code? And, is there any limitation on the number of blocks and threads that can be launched in parallel?

The answers to these questions are pivotal. As we will see later on in this chapter, threads in the same blocks can communicate with each other via shared memory. So, there is an advantage to launching...

The rest of the chapter is locked

A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.

Unlock this book and the full library FREE for 7 days

Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of

Start free trial

Renews at $19.99/month. Cancel anytime

Authors (1)

Vaidya

Bhaumik Vaidya is an experienced computer vision engineer and mentor. He has worked extensively on OpenCV Library in solving computer vision problems. He is a University gold medalist in masters and is now doing a PhD in the acceleration of computer vision algorithms built using OpenCV and deep learning libraries on GPUs. He has a background in teaching and has guided many projects in computer vision and VLSI(Very-large-scale integration). He has worked in the VLSI domain previously as an ASIC verification engineer, so he has very good knowledge of hardware architectures also. He has published many research papers in reputable journals to his credit. He, along with his PhD mentor, has also received an NVIDIA Jetson TX1 embedded development platform as a research grant from NVIDIA.

See other products by Vaidya