Subscription

Explore Products

Best Sellers

New Releases

Books

Videos

Audiobooks

Learning Hub

Conferences

Free Learning

You're reading from Hands-On GPU-Accelerated Computer Vision with OpenCV and CUDA Effective techniques for processing complex image data in real time using GPUs

Product type Paperback

Published in Sep 2018

Publisher Packt

ISBN-13 9781789348293

Length 380 pages

Edition 1st Edition

Languages

C++

Tools

CUDA

Concepts

Computer Vision

Author (1):

Bhaumik Vaidya

View More author details

Table of Contents (15) Chapters

Preface

1. Introducing CUDA and Getting Started with CUDA

2. Parallel Programming using CUDA C FREE CHAPTER

3. Threads, Synchronization, and Memory

4. Advanced Concepts in CUDA

5. Getting Started with OpenCV with CUDA Support

6. Basic Computer Vision Operations Using OpenCV and CUDA

7. Object Detection and Tracking Using OpenCV and CUDA

8. Introduction to the Jetson TX1 Development Board and Installing OpenCV on Jetson TX1

9. Deploying Computer Vision Applications on Jetson TX1

10. Getting Started with PyCUDA

11. Working with PyCUDA

12. Basic Computer Vision Applications Using PyCUDA

13. Assessments

14. Other Books You May Enjoy

Leave a review - let other readers know what you think

Executing threads on a device

We have seen that, while configuring kernel parameters, we can start multiple blocks and multiple threads in parallel. So, in which order do these blocks and threads start and finish their execution? It is important to know this if we want to use the output of one thread in other threads. To understand this, we have modified the kernel in the hello,CUDA! program we saw in the first chapter, by including a print statement in the kernel call, which prints the block number. The modified code is as follows:

#include <iostream>
#include <stdio.h>
__global__ void myfirstkernel(void) 
{
  //blockIdx.x gives the block number of current kernel
   printf("Hello!!!I'm thread in block: %d\n", blockIdx.x);
}
int main(void) 
{
   //A kernel call with 16 blocks and 1 thread per block
   myfirstkernel << <16,1>> >();
 
   //Function...

You have been reading a chapter from

Hands-On GPU-Accelerated Computer Vision with OpenCV and CUDA

Published in: Sep 2018

Publisher: Packt

ISBN-13: 9781789348293

A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.

Unlock this book and the full library FREE for 7 days

Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of

Start free trial

Renews at $19.99/month. Cancel anytime

Authors (1)

Vaidya

Bhaumik Vaidya is an experienced computer vision engineer and mentor. He has worked extensively on OpenCV Library in solving computer vision problems. He is a University gold medalist in masters and is now doing a PhD in the acceleration of computer vision algorithms built using OpenCV and deep learning libraries on GPUs. He has a background in teaching and has guided many projects in computer vision and VLSI(Very-large-scale integration). He has worked in the VLSI domain previously as an ASIC verification engineer, so he has very good knowledge of hardware architectures also. He has published many research papers in reputable journals to his credit. He, along with his PhD mentor, has also received an NVIDIA Jetson TX1 embedded development platform as a research grant from NVIDIA.

See other products by Vaidya