Explore Products

Best Sellers

New Releases

Books

Videos

Audiobooks

Learning Hub

Newsletter Hub

Free Learning

You're reading from Hands-On GPU-Accelerated Computer Vision with OpenCV and CUDA Effective techniques for processing complex image data in real time using GPUs

Product type Paperback

Published in Sep 2018

Publisher Packt

ISBN-13 9781789348293

Length 380 pages

Edition 1st Edition

Languages

C++

Tools

CUDA

Concepts

Computer Vision

Author (1):

Bhaumik Vaidya

View More author details

Table of Contents (15) Chapters

Preface

1. Introducing CUDA and Getting Started with CUDA FREE CHAPTER

2. Parallel Programming using CUDA C

3. Threads, Synchronization, and Memory

4. Advanced Concepts in CUDA

5. Getting Started with OpenCV with CUDA Support

6. Basic Computer Vision Operations Using OpenCV and CUDA

7. Object Detection and Tracking Using OpenCV and CUDA

8. Introduction to the Jetson TX1 Development Board and Installing OpenCV on Jetson TX1

9. Deploying Computer Vision Applications on Jetson TX1

10. Getting Started with PyCUDA

11. Working with PyCUDA

12. Basic Computer Vision Applications Using PyCUDA

13. Assessments

14. Other Books You May Enjoy

Leave a review - let other readers know what you think

Thread and block execution in PyCUDA

We saw in the A kernel call section that we can start multiple blocks and multiple threads in parallel. So, in which order do these blocks and threads start and finish their execution? It is important to know this if we want to use the output of one thread in other threads. To understand this, we have modified the kernel in the hello,PyCUDA! program, seen in the earlier section, by including a print statement in a kernel call, which prints the block number. The modified code is shown as follows:


import pycuda.driver as drv
import pycuda.autoinit
from pycuda.compiler import SourceModule

mod = SourceModule("""
  #include <stdio.h>
  __global__ void myfirst_kernel()
  {
    printf("I am in block no: %d \\n", blockIdx.x);
  }
""")
 
function = mod.get_function("myfirst_kernel")
function(grid=(4,1),block...

The rest of the chapter is locked

A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.

Unlock this book and the full library FREE for 7 days

Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of

Start free trial

Renews at €18.99/month. Cancel anytime

Authors (1)

Vaidya

Bhaumik Vaidya is an experienced computer vision engineer and mentor. He has worked extensively on OpenCV Library in solving computer vision problems. He is a University gold medalist in masters and is now doing a PhD in the acceleration of computer vision algorithms built using OpenCV and deep learning libraries on GPUs. He has a background in teaching and has guided many projects in computer vision and VLSI(Very-large-scale integration). He has worked in the VLSI domain previously as an ASIC verification engineer, so he has very good knowledge of hardware architectures also. He has published many research papers in reputable journals to his credit. He, along with his PhD mentor, has also received an NVIDIA Jetson TX1 embedded development platform as a research grant from NVIDIA.

See other products by Vaidya