Subscription

Explore Products

Best Sellers

New Releases

Books

Videos

Audiobooks

Learning Hub

Newsletter Hub

Free Learning

You're reading from Hands-On GPU Programming with Python and CUDA Explore high-performance parallel computing with CUDA

Product type Paperback

Published in Nov 2018

Publisher Packt

ISBN-13 9781788993913

Length 310 pages

Edition 1st Edition

Languages

Python

Tools

CUDA

Concepts

Graphics Programming

Author (1):

Tuomanen

View More author details

Table of Contents (15) Chapters

Preface

1. Why GPU Programming? FREE CHAPTER

2. Setting Up Your GPU Programming Environment

3. Getting Started with PyCUDA

4. Kernels, Threads, Blocks, and Grids

5. Streams, Events, Contexts, and Concurrency

6. Debugging and Profiling Your CUDA Code

7. Using the CUDA Libraries with Scikit-CUDA

8. The CUDA Device Function Libraries and Thrust

9. Implementation of a Deep Neural Network

10. Working with Compiled GPU Code

11. Performance Optimization in CUDA

12. Where to Go from Here

13. Assessment

14. Other Books You May Enjoy

Leave a review - let other readers know what you think

Summary

We started with an implementation of Conway's Game of Life, which gave us an idea of how the many threads of a CUDA kernel are organized in a block-grid tensor-type structure. We then delved into block-level synchronization by way of the CUDA function, __syncthreads(), as well as block-level thread intercommunication by using shared memory; we also saw that single blocks have a limited number of threads that we can operate over, so we'll have to be careful in using these features when we create kernels that will use more than one block across a larger grid.

We gave an overview of the theory of parallel prefix algorithms, and we ended by implementing a naive parallel prefix algorithm as a single kernel that could operate on arrays limited by a size of 1,024 (which was synchronized with ___syncthreads and performed both the for and parfor loops internally), and...

The rest of the chapter is locked

Tech Concepts

Programming languages

Tech Tools

Unlimited access to the largest independent learning library in tech of over 8,000 expert-authored tech books and videos.

Innovative learning tools, including AI book assistants, code context explainers, and text-to-speech.

50+ new titles added per month and exclusive early access to books as they are being written.

A free Packt account unlocks extra newsletters, articles, discounted offers, and much more. Start advancing your knowledge today.

Unlock this book and the full library FREE for 7 days

Get unlimited access to 7000+ expert-authored eBooks and videos courses covering every tech area you can think of

Start free trial

Renews at $19.99/month. Cancel anytime

Authors (1)

Tuomanen

Dr. Brian Tuomanen has been working with CUDA and General-Purpose GPU Programming since 2014. He received his Bachelor of Science in Electrical Engineering from the University of Washington in Seattle, and briefly worked as a Software Engineer before switching to Mathematics for Graduate School. He completed his Ph.D. in Mathematics at the University of Missouri in Columbia, where he first encountered GPU programming as a means for studying scientific problems. Dr. Tuomanen has spoken at the US Army Research Lab about General Purpose GPU programming, and has recently lead GPU integration and development at a Maryland based start-up company. He currently lives and works in the Seattle area.

See other products by Tuomanen