GPU Programming using Rust and CUDA (Paperback)
Language: English
Published by Gitforgits, 2026
- Softcover
- New

Seller: Grand Eagle Retail, Bensenville, IL, U.S.A.Grand Eagle Retail
AbeBooks seller since October 12, 2005
Condition: New
US$ 69.99
Quantity: 1 available
Add to basketItem description from seller
Paperback. C++ has been the go-to for GPU programming for almost 20 years. Can Rust do the job, and how well?This book is all about getting hands-on with different toolchains that connect Rust to NVIDIA hardware. There's RustaCUDA for safe host-side control, the Rust-CUDA project for writing kernels in pure Rust, and NVIDIA's experimental cuda-oxide compiler with its typed launches and async execution graphs.We're going to build one Cargo workspace that keeps on growing. It'll include device queries, launch planning, Rust-written kernels, memory optimization, parallel reductions and scans, multi-stream pipelines, matrix multiplication benchmarked against cuBLAS, a Monte Carlo option pricer validated against a closed formula, and a complete batched inference application measured against a Python baseline. We'll check every result against a CPU reference, and the reports will give accurate numbers, including where libraries outperform hand-written kernels and where experimental toolchains are still a work in progress.Key LearningsLaunch, synchronize, and verify GPU kernels with ownership-managed device memory.Write real CUDA kernels using Rust-CUDA and cuda-oxide.Plan grids, blocks, and warps for 2D workloads.Accelerate transfer speeds with pinned memory and coalesced access patterns.Build race-free thread cooperation using shared memory, barriers, and atomics.Overlap transfers with computation using streams, events, and async Rust pipelines.Optimize matrix multiplication and benchmark against cuBLAS ceiling.Wrap CUDA C library safely with handles, error enums, and Drop.Ship complete batched GPU inference application against Python baselines.Diagnose performance with Nsight Systems, Nsight Compute, and compute-sanitizer.Table of ContentNew Beneficiary of GPU ComputingThinking in ThreadsCommanding GPUWriting GPU KernelsCleaner Kernels with cuda-oxideMastering GPU MemoryMaking Threads CooperateKeeping GPU BusyDelivering Real MathBorrowing NVIDIA's MuscleShipping Complete GPU ApplicationProving Performance This item is printed on demand. Shipping may be from multiple locations in the US or from the UK, depending on stock availability.…
Seller Inventory # 9789349174375
- Title
- GPU Programming using Rust and CUDA (Paperback)
- Author
- Maris Fenlor
- Publisher
- Gitforgits
- Publication year
- 2026
- Condition
- new
- Binding
- Paperback
- Language
- English
- ISBN 10
- 9349174375
- ISBN 13
- 9789349174375
C++ has been the go-to for GPU programming for almost 20 years. Can Rust do the job, and how well?
This book is all about getting hands-on with different toolchains that connect Rust to NVIDIA hardware. There's RustaCUDA for safe host-side control, the Rust-CUDA project for writing kernels in pure Rust, and NVIDIA's experimental cuda-oxide compiler with its typed launches and async execution graphs.
We're going to build one Cargo workspace that keeps on growing. It'll include device queries, launch planning, Rust-written kernels, memory optimization, parallel reductions and scans, multi-stream pipelines, matrix multiplication benchmarked against cuBLAS, a Monte Carlo option pricer validated against a closed formula, and a complete batched inference application measured against a Python baseline. We'll check every result against a CPU reference, and the reports will give accurate numbers, including where libraries outperform hand-written kernels and where experimental toolchains are still a work in progress.
Key Learnings
- Launch, synchronize, and verify GPU kernels with ownership-managed device memory.
- Write real CUDA kernels using Rust-CUDA and cuda-oxide.
- Plan grids, blocks, and warps for 2D workloads.
- Accelerate transfer speeds with pinned memory and coalesced access patterns.
- Build race-free thread cooperation using shared memory, barriers, and atomics.
- Overlap transfers with computation using streams, events, and async Rust pipelines.
- Optimize matrix multiplication and benchmark against cuBLAS ceiling.
- Wrap CUDA C library safely with handles, error enums, and Drop.
- Ship complete batched GPU inference application against Python baselines.
- Diagnose performance with Nsight Systems, Nsight Compute, and compute-sanitizer.
Table of Content
- New Beneficiary of GPU Computing
- Thinking in Threads
- Commanding GPU
- Writing GPU Kernels
- Cleaner Kernels with cuda-oxide
- Mastering GPU Memory
- Making Threads Cooperate
- Keeping GPU Busy
- Delivering Real Math
- Borrowing NVIDIA's Muscle
- Shipping Complete GPU Application
- Proving Performance
"Synopsis" may belong to another edition of this title.
Grand Eagle Retail
Bensenville, IL, U.S.A.
AbeBooks seller since October 12, 2005
Shipping rates within U.S.A.
| Item | 6 to 14 business days | 6 to 16 business days |
|---|---|---|
| First item | US$ 0.00 | US$ 0.00 |
Payment methods
Seller's business information
APOLLO ONLINE CORP.
605 Geddes Street
Wilmington, DE U.S.A. 19805
Terms of sale
We guarantee the condition of every book as it¿s described on the Abebooks web sites. If you¿ve changed
your mind about a book that you¿ve ordered, please use the Ask bookseller a question link to contact us
and we¿ll respond within 2 business days.
Books ship from California and Michigan.
Shipping terms
Orders usually ship within 2 business days. All books within the US ship free of charge. Delivery is 4-14 business days anywhere in the United States.
Books ship from California and Michigan.
If your book order is heavy or oversized, we may contact you to let you know extra shipping is required.