Fenlor Maris (23 results)

- Softcover
Seller: BargainBookStores, Grand Rapids, MI, U.S.A.BargainBookStores
Contact seller5-star sellerCondition: New
US$ 52.66
Free ShippingShips within U.S.A.Quantity: 5 available
Paperback or Softback. Condition: New. Practical GPU Programming: High-performance computing with CUDA, CuPy, and Python on modern GPUs. Book.

- Softcover
Seller: Rarewaves USA, HEBRON, KY, U.S.A.Rarewaves USA
Contact seller5-star sellerCondition: New
US$ 56.95
Free ShippingShips within U.S.A.Quantity: Over 20 available
Paperback. Condition: New.

- Softcover
Seller: California Books, Miami, FL, U.S.A.California Books
Contact seller4-star sellerCondition: New
US$ 58.00
Free ShippingShips within U.S.A.Quantity: Over 20 available
Condition: New.
More images- Softcover
Seller: Rarewaves.com USA, London, LONDO, United KingdomRarewaves.com USA
Contact seller5-star sellerCondition: New
US$ 61.48
Free ShippingShips from United Kingdom to U.S.A.Quantity: Over 20 available
Paperback. Condition: New.

- Softcover
Seller: PBShop.store US, Wood Dale, IL, U.S.A.PBShop.store US
Contact seller5-star sellerCondition: New
US$ 68.22
Free ShippingShips within U.S.A.Quantity: Over 20 available
PAP. Condition: New. New Book. Shipped from UK. Established seller since 2000.

- Softcover
Seller: California Books, Miami, FL, U.S.A.California Books
Contact seller4-star sellerCondition: New
US$ 70.00
Free ShippingShips within U.S.A.Quantity: Over 20 available
Condition: New.

- Softcover
Seller: PBShop.store UK, Fairford, GLOS, United KingdomPBShop.store UK
Contact seller5-star sellerCondition: New
US$ 67.58
US$ 4.40 shippingShips from United Kingdom to U.S.A.Quantity: Over 20 available
PAP. Condition: New. New Book. Shipped from UK. Established seller since 2000.

- Softcover
Seller: PBShop.store UK, Fairford, GLOS, United KingdomPBShop.store UK
Contact seller5-star sellerCondition: New
US$ 81.35
US$ 5.57 shippingShips from United Kingdom to U.S.A.Quantity: Over 20 available
PAP. Condition: New. New Book. Shipped from UK. Established seller since 2000.
More images- Softcover
Seller: Rarewaves USA United, HEBRON, KY, U.S.A.Rarewaves USA United
Contact seller5-star sellerCondition: New
US$ 63.15
US$ 50.00 shippingShips within U.S.A.Quantity: Over 20 available
Paperback. Condition: New.
More images- Softcover
Seller: Rarewaves.com UK, London, United KingdomRarewaves.com UK
Contact seller5-star sellerCondition: New
US$ 62.64
US$ 86.98 shippingShips from United Kingdom to U.S.A.Quantity: Over 20 available
Paperback. Condition: New.

- Softcover
- Print on Demand
Seller: Grand Eagle Retail, Bensenville, IL, U.S.A.Grand Eagle Retail
Contact seller5-star sellerCondition: New
US$ 69.99
Free ShippingShips within U.S.A.Quantity: 1 available
Paperback. Condition: new. Paperback. C++ has been the go-to for GPU programming for almost 20 years. Can Rust do the job, and how well?This book is all about getting hands-on with different toolchains that connect Rust to NVIDIA hardware. There's RustaCUDA for safe host-side control, the Rust-CUDA project for writing kernels in pure Rust, and NVIDIA's experimental cuda-oxide compiler with its typed launches and async execution graphs.We're going to build one Cargo workspace that keeps on growing. It'll include device queries, launch planning, Rust-written kernels, memory optimization, parallel reductions and scans, multi-stream pipelines, matrix multiplication benchmarked against cuBLAS, a Monte Carlo option pricer validated against a closed formula, and a complete batched inference application measured against a Python baseline. We'll check every result against a CPU reference, and the reports will give accurate numbers, including where libraries outperform hand-written kernels and where experimental toolchains are still a work in progress.Key LearningsLaunch, synchronize, and verify GPU kernels with ownership-managed device memory.Write real CUDA kernels using Rust-CUDA and cuda-oxide.Plan grids, blocks, and warps for 2D workloads.Accelerate transfer speeds with pinned memory and coalesced access patterns.Build race-free thread cooperation using shared memory, barriers, and atomics.Overlap transfers with computation using streams, events, and async Rust pipelines.Optimize matrix multiplication and benchmark against cuBLAS ceiling.Wrap CUDA C library safely with handles, error enums, and Drop.Ship complete batched GPU inference application against Python baselines.Diagnose performance with Nsight Systems, Nsight Compute, and compute-sanitizer.Table of ContentNew Beneficiary of GPU ComputingThinking in ThreadsCommanding GPUWriting GPU KernelsCleaner Kernels with cuda-oxideMastering GPU MemoryMaking Threads CooperateKeeping GPU BusyDelivering Real MathBorrowing NVIDIA's MuscleShipping Complete GPU ApplicationProving Performance This item is printed on demand. Shipping may be from multiple locations in the US or from the UK, depending on stock availability.…

- Softcover
- Print on Demand
Seller: Majestic Books, Hounslow, United KingdomMajestic Books
Contact seller4-star sellerCondition: New
US$ 82.39
US$ 8.70 shippingShips from United Kingdom to U.S.A.Quantity: 4 available
Condition: New. Print on Demand.

- Softcover
- Print on Demand
Seller: Books Puddle, New York, NY, U.S.A.Books Puddle
Contact seller4-star sellerCondition: New
US$ 88.79
US$ 3.99 shippingShips within U.S.A.Quantity: 4 available
Condition: New. Print on Demand.

- Softcover
- Print on Demand
Seller: Biblios, frankfurt am main, HESSE, GermanyBiblios
Contact seller4-star sellerCondition: New
US$ 89.10
US$ 11.41 shippingShips from Germany to U.S.A.Quantity: 4 available
Condition: New. PRINT ON DEMAND.

- Softcover
- Print on Demand
Seller: BuchWeltWeit Ludwig Meier e.K., Bergisch Gladbach, GermanyBuchWeltWeit Ludwig Meier e.K.
Contact seller5-star sellerCondition: New
US$ 75.49
US$ 26.38 shippingShips from Germany to U.S.A.Quantity: 2 available
Taschenbuch. Condition: Neu. This item is printed on demand - it takes 3-4 days longer - Neuware 130 pp. Englisch.

- Softcover
- Print on Demand
Seller: BuchWeltWeit Ludwig Meier e.K., Bergisch Gladbach, GermanyBuchWeltWeit Ludwig Meier e.K.
Contact seller5-star sellerCondition: New
US$ 88.01
US$ 26.38 shippingShips from Germany to U.S.A.Quantity: 2 available
Taschenbuch. Condition: Neu. This item is printed on demand - it takes 3-4 days longer - Neuware 166 pp. Englisch.

- Softcover
- Print on Demand
Seller: CitiRetail, Stevenage, United KingdomCitiRetail
Contact seller5-star sellerCondition: New
US$ 79.24
US$ 49.51 shippingShips from United Kingdom to U.S.A.Quantity: 1 available
Paperback. Condition: new. Paperback. C++ has been the go-to for GPU programming for almost 20 years. Can Rust do the job, and how well?This book is all about getting hands-on with different toolchains that connect Rust to NVIDIA hardware. There's RustaCUDA for safe host-side control, the Rust-CUDA project for writing kernels in pure Rust, and NVIDIA's experimental cuda-oxide compiler with its typed launches and async execution graphs.We're going to build one Cargo workspace that keeps on growing. It'll include device queries, launch planning, Rust-written kernels, memory optimization, parallel reductions and scans, multi-stream pipelines, matrix multiplication benchmarked against cuBLAS, a Monte Carlo option pricer validated against a closed formula, and a complete batched inference application measured against a Python baseline. We'll check every result against a CPU reference, and the reports will give accurate numbers, including where libraries outperform hand-written kernels and where experimental toolchains are still a work in progress.Key LearningsLaunch, synchronize, and verify GPU kernels with ownership-managed device memory.Write real CUDA kernels using Rust-CUDA and cuda-oxide.Plan grids, blocks, and warps for 2D workloads.Accelerate transfer speeds with pinned memory and coalesced access patterns.Build race-free thread cooperation using shared memory, barriers, and atomics.Overlap transfers with computation using streams, events, and async Rust pipelines.Optimize matrix multiplication and benchmark against cuBLAS ceiling.Wrap CUDA C library safely with handles, error enums, and Drop.Ship complete batched GPU inference application against Python baselines.Diagnose performance with Nsight Systems, Nsight Compute, and compute-sanitizer.Table of ContentNew Beneficiary of GPU ComputingThinking in ThreadsCommanding GPUWriting GPU KernelsCleaner Kernels with cuda-oxideMastering GPU MemoryMaking Threads CooperateKeeping GPU BusyDelivering Real MathBorrowing NVIDIA's MuscleShipping Complete GPU ApplicationProving Performance This item is printed on demand. Shipping may be from our UK warehouse or from our Australian or US warehouses, depending on stock availability.…

- Softcover
- Print on Demand
Seller: AussieBookSeller, Truganina, VIC, AustraliaAussieBookSeller
Contact seller5-star sellerCondition: New
US$ 102.44
US$ 37.00 shippingShips from Australia to U.S.A.Quantity: 1 available
Paperback. Condition: new. Paperback. C++ has been the go-to for GPU programming for almost 20 years. Can Rust do the job, and how well?This book is all about getting hands-on with different toolchains that connect Rust to NVIDIA hardware. There's RustaCUDA for safe host-side control, the Rust-CUDA project for writing kernels in pure Rust, and NVIDIA's experimental cuda-oxide compiler with its typed launches and async execution graphs.We're going to build one Cargo workspace that keeps on growing. It'll include device queries, launch planning, Rust-written kernels, memory optimization, parallel reductions and scans, multi-stream pipelines, matrix multiplication benchmarked against cuBLAS, a Monte Carlo option pricer validated against a closed formula, and a complete batched inference application measured against a Python baseline. We'll check every result against a CPU reference, and the reports will give accurate numbers, including where libraries outperform hand-written kernels and where experimental toolchains are still a work in progress.Key LearningsLaunch, synchronize, and verify GPU kernels with ownership-managed device memory.Write real CUDA kernels using Rust-CUDA and cuda-oxide.Plan grids, blocks, and warps for 2D workloads.Accelerate transfer speeds with pinned memory and coalesced access patterns.Build race-free thread cooperation using shared memory, barriers, and atomics.Overlap transfers with computation using streams, events, and async Rust pipelines.Optimize matrix multiplication and benchmark against cuBLAS ceiling.Wrap CUDA C library safely with handles, error enums, and Drop.Ship complete batched GPU inference application against Python baselines.Diagnose performance with Nsight Systems, Nsight Compute, and compute-sanitizer.Table of ContentNew Beneficiary of GPU ComputingThinking in ThreadsCommanding GPUWriting GPU KernelsCleaner Kernels with cuda-oxideMastering GPU MemoryMaking Threads CooperateKeeping GPU BusyDelivering Real MathBorrowing NVIDIA's MuscleShipping Complete GPU ApplicationProving Performance This item is printed on demand. Shipping may be from our Sydney, NSW warehouse or from our UK or US warehouse, depending on stock availability.…

- Softcover
- Print on Demand
Seller: buchversandmimpf2000, Emtmannsberg, BAYE, Germanybuchversandmimpf2000
Contact seller5-star sellerCondition: New
US$ 75.49
US$ 68.81 shippingShips from Germany to U.S.A.Quantity: 1 available
Taschenbuch. Condition: Neu. This item is printed on demand - Print on Demand Titel. Neuware -If you're a Python pro looking to get the most out of your code with GPUs, then Practical GPU Programming is the right book for you. This book will walk you through the basics of GPU architectures, show you hands-on parallel programming techniques, and give you the know-how to confidently speed up real workloads in data processing, analytics, and engineering.The first thing you'll do is set up the environment, install CUDA, and get a handle on using Python libraries like PyCUDA and CuPy. You'll then dive into memory management, kernel execution, and parallel patterns like reductions and histogram computations. Then, we'll dive into sorting and search techniques, but with a focus on how GPU acceleration transforms business data processing. We'll also put a strong emphasis on linear algebra to show you how to supercharge classic vector and matrix operations with cuBLAS and CuPy. Plus, with batched computations, efficient broadcasting, custom kernels, and mixed-library workflows, you can tackle both standard and advanced problems with ease.Throughout, we evaluate numerical accuracy and performance side by side, so you can understand both the strengths and limitations of GPU-based solutions. The book covers nearly every essential skill and modern toolkit for practical GPU programming, but it's not going to turn you into a master overnight.Key LearningsBoost processing speed and efficiency for data-intensive tasks.Use CuPy and PyCUDA to write and execute custom CUDA kernels.Maximize GPU occupancy and throughput efficiency by using optimal thread block and grid configuration.Reduce global memory bottlenecks in kernels by using shared memory and coalesced access patterns.Perform dynamic kernel compilation to ensure tailored performance.Use CuPy to carry out custom, high-speed elementwise GPU operations and expressions.Implement bitonic and radix sort algorithms for large or batch integer datasets.Execute parallel linear search kernels to detect patterns rapidly.Scale matrix operations using Batched GEMM and high-level cuBLAS routines.Table of ContentIntroduction to GPU FundamentalsSetting up GPU Programming EnvironmentBasic Data Transfers and Memory TypesSimple Parallel PatternsIntroduction to Kernel OptimizationWorking with PyCUDA and CuPy FeaturesPractical Sorting and SearchLinear Algebra Essentials on GPULibri GmbH, Europaallee 1, 36244 Bad Hersfeld 130 pp. Englisch.…

- Softcover
- Print on Demand
Seller: AHA-BUCH GmbH, Einbeck, GermanyAHA-BUCH GmbH
Contact seller5-star sellerCondition: New
US$ 114.48
US$ 34.98 shippingShips from Germany to U.S.A.Quantity: 1 available
Taschenbuch. Condition: Neu. nach der Bestellung gedruckt Neuware - Printed after ordering - If you're a Python pro looking to get the most out of your code with GPUs, then Practical GPU Programming is the right book for you. This book will walk you through the basics of GPU architectures, show you hands-on parallel programming techniques, and give you the know-how to confidently speed up real workloads in data processing, analytics, and engineering.The first thing you'll do is set up the environment, install CUDA, and get a handle on using Python libraries like PyCUDA and CuPy. You'll then dive into memory management, kernel execution, and parallel patterns like reductions and histogram computations. Then, we'll dive into sorting and search techniques, but with a focus on how GPU acceleration transforms business data processing. We'll also put a strong emphasis on linear algebra to show you how to supercharge classic vector and matrix operations with cuBLAS and CuPy. Plus, with batched computations, efficient broadcasting, custom kernels, and mixed-library workflows, you can tackle both standard and advanced problems with ease.Throughout, we evaluate numerical accuracy and performance side by side, so you can understand both the strengths and limitations of GPU-based solutions. The book covers nearly every essential skill and modern toolkit for practical GPU programming, but it's not going to turn you into a master overnight.Key LearningsBoost processing speed and efficiency for data-intensive tasks.Use CuPy and PyCUDA to write and execute custom CUDA kernels.Maximize GPU occupancy and throughput efficiency by using optimal thread block and grid configuration.Reduce global memory bottlenecks in kernels by using shared memory and coalesced access patterns.Perform dynamic kernel compilation to ensure tailored performance.Use CuPy to carry out custom, high-speed elementwise GPU operations and expressions.Implement bitonic and radix sort algorithms for large or batch integer datasets.Execute parallel linear search kernels to detect patterns rapidly.Scale matrix operations using Batched GEMM and high-level cuBLAS routines.Table of ContentIntroduction to GPU FundamentalsSetting up GPU Programming EnvironmentBasic Data Transfers and Memory TypesSimple Parallel PatternsIntroduction to Kernel OptimizationWorking with PyCUDA and CuPy FeaturesPractical Sorting and SearchLinear Algebra Essentials on GPU.…

- Softcover
- Print on Demand
Seller: preigu, Osnabrück, Germanypreigu
Contact seller5-star sellerCondition: New
US$ 72.30
US$ 80.28 shippingShips from Germany to U.S.A.Quantity: 5 available
Taschenbuch. Condition: Neu. Practical GPU Programming | High-performance computing with CUDA, CuPy, and Python on modern GPUs | Maris Fenlor | Taschenbuch | Englisch | 2025 | GitforGits | EAN 9789349174795 | Verantwortliche Person für die EU: Libri GmbH, Europaallee 1, 36244 Bad Hersfeld, gpsr[at]libri[dot]de | Anbieter: preigu Print on Demand. …

- Softcover
- Print on Demand
Seller: preigu, Osnabrück, Germanypreigu
Contact seller5-star sellerCondition: New
US$ 78.49
US$ 80.28 shippingShips from Germany to U.S.A.Quantity: 5 available
Taschenbuch. Condition: Neu. GPU Programming using Rust and CUDA | Exploring Rust's potential in GPU and parallel computing using Rust-CUDA, cuda-oxide, and RustaCUDA | Maris Fenlor | Taschenbuch | Englisch | 2026 | GitforGits | EAN 9789349174375 | Verantwortliche Person für die EU: Libri GmbH, Europaallee 1, 36244 Bad Hersfeld, gpsr[at]libri[dot]de | Anbieter: preigu Print on Demand. …

- Softcover
- Print on Demand
Seller: AHA-BUCH GmbH, Einbeck, GermanyAHA-BUCH GmbH
Contact seller5-star sellerCondition: New
US$ 130.59
US$ 34.98 shippingShips from Germany to U.S.A.Quantity: 2 available
Taschenbuch. Condition: Neu. nach der Bestellung gedruckt Neuware - Printed after ordering - C++ has been the go-to for GPU programming for almost 20 years. Can Rust do the job, and how well This book is all about getting hands-on with different toolchains that connect Rust to NVIDIA hardware. There's RustaCUDA for safe host-side control, the Rust-CUDA project for writing kernels in pure Rust, and NVIDIA's experimental cuda-oxide compiler with its typed launches and async execution graphs.We're going to build one Cargo workspace that keeps on growing. It'll include device queries, launch planning, Rust-written kernels, memory optimization, parallel reductions and scans, multi-stream pipelines, matrix multiplication benchmarked against cuBLAS, a Monte Carlo option pricer validated against a closed formula, and a complete batched inference application measured against a Python baseline. We'll check every result against a CPU reference, and the reports will give accurate numbers, including where libraries outperform hand-written kernels and where experimental toolchains are still a work in progress.Key LearningsLaunch, synchronize, and verify GPU kernels with ownership-managed device memory.Write real CUDA kernels using Rust-CUDA and cuda-oxide.Plan grids, blocks, and warps for 2D workloads.Accelerate transfer speeds with pinned memory and coalesced access patterns.Build race-free thread cooperation using shared memory, barriers, and atomics.Overlap transfers with computation using streams, events, and async Rust pipelines.Optimize matrix multiplication and benchmark against cuBLAS ceiling.Wrap CUDA C library safely with handles, error enums, and Drop.Ship complete batched GPU inference application against Python baselines.Diagnose performance with Nsight Systems, Nsight Compute, and compute-sanitizer.Table of ContentNew Beneficiary of GPU ComputingThinking in ThreadsCommanding GPUWriting GPU KernelsCleaner Kernels with cuda-oxideMastering GPU MemoryMaking Threads CooperateKeeping GPU BusyDelivering Real MathBorrowing NVIDIA's MuscleShipping Complete GPU ApplicationProving Performance. …