LIBRISTO
LIBROAMANTO
obowiązkowe
Zostań członkiem wspólnoty miłośników książek z całego świata i zyskaj mnóstwo korzyści. Załóż konto bezpłatnie
0
Darmowa dostawa z usługą Inpost oraz Orlen od 299.00 zł
DPD Kurier 12.99 Poczta Polska 18.99 Paczkomat 13.99 InPost Kurier 12.99 Punkt DPD 11.99

Darmowa dostawa dla zamówień powyżej 299,00 zł.

CUDA Systems Engineering

Bare-Metal Kernels, Memory Hierarchies, and Enterprise GPU Optimization

Język AngielskiAngielski
Książka Miękka
Książka CUDA Systems Engineering Elmer Robinson
Kod Libristo: 53273169
Wydawnictwo Independently published, lipiec 2026
Stop leaving teraflops on the table. Engineer bare-metal kernels and deploy high-throughput GPU infr... Cały opis
? points 82 b Zapowiedź Zapowiedź Nowość Nowość
147.84
Zapowiedź Wydanie 24. 07. 2026

Nawet do 30 dni na zwrot

Stop leaving teraflops on the table. Engineer bare-metal kernels and deploy high-throughput GPU infrastructure at enterprise scale.

Writing CUDA code that successfully compiles is a baseline skill. Writing CUDA code that commands modern NVIDIA architectures and scaling it across multi-tenant enterprise clusters is a hardcore systems engineering discipline. When infrastructure compute costs millions, uncoalesced memory reads, warp divergence, and inefficient host-to-device transfers are catastrophic failures of design.

CUDA Systems Engineering is the definitive operational manual for infrastructure architects and bare-metal programmers. We bypass the introductory tutorials and dive straight into the brutal realities of the GPU memory wall, instruction-level parallelism, and large-scale deployment.

You will learn to tear down high-level abstractions and control the hardware at the atomic level. From orchestrating data through the Hopper Tensor Memory Accelerator (TMA) using raw PTX to hard-partitioning multi-tenant workloads via Multi-Instance GPU (MIG), this playbook gives you the power to write and deploy code that executes at the theoretical limit of the silicon.

Inside this manual, you will execute:
Bare-Metal Kernel Optimization: Mastering the nvcc pipeline, PTX as a virtual ISA, and SASS profiling to keep execution pipelines completely saturated.

Defeating the Memory Wall: Forcing perfect memory coalescing, eliminating distributed shared memory bank conflicts, and leveraging the TMA engine for bulk asynchronous transfers.

Lock-Free GPU Architecture: Implementing device-side queues, warp-level atomics, and cooperative groups to bypass host serialization constraints.

Tensor Core Weaponization: Exploiting mixed-precision arithmetic, FP8 pipelines, and MMA instructions to push matrix workloads to maximum throughput.

Enterprise Infrastructure Deployment: Scaling your optimized kernels into production using MIG slicing, NVLink peer-to-peer DMA, and Kubernetes container passthrough.

Who is this for?

This manual is built exclusively for HPC Engineers, AI Infrastructure Architects, Low-Latency Systems Programmers, and Technical Leads building mission-critical computing clusters. If your software runs on enterprise hardware and every wasted clock cycle is a massive financial leak, this is your blueprint for survival.
Stop treating the GPU like a black box. Grab your copy, saturate your pipelines, and dominate the hardware today.

Aktorka & Poliglotka
EWA KASP dla
Odtworzyć wideo
Ewa Kasp
Libristo ma największy wybór literatury obcojęzycznej. Dlatego tutaj kupuję swoje książki.
Podaruj tę książkę jeszcze dziś
To łatwe
1 Dodaj książkę do koszyka i wybierz „dostarczyć jako prezent” 2 W odpowiedzi wyślemy Ci bon 3 Książka dotrze na adres obdarowanego

Logowanie

Zaloguj się do swojego konta. Nie masz jeszcze konta Libristo? Utwórz je teraz!

 
obowiązkowe
obowiązkowe

Nie masz konta? Zyskaj korzyści konta Libristo!

Dzięki kontu Libristo będziesz mieć wszystko pod kontrolą.

Utwórz konto Libristo
Doradca książkowy Libroamiko
Cześć, jestem Libroamiko, w czym mogę pomóc?