Published on February 3, 2026, the Manning book covers CUDA programming for NVIDIA GPUs, from initial kernels to Flash Attention, and mentions profiling with Nsight Compute.
Published on February 3, 2026, Manning’s “CUDA for LLMs” introduces CUDA programming for NVIDIA GPUs, starting with kernels and progressing to LLM features such as Flash Attention. Its preview also mentions profiling with Nsight Compute; the material is aimed at engineers interested in GPU-level programming and performance analysis.
To verify the scope and topics, consult the book description and preview on the publisher’s page and compare them with the available edition. If you use AI to study the material, do not submit internal code, data, or documents without authorization: follow your organization’s policy and use synthetic or anonymized data where possible.