Intelligent systems engineering library - 9798187898176 - the art of ptx and sass programming: advanced cuda performance engineering: 13 di heller, dion m. (5 risultati)

Lingua: Inglese
Editore: Independently published, 2026
Serie: Libro 2 - GPU & High-Performance Computing Programming Library
- Brossura
Da: PBShop.store US, Wood Dale, IL, U.S.A.PBShop.store US
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 21,24
Spedizione gratuitaSpedito in U.S.A.Quantità: Più di 20 disponibili
PAP. Condizione: New. New Book. Shipped from UK. Established seller since 2000.

Lingua: Inglese
Editore: Independently published, 2026
Serie: Libro 2 - GPU & High-Performance Computing Programming Library
- Brossura
Da: PBShop.store UK, Fairford, GLOS, Regno UnitoPBShop.store UK
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 18,57
EUR 4,85 spedizioneSpedito da Regno Unito a U.S.A.Quantità: Più di 20 disponibili
PAP. Condizione: New. New Book. Shipped from UK. Established seller since 2000.

Lingua: Inglese
Editore: Amazon Digital Services LLC - Kdp Jul 2026, 2026
Serie: Libro 2 - GPU & High-Performance Computing Programming Library
- Brossura
Da: AHA-BUCH GmbH, Einbeck, GermaniaAHA-BUCH GmbH
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 27,81
EUR 30,50 spedizioneSpedito da Germania a U.S.A.Quantità: 2 disponibili
Taschenbuch. Condizione: Neu. Neuware.

Lingua: Inglese
Editore: Independently published, 2026
Serie: Libro 2 - GPU & High-Performance Computing Programming Library
- Brossura
- Print on Demand
Da: California Books, Miami, FL, U.S.A.California Books
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 18,60
Spedizione gratuitaSpedito in U.S.A.Quantità: Più di 20 disponibili
Condizione: New. Print on Demand.

Lingua: Inglese
Editore: Independently Published, 2026
Serie: Libro 2 - GPU & High-Performance Computing Programming Library
- Brossura
- Print on Demand
Da: CitiRetail, Stevenage, Regno UnitoCitiRetail
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 22,18
EUR 43,11 spedizioneSpedito da Regno Unito a U.S.A.Quantità: 1 disponibili
Paperback. Condizione: new. Paperback. Have you ever optimized a CUDA kernel only to discover it still runs far slower than expected with no clear explanation why?If you've experienced that frustration, you're not alone. The answer often isn't hidden in your C++ source code, it's buried in the PTX and SASS instructions your comp…iler generates.Most CUDA programming books teach kernel syntax, memory models, and launch configurations. Those are essential skills, but they only tell part of the story. True GPU optimization begins below the source code, where register allocation, instruction scheduling, memory access, and warp execution determine whether your application fully utilizes the hardware or leaves performance on the table.This book takes you beyond CUDA programming and into CUDA performance engineering.Through fifteen comprehensive chapters, you'll follow the complete journey of a CUDA kernel from high-level source code to PTX intermediate representation and finally to the native machine instructions executed by NVIDIA GPUs. You'll learn how to read GPU disassembly with confidence, identify performance bottlenecks, and make optimization decisions based on evidence rather than guesswork.Inside this book, you'll learn how to: Understand the complete CUDA compilation pipeline, from source code to native GPU instructions.Read and interpret PTX and SASS output to understand exactly what your compiler is generating.Diagnose register pressure, occupancy limitations, and instruction-level bottlenecks.Optimize memory access patterns for maximum bandwidth and cache efficiency.Identify, measure, and eliminate warp divergence that reduces execution efficiency.Understand Tensor Core and matrix acceleration instructions and verify when your kernels are using specialized hardware.Master professional profiling and performance analysis tools used in production GPU development.Build a systematic, repeatable workflow for diagnosing and optimizing CUDA applications with confidence.Learn Through Real Performance InvestigationsEvery chapter is built around practical, real-world examples rather than isolated code fragments. You'll examine kernels before optimization, analyze their generated instructions, identify performance issues, implement targeted improvements, and verify measurable results using industry-standard tools and methodologies.Rather than relying on trial and error, you'll develop the ability to explain why a kernel performs the way it does and how to improve it with precision.Who Should Read This Book?This guide is written for GPU programmers, systems engineers, high-performance computing professionals, machine learning infrastructure engineers, compiler enthusiasts, graduate students, and software developers who want a deeper understanding of CUDA performance. Whether you're building scientific applications, AI frameworks, graphics engines, or large-scale compute systems, the techniques in this book will help you optimize with greater confidence and accuracy.Stop Guessing. Start Understanding.The fastest CUDA developers aren't the ones who memorize optimization tricks they're the ones who understand what the GPU is actually executing.If you're ready to move beyond surface-level optimization and learn how to analyze, diagnose, and maximize GPU performance from the instruction level upward, this book is your definitive guide. This item is printed on demand. Shipping may be from our UK warehouse or from our Australian or US warehouses, depending on stock availability.