GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems (AI Infrastructure, Hardware & Compiler Series)

Prices from
7.48

Sponsored links

COMPARE ALL WEBSHOPS (2)

Description

GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems (AI Infrastructure, Hardware & Compiler Series)

Compare webshops (2)

Sponsored links · Some shops pay us a fee

Sort by:

£ 7.48

£ 7.48

Description (0)

GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems (AI Infrastructure, Hardware & Compiler Series)


Product specifications

Brand Independently Published
EAN
  • 9798185800379

Featured Choice
£ 7.48
To Shop