GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems (AI Infrastructure, Hardware & Compiler Series)

Preise von
9,34

Gesponserte Links

ALLE WEBSHOPS VERGLEICHEN (2)

Beschreibung

GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems (AI Infrastructure, Hardware & Compiler Series)

Webshops vergleichen (2)

Gesponserte Links · Einige Shops zahlen uns eine Vergütung

Sortieren nach:

9,34 € Gratis Versand

9,34 € Gratis Versand

Beschreibung (0)

GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems (AI Infrastructure, Hardware & Compiler Series)


Produktspezifikationen

Marken Independently Published
EAN
  • 9798185800379

Hervorgehobene Wahl
9,34 €
Zum Shop