Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding (Scaling AI Systems Series)

Preise von
9,28

Gesponserte Links

ALLE WEBSHOPS VERGLEICHEN (2)

Beschreibung

Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding (Scaling AI Systems Series)

Webshops vergleichen (2)

Gesponserte Links · Einige Shops zahlen uns eine Vergütung

Sortieren nach:

9,28 € Gratis Versand

9,28 € Gratis Versand

Beschreibung (0)

Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding (Scaling AI Systems Series)


Produktspezifikationen

Marken Independently Published
EAN
  • 9798192412626

Hervorgehobene Wahl
9,28 €
Zum Shop