Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding (Scaling AI Systems Series)

Preise von
9,28

Hervorgehoben

ALLE WEBSHOPS VERGLEICHEN (2)

Beschreibung

Amazon Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding (Scaling AI Systems Series)

Webshops vergleichen (2)

Shop
Preis
Versandkosten
Total price
9,28 
Gratis
9,28 
Zum Shop
Gratis Shipping Costs
9,28 
Gratis
9,28 
Zum Shop
Gratis Shipping Costs
Beschreibung (1)

Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding (Scaling AI Systems Series)


Produktspezifikationen

Marken Independently Published
EAN
  • 9798192412626

Hervorgehobene Wahl
9,28 
Zum Shop