Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding (Scaling AI Systems Series)

Prices from
7.40

Featured

COMPARE ALL WEBSHOPS (2)

Description

Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding (Scaling AI Systems Series)

Compare webshops (2)

Shop
Price
£ 7.40
£ 7.40
Description (0)

Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding (Scaling AI Systems Series)


Product specifications

Brand Independently Published
EAN
  • 9798192412626

Featured Choice
£ 7.40
To Shop