Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding (Scaling AI Systems Series)

Precios a partir de
9,02

Destacado

COMPARAR TODAS LAS TIENDAS WEB (2)

Descripción

Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding (Scaling AI Systems Series)

Comparar tiendas web (2)

Shop
Precio
9,02 
9,02 
Descripción (0)

Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding (Scaling AI Systems Series)


Especificaciones del producto

Marca Independently Published
EAN
  • 9798192412626

Elección Destacada
9,02 
Ver oferta