Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding

Prezzi da
9,02

Link sponsorizzati

CONFRONTA TUTTI I WEBSHOP (2)

Descrizione

Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding

Confronta i webshop (2)

Link sponsorizzati · Alcuni negozi ci pagano un compenso

Ordina per:

9,02 €

9,02 €

Descrizione (0)

Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding


Specifiche del prodotto

Marchio Independently Published
EAN
  • 9798192412626

Scelta in evidenza
9,02 €
Vai allo shop