GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems

Prezzi da
9,08

Link sponsorizzati

CONFRONTA TUTTI I WEBSHOP (2)

Descrizione

GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems

Confronta i webshop (2)

Link sponsorizzati · Alcuni negozi ci pagano un compenso

Ordina per:

9,08 €

9,08 €

Descrizione (0)

GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems


Specifiche del prodotto

Marchio Independently Published
EAN
  • 9798185800379

Scelta in evidenza
9,08 €
Vai allo shop