GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems

Prijzen vanaf
9,34

Uitgelicht

VERGELIJK ALLE AANBIEDERS (2)

Beschrijving

GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems

Vergelijk aanbieders (2)

Shop
Prijs
Verzendkosten
Totale prijs
9,34
Gratis
9,34
Naar shop
Gratis Shipping Costs
9,34
Gratis
9,34
Naar shop
Gratis Shipping Costs
Beschrijving (0)

GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems


Productspecificaties

Merk Independently Published
EAN
  • 9798185800379

Uitgelichte Keuze
9,34
Naar shop