Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding

Prijzen vanaf
9,28

Uitgelicht

VERGELIJK ALLE AANBIEDERS (2)

Beschrijving

Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding

Vergelijk aanbieders (2)

Sorteren op:

€ 9,28 Gratis verzending

€ 9,28 Gratis verzending

Beschrijving (0)

Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding


Productspecificaties

Merk Independently Published
EAN
  • 9798192412626

Uitgelichte Keuze
9,28
Naar shop