Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding

Prix à partir de
9,19

En vedette

COMPARER TOUS LES MAGASINS EN LIGNE (2)

Description

Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding

Comparer les boutiques en ligne (2)

Trier par:

9,19 € Livraison gratuite

9,19 € Livraison gratuite

Description (0)

Inference at Full Throttle: LLM serving performance with vLLM, quantization, KV cache tuning and speculative decoding


Spécifications du produit

Marque Independently Published
EAN
  • 9798192412626

Choix en vedette
9,19 €
Voir l’offre