GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems

Prijzen vanaf
9,25

Uitgelicht

VERGELIJK ALLE AANBIEDERS (2)

Beschrijving

GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems

Vergelijk aanbieders (2)

Sorteren op:

€ 9,25 Gratis verzending

€ 9,25 Gratis verzending

Beschrijving (0)

GPU Kernel Engineering for LLM Inference: CUDA, Triton, and Flash Attention Optimization High-Throughput AI Production Systems


Productspecificaties

Merk Independently Published
EAN
  • 9798185800379

Uitgelichte Keuze
9,25
Naar shop