vLLM Serving: High‑Throughput LLM APIs with PagedAttention and KV Cache Tuning

vLLM Serving: High‑Throughput LLM APIs with PagedAttention and KV Cache Tuning

NobleTrex Press

Pages: 238, Paperback, NobleTrex Press

Compare prices (1 shop)

shop Price Action
30,31 GBP Go to shop

Similar products