AWS Machine Learning Blog Models & Releases · Июл 10, 2026 15:20 imp:40 Disaggregated prefill and decode for LLM inference on SageMaker HyperPod In this post, we show how to implement DPD with vLLM on Amazon SageMaker HyperPod using the HyperPod Inference Operator. Читать оригинал на AWS Machine Learning Blog →