AWS Machine Learning Blog Follow Disaggregated prefill and decode for LLM inference on SageMaker HyperPod In this post, we show how to implement DPD with vLLM on Amazon SageMaker HyperPod using the HyperPod Inference Operator. https://aws.amazon.com/blogs/machine-learning/disaggregated-prefill-and-decode-for-llm-inference-on-sagemaker-hyperpod/ aws.amazon.com