
5/12/2026
What this post added
Introduces a deployment blueprint for running Red Hat AI Inference on CoreWeave Kubernetes Service (CKS) to enable hybrid inference across on-premises and cloud environments. This reference architecture allows enterprises to run the same open-source inference stack on-premises and on CoreWeave, leveraging Kubernetes-native control and open runtimes. It complements CoreWeave's existing inference portfolio by offering a supported path for self-managed inference on CKS using Red Hat's stack, building on the collaboration around the llm-d project.