<?xml version="1.0" encoding="UTF-8"?>
<feed xmlns="http://www.w3.org/2005/Atom" xml:lang="en">
  <title>prokube.ai Blog</title>
  <id>https://prokube.ai/en/blog/</id>
  <updated>2026-08-19T00:00:00.000Z</updated>
  <link href="https://prokube.ai/en/blog/" />
  <link href="https://prokube.ai/blog/atom.xml" rel="self" type="application/atom+xml" />
  <subtitle>Articles about sovereign AI infrastructure, Kubernetes, MLOps, and agentic workloads.</subtitle>
  <entry>
    <title>How to Run Qwen 3.8 27B on Kubernetes with KServe</title>
    <id>https://prokube.ai/en/blog/serving-qwen38-27b-kserve-vllm/</id>
    <link href="https://prokube.ai/en/blog/serving-qwen38-27b-kserve-vllm/" />
    <published>2026-08-19T00:00:00.000Z</published>
    <updated>2026-08-19T00:00:00.000Z</updated>
    <author><name>Dr. Christian Geier</name></author>
    <summary>A complete KServe setup for Qwen3.8 27B, including a reusable vLLM runtime, BF16 and FP8 deployment settings, H100 NVL validation, and the most important tuning knobs.</summary>
    <category term="Qwen3.8" />
    <category term="KServe" />
    <category term="vLLM" />
    <category term="Kubernetes" />
    <category term="LLMs" />
  </entry>
  <entry>
    <title>AI Gateway, API Gateway, Gateway API, and friends: A Map Through the Gateway Confusion</title>
    <id>https://prokube.ai/en/blog/ai-gateway-api-gateway-gateway-api/</id>
    <link href="https://prokube.ai/en/blog/ai-gateway-api-gateway-gateway-api/" />
    <published>2026-06-21T00:00:00.000Z</published>
    <updated>2026-06-21T00:00:00.000Z</updated>
    <author><name>Dr.-Ing. Henrik Steude</name></author>
    <summary>Why &quot;gateway&quot; in cloud native and AI now means almost anything, and how to avoid losing the plot completely.</summary>
    <category term="AI Gateway" />
    <category term="Kubernetes" />
    <category term="Gateway API" />
    <category term="LLMs" />
    <category term="Agentic AI" />
  </entry>
  <entry>
    <title>Self-Hosting Gemma 4 on Kubernetes with KServe and vLLM</title>
    <id>https://prokube.ai/en/blog/self-hosting-gemma-4-kubernetes-kserve-vllm/</id>
    <link href="https://prokube.ai/en/blog/self-hosting-gemma-4-kubernetes-kserve-vllm/" />
    <published>2026-04-10T00:00:00.000Z</published>
    <updated>2026-04-10T00:00:00.000Z</updated>
    <author><name>Reyan Korel Erben</name></author>
    <author><name>Dr. Christian Geier</name></author>
    <summary>How we ran the Gemma 4 31B instruction-tuned model on a single A100 80GB GPU using KServe and vLLM, including the custom runtime, PVC setup, tuning flags, verification, monitoring, and troubleshooting notes.</summary>
    <category term="Gemma 4" />
    <category term="KServe" />
    <category term="vLLM" />
    <category term="Kubernetes" />
    <category term="LLMs" />
  </entry>
</feed>
