<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Self-Hosting on CuraDevOps</title><link>https://curadevops.metacog.co.kr/tags/self-hosting/</link><description>Recent content in Self-Hosting on CuraDevOps</description><generator>Hugo</generator><language>en-us</language><lastBuildDate>Thu, 16 Jul 2026 12:15:50 +0000</lastBuildDate><atom:link href="https://curadevops.metacog.co.kr/tags/self-hosting/index.xml" rel="self" type="application/rss+xml"/><item><title>Self-hosting LLMs on Kubernetes with vLLM (CNCF guide)</title><link>https://curadevops.metacog.co.kr/insights/2026-07-16-running-a-self-hosted-llm-in-kubernetes-with-vllm/</link><pubDate>Thu, 16 Jul 2026 12:15:50 +0000</pubDate><guid>https://curadevops.metacog.co.kr/insights/2026-07-16-running-a-self-hosted-llm-in-kubernetes-with-vllm/</guid><description>&lt;ul>
&lt;li>&lt;strong>Platform/SRE — Learn:&lt;/strong> Useful reference for platform engineers evaluating GPU workload patterns on Kubernetes; no forced migration or deadline, but shapes how you&amp;rsquo;d design node pools and scheduling for LLM inference.&lt;/li>
&lt;li>&lt;strong>CI/CD — Skip&lt;/strong>&lt;/li>
&lt;li>&lt;strong>Leader — Learn:&lt;/strong> Relevant context for build-vs-buy decisions on LLM inference — self-hosting via vLLM vs managed API services — but no concrete strategic decision is forced by this content.&lt;/li>
&lt;/ul></description></item></channel></rss>