<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Llm-Inference on CuraDevOps</title><link>https://curadevops.metacog.co.kr/tags/llm-inference/</link><description>Recent content in Llm-Inference on CuraDevOps</description><generator>Hugo</generator><language>en-us</language><lastBuildDate>Thu, 16 Jul 2026 12:15:50 +0000</lastBuildDate><atom:link href="https://curadevops.metacog.co.kr/tags/llm-inference/index.xml" rel="self" type="application/rss+xml"/><item><title>LLM-D: Kubernetes-Native Distributed LLM Inference Announced</title><link>https://curadevops.metacog.co.kr/insights/2026-07-16-llm-d-kubernetes-native-distributed-inference/</link><pubDate>Thu, 16 Jul 2026 12:15:50 +0000</pubDate><guid>https://curadevops.metacog.co.kr/insights/2026-07-16-llm-d-kubernetes-native-distributed-inference/</guid><description>&lt;ul>
&lt;li>&lt;strong>Platform/SRE — Learn:&lt;/strong> A new project enabling distributed LLM inference natively on Kubernetes—worth evaluating for teams planning AI/ML serving infrastructure, but no GA status confirmed and no enrichment signals to anchor a higher verdict.&lt;/li>
&lt;li>&lt;strong>CI/CD — Skip&lt;/strong>&lt;/li>
&lt;li>&lt;strong>Leader — Learn:&lt;/strong> Signals a maturing ecosystem for running LLM inference on existing Kubernetes infrastructure, relevant to strategy around AI workload hosting; no near-term decision required.&lt;/li>
&lt;/ul></description></item></channel></rss>