<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0"><channel><title>GPU</title><link>https://sreai.net/en/tags/gpu/</link><description>Practical notes on SRE, cloud-native reliability, and AI infrastructure.</description><language>en</language><lastBuildDate>Thu, 13 Aug 2026 23:52:33 +0800</lastBuildDate><item><title>LLM Infrastructure Setup and GPU Cluster Performance Tuning</title><link>https://sreai.net/en/posts/llm-gpu-infra/</link><guid isPermaLink="true">https://sreai.net/en/posts/llm-gpu-infra/</guid><pubDate>Wed, 22 Jul 2026 00:00:00 +0000</pubDate><description>Build a large language model inference infrastructure from scratch, covering vLLM distributed deployment, GPU memory optimization strategies, and NVIDIA MIG partitioning practices.</description><category>Artificial Intelligence</category><category>LLM</category><category>GPU</category><category>AI</category></item></channel></rss>