<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>On-Prem AI · privateai.blog</title><description>Building and running AI on hardware you own — GPUs, clusters, interconnects, power, and the measured numbers behind them.</description><link>https://www.privateai.blog/</link><language>en</language><item><title>How I make AI videos with my own GPUs: 3090 vs 4090, first results</title><link>https://www.privateai.blog/en/posts/teaching-my-logo-to-wave</link><guid isPermaLink="true">https://www.privateai.blog/en/posts/teaching-my-logo-to-wave</guid><description>A fully self-hosted pipeline: script to voiceover to render, no cloud in the path. Two mascots, worst-to-best, the 3090/4090/DGX numbers, and why the right model depends on your character.</description><pubDate>Sun, 30 Aug 2026 00:00:00 GMT</pubDate><category>on-prem-ai</category><category>video-generation</category><category>wan</category><category>comfyui</category><category>rtx-4090</category><category>rtx-3090</category><category>benchmarks</category><category>open-source</category><category>mascot</category><category>campaign:launch</category></item><item><title>The weights you run were probably not made by the lab that trained them</title><link>https://www.privateai.blog/en/posts/who-made-the-weights-you-run</link><guid isPermaLink="true">https://www.privateai.blog/en/posts/who-made-the-weights-you-run</guid><description>Nvidia is reported to be buying Hugging Face. A good week to learn what is really on the shelf: master weights, vendor builds, and the community quants most of us run.</description><pubDate>Sun, 30 Aug 2026 00:00:00 GMT</pubDate><category>on-prem-ai</category><category>hugging-face</category><category>quantization</category><category>nvfp4</category><category>fp8</category><category>bf16</category><category>glm</category><category>vllm</category><category>open-weights</category><category>model-archive</category><category>supply-chain</category></item><item><title>GPU passthrough cost me nothing — until the GPUs had to talk to each other</title><link>https://www.privateai.blog/en/posts/gpu-passthrough-cost-me-nothing</link><guid isPermaLink="true">https://www.privateai.blog/en/posts/gpu-passthrough-cost-me-nothing</guid><description>A year after virtualizing my edge AI supercomputer, I finally measured what the hypervisor costs: single-GPU work is free — the VM even won — and four-way tensor-parallel serving pays a third.</description><pubDate>Sun, 23 Aug 2026 00:00:00 GMT</pubDate><category>on-prem-ai</category><category>vfio</category><category>gpu-passthrough</category><category>proxmox</category><category>vllm</category><category>nccl</category><category>tensor-parallel</category><category>sdxl</category><category>rtx-3090</category><category>benchmarks</category><category>virtualization</category><category>campaign:launch</category></item></channel></rss>