<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
<channel>
<title>Deep Variance blog</title>
<link>https://www.deepvariance.com/blog</link>
<description>Notes from Deep Variance on verifying and optimizing AI inference infrastructure, and on the products we build.</description>
<language>en</language>
<lastBuildDate>Wed, 07 Oct 2026 00:00:00 GMT</lastBuildDate>
<atom:link href="https://www.deepvariance.com/blog/rss.xml" rel="self" type="application/rss+xml"/>
<item><title>Why our own products run on our inference stack first</title><link>https://www.deepvariance.com/blog/our-products-first</link><guid isPermaLink="true">https://www.deepvariance.com/blog/our-products-first</guid><pubDate>Wed, 07 Oct 2026 00:00:00 GMT</pubDate><description>Deep Variance optimizes inference for open models. Every change runs in our own production before it reaches a customer&apos;s GPUs. Here is why.</description><category>inference optimization</category><category>open models</category><category>how we work</category></item>
<item><title>Where the cost goes when you serve an open video model</title><link>https://www.deepvariance.com/blog/where-video-inference-cost-goes</link><guid isPermaLink="true">https://www.deepvariance.com/blog/where-video-inference-cost-goes</guid><pubDate>Wed, 07 Oct 2026 00:00:00 GMT</pubDate><description>Open video models cost far more to serve than text models. A plain-English walk through where the GPU time goes, and which levers move it.</description><category>video generation</category><category>inference optimization</category><category>GPUs</category><category>open models</category></item>
</channel>
</rss>
