<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Smart-Inference on Ilscipio - Enterprise Commerce Experts</title>
    <link>https://www.ilscipio.com/en/tags/smart-inference/</link>
    <description>Recent content in Smart-Inference on Ilscipio - Enterprise Commerce Experts</description>
    <generator>Hugo</generator>
    <language>en-US</language>
    <lastBuildDate>Tue, 18 Aug 2026 00:00:00 +0000</lastBuildDate>
    <atom:link href="https://www.ilscipio.com/en/tags/smart-inference/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>I Forgot About My AI Inference Bill for a Month</title>
      <link>https://www.ilscipio.com/en/blog/real-cost-ai-inference/</link>
      <pubDate>Tue, 18 Aug 2026 00:00:00 +0000</pubDate>
      <guid>https://www.ilscipio.com/en/blog/real-cost-ai-inference/</guid>
      <description>&lt;p&gt;I run a service that routes AI inference to the cheapest available GPU. So I spend an unreasonable amount of time staring at provider pricing dashboards. And the thing I keep learning is that the sticker price on an AI model tells you almost nothing about what inference actually costs you.&lt;/p&gt;&#xA;&lt;p&gt;Take a 70B parameter model. Provider A charges $0.60 per million input tokens. Provider B charges $0.90. Provider A is cheaper, right?&lt;/p&gt;</description>
    </item>
    <item>
      <title>How a Spreadsheet Turned Into an AI Inference Router</title>
      <link>https://www.ilscipio.com/en/blog/route-ai-cheapest-gpu/</link>
      <pubDate>Fri, 14 Aug 2026 00:00:00 +0000</pubDate>
      <guid>https://www.ilscipio.com/en/blog/route-ai-cheapest-gpu/</guid>
      <description>&lt;p&gt;I accidentally stumbled on spot market pricing for AI compute. So naturally, I built a service around it.&lt;/p&gt;&#xA;&lt;p&gt;The short version: GPU prices vary wildly between providers. Same model, same output, different costs depending on when and where you run it. An H100 on one provider costs $2.10/hr right now. Same GPU on another, $3.80/hr. Tomorrow those numbers will be different again.&lt;/p&gt;&#xA;&lt;p&gt;I found out the hard way. I was running AIVory Guard on a single provider, picked because it was cheapest when I checked. A month later the bill was 40% higher than I expected. The prices had moved and I had not noticed.&lt;/p&gt;</description>
    </item>
  </channel>
</rss>
