<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Token on Enterium</title>
    <link>https://aienterium.top/tags/token/</link>
    <description>Recent content in Token on Enterium</description>
    <generator>Hugo</generator>
    <language>en</language>
    <lastBuildDate>Sat, 25 Jul 2026 12:47:42 +0000</lastBuildDate>
    <atom:link href="https://aienterium.top/tags/token/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Token pricing models: calculate real API costs</title>
      <link>https://aienterium.top/posts/token-pricing-models-calculate-real-api-costs/</link>
      <pubDate>Thu, 23 Jul 2026 00:00:00 +0000</pubDate>
      <author>Enterium Editorial</author>
      <guid>https://aienterium.top/posts/token-pricing-models-calculate-real-api-costs/</guid>
      <description>DeepSeek V4 Flash lists at $0.14 per million input tokens, yet most enterprises still guess at their total spend.</description>
    </item>
    <item>
      <title>LLM cost compounding: Stop the 35x budget spike</title>
      <link>https://aienterium.top/posts/llm-cost-compounding-stop-the-35x-budget-spike/</link>
      <pubDate>Sun, 19 Jul 2026 00:00:00 +0000</pubDate>
      <author>Enterium Editorial</author>
      <guid>https://aienterium.top/posts/llm-cost-compounding-stop-the-35x-budget-spike/</guid>
      <description>A single coding agent session on Claude Opus 4.6 can burn through $7 if it makes 200 API calls. This isn&#39;t a bug; it&#39;s the math of agentic systems.</description>
    </item>
    <item>
      <title>LLM pricing traps: Stop burning monthly</title>
      <link>https://aienterium.top/posts/llm-pricing-traps-stop-burning-monthly/</link>
      <pubDate>Tue, 07 Jul 2026 00:00:00 +0000</pubDate>
      <author>Enterium Editorial</author>
      <guid>https://aienterium.top/posts/llm-pricing-traps-stop-burning-monthly/</guid>
      <description>A support bot generated a $14,000 bill answering 30 questions. Learn how model routing and caching cut API spend by 70% without quality loss.</description>
    </item>
    <item>
      <title>AI model endpoints: measure real token speed</title>
      <link>https://aienterium.top/posts/ai-model-endpoints-measure-real-token-speed/</link>
      <pubDate>Thu, 02 Jul 2026 00:00:00 +0000</pubDate>
      <author>Enterium Editorial</author>
      <guid>https://aienterium.top/posts/ai-model-endpoints-measure-real-token-speed/</guid>
      <description>Compare 500+ AI model endpoints to spot 10x pricing gaps. Learn to benchmark latency and context windows using live data from Artificial Analysis.</description>
    </item>
    <item>
      <title>Markdown extraction stops RAG pipeline waste</title>
      <link>https://aienterium.top/posts/markdown-extraction-stops-rag-pipeline-waste/</link>
      <pubDate>Mon, 29 Jun 2026 00:00:00 +0000</pubDate>
      <author>Enterium Editorial</author>
      <guid>https://aienterium.top/posts/markdown-extraction-stops-rag-pipeline-waste/</guid>
      <description>Cut token consumption by 67% by stripping HTML noise. Learn why clean markdown is essential for efficient RAG pipelines and lower costs.</description>
    </item>
    <item>
      <title>Token consumption spikes: Why agents burn 50x</title>
      <link>https://aienterium.top/posts/token-consumption-spikes-why-agents-burn-50x/</link>
      <pubDate>Tue, 16 Jun 2026 00:00:00 +0000</pubDate>
      <author>Enterium Editorial</author>
      <guid>https://aienterium.top/posts/token-consumption-spikes-why-agents-burn-50x/</guid>
      <description>Agentic workflows multiply compute usage by 50x per task. Enterium provides the governance layer to visualize spend before budgets vanish.</description>
    </item>
  </channel>
</rss>
