<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Metering on Rik Kisnah - Blog</title><link>https://www.rik-kisnah.ai/tags/metering/</link><description>Recent content in Metering on Rik Kisnah - Blog</description><generator>Hugo</generator><language>en</language><lastBuildDate>Tue, 16 Jun 2026 09:00:00 -0700</lastBuildDate><atom:link href="https://www.rik-kisnah.ai/tags/metering/feed.xml" rel="self" type="application/rss+xml"/><item><title>Design a Token Usage and Limits Service</title><link>https://www.rik-kisnah.ai/teach/systems/design-a-token-usage-and-limits-service/</link><pubDate>Tue, 16 Jun 2026 09:00:00 -0700</pubDate><guid>https://www.rik-kisnah.ai/teach/systems/design-a-token-usage-and-limits-service/</guid><description>Every request to a language model burns tokens. Somebody has to answer &amp;lsquo;may this user spend more?&amp;rsquo; in a few milliseconds, for personal accounts and for companies with a hundred engineers under one bill, and then record what was actually spent. Design that gate: the entities, the three endpoints, the counters, and the one rule that keeps it from ever blocking the model by accident.</description></item></channel></rss>