Last updated: 2026-06-10
Your lm9 LLM plan includes a monthly token allowance. This page explains what tokens are, what counts toward your limit, and how to get the most from your plan. Plan prices and allowances are on our pricing page.
Tokens are small pieces of text the AI model reads and writes — roughly a word or part of a word. We measure usage in tokens because that is how the underlying models bill compute. Your monthly allowance is a cap on how many tokens you can use across chat and API access in a billing period.
Rough guide for English: about 4 characters ≈ 1 token. A short question might use a few hundred tokens total; a long conversation or large pasted document can use thousands or more in a single request.
Each AI request uses two kinds of tokens:
Your usage = input tokens + output tokens, added together across all qualifying requests in the month.
Usage is counted when you:
Counts come from the model runtime after each successful response. Your balance resets on the first day of each calendar month at 00:00 UTC (not necessarily your local timezone or subscription anniversary date).
Token limits were increased by 50% in June 2026. Current monthly allowances:
When you reach your monthly limit, new chat and API requests pause until the next UTC month or until you upgrade your plan. If that happens, you will see a message showing how many tokens you have used.
Model size does not multiply tokens. A smaller model and a larger one use roughly the same token count for the same text. Plans differ mainly in which models you can access, speed, and priority — not in a separate per-model token multiplier.
See our Terms of Service for billing and trial details, or contact us if something looks wrong with your usage.