Token
A token is the unit language models read and generate — roughly three-quarters of a word in English. Model pricing is per token, so tokens are the unit AI features are actually billed in.
Both input and output are charged, usually at different rates. A support bot that sends ten pages of retrieved documents with every question pays for those ten pages on every single question, which is how a feature that seemed cheap in testing becomes expensive at volume.
Cost control is mostly retrieval discipline: send the three relevant paragraphs rather than the whole manual, cache what repeats, and pick the smallest model that does the job.
Why it matters
If you're quoted a monthly AI running cost, tokens are what's being estimated. Ask what assumptions about volume and context size sit behind the number.
Related terms
More in AI Automation & Agents
Need this built rather than explained?
We publish every price we charge, and you get a quote in writing before anything starts.