Concept
The model has no memory — only whatever fits in the window.
Every turn re-sends the conversation. When it no longer fits, something has to be dropped.
Learn what a token is, how context windows fill up, and why cost and quality both hinge on context. Free, one sitting.
Why this matters
Teams routinely ship AI features without knowing what a request costs or why quality falls apart in long conversations. Both trace to the same thing: the context window. Understand tokens and you can forecast spend, spot the moment a chat starts losing the thread, and design around it.
What you'll cover
What you'll understand
A look inside
The real thing — not a mockup of it.
Concept
Every turn re-sends the conversation. When it no longer fits, something has to be dropped.
How it fits
Both are metered, usually at different rates, on every single call.
Apply
How it works
A plain-language walkthrough of the idea itself, no prior context assumed.
A simple diagram or example showing how it actually fits together.
One quick check that you can recognise it, not just recall it.
Useful for
Want to go deeper? Explore AI Product Manager →
Frequently asked
No. A token is roughly three-quarters of an English word on average, but punctuation, code, and non-English text tokenize very differently.
How tokens are counted, how context windows fill and truncate, and how to forecast cost and quality from both.
Yes. Quick Lessons are free, short, and do not require a paid plan — sign in only to save your progress.