Context Window
How much text an AI model can take into account at once, measured in tokens: the conversation so far, any documents supplied and the answer being written. A longer window lets one conversation run longer and lets the model work across longer documents.
What it means in practice
The GenAI FAQ gives a practical case: NotebookLM runs on Gemini, which has one of the largest windows of the public tools, so it handles many long documents while staying within the sources it is given. The FAQ on paid plans notes that free plans hit a usage limit and make you wait for it to reset, and that paid plans give a longer context window. The figures themselves are on the large language model page.
How we use it
The workflow running cost article shows why conversation length matters for cost: providers bill by the amount of text a model reads and writes, and in a long conversation the earlier turns are sent again with each new one. Keeping a task narrow keeps what the model has to read small.
FAQ
Questions about Context window
Do paid AI plans have a bigger Context window?
Generally yes. The GenAI FAQ says paid plans give a longer context window, so one conversation can run much longer, and that Gemini is the most generous with context on its free plan.
What do you get with paid ChatGPT, Claude or Gemini that the free plan does not give you?
Does a longer Conversation cost more to run?
Yes, where use is billed by text. The workflow running cost article explains that the earlier turns of a conversation are sent again with each new one, so a conversation of twenty turns reads its own history twenty times.
What it means for a Business
A bigger window lets a model work across a longer document or conversation without losing track of the earlier part. Providers do not all publish the figure, so two tools cannot always be compared on it like for like.
Ready to put this to work?
Tell us where your team is with AI and we will tell you honestly what would make the biggest difference.