Both your input and the model's output count as tokens, and you pay for both. A long document you paste in plus a long answer back can add up fast. Understanding tokens is how you predict and control what an AI feature will cost to run.
Tokens also define how much a model can consider at once. That limit is called the context window, and it's measured in tokens too.
Updated July 2026