Context Window
The context window is how much text – measured in tokens – a model can keep in mind at once. Think of it as the model’s short-term memory size.
Larger windows (64K, 128K, 200K+ tokens) let a model work on whole books or codebases, but cost more memory and slow things down. In practice, ~64K covers most everyday work.
