Models & limits
Pick the right Claude model and understand what it can and can't hold in mind.
Pick the model for the task
Claude comes in a range of sizes that trade speed against depth. A fast model is right for quick edits, simple questions, and high-volume work where latency matters. A deeper model earns its cost on architecture, hard reasoning, and long-document analysis where getting it right is the point. Defaulting to the strongest model for everything is slow and wasteful; defaulting to the fastest for everything gets you fluent mistakes on the hard tasks. The habit worth building is choosing deliberately based on what the task actually demands.
Know what fits in mind
Claude's context window — everything it can hold in mind at once, including your prompt, the conversation, and any files — is large, but it is not infinite, and relevance matters more than raw size. You can give it long documents, but padding the window with unrelated material dilutes the signal and can degrade the answer. Understanding this limit is what lets you use the big window well: give Claude the material that bears on the task, and leave out the rest.
Which Claude model should I use?
A fast model for quick edits and high-volume work, a deeper model for architecture, hard reasoning, and long documents. Choose based on what the task demands rather than always using one.
What is Claude's context window?
It is everything Claude can consider at once — your prompt, the conversation, and any files. Claude's window is large, letting it reason over long documents, but relevance still matters more than size.
Does giving Claude more context always help?
No. Padding the context window with unrelated material dilutes the signal and can worsen the answer. Include what bears on the task and leave the rest out.