Product and Model Selection — study note
Product and Model Selection — study note
Domain 3 of the Claude Certified Associate – Foundations exam. Its four objectives ask you to choose the right Claude feature, tell the model families apart, match a model to the task's needs for cost, speed and quality, and manage context limits and memory. This note summarizes each topic and the Anthropic pages behind it. Model versions, context sizes and plans change often; the note deliberately teaches the relative positions and the judgement instead.
Choosing a feature
| You need… | Use |
|---|---|
| A one-off question or quick task | A plain chat |
| The same documents or instructions in many chats | A project |
| One current fact | Web search |
| A cited report from many sources | Research (paid plans; needs web search on; can use limits faster) |
| Output to edit, reuse or share | An artifact (needs code execution and file creation on; starts private to you) |
The model families
| Of these three | Position |
|---|---|
| Haiku | Lowest latency and price |
| Sonnet | Balance of speed and intelligence, for everyday work |
| Opus | The capability end, for complex work |
The lineup also includes other models; the exam's objective names these three.
Choosing a model means balancing capability, speed and cost. You can see the current model by the message box and change it — or the effort and thinking settings — at any point in a conversation; changes apply starting with Claude's next response. On organization plans, an administrator may limit which models appear.
Matching the model to the task
- Efficiency first for high-volume, straightforward work: quick iteration at lower cost, often sufficient; upgrade only for a specific capability gap.
- Capability first for complex work where accuracy outweighs cost: start with the strongest starting point for the task, and optimize to more efficient models later.
- Effort before switching (on models that offer an effort setting; Haiku does not, though you can still switch extended thinking on or off): effort controls how much thinking Claude applies. Lower effort suits routine tasks and stretches usage; raise effort or turn on thinking for complex work, at the cost of time and usage.
- Test it: a good evaluation set specific to your task, run with your actual prompts and data, is the most important step. Speed depends on prompt length, output length and effort, not on the model alone.
Context limits and memory
| Aspect | Usage limit | Length limit |
|---|---|---|
| Controls | How much you use Claude over time | How long one conversation can grow |
| When reached | Wait for the reset, or add usage depending on your plan | Start a new chat, or use a project (a message too long even in a new chat must be shortened) |
- Part of the context window is reserved for Claude's reply. On paid plans with code execution enabled, Claude summarizes earlier messages near the limit — but longer conversations use more of your usage.
- Tools and connectors take up context space and usage; switch off ones a chat does not need.
- Restart well: summarize decisions and open points, then start a new chat with that summary.
- Carrying context: a new chat starts without earlier context unless memory, chat search or project knowledge carries it, depending on your plan and settings (memory and chat search are on by default for some users). Project knowledge is shared across a project's chats, and on paid plans, once it nears the context limit, Claude retrieves only what each question needs. Memory carries your role and preferences into new chats, with each project keeping its own. Chat search, on paid plans, finds earlier conversations. Incognito chats, started outside a project, stay out of memory and history.
Sources
- What are projects? — https://support.claude.com/en/articles/9517075-what-are-projects
- How can I create and manage projects? — https://support.claude.com/en/articles/9519177-how-can-i-create-and-manage-projects
- Retrieval augmented generation (RAG) for projects — https://support.claude.com/en/articles/11473015-retrieval-augmented-generation-rag-for-projects
- When should I use web search, extended thinking, and research? — https://support.claude.com/en/articles/11095361-when-should-i-use-web-search-extended-thinking-and-research
- Use research on Claude — https://support.claude.com/en/articles/11088861-use-research-on-claude
- What are artifacts and how do I use them? — https://support.claude.com/en/articles/17153992-what-are-artifacts-and-how-do-i-use-them
- Models overview — https://platform.claude.com/docs/en/models/overview
- Choosing the right model — https://platform.claude.com/docs/en/about-claude/models/choosing-a-model
- Change the model, effort, and thinking settings — https://support.claude.com/en/articles/8664678-change-the-model-effort-and-thinking-settings
- Get started with Claude — https://support.claude.com/en/articles/8114491-get-started-with-claude
- How do usage and length limits work? — https://support.claude.com/en/articles/11647753-how-do-usage-and-length-limits-work
- How large is the context window on paid Claude plans? — https://support.claude.com/en/articles/8606394-how-large-is-the-context-window-on-paid-claude-plans
- Troubleshoot Claude error messages — https://support.claude.com/en/articles/12466728-troubleshoot-claude-error-messages
- Use Claude’s chat search and memory to build on previous context — https://support.claude.com/en/articles/11817273-use-claude-s-chat-search-and-memory-to-build-on-previous-context