How to Use Claude Opus 5.5 Effectively: 10 Practical Tips
Get better results from Claude Opus 5.5: pick the right effort level, write briefs it can finish, cut token costs, and know when to switch models.
Claude Opus 5.5 is cheaper and faster than the model it replaces, and Anthropic says it performs at the level of its top model, Fable 5.1, on most work. But a strong model does not guarantee strong results. How you set it up and how you brief it decide whether you get a finished job or three rounds of "that's not what I meant."
This guide covers how to use Claude Opus 5.5 well: the settings that matter, the habits that save the most time, and the moments when a different model is the smarter call. It comes from daily use: we build and ship our own apps with Claude.
What's new in Opus 5.5, in one minute
Before the tips, the facts that change how you should work with it:
- Price: $4 input and $20 output per million tokens through the API, 20% less than Opus 5. Cached input reads cost $0.20 per million, 60% less.
- Speed: output is generated more than 30% faster than Opus 5.
- Default effort is medium, one step lower than Opus 5. This is the setting people most often miss.
- Thinking is always on. You control depth with the effort level, not an on/off switch.
- 1M-token context window and up to 128K tokens of output in a single response.
- Where you get it: the Claude apps on Pro ($20 per month) and Max (from $100 per month) plans, Claude Code, the Claude API as
claude-opus-5-5, and AWS, Google Cloud and Microsoft Azure.
If you are still deciding between models, start with our comparison of Claude Opus 5.5, Fable 5.1 and Sonnet 5.5.
1. Start at the default effort, then move one step at a time
Effort controls how hard Opus 5.5 thinks, how many tool calls it makes, and how many tokens it spends. Anthropic says that at the default medium effort, Opus 5.5 matches what earlier models needed higher settings to reach.
A practical ladder:
| Effort | Use it for |
|---|---|
| low | Quick questions, chat, sorting and tagging, short rewrites |
| medium (default) | Most everyday work: drafts, summaries, single-file code changes |
| high | Important analysis, code that touches several files, anything you will publish |
| xhigh | Long agent runs, hard debugging, big refactors |
| max | Rare cases where correctness matters more than time or cost |
If an answer feels shallow, raise effort by one step before switching to a more expensive model. It is the cheapest upgrade you have. Claude Code and the API both let you set it.
2. Give it the whole job up front
Opus 5.5 is built for long, multi-step work, and it does that best when the full task is on the table from the first message. Drip-feeding requirements ("oh, and also…") forces it to redo work and burns tokens.
A brief that works has four parts:
- Goal: what you want and why it matters.
- Context: the files, data, audience or background it needs.
- Constraints: tone, length, format, and what it must not touch.
- Done means: how you, or it, will check the result.
Here is the template we reuse:
Goal: Rewrite our app's onboarding screen copy so new users understand the
main benefit in under 5 seconds. We are losing people on screen 1.
Context: Current copy is below. Audience: busy US office workers, not techy.
The app turns a folder of screenshots into a clean PDF in one click.
Constraints: Max 12 words for the headline, 25 for the subtext. Plain English,
no exclamation marks. Keep the "Get started" button label.
Done means: 3 options, each with headline + subtext, plus one sentence on
why each works. Flag anything you had to assume.
3. Explain the reason, skip the shouting
Older prompting habits (ALL CAPS, "you MUST", long lists of rigid rules) tend to backfire with current models. They follow instructions closely, so over-specified prompts make them stiff and literal.
Instead, give the reason behind a rule. "Keep it under 150 words because it goes in a push notification" beats "KEEP IT SHORT!!!" because the model can make sensible trade-offs when it knows what the limit is protecting.
4. Ask for a plan before a big task
For anything that will take many steps, such as a multi-file feature, a research report or a data clean-up, ask Opus 5.5 to outline its plan first and wait for your OK. It costs one short message and catches misunderstandings before they turn into an hour of wrong work.
A simple line does it: "Before you start, list the steps you plan to take and any assumptions. Don't make changes until I confirm."
5. Give it a way to check its own work
The single biggest quality jump comes from telling the model how to verify the result. For code, that means tests it can run. For writing, a checklist: word limits, required points, banned phrases. For data, expected totals or row counts.
Anthropic's launch notes include a tester reporting that Opus 5.5 solved more terminal tasks than Opus 5 in less than half the steps. That kind of efficiency shows up most when the model can tell for itself whether it is done.
6. Use the big context window, but curate it
With a 1M-token context window you can paste entire documents, long transcripts or large parts of a codebase. That is powerful for questions like "where do these two contracts disagree?"
Still, more is not always better. Irrelevant material costs tokens and gives the model more to trip over. Include what the task needs, and say which parts matter most.
7. Reuse instructions instead of retyping them
If you give Claude the same background every day, such as your brand voice, product facts or coding conventions, store it once:
- In the Claude apps, keep it in a Project so every chat starts with it.
- In Claude Code, put it in the project's
CLAUDE.mdfile. - Through the API, use prompt caching. Cached input reads on Opus 5.5 cost $0.20 per million tokens instead of $4, so a long, stable system prompt becomes almost free to reuse.
8. Hand routine work to Sonnet 5.5
Opus 5.5 is the right default for complex work, but not every job is complex. Anthropic positions Sonnet 5.5 as a faster, cheaper complement that is strongest at well-scoped everyday tasks, bug fixes, and polished documents, slides and spreadsheets.
At $2 input and $10 output per million tokens it costs half as much as Opus. If you can describe exactly what "done" looks like in a sentence, Sonnet will usually get there.
9. Know when to escalate to Fable 5.1
On Anthropic's published benchmarks Opus 5.5 beats Fable 5.1, yet Fable 5.1 is still Anthropic's pick for the most demanding reasoning and long-horizon agent work. It costs $10 input and $50 output per million tokens.
Escalate when Opus at high or xhigh effort keeps missing the same thing: a buried condition, a contradiction between documents, a bug it cannot crack. Do not start there.
10. Expect a different model on security and biology topics
Anthropic says that on Opus 5.5, cybersecurity and biology requests are routed to Claude Opus 4.8 by default as a safeguard. If answers in those areas feel different, that is why. Vetted professionals can apply through Anthropic's verification programs for access.
Common mistakes that waste the most time
- Jumping straight to the most expensive model instead of raising effort on Opus 5.5 first.
- Leaving effort at medium for high-stakes work because you assumed it was still high, as it was on Opus 5.
- Vague asks like "make this better" with no goal, audience or definition of done.
- Correcting in ten small messages what one complete brief would have covered.
- Pasting everything you have instead of what the task needs.
The bottom line
Treat Opus 5.5 like a capable colleague: give it the whole picture, tell it why, let it plan, and give it a way to check its own work. Start at the default effort and step up only when the output shows you need to. Hand the routine jobs to Sonnet 5.5 and keep Fable 5.1 for the rare problem that beats everything else.
Want the full model-by-model breakdown, with prices and benchmarks side by side? Read Claude Opus 5.5 vs Fable 5.1 vs Sonnet 5.5.
If you write code with it, our Claude Code vs Codex comparison shows how Claude's coding agent stacks up, and the complete Claude guide covers plans, features and costs in one place.
Frequently asked questions
What effort level should I use with Claude Opus 5.5?
Start with the default, which is medium on Opus 5.5. Use low for quick chat, sorting and short answers. Move to high for important analysis or code that spans several files, and xhigh for long agent runs and hard coding tasks. Save max for work where getting it right matters more than cost or speed.
Can I turn off thinking on Claude Opus 5.5?
No. Through the API, Opus 5.5 rejects requests that try to disable thinking. The way to make it faster and cheaper is to lower the effort level, which reduces how much it thinks and how many tool calls it makes.
How much does Claude Opus 5.5 cost?
Through the API, Opus 5.5 costs $4 per million input tokens and $20 per million output tokens, with cached input reads at $0.20 per million. In the Claude apps it is included in the Pro plan ($20 per month) and Max plans (from $100 per month), but not the Free plan.
Is Claude Opus 5.5 good for coding?
Yes. Anthropic reports 66.4% on Terminal-Bench 4.0 and 57.8% on CursorBench 4.0, both well above Opus 5. It is also available in Claude Code. For long coding sessions, give it the whole task up front and a way to check its work, such as tests it can run.
Why did Claude answer my security question with a different model?
Anthropic says cybersecurity and biology requests on Opus 5.5 are routed to Claude Opus 4.8 by default as a safeguard. Vetted users can apply through Anthropic's verification programs for access.