Enjoying this issue?
Get tomorrow's AI & engineering digest in your inbox — hand-picked, summarized, and always spam-free.
TLDR
Anthropic's official Opus 5.5 prompt guide is built around one practical shift: the model now defaults to medium effort instead of high, and Anthropic explicitly says not to carry over old settings, because the same effort level no longer means the same amount of thinking. The rest of the guide is mostly reasoned default-breaking advice, like removing vague thinking instructions, telling Claude to check multiple sources before acting, and separating pasted instructions from your own request.
Key points
Opus 5.5 defaults to medium effort while Opus 5 defaulted to high.
Anthropic says removing thinking instructions like "think carefully" sped up replies without clear quality decline.
Opus 5.5 may revisit previous answers while thinking, so Anthropic suggests instructing it to treat earlier answers as settled.
Anthropic recommends asking Claude to explore relevant emails, documents, and spreadsheet tabs beyond the first file.
Anthropic reports Opus 5.5 reads visual material more accurately than Opus 5 without extra image tools.
Tools mentioned
Techniques
- Effort-level tuning
- Comparison testing across effort levels
- Relevant context sourcing
- Prompt boundary marking
- Progress update scheduling
- Completed-task output definition
Stop scrolling. Start reading smarter.
Receive the day's most important AI & engineering updates in one concise email. No spam.
Transcript (captions)
Enthropic just dropped this guide on how to use Opus 5.5. And if you want to get the most out of it, I'm going to speedrun eight tips covering the key takeaways, what's different, what to
change, and practical examples you can use. And a few of these I had absolutely no idea about. This should not only help you get better results with this model, but also save you money on token costs.
So, without further ado, let's dive right in. The first concrete change is the default effort setting. Opus 5 defaulted to the high effort level while Opus 5.5 now defaults to medium.
Basically, effort controls how much work the model puts into thinking before it responds. And Enthropic recommends starting at medium and testing it on your own tasks. And don't assume that
you need to carry over the higher setting that you used with the previous models. The same effort level doesn't necessarily mean the same amount of thinking across these models. If you
keep your old configuration, you could get longer turns and more output tokens. Enthropic recommends lowering the effort level first when you want less thinking and then only using the highest settings
where you've actually measured a benefit. This way, it's going to save you a lot on token usage. Otherwise, you're just going to use a higher model even though it might not be necessary
with Opus 5.5. And then for developers, the guide also covers leaving enough room for thinking tokens and changing effort without invalidating the prompt cache. But if you're not a developer,
this probably sounds like gibberish, so it probably doesn't apply to you. A useful way to choose your setting is to compare the same task at different effort levels. For example, you can give
Claude two proposals in the same list of requirements, then compare both medium and the higher effort. And then you could ask yourself, did one miss an important requirement? Was a
recommendation more useful? How long did it take? A longer answer alone doesn't necessarily tell you whether the effort level was worthwhile. And this is going to be custom to you and your individual
use cases. All right, moving on to tip number two, and this is thinking instructions. This is worth checking if you've built up a longer system prompt over time inside of Claude. Enthropic
says that instructions like thinking carefully before answering might be unnecessary for Opus 5.5 because the model already decides how much to think and the effort is the main control. In
the specific test that Enthropic ran, removing those kinds of instructions made replies start sooner, and this meant you're getting quicker replies, and it even did so with no clear decline
in the quality of the reply. Another behavior that Enthropic documents is that Opus 5.5 can go back over an earlier answer while thinking about a new message, even a simple follow-up.
The guide gives an instruction that tells it to treat previous answers as settled unless you question them. that can reduce unnecessary work. But there is a trade-off. You may want Claude to
revisit earlier conclusions during research, analysis, or a task where new evidence appears. Now, right here is a perfect example. When you write out a vague prompt, something like thinking
deeply and carefully, it doesn't necessarily tell Claude what makes that answer useful to you, and that's probably one of the reasons you're not getting the outputs that you're wanting.
Whereas, if you give it a specific prompt, like asking it to compare two proposals against your requirements, recommend one, and explain the made trade-off, gives it a concrete target to
work towards. You're basically still asking for thoughtful work, but you're defining the result that you need. Next up is tip number three, and this is relevant context. Enthropic says that
Opus 5.5 tends to get to work very quickly. This can be very useful, but in a workflow involving several apps, the task might depend on information that the request doesn't explicitly mention.
This could be something like a deadline that could have changed in an email or a rule that could be on another spreadsheet tab, for example. Their recommendation is to ask Claude to look
through relevant sources before acting rather than assuming the first document contains everything. This way, it has all of that relevant context. This screenshot right here shows the actual
instruction that Enthropic provides. It tells the agent to explore the emails, documents, spreadsheet tabs, and records that could be relevant, including ones the task didn't specifically name.
Enthropic found that this improved completion on its own multi-app test with slightly more tool calls and tokens. The important conditions are that the sources are accessible and that
the agent isn't necessarily blindly acting on untrusted material. Now, let's move on to an example. Imagine you ask Claude to write a project update from an original brief. And let's say that this
brief says the deadline is Monday, but a later email moves it to Friday. If Claude only reads the brief, it can produce a polished update with the wrong date. Obviously, this prompt that you
see right here asks it to check both sources before drafting. So, this way, you're giving it a specific reason to look beyond that first file that we gave it. Moving on to tip number four, and
this is pasted material. When you paste an email or web page into a message, there are really two different things present. your request and somebody else's content. That content might
include instructions, but those instructions aren't necessarily yours. Enthropic recommends marking the boundary clearly. The everyday principle is simple. Make it obvious what you want
Claude to do and which material you're asking it to read or analyze. So, just start with a job basically like saying summarize this email. Then, paste the email in separately so the request and
the source material are clearly separated. This is very simple but just really important when prompting Opus 5.5. All right, so moving on to tip number five. Why Claude can seem silent.
Opus 5.5 can produce progress updates that a custom app does not display. If you built the app, check that it can receive and show those updates. Asking Claude to talk more will not fix updates
the app is hiding. You can tell Claude when you want an update rather than leaving the timing open. Just make sure to choose useful moments when you're starting the work, finishing the review,
and delivering that result. for automated setups. The guide also explains how to send a limited number of reminders. All right, so here is another example. Here we ask for an update after
the files are reviewed in when the draft is ready to send. Claude should keep working between these different updates. This gives you visibility without turning every step into an approval
request. Now, moving on to tip number six. A finished reply is not a finished task. With Opus 5.5, a progress update can end a turn while work is still outstanding. List what you expect to
receive and keep track of the unfinished items. For an automated agent, the app must also check for completion. And a checklist alone is not a guarantee. Now, if a background task is still running,
wait for its results before calling the job done. Anthropic provides extra instructions for agents that run without someone watching. Those instructions can use more time and tokens, and they
should still preserve necessary approvals. This longer instruction that you see targets a simple problem, reporting the next step instead of taking it. It asks an unattended agent
to continue work that does not need your input. Real blockers and required approvals still matter. Now, right here is an example. We need to tell Claude what done includes first. We can name
the outputs, whether that's a report, source list, and a summary. If one of those is missing, point to it directly instead of just saying keep going. Ask Claude to finish the item or just
explain the blocker. Moving on to tip number seven, and this is giving Claude a clear design direction. Opus 5.5 still falls back on default website styles. Anthropic says less generic often swaps
out one default for another. So give concrete visual feedback and refine what comes back. This is obviously just an existing best practice and not a new feature baked into Opus 5.5, but I do
think it is relevant for you to understand when prompting this model. Now moving on to a very simple example. We need to be specific about the design. And so this is a perfect example of
this. When we say make it less generic, this leaves most of the choices open and it's going to default to the next design style. The stronger example names a background, headline style, button
shape, and specific spacing. We can use our own preferences or a reference image as an example. Moving on to tip number eight, making small details easier to read. Anthropic reports that Opus 5.5
reads visual material more accurately than Opus 5 without extra tools. Dense charts and tiny labels can still benefit from a sharp original and a close-up. If image tools are available, Claude can
crop and inspect the details itself. Now, here's a quick example. Point Claude to the exact label or part of the image that matters. Let it crop if tools are available or attach your own
close-up alongside that original image. Then, you could ask it to say when the detail is unreadable rather than guess, which is going to help provide more context here. All right, so I know we
talked about a lot of different tips here, and this might be overwhelming. You don't need to add every single recommendation to one enormous prompt or you don't need to test out every single
one of these. But hopefully this helps you prompt Opus 5.5 in order to get better responses and also save you money on token usage since that is obviously something that you're wanting to do. If
you enjoyed this video, subscribe to the channel for more content like this. Let me know what you think in the comments and I'll see you in the next video. Cheers.