Anthropic released it on 7 October 2026, and its Haiku page (read 8 October 2026) lists it on the web, in the iOS and Android apps and in Claude Code. The right mix of model and effort setting can make your usage limit last longer.
What is new in Haiku 5.5
- A big step up for the small model. The announcement calls it the cheapest, fastest and most capable small model Anthropic has released, built for quick, repetitive jobs like summaries and sorting.
- An effort setting. Haiku 4.5 had none. Haiku 5.5 has five levels, from Low to Max, explained in plain terms below.
Which jobs to hand to Haiku, Sonnet or Opus
Here is how Anthropic's model guide and launch posts divide the work (read 8 October 2026), turned into everyday jobs.
Haiku 5.5
- Built for: quick, repetitive jobs with clear instructions, such as summaries, sorting, labelling and pulling one fact out of a document.
- Everyday examples: three bullet points from a long email thread, 60 survey replies labelled happy, unhappy or mixed, a list turned into a table, a quick look-up in a file you uploaded.
- Verdict: start here for anything short where you would spot a wrong answer at a glance.
Sonnet 5.5
- Built for: writing, documents and data analysis.
- Everyday examples: a first draft of a report, a reply that needs the right tone, a spreadsheet you want explained.
- Verdict: the everyday default for anything where two good answers could differ.
Opus 5.5
- Built for: long, multi-step work that needs judgement the whole way through.
- Everyday examples: a project plan where each step depends on the last, a contract you want checked for anything unusual, a decision with real trade-offs.
- Verdict: worth the extra usage when a mistake would be costly or hard to spot. The Free plan does not include it.
What the effort levels change
Thinking cannot be switched off on Haiku 5.5, so effort is the one dial you have. Anthropic's help page (read 8 October 2026) puts it in the same menu as the model:
- Click the model name next to the send button. The menu that opens holds the model, the effort level and the thinking setting together.
- Click "Effort". The level Anthropic recommends for that model is marked "Default", so you always have a safe place to return to.
- Choose a level. The change applies from Claude's next reply, so you can drop to Low for a quick job and raise it again in the same chat.
If Haiku 5.5 is not listed, click "More models". If "Effort" is missing, you are on an older model without it. On an Enterprise plan your admin can also hide a model or an effort level.
- Low and Medium: routine jobs, and they stretch your usage furthest. Anthropic's Haiku 5.5 prompting guide warns that on Low it is more likely to stop early or skip a check on longer tasks.
- High: the best balance of quality and speed for harder work.
- Extra high: built for long coding jobs and agent tasks, where Claude works through many steps on its own. The prompting guide notes that in back-and-forth chats at this level Haiku sometimes ends a turn with no visible reply.
- Max: the deepest reasoning, the slowest replies and the most usage.
Simon Willison, a developer who tests new models on launch day, found a wide gap when he asked Haiku 5.5 for the same drawing at every level on 7 October: Low got the bicycle frame wrong in 7 seconds, Medium got it right, and Max took over five minutes and cost about 36 times as much as Low.
Writing "keep it short" in your message will not save you that, either, because the prompting guide says asking Haiku to answer directly did not stop it thinking.
The prompt: sort my tasks across Claude's models
Paste this into Claude and fill in the About me lines with the jobs you hand Claude in a normal week. It gives each one a model, an effort level and a reason, then flags the ones where the cheap choice would cost you more in fixing. Run it on Sonnet 5.5 or above, since sorting is judgement work.
You are my Claude model planner: the person who makes my usage limit last the whole week without ever handing me a sloppy answer on something that matters. I pay for my Claude plan with a usage limit that runs out, and I waste it in two ways: I send small jobs to the biggest model at a high effort setting, and I send jobs that need real judgement to the smallest model, then spend longer fixing the result than the job would have taken. Your job is to sort my own list of tasks into the right model and effort level, give a reason for every call, and warn me where the cheap choice is a false saving. You have no reason to favour any model. Think in stages, show your reasoning briefly, and ask before you assume. WHAT YOU NEED TO KNOW ABOUT THE CHOICES - The models, from smallest to largest: Haiku 5.5 (the quickest and lowest cost, built for short, repetitive, clearly described jobs such as summaries, sorting, labelling and pulling facts out of a document), Sonnet 5.5 (everyday writing, documents, data analysis, slides and spreadsheets), Opus 5.5 (long, multi-step work that needs judgement across many steps). If my menu shows a model not described here, such as Fable 5.1 or an older version, do not guess what it is good at. Place my tasks on the models described here and say in one line that you left it out. - The effort levels, from lightest to heaviest: Low, Medium, High, Extra high, Max. Higher effort means more thinking, slower replies and more of my usage limit. Low and Medium suit routine jobs. High is the balanced choice for harder work. Extra high is meant for long coding and agent tasks. Max is for the deepest reasoning and costs the most. - On these models, thinking cannot be switched off, and telling the model "answer briefly" does not stop it thinking. The effort level is the dial that controls how hard it works. - At Extra high and Max, Haiku 5.5's thinking and replies get much longer, and Anthropic's own advice is to compare it against Sonnet 5.5 at those levels. Treat Haiku on Extra high or Max as a warning sign, never a bargain. - On Low, Haiku 5.5 is more likely to stop early or skip a check on longer tasks. ABOUT ME (I fill this in once; any line can say "not stated") - My plan: [e.g. "Free", "Pro", "Max", "Team at work". If you do not know which models your plan shows, open the model menu next to the send button and list what you see, e.g. "I see Haiku 5.5, Sonnet 5.5 and Opus 5.5".] - What I use now: [e.g. "Opus 5.5 on the default effort for everything", "Sonnet 5.5 on High", "whatever opens when I start a chat".] - The effort level marked Default in my menu: [open the model menu, click Effort and see which level says Default, e.g. "Medium", or "not stated".] - Where I use Claude: [e.g. "the website and the phone app", "the desktop app", "Claude Code in my terminal for a side project", or a mix.] - What I run out of first: [e.g. "I hit the five-hour limit most afternoons", "I hit the weekly limit by Thursday", "I never hit a limit, I just want faster replies", "not stated".] - What hurts most when a job goes wrong: [e.g. "a client sees a wrong number", "I lose half an hour redoing it", "nothing much, it is just notes for me".] - My task list: [List every job you hand to Claude in a normal week, one per line, with roughly how often and how long the input is. The more specific, the better the sorting, e.g. "summarise the team's weekly update email, 1 page, every Monday"; "label 60 customer survey replies as happy, unhappy or mixed, twice a month"; "draft a tricky reply to a landlord about a deposit, once"; "turn my meeting notes into actions with owners, 3 a week"; "plan a two-week project with dependencies, once a month"; "check a 20-page contract for anything unusual, rarely"; "rename files in my project and fix the imports, in Claude Code".] STAGE 1: UNDERSTAND MY LIST Read everything above. In one or two lines, say what kind of user I look like (mostly quick jobs, mostly thinking jobs, or a mix) and which limit matters most to me. If a task is too vague to place, or a detail would change the answer, ask me up to four short questions with choices where possible, e.g. "Is the landlord reply something you will send as is, or edit first?", "Does the survey labelling need a reason for each label, yes or no?". Then stop and wait for my answers. If nothing important is missing, say so and carry on. If a line above still shows its bracketed example, treat it as not stated. STAGE 2: WEIGH EACH TASK BEFORE YOU PLACE IT For every task, think through these five questions, briefly: 1. Are the instructions clear and the right answer easy to recognise? (Favours Haiku.) 2. Does it need judgement, tone or trade-offs, where two good answers could differ? (Favours Sonnet or Opus.) 3. How bad is a wrong answer, and would I spot it? (A mistake I would not notice, or one that reaches another person, pushes the task up a model.) 4. How many steps does it take, and does each depend on the last? (Long chains favour Opus.) 5. How often do I do it? (A frequent job saves the most from a smaller model, so it is worth testing.) STAGE 3: PLACE EACH TASK Give every task one model and one effort level, using these rules: - Start every task at the smallest model and lowest effort that you expect to get it right first time. Move up only with a reason. - Prefer raising the effort level over switching to a bigger model when the task is the right size for the model but needs more care. - If my plan does not include a model, never route to it. Give the best choice I do have and say what I am missing. - If I use Claude Code and the task is a helper job inside a bigger piece of work (searching files, reading logs, checking a result), say whether it suits a Haiku helper running under a bigger main model. - If a task has several steps or a long input, do not put it on Low. Start it on Medium. - If my pick for a task matches the Default in my menu, say "leave it on Default" so I do not have to change anything. - If you put a Haiku chat task (not coding) on Extra high, warn me that a reply can sometimes come back blank, and to try High instead. - If I run out of the weekly limit first, protect it: move my most frequent jobs to the lowest safe pair first. - If I hit the five-hour limit first, also tell me which heavy jobs to move to a different part of the day. - If I never hit a limit and want speed, favour lower effort for speed, and say where that risks a redo. - If what hurts most is a mistake another person sees, move any task that leaves my hands up one step, model or effort, and say which. - Never put a cheap model on a job where the result is hard to undo, such as something sent to another person without me reading it, or a change to files I cannot roll back. STAGE 4: FLAG THE FALSE SAVINGS Go back over your Haiku and Low-effort choices and mark any that are a false saving: where a likely redo, a careful check or a mistake I might miss would cost me more time or trust than the bigger model would cost in usage. Mark these FALSE SAVING with one line on why and what to pick instead. Also flag any task where I would need Haiku on High or above to get it right, because Sonnet on Low or Medium is likely the better pair there. STAGE 5: SET ME A QUICK TEST Pick the two tasks where your call is least certain. For each, tell me exactly how to test it: run the same input on the two candidate pairs in two new chats, and what to look for in the answers to decide. WHAT YOU HAND BACK 1. A table with these columns: Task | Model | Effort | Why, in one line | Move up if. 2. The FALSE SAVING list. 3. The two tests. 4. One line on the single change from what I use now that would make my usage limit last longer, based only on my list. WHEN I COME BACK TO THIS CHAT - When I say "quick call: [task]", give me one model and one effort level, a one-line reason, and the sign that means I should move up. No table. - When I say "this came back wrong: [task, the pair I used, what was wrong]", decide which kind of failure it was. If it stopped early, skipped a step or missed something obvious, tell me to raise the effort one level first. If it got the judgement or tone wrong, tell me to move up a model. If my instructions were vague, rewrite them and keep the pair. - When I say "I hit my limit", look back at the table, name the two rows that most likely use the most, and give the cheaper pair for each that you still expect to get them right. If no cheaper pair is safe, say so. - When I add or drop a task, update only the rows that change. RULES FOR THIS WHOLE CONVERSATION - Never invent a price, a usage figure, a benchmark score or a feature. If you do not know something about my plan or a model, write "not stated" and say how it changes the advice. - Never promise how much of my usage limit a model will save. Say "likely uses less" or "likely uses more", never a number. - If a task needs current facts or a web search, say so in its row, since every model can get those wrong. - Keep the table rows short. Detail goes under the table, only where a call is borderline. CHECK YOURSELF BEFORE YOU ANSWER 1. Did every task in my list get a model, an effort level and a reason? 2. Did you route to any model my plan does not show? If so, fix it. 3. Is any Haiku task one where a mistake would reach another person unchecked? If so, move it up or flag it. 4. Did you invent any number? Remove it. 5. Is any task with several steps sitting on Low? Move it to Medium or say why not. 6. Is any row Haiku on Extra high or Max without a FALSE SAVING flag? Flag it. 7. Name the one call in your table you are least sure of, and why.
For the example list in the prompt, we would expect the weekly summary on Haiku at Low and the survey labelling on Haiku at Medium. The landlord reply would go to Sonnet, the contract check and project plan to Opus, and the meeting actions would come back as a borderline call to test.
Does Haiku make your usage limit last longer
Anthropic's usage page (read 8 October 2026) names "which Claude model you're chatting with" and "the effort level you've selected" among the things that use up your limit, and says a lower effort stretches it. It publishes no figure for how much less Haiku uses, so measure it yourself:
- On a normal day, note your weekly bar in Settings > Usage, morning and evening. Your chats, Claude Code and the desktop app all draw on the same limit, so this one bar covers everything.
- On the next similar day, move the jobs the prompt put on Haiku. Keep your thinking jobs where they were, so the only change is the small stuff.
- Compare how far the bar moved on each day. If it moved less and you did not redo anything, keep the habit. Our guide to why you keep hitting your limit covers the other things that eat it.
The Free plan has no weekly bar, so judge by whether you hit the five-hour limit later than usual.
The honest bit
- It is a day old. Early testers on the launch's Hacker News thread split: one uses it to sort bugs before handing them to Sonnet, another found it could not build a complex web page and watched it hand the job to Opus.
- Low can show its working. The prompting guide says that on Low it sometimes puts reasoning into the answer itself, so raise it to Medium.
- Some refusals are new. It refuses some requests Haiku 4.5 answered, and resending usually gets the same refusal, so rephrase or switch to Sonnet.
- Big builds still belong to the bigger models. Anthropic's own announcement says Sonnet 5.5 and Opus 5.5 remain better choices for complex coding work.
- The prompt and the Claude Code helper file are untested by us. We built both from Anthropic's published guidance and have not yet run them on a real week of work. Treat the prompt's table as a first draft and trust your own test.
Try it on your next quick job
The next time you ask Claude to summarise, sort or tidy something, switch to Haiku 5.5 on Low before you send it and see whether you would have kept the answer. If you would, you have found your first job to move, and the prompt above will find the rest.
For builders
This part is for people who use Claude Code or pay per token on Anthropic's developer platform.
Haiku as the helper in Claude Code
Developers trying Haiku 5.5 this week are pairing a big model that does the thinking with Haiku as the cheap helper for small sub-jobs, such as searching files or checking a result. In Claude Code that helper is a sub-agent, a file with its own instructions that the main model hands searches and checks to.
- Update first with
claude update. Anthropic's model settings page says Haiku 5.5 needs version 2.1.293 or later. - Check where the alias points.
haikumeans Haiku 5.5 only when you sign in through Anthropic; on Amazon Bedrock, Google Cloud or Microsoft Foundry it still means Haiku 4.5. - Save the helper below as
.claude/agents/haiku-scout.mdin your project. Its settings at the top setmodel: haikuandeffort: low, as the sub-agents page shows. The built-in Explore helper runs on your main model, so a Haiku helper of your own is how you move searching off it. - Watch it run with
/tasks. You can see which helper is working and on which model, so a helper quietly running on your main model is easy to catch.
--- name: haiku-scout description: Finds files, reads logs and reports facts back. Use for searches and look-ups, never for edits. model: haiku effort: low --- You find things and report them. Search the files or logs you are pointed at and report exactly what you found, with file paths and line numbers. Do not edit, create or delete anything. If you cannot find what was asked for, say "not found" and list where you looked. Never guess at the contents of a file you did not read. If the question needs a judgement call rather than a fact, say so and hand it back to the main model.
To run a whole session on Haiku instead, type /model, pick Haiku, then /effort low.
If you use Fable 5, fable-baton is a free plugin by a solo developer, updated 7 October, that makes Fable 5 the lead model in every session and hands the legwork to helpers on Opus, Sonnet and Haiku.
Its setup changes your default model to Fable 5, so it needs a plan or API access that includes Fable 5, and without that Opus leads instead. Its two Haiku helpers ship with effort: max, so change them to low or medium.
Prices on the developer platform
- Haiku 5.5 is cheaper per token. Per the announcement, it costs $0.10 per million tokens in and $0.50 out for prompts up to 100,000 tokens, and Anthropic says it costs around 75% less to run than Haiku 4.5. A token is a chunk of text, roughly three quarters of a word.
- The same text counts as more tokens. Anthropic's model page says the same text counts as roughly 30% more tokens than on Haiku 4.5, so the saving is smaller than the price drop suggests.
- Sonnet 5.5 cache reads are half price. The same announcement halves its cache reads (text the model has already seen once) to $0.10 per million tokens, which Anthropic says makes most agent work around 20% cheaper.
- Monthly API credits for Max and Team. Max 5x gets $100, Max 20x $200 and Team up to $500 pooled, to spend on building apps and agents with the developer platform, per the help article.
- On every cloud platform. Haiku 5.5 launched on 7 October on Amazon Web Services, Google Cloud and Microsoft Azure as well as Anthropic's own platform, per the announcement.
A few quick questions
What effort level does Haiku 5.5 start on?
Medium in Claude Code and on the developer platform, according to Anthropic's model settings page. In the app, open the model menu and look for the level marked Default.
Do the new monthly API credits give me more Claude usage?
No. Anthropic's help article says they are for building with the developer platform, they do not change your usage limits in Claude, Claude Code or Cowork, Free, Pro and Enterprise plans are not eligible, and unused credits expire at the end of each billing cycle.