Claude GuideDevantia × Executive Partners Group
DevantiaExecutive Partners Group
0/23 Jean-Christophe Leroy
Jean-Christophe Leroy
Guide author
✓ Verified on September 3, 2026⏱ 4 min read
Module 15 ter

Scaling up — deployment, ROI & cost control

Adopting Claude on your own is easy; getting a measurable gain across a team takes a method. Here is a 4-week deployment framework, followed by the habits that keep you from wasting your credits.

The 4-week method

W1

Map it out. List everything that recurs more than once a week. Separate thinking (creating, deciding) from execution (organizing, extracting, filling in). Spot the tasks where quality depends on human focus.

W2

Automate 3 tasks — no more. Meeting notes, CRM-driven email follow-ups, document summaries. A lasting system starts with 3 workflows that are 100% reliable, not 10 shaky ones.

W3

Create Skills. Turn each recurring task into a pre-configured command (e.g. /followup, /notes). Your workflow, your voice, reusable across the whole team.

W4

Measure and iterate. Calculate the time saved per person per week on your data, identify the next task to automate, and document the system so it can be handed off.

⚠️ Measure against your own numbers Gains vary widely by role and by team maturity. Set a baseline (time spent today) before automating, then compare — that's the only credible ROI in front of a skeptical executive committee. [The percentages circulating on social media are not verified benchmarks.]

Controlling your costs — the right anti-waste habits

Technical note: prompt caching also cuts costs on repeated prompts.

💡 The principle that explains everything else Counterintuitive but proven: the big savings don't come from a better-worded prompt, but from what you let into the conversation. Claude rereads everything you give it, with every message. Cleaning up content before pasting it, trimming a file, or loading only what's useful can cut the cost by a factor of 10 — where a prompt "magic formula" changes almost nothing. Remember the rule: the best token is the one you don't send.

📄 PDF → Markdown

Converting a heavy PDF to .md before analysis dramatically lightens the context, for the same result.

🧠 Plan in Chat, execute in Cowork

Brainstorming costs little in Chat; save Cowork for handling files.

🔎 Search / Connectors on demand

Turn them on only when the task requires it — an unnecessary search weighs down every exchange.

✏️ Edit rather than resend

Edit your message instead of piling on follow-ups: you start again from a clean context.

🎯 One task per message

Three clear requests beat one confused message — for you as much as for Claude.

⚡ The right model

Sonnet 5 every day; Opus 5 for heavy work; Fable 5.1 for the summit (see the single rule, module 02).

🗂️ Reuse Projects

Load your files once into a Project; no need to re-paste them into every conversation.

🔁 New thread regularly

Over a very long conversation, Claude rereads everything: open a new thread per topic.

📝 Summarize before starting over

Session dragging on? Ask "summarize our exchange in key points (decisions, figures, to-dos)," then start fresh in a new thread with that summary. Claude no longer rereads 40 messages, just 10 lines. Near the limit, Claude does it automatically — but doing it yourself keeps it from losing a figure or an important detail.

🧹 Clean up before pasting

Don't paste an entire web page (menus, ads, cookies) or an email thread with signatures: keep the useful passage. On raw content, this is the most spectacular saving.

🖼️ Lighten your images

A very high-resolution screenshot or photo costs a lot, often for nothing. If the text stays legible, shrink it before sending.

✂️ Ask for brief

What Claude writes counts too. "Answer directly, no preamble" or "in 5 bullets" avoids needless pages when you just want the essentials.

📋 Concise Project instructions

A Project's instructions are reread with every message. Keep them short and sharp: a rambling instruction is paid for continuously and Claude follows it less well.

⭐ The one rule to remember It's not the plan that makes the difference between Pro and Max, it's discipline: a clear 30-word prompt beats a confused 500-word one. For the vast majority of individual use cases, a disciplined Pro plan is enough.

Optimizing your tokens — the playbook for recent models

A token is Claude's unit of account: a piece of a word (an English word ≈ 1.3 tokens). Everything Claude reads — your messages, your files, the conversation history — and everything it writes is counted in tokens. Understanding how that meter runs means doubling the value of your plan without spending a euro more.

How your budget is counted

1

Two counters run in parallel: a limit per 5-hour session and a weekly limit. Track them in real time in Settings → Usage — two progress bars, no surprises.

2

Every message rereads the whole conversation. On the 40th exchange of the same thread, you pay to reread the previous 39. That's the reason for the "new thread per topic" habit seen above.

3

Everything counts: message length, attachments (one PDF page ≈ 1,500 to 3,000 tokens!), tools turned on (Search, connectors), the model chosen, the effort level, and creating artifacts.

4

Projects work in your favor: their content is cached — a document loaded into a Project doesn't "reweigh" with every message, unlike the same document re-pasted into each conversation.

The levers specific to Opus 5 and Fable 5.1

⚙️ Set the effort level

Fable 5.1 always thinks ("adaptive" reasoning): the effort slider is your main lever. Low effort for the routine, high for the strategic — it's the number-one factor in consumption.

💰 A premium model — dose it

Fable 5.1 costs twice as much as Opus 5 per token. Save it for the summit; Opus 5 for heavy work, Sonnet 5 for most everyday needs.

🧠 1 million tokens of context

Enough to load a whole case file at once. But a massive context is paid for with every exchange: load what's useful, not the full archive "just in case."

🛟 Refusals aren't billed

If a safety guardrail declines a request (fewer than 5% of sessions), nothing is counted and the automatic fallback to Opus takes over.

⭐ The winning trio for everyday use (1) Check Settings → Usage once a week to spot what's consuming. (2) Work at standard effort by default and raise it occasionally. (3) Let memory and conversation search recall context ("find our exchange about X") instead of re-pasting everything. Need an occasional top-up? Usage credits can be bought from the same page — often smarter than switching plans for a spike in activity.
🚫 Misconceptions — what saves (almost) nothing To avoid wasting time on false tricks: watching the answer write itself word by word is a display comfort, not a saving. Being very polite ("thanks," "please") is nice but negligible. And caching lightens the bill on repeated content — it doesn't reduce what Claude rereads. The real levers stay the ones above: less context, clean inputs, short answers.
🔄 A benchmark that shifts with every model Each new model counts tokens a little differently (Sonnet 5 uses noticeably more than the previous generation for the same text). So don't rely on old estimates: the only reliable judge remains Settings → Usage, on your own exchanges.
🛠️ Your turn — 5 minutes

Time a recurring task this week, then redo it with Claude and compare.

Expected result : A before / after reference figure, on your own data.