Skip to content

Claude settings guide

Claude (Anthropic) · Long context · Reasoning · Coding

Chat / LLMFree tierAPI
Token burn
Low burn
Paid plans from
$20/mo (Pro)

Token-efficient; free tiers go a long way. Last reviewed September 2026; plans and model names change often, so check the tool's own pricing page before you buy.

1

Put standing context in a Project: upload reference files and write project instructions once, so you stop pasting the same background into every chat.

2

Match the model to the job: Haiku 4.5 for quick, cheap tasks, Sonnet 5.5 for everyday work and coding, Opus 5.5 or Fable 5.1 only for the hardest multi-step problems. On the API, Haiku costs $1/$5 per million tokens against $10/$50 for Fable.

3

Ask for the format you want up front ("Answer in 5 bullets, no preamble"). Shorter outputs cost fewer tokens and are easier to check.

4

Structure long prompts with XML-style tags such as <context>, <task> and <format>. Anthropic recommends this in its prompting docs because it keeps instructions separate from material.

5

Turn extended thinking on for analysis, maths and planning, and off for simple questions. Thinking tokens are billed as output.

6

On the API, always set max_tokens and use prompt caching for long, repeated system prompts; cached input is billed at a fraction of the normal rate.

7

Start a new chat when the topic changes. Long threads resend history with every message, which costs more and can blur the answer.

API cost estimate

Monthly cost at list prices (September 2026), before and after a token cut you choose.

Now
$1.92
₹161
After a 25% cut
$1.44
₹121
Saved per month
$0.48
₹40

Assumes about 600 input and 200 output tokens per request (low burn). Excludes caching discounts, search or request fees and taxes. Check the vendor's pricing page before budgeting.

Reveaxa guide reference: RVX-T-claude