Claude settings guide
Claude (Anthropic) · Long context · Reasoning · Coding
- Token burn
- Low burn
- Paid plans from
- $20/mo (Pro)
Token-efficient; free tiers go a long way. Last reviewed September 2026; plans and model names change often, so check the tool's own pricing page before you buy.
Put standing context in a Project: upload reference files and write project instructions once, so you stop pasting the same background into every chat.
Match the model to the job: Haiku 4.5 for quick, cheap tasks, Sonnet 5.5 for everyday work and coding, Opus 5.5 or Fable 5.1 only for the hardest multi-step problems. On the API, Haiku costs $1/$5 per million tokens against $10/$50 for Fable.
Ask for the format you want up front ("Answer in 5 bullets, no preamble"). Shorter outputs cost fewer tokens and are easier to check.
Structure long prompts with XML-style tags such as <context>, <task> and <format>. Anthropic recommends this in its prompting docs because it keeps instructions separate from material.
Turn extended thinking on for analysis, maths and planning, and off for simple questions. Thinking tokens are billed as output.
On the API, always set max_tokens and use prompt caching for long, repeated system prompts; cached input is billed at a fraction of the normal rate.
Start a new chat when the topic changes. Long threads resend history with every message, which costs more and can blur the answer.
API cost estimate
Monthly cost at list prices (September 2026), before and after a token cut you choose.
- Now
- $1.92
- ₹161
- After a 25% cut
- $1.44
- ₹121
- Saved per month
- $0.48
- ₹40
Assumes about 600 input and 200 output tokens per request (low burn). Excludes caching discounts, search or request fees and taxes. Check the vendor's pricing page before budgeting.
Reveaxa guide reference: RVX-T-claude
More Chat / LLM guides
ChatGPT
Chat / LLM · General purpose · Multimodal
Fill in Custom Instructions (Settings → Personalization) with who you are and how you want answers. It applies to every chat, so you stop repeating it.
Gemini
Chat / LLM · Multimodal · Google Workspace
The free plan includes Gemini 3.6 Flash, with Gemini 3.1 Pro when capacity allows. Google AI Plus is ₹399/month and AI Pro ₹1,950/month in India (gemini.google, September 2026).
Groq / Llama 3
Chat / LLM · Ultra-fast · Free LLM
Groq serves open models very fast on its own LPU hardware; use it when response speed matters (real-time apps, prototyping, streaming).