Gemini settings guide
Google Gemini · Multimodal · Google Workspace
- Token burn
- Low burn
- Paid plans from
- ₹1,950/mo (Google AI Pro, India)
Token-efficient; free tiers go a long way. Last reviewed September 2026; plans and model names change often, so check the tool's own pricing page before you buy.
The free plan includes Gemini 3.6 Flash, with Gemini 3.1 Pro when capacity allows. Google AI Plus is ₹399/month and AI Pro ₹1,950/month in India (gemini.google, September 2026).
Use Google AI Studio for System Instructions and to test prompts before you build anything; it shows token counts as you go.
Save repeat assistants as Gems so a system prompt is stored once instead of pasted each time.
Connect Google Workspace so Gemini can read Docs, Sheets and Gmail you point it at, without copy-paste.
Use the long context window for whole reports or codebases, but still ask about specific sections; focused questions give better answers.
Turn on Grounding with Google Search in AI Studio when you need current facts, and check the cited links.
Gemini handles Hindi well: set "Respond in Hindi" or "Respond in Hinglish" in the system instruction for regional content.
API cost estimate
Monthly cost at list prices (September 2026), before and after a token cut you choose.
- Now
- $1.92
- ₹161
- After a 25% cut
- $1.44
- ₹121
- Saved per month
- $0.48
- ₹40
Assumes about 600 input and 200 output tokens per request (low burn). Excludes caching discounts, search or request fees and taxes. Check the vendor's pricing page before budgeting.
Reveaxa guide reference: RVX-T-gemini
More Chat / LLM guides
Claude
Chat / LLM · Long context · Reasoning · Coding
Put standing context in a Project: upload reference files and write project instructions once, so you stop pasting the same background into every chat.
ChatGPT
Chat / LLM · General purpose · Multimodal
Fill in Custom Instructions (Settings → Personalization) with who you are and how you want answers. It applies to every chat, so you stop repeating it.
Groq / Llama 3
Chat / LLM · Ultra-fast · Free LLM
Groq serves open models very fast on its own LPU hardware; use it when response speed matters (real-time apps, prototyping, streaming).