Gemma Pricing in India (₹ per Token)
Gemma 4 31B costs ₹13.10 per 1M input tokens and ₹38.30 per 1M output tokens on unoblox, billed monthly in rupees with GST invoicing.
Google's Gemma line is one of the more budget-friendly open-weight families on unoblox, and like every other model on the catalog it's billed in rupees on a single monthly GST invoice — not through an overseas account, and not through a separate billing relationship with Google.
Gemma pricing on unoblox
| Model | Input (₹ / 1M tokens) | Output (₹ / 1M tokens) |
|---|---|---|
Gemma 4 31B (google/gemma-4-31b-it) | ₹13.10 | ₹38.30 |
Gemma 4 26B-A4B (google/gemma-4-26b-a4b-it) | see live ₹ pricing | see live ₹ pricing |
Gemma 4 31B's published rate makes it one of the cheaper general-purpose options on the whole catalog — only a handful of models, like DeepSeek V4 Flash, undercut it on input price. The 26B-A4B variant is also on the catalog; check /models/google/gemma-4-26b-a4b-it for its current live rate rather than assuming it matches the 31B number.
A worked example
A request with 3,000 input tokens and 1,000 output tokens on Gemma 4 31B costs: (3,000/1,000,000 × ₹13.10) + (1,000/1,000,000 × ₹38.30) = ₹0.039 + ₹0.038 ≈ ₹0.077. Scaled to 20,000 such requests a month, that's roughly ₹1,540 — a plannable rupee number with no lookup step in between.
Calling Gemma from your existing code
Gemma sits behind the same OpenAI-compatible endpoint as every other model — only the model field changes.
base_url: https://api.unoblox.ai/v1
api_key: ub-gw-xxxxxxxxxxxxxxxxxxxx
model: google/gemma-4-31b-it
Gemma vs other budget open-weight models
- DeepSeek V4 Flash — ₹9.07 / ₹18.14, the cheapest general model on the catalog today.
- Llama 4 Scout — ₹10.08 / ₹30.24, a close second on input price.
- Gemma 4 31B — ₹13.10 / ₹38.30, competitive on both input and output.
- Qwen3 235B-A22B — ₹9.07 / ₹55.44, cheap on input with a larger model behind it.
If your workload is genuinely price-sensitive and quality-tolerant, it's worth benchmarking your own prompts across two or three of these rather than picking on price alone — per-token rate is only half the cost equation; how many tokens a model needs to do the job well is the other half.
Frequently asked questions
Is Gemma 4 31B the cheapest model on unoblox?
It's among the cheapest, but DeepSeek V4 Flash is currently lower on both input and output. See /models for the full, current list sorted by rate.
Does Gemma run on infrastructure in India? Gemma 4 31B and 26B-A4B run on upstream infrastructure, not unoblox's own India-hosted boxes. The India-specific benefit here is rupee billing and GST invoicing, not data residency — that only applies to unoblox's own small hosted models.
Can I use both Gemma variants under one key?
Yes — google/gemma-4-31b-it and google/gemma-4-26b-a4b-it are both reachable with the same ub-gw-... key; only the model field changes between calls.
Why is Gemma 4 26B-A4B not priced on this page?
Its rate is live on its own model page rather than fixed here, since we only quote numbers we can stand behind at the time of writing. Check /models/google/gemma-4-26b-a4b-it for the current figure.
Is there a free Gemma option for testing? Gemma itself isn't the free-tier model on unoblox, but Qwen3 1.7B is available at ₹0 for prototyping before you commit spend to a paid model.
How is Gemma usage billed alongside other models I call? Together, on one monthly GST invoice — there's no separate bill per model or per upstream provider.
Get started in rupees → https://unoblox.ai/sign-in
More from unoblox
API Pricing Guides
Vision LLM API pricing in India
API Pricing Guides
Batch AI API pricing in India
API Pricing Guides
RAG pipeline cost in India
API Pricing Guides
AI agent running cost in India
API Pricing Guides
OpenAI o3 Pricing in India (Live ₹ Rate)
API Pricing Guides
Cheapest Vision LLM in India (₹ Pricing)
Start building in rupees
Call every major model through one OpenAI-compatible endpoint, billed in ₹ on a GST invoice.