OpenAI o4-mini API in India
Call OpenAI o4-mini for step-by-step reasoning in India — ₹ billing, one GST invoice, no international card, live pricing included.
o4-mini is OpenAI's smaller reasoning model — the o-series' answer to running step-by-step problem solving at a lower cost than a full-size reasoning model — and it's callable from India through unoblox on the same ₹-billed, OpenAI-compatible endpoint as everything else in the catalogue.
Where o4-mini sits
Like other reasoning models, o4-mini works through a problem internally before producing its final answer, which is why it tends to do better than same-sized non-reasoning models on math, multi-step logic and code debugging. The "mini" in its name signals that it's built to bring that reasoning behaviour to a wider set of workloads at a lower cost than a full flagship reasoning model, without giving up the step-by-step approach entirely.
₹ pricing for o4-mini
o4-mini's live ₹ price is on its model page rather than fixed here — reasoning models are worth pricing out per workload since output-token usage per answer can run higher than a standard chat model answering the same question. For general context, here's where two non-reasoning OpenAI tiers sit:
| Model | Input (₹ / 1M tokens) | Output (₹ / 1M tokens) |
|---|---|---|
| GPT-5 mini | 25.2 | 201.6 |
| GPT-4.1 | 201.6 | 806.4 |
Calling o4-mini from the unoblox endpoint
curl https://api.unoblox.ai/v1/chat/completions \
-H "Authorization: Bearer ub-gw-xxxxxxxxxxxxxxxx" \
-H "Content-Type: application/json" \
-d '{
"model": "openai/o4-mini",
"messages": [{"role": "user", "content": "Debug why this function returns the wrong total."}],
"max_tokens": 1000
}'
As with any o-series model, give the completion budget enough headroom for internal reasoning — a max_tokens that's too tight is the most common reason a reasoning-model call ends without a usable answer.
o4-mini vs o3 vs GPT-5 mini
- o4-mini — reasoning behaviour at a lower cost than a full reasoning model; a fit for debugging, structured logic and multi-step tasks at higher volume than o3 would be economical for.
- o3 — reserve for the harder subset of reasoning tasks where o4-mini's answers aren't reliable enough.
- GPT-5 mini — for tasks that don't need explicit step-by-step reasoning at all, a non-reasoning small model is usually faster and cheaper for the same quality bar.
Frequently asked questions
Is o4-mini available in India through unoblox?
Yes, openai/o4-mini is callable on https://api.unoblox.ai/v1 right after signup, billed in rupees.
What does o4-mini cost on unoblox? It isn't in our fixed indicative price list; check the live ₹ rate on its model page.
How is o4-mini different from GPT-5 mini? o4-mini reasons through a problem step by step before answering; GPT-5 mini answers directly. Reasoning helps most on math, logic and debugging tasks.
Why is my o4-mini response getting cut off?
Increase max_tokens — part of the budget is used for internal reasoning before the visible answer is written.
Does o4-mini run on infrastructure in India? No, it runs on OpenAI's infrastructure. unoblox adds ₹ billing, a GST invoice and one endpoint, not data residency.
Is there one GST invoice covering o4-mini and other models? Yes, all usage is combined into a single monthly GST invoice from an Indian entity, with input tax credit claimable.
Get started in rupees → https://unoblox.ai/sign-in
Start building in rupees
Call every major model through one OpenAI-compatible endpoint, billed in ₹ on a GST invoice.