Skip to content
Start freeGPT-5 · Claude · DeepSeek V4 · Qwen3 — in ₹One OpenAI-compatible endpointBilled in rupeesGST invoiceNo international cardGet started →Start freeGPT-5 · Claude · DeepSeek V4 · Qwen3 — in ₹One OpenAI-compatible endpointBilled in rupeesGST invoiceNo international cardGet started →
Model API Guides

OpenAI o4-mini API in India

Call OpenAI o4-mini for step-by-step reasoning in India — ₹ billing, one GST invoice, no international card, live pricing included.

o4-mini is OpenAI's smaller reasoning model — the o-series' answer to running step-by-step problem solving at a lower cost than a full-size reasoning model — and it's callable from India through unoblox on the same ₹-billed, OpenAI-compatible endpoint as everything else in the catalogue.

Where o4-mini sits

Like other reasoning models, o4-mini works through a problem internally before producing its final answer, which is why it tends to do better than same-sized non-reasoning models on math, multi-step logic and code debugging. The "mini" in its name signals that it's built to bring that reasoning behaviour to a wider set of workloads at a lower cost than a full flagship reasoning model, without giving up the step-by-step approach entirely.

₹ pricing for o4-mini

o4-mini's live ₹ price is on its model page rather than fixed here — reasoning models are worth pricing out per workload since output-token usage per answer can run higher than a standard chat model answering the same question. For general context, here's where two non-reasoning OpenAI tiers sit:

ModelInput (₹ / 1M tokens)Output (₹ / 1M tokens)
GPT-5 mini25.2201.6
GPT-4.1201.6806.4

Calling o4-mini from the unoblox endpoint

curl https://api.unoblox.ai/v1/chat/completions \
  -H "Authorization: Bearer ub-gw-xxxxxxxxxxxxxxxx" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "openai/o4-mini",
    "messages": [{"role": "user", "content": "Debug why this function returns the wrong total."}],
    "max_tokens": 1000
  }'

As with any o-series model, give the completion budget enough headroom for internal reasoning — a max_tokens that's too tight is the most common reason a reasoning-model call ends without a usable answer.

o4-mini vs o3 vs GPT-5 mini

  • o4-mini — reasoning behaviour at a lower cost than a full reasoning model; a fit for debugging, structured logic and multi-step tasks at higher volume than o3 would be economical for.
  • o3 — reserve for the harder subset of reasoning tasks where o4-mini's answers aren't reliable enough.
  • GPT-5 mini — for tasks that don't need explicit step-by-step reasoning at all, a non-reasoning small model is usually faster and cheaper for the same quality bar.

Frequently asked questions

Is o4-mini available in India through unoblox? Yes, openai/o4-mini is callable on https://api.unoblox.ai/v1 right after signup, billed in rupees.

What does o4-mini cost on unoblox? It isn't in our fixed indicative price list; check the live ₹ rate on its model page.

How is o4-mini different from GPT-5 mini? o4-mini reasons through a problem step by step before answering; GPT-5 mini answers directly. Reasoning helps most on math, logic and debugging tasks.

Why is my o4-mini response getting cut off? Increase max_tokens — part of the budget is used for internal reasoning before the visible answer is written.

Does o4-mini run on infrastructure in India? No, it runs on OpenAI's infrastructure. unoblox adds ₹ billing, a GST invoice and one endpoint, not data residency.

Is there one GST invoice covering o4-mini and other models? Yes, all usage is combined into a single monthly GST invoice from an Indian entity, with input tax credit claimable.

Get started in rupees → https://unoblox.ai/sign-in

o4-mini api indiallmopenaio4-mini
ShareLinkedInWhatsAppTelegram

Start building in rupees

Call every major model through one OpenAI-compatible endpoint, billed in ₹ on a GST invoice.