Skip to content
Start freeGPT-5 · Claude · DeepSeek V4 · Qwen3 — in ₹One OpenAI-compatible endpointBilled in rupeesGST invoiceNo international cardGet started →Start freeGPT-5 · Claude · DeepSeek V4 · Qwen3 — in ₹One OpenAI-compatible endpointBilled in rupeesGST invoiceNo international cardGet started →
Model API Guides

Qwen3.8 Flash API in India

Qwen3.8 Flash: ₹10.88 input, ₹36.78 output per 1M tokens in India. Alibaba's low-cost, fast Qwen3.8 tier with a 1M-token context, billed in rupees.

Qwen3.8 Flash is Alibaba's fast, low-cost tier of the Qwen3.8 generation, live on unoblox as qwen/qwen3.8-flash. It offers a one-million-token context, reasoning and tool calling, billed in rupees.

What Qwen3.8 Flash is

Qwen3.8 Flash is the economical option in the Qwen3.8 generation, next to the flagship Qwen3.8 Max. unoblox lists a context length of 1,000,000 tokens, reasoning support and tool calling, with text in and text out.

Flash is meant for the bulk of everyday traffic: assistants, extraction, routing and agent steps where per-token cost and latency dominate.

Where it fits

  • Chatbots and support assistants with long conversation memory.
  • Structured extraction from large batches of documents.
  • Agent sub-tasks such as planning, routing and tool calls.

Use Flash for volume and Qwen3.8 Max when a task needs the flagship. Compare it with MiMo V2.6 Flash and DeepSeek V4 Flash (0731) on your own prompts.

₹ pricing for Qwen3.8 Flash

On unoblox, Qwen3.8 Flash costs ₹10.88 per million input tokens and ₹36.78 per million output tokens. That displayed price is what you are billed per token; nothing is added on top. Cached input tokens are listed at ₹1.36 per million. The model's context length on unoblox is 1,000,000 tokens.

ModelInput (₹ / 1M tokens)Output (₹ / 1M tokens)
Qwen3.8 Flash₹10.88₹36.78
Qwen3.8 Max₹158.87₹476.69
DeepSeek V4 Flash (0731)₹5.78₹17.33
MiMo V2.6 Flash₹13.48₹26.96

As a worked example, a month with 20 million input tokens and 4 million output tokens comes to about ₹364.72 on Qwen3.8 Flash (₹217.60 for input plus ₹147.12 for output). Live rates are always on the model page, and each request is rounded up to the nearest paisa.

Calling Qwen3.8 Flash from the unoblox endpoint

unoblox is OpenAI-compatible, so any OpenAI SDK works by changing the base URL and key.

curl https://api.unoblox.ai/v1/chat/completions \
  -H "Authorization: Bearer YOUR_UNOBLOX_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "qwen/qwen3.8-flash",
    "messages": [{"role": "user", "content": "Summarise this support ticket in two lines."}],
    "max_tokens": 512
  }'
from openai import OpenAI

client = OpenAI(
    api_key="YOUR_UNOBLOX_API_KEY",
    base_url="https://api.unoblox.ai/v1",
)
resp = client.chat.completions.create(
    model="qwen/qwen3.8-flash",
    messages=[{"role": "user", "content": "Summarise this support ticket in two lines."}],
    max_tokens=512,
)
print(resp.choices[0].message.content)

Get a key at https://unoblox.ai/sign-in and top up your wallet in rupees.

Notes on parameters, billing and data

Use max_tokens to cap output. The model is served by an external provider rather than from India, so no data-residency claim is made. unoblox provides rupee pricing, GST invoicing and a single endpoint.

Frequently asked questions

Is Qwen3.8 Flash available in India? Yes, as qwen/qwen3.8-flash on https://api.unoblox.ai/v1.

What is the context length? unoblox lists 1,000,000 tokens.

Does it support reasoning and tools? Yes, the catalogue lists both.

How does it differ from Qwen3.8 Max? Max is the flagship with higher per-token prices; Flash is the economical tier.

What currency is billing in? Rupees, with a GST invoice.

Does it accept images? unoblox lists it as text in, text out.

Get started in rupees → https://unoblox.ai/sign-in

qwen3.8 flash api indiallmqwenqwen3.8-flash
ShareLinkedInWhatsAppTelegram

Start building in rupees

Call every major model through one OpenAI-compatible endpoint, billed in ₹ on a GST invoice.