Skip to content
Start freeGPT-5 · Claude · DeepSeek V4 · Qwen3 — in ₹One OpenAI-compatible endpointBilled in rupeesGST invoiceNo international cardGet started →Start freeGPT-5 · Claude · DeepSeek V4 · Qwen3 — in ₹One OpenAI-compatible endpointBilled in rupeesGST invoiceNo international cardGet started →
API Pricing Guides

Cheapest Vision LLM in India (₹ Pricing)

Qwen3.8-27B is unoblox's cheapest vision-capable model at ₹16.32 per 1M input tokens, billed in rupees with a monthly GST invoice.

If you need a model that can read images as well as text, Qwen3.8-27B is the cheapest vision-capable option on unoblox's catalog — ₹16.32 per 1 million input tokens, ₹48.96 per 1 million output tokens, billed the same rupee-and-GST way as every other model.

The cheapest vision-capable model on unoblox

ModelInput (₹ / 1M tokens)Output (₹ / 1M tokens)Notes
Qwen3.8-27B₹16.32₹48.96Open weights, accepts image input

At roughly a twelfth of Claude Opus's input rate and well under half of Claude Sonnet's, Qwen3.8-27B is the practical starting point if your workload needs to look at images — receipts, screenshots, product photos, scanned documents — and you're price-sensitive.

What "vision" means in practice

A vision-capable model accepts image content in the same request as your text prompt — you're not running a separate OCR step and feeding text back in; the model reads the image directly alongside your instructions. This is useful for tasks like describing an image, extracting structured data from a photographed document, or answering questions about a chart or screenshot.

If you need more capability than Qwen3.8-27B offers

Several of the larger closed models on the catalog are also generally multimodal and accept image input, at a higher per-token rate:

ModelInput (₹ / 1M tokens)Output (₹ / 1M tokens)
GPT-4.1₹201.6₹806.4
Claude Sonnet₹201.6₹1,008
GPT-4o₹252₹1,008
Claude Opus₹504₹2,520

These are worth the extra cost when a task genuinely needs stronger reasoning over what's in the image, not just recognition — for example, reasoning about relationships between multiple charts rather than reading one clearly-labelled receipt. Always confirm image-input support and the exact rate on the model's own page before building against it.

A worked example

Image-heavy requests tend to be input-heavy, because an image encodes to a meaningful number of tokens before your text prompt is even counted. A request that works out to 4,000 input tokens (image plus prompt) and 300 output tokens on Qwen3.8-27B costs: (4,000/1,000,000 × ₹16.32) + (300/1,000,000 × ₹48.96) = ₹0.065 + ₹0.015 ≈ ₹0.08 per request. At 10,000 such requests a month, that's roughly ₹800.

Calling a vision model from your code

The request shape is the same OpenAI-compatible format you already use for text-only calls — you add image content to the message, and the model field points at a vision-capable id.

base_url: https://api.unoblox.ai/v1
api_key: ub-gw-xxxxxxxxxxxxxxxxxxxx
model: qwen/qwen3.8-27b

Frequently asked questions

Is Qwen3.8-27B the only vision-capable model on unoblox? It's the cheapest one, and the one explicitly listed with vision support in the current catalog. Larger models like GPT-4o, GPT-4.1, and the Claude Sonnet/Opus family are also generally multimodal — check each model's page to confirm before you build against it.

Does "open weights" mean anything for how I call Qwen3.8-27B? No — you call it through the same OpenAI-compatible endpoint and key as every other model on unoblox. "Open weights" describes how the model itself is distributed, not how you access it here.

Is image processing billed differently from text? No separate line item — an image is encoded into tokens the same way text is, and counted as part of your input tokens at the model's published input rate.

Can I send multiple images in one request? That depends on the model's own limits rather than anything unoblox adds on top. Check the model's page and your SDK's documentation for the exact constraints.

Does using a vision model mean my images are stored or used for training? That depends on the underlying model provider's own policy — check the relevant provider terms linked from the model's page rather than assuming either way.

Is there a free way to test a vision workload before paying? Qwen3.8-27B's rate is already low enough for cheap prototyping. unoblox's free-tier model, Qwen3 1.7B, is not listed as vision-capable, so for image workloads Qwen3.8-27B is the better starting point even while testing.

Get started in rupees → https://unoblox.ai/sign-in

cheapest vision llm indiacostvisionpricing
ShareLinkedInWhatsAppTelegram

Start building in rupees

Call every major model through one OpenAI-compatible endpoint, billed in ₹ on a GST invoice.