Build a chatbot with an AI API in India (₹ rupees)
Build an AI chatbot in India using unoblox. ₹-native API, GST invoice, one endpoint for 40+ models. See the 3-step setup and real ₹ costs.
Build a chatbot with an AI API in India
Building a chatbot? In India, use an OpenAI-compatible API billed in ₹ rupees with a GST invoice. No international payment friction, input-tax-credit claimable. Here's how to ship in 3 steps.
Why unoblox for India chatbots
- ₹ billing: Invoice in Indian rupees (no conversion)
- GST invoice: Input-tax-credit claimable for your business
- OpenAI-compatible: Drop-in replacement for OpenAI SDK
- 40+ models: GPT, Claude, DeepSeek, Qwen, all at one endpoint
- Data residency: India-hosted compliance with DPDP
Recommended model: Qwen3 Max or Claude Sonnet
- Qwen3 Max (₹120.95 input / ₹604.77 output per 1M tokens): Balanced quality and speed, best for high-volume chats
- Claude Sonnet (₹201.6 input / ₹1,008 output per 1M tokens): Superior tone and context, best for customer-facing bots
Step 1: Create an API key
- Sign up at unoblox.ai/sign-in
- Go to Keys in the dashboard
- Create a key
ub-gw-… - Copy and store in your
.env
Step 2: Install OpenAI SDK (swap endpoint)
Your code stays unchanged. Just swap the endpoint and key.
from openai import OpenAI
client = OpenAI(
base_url="https://api.unoblox.ai/v1",
api_key="ub-gw-..."
)
response = client.chat.completions.create(
model="qwen/qwen3-max",
messages=[{"role": "user", "content": "Hello! Can you help?"}],
stream=True
)
for chunk in response:
print(chunk.choices[0].delta.content, end="")
Step 3: Deploy and monitor ₹ costs
- In your app: Point to
https://api.unoblox.ai/v1 - Monitor dashboard: See real-time usage in ₹ rupees
- Receive invoice: Monthly GST invoice; claim input-tax credit
Real ₹ cost: 1000 users, 5 queries/day
Assumptions:
- 5,000 queries/day
- Avg: 100 input tokens, 150 output tokens
- Using Qwen3 Max (₹120.95 input / ₹604.77 output per 1M tokens)
Daily cost (Qwen3 Max):
- Input: ₹60.48
- Output: ₹453.58
- Total: ₹514/day ≈ ₹15,420/month
Using Claude Sonnet (₹201.6 input / ₹1,008 output per 1M tokens):
- Input: ₹100.80
- Output: ₹756.00
- Total: ₹856.80/day ≈ ₹25,700/month
Qwen saves roughly ₹10,280/month for the same volume.
Features: Streaming, tools, JSON mode
All OpenAI SDK features work unchanged:
- Streaming: Real-time responses
- Function calling: Let the bot call APIs
- JSON mode: Structured outputs for automation
- Vision: Upload images (GPT-4o only)
Production tips
- Set model per workspace: Change model= param to test Qwen3, Claude, GPT
- Track ₹ spend: Dashboard shows real-time costs; set budget alerts
- Claim GST: Save invoices; consult your CA on input-tax credit
- Scale limits: Request higher rate limits via dashboard
Frequently asked questions
Q: Can I switch models mid-chat? Yes. Change the model param. Both are OpenAI-compatible; no refactoring needed.
Q: What if rate limits are hit? Contact support or request higher limits in the dashboard. Limits start at 100 req/min.
Q: How do I handle conversation memory? Store messages in Postgres or SQLite. Send recent history in messages=[] array with each request.
Q: Does unoblox comply with DPDP and RBI? Yes. Data hosted in India, ₹ billing, monthly compliance reports available.
Q: Can I use LangChain or LlamaIndex? Absolutely. Both support OpenAI-compatible endpoints. Swap base_url and api_key—everything works unchanged.
Q: What about conversation rate limits? No limits on conversation length. Embed old messages to compress context if needed.
Get started in rupees
Get started → https://unoblox.ai/sign-in
Start building in rupees
Call every major model through one OpenAI-compatible endpoint, billed in ₹ on a GST invoice.