Gemini 2.5 Flash vs GPT-4o Mini: API Price & Benchmark Comparison (INR)
Side-by-side pricing
| Gemini 2.5 Flash | GPT-4o Mini | |
|---|---|---|
| Input price / 1M tokens | ₹30.95 | ₹15.48 |
| Output price / 1M tokens | ₹257.92 | ₹61.90 |
| Context window | 1,048,576 | 128,000 |
| Provider | OpenAI |
Velona effective INR prices, refreshed every 6 hours (last: 2026-07-28T17:32:59Z).
When to pick which
GPT-4o Mini is half the input price ($0.15 vs $0.30 per million) but Gemini 2.5 Flash pulls ahead almost everywhere else: a 1M-token context window versus 128K, stronger multimodality, and generally better reasoning for a small model. Output pricing favours GPT-4o Mini ($0.60 vs $2.50), so generation-heavy workloads like long summaries can still be cheaper on OpenAI. Use GPT-4o Mini for short, high-volume, output-heavy calls. Use Flash when context length, vision, or answer quality is the constraint. It punches a clear class above.
Benchmarks
Gemini 2.5 Flash
Benchmarks
Independent scores from Artificial Analysis.
GPT-4o Mini
Benchmarks
Independent scores from Artificial Analysis.
Cost calculator for both models
Frequently asked questions
Which is cheaper: Gemini 2.5 Flash or GPT-4o Mini?
On input tokens, Gemini 2.5 Flash costs ₹30.95/1M and GPT-4o Mini costs ₹15.48/1M through Velona. Output tokens: ₹257.92 vs ₹61.90 per 1M.
Can I switch between Gemini 2.5 Flash and GPT-4o Mini without changing code?
Yes. Both are behind Velona's single gateway API. Change the model id in your request body and everything else stays the same, including your INR wallet.
Are these prices current?
Prices are regenerated every 6 hours from live provider pricing and the current USD→INR rate, so the figures on this page track what you'd actually be billed.
Velona is India's INR-native AI API gateway with 300+ models behind one API key, UPI top-up from ₹10, no foreign card needed.