Why we show you this
Most software companies hide their costs. They don't want you to know the margin, or how much the AI APIs actually cost, or how much of your subscription goes to cloud infrastructure versus the vendor's pocket. Opacity is the default, and it is usually intentional.
We do the opposite. This page shows, line by line, what it costs to run OmniThink — the AI API bills, the infrastructure, the payment fees — and exactly how the prices are derived from those costs. We do it because we think you deserve to know, and because the only way to ask for your support honestly is to first show you what you are supporting.
What one question costs us
When you ask a question, it is sent to all five models at the same moment. Every token each model reads and writes is billed to us at that provider's published API rate. A typical question is about 800 tokens in (your question plus any active Answer Modes) and 600 tokens out per model. At the current rates (verified July 2026):
| Model | Provider | Input /1M | Output /1M | Per question |
|---|---|---|---|---|
| ChatGPT (GPT-5.5) | OpenAI | $5.00 | $30.00 | $0.0220 |
| Claude (Opus 4.8) | Anthropic | $5.00 | $25.00 | $0.0190 |
| Gemini (3.1 Pro) | $2.00 | $12.00 | $0.0088 | |
| Grok (4.3) | xAI | $1.25 | $2.50 | $0.0025 |
| DeepSeek (V4 Pro) | DeepSeek | $0.435 | $0.87 | $0.0008 |
| All five models, one question | ≈ $0.053 | |||
Running Compare or Aggregate afterwards is a sixth call — Claude reads all five answers (≈4,800 tokens in, 800 out) and synthesises them: about $0.044 more. So a full session — five models plus a Compare — costs us roughly 10 cents in raw AI fees. Answer Modes lengthen the input; multi-round Debate re-sends the growing transcript each round and can cost several times more per question.
What it costs to keep the lights on
Behind every question there is production infrastructure running continuously, whether anyone is using it or not:
| Component | Monthly |
|---|---|
| API backend, security gateway, cache, database, CDN, monitoring, key vault (Azure) | $62–102 |
| Analytics, email, domains, certificates | $8–13 |
| Baseline before a single question is asked | ≈ $70–115 |
Plus payment processing (Stripe keeps 2.9% + $0.30 of every charge: 53¢ of a Plus payment, 73¢ of a Pro payment) and tax on revenue.
The plans, and how the prices were derived
| Plan | Price | Questions | Who it's for |
|---|---|---|---|
| Guest | $0 | 3 / day | Try it with zero friction — no account. |
| Free account | $0 | 20 / month | Occasional use, full feature access. |
| Plus | $7.99 / mo | 150 / month | Regular use across all five models. |
| Pro | $14.99 / mo | 300 / month | Daily serious use; double the headroom. |
The derivation is straightforward. Take Pro: of $14.99, Stripe takes $0.73 and we provision 15% for tax, leaving about $12. A subscriber who uses their entire 300-question quota at realistic usage (three models per question on average, Compare on some) costs about $13–16 in AI fees plus an infrastructure share. In other words: a subscriber who maxes their plan costs us slightly more than they pay, and everyone who uses less than the maximum is where the sustainability comes from. The price tracks the cost — there is no hidden multiple on top.
The same is true of Plus at its 150-question scale. Free tiers are funded by the paid ones: a free account costs us at most about a dollar a month, and a guest a few cents a day — that is the marketing budget, spent on you trying the product instead of on ads.
Compared with subscribing directly
| What you could pay instead | Monthly |
|---|---|
| ChatGPT Plus + Claude Pro + Gemini Advanced, separately | $60 |
| OmniThink Plus — all five models, side by side | $7.99 |
| OmniThink Pro — all five models, double quota | $14.99 |
The catch, stated honestly: those direct subscriptions include unlimited-ish chat with one model. OmniThink gives you a monthly budget of questions answered by all five at once, with Compare, Debate, and 46 Answer Modes on top. Different shape, radically lower price for what it does.
Fair use, and what keeps this sustainable
- One quota pool. Your plan's monthly quota is shared between the OmniThink app and partner apps that use your account — one subscription, one pool.
- Personal use only. No automation, scripting, or resale. Quotas and rate limits enforce this technically, not just contractually.
- Bounded answers. Model responses are capped at generous lengths so a single request can't consume unbounded compute.
- We eat price changes first. When providers change API prices (OpenAI doubled theirs in April 2026), we update this page and absorb the difference until a considered re-pricing — we don't silently degrade the product.
The bottom line
| Item | Number |
|---|---|
| One question, all five models | ≈ $0.053 |
| Full session with Compare | ≈ $0.10 |
| Serving a quota-maxing Pro subscriber, all-in | ≈ $16–20 / mo |
| What that subscriber pays | $14.99 / mo |
| Margin padding hidden in the price | $0 |
The math is honest. The price is real. If you use OmniThink regularly and find it valuable, subscribing is the most direct way to say: this should continue to exist.
Rates verified against provider price lists July 2026. This page is updated whenever provider pricing, infrastructure costs, or the plans change. © 2025–2026 Belvantis. All rights reserved.