Pricing

Free to use. Pay for what you run.

MindsHub Cowork and Unified Inference use the same rate card. There is no subscription or plan.

Rate card

Model rates.

Prices are in US dollars per million tokens. Expand a row for long-prompt rates, web search, cache writes, reasoning levels, and failover order.

Compare models
Model Provider Input Output Cached input Cache write
Free allowance: the first 5M Air tokens each month are on us MindsHub $0.20 $1.20 $0.02 $0.25
Free allowance
5M tokens a month at no charge, in Cowork and through the API alike. The rates here apply beyond the allowance. Air is an alias we keep pointed at the model that strikes the best balance of cost, speed and intelligence, so both the target and these rates can move as better options ship.
Long prompts — 272K+ tokens
$0.40 input $1.80 output $0.04 cached input $0.50 cache write Per 1M tokens. Once a prompt crosses 272K tokens, these rates replace the standard ones for the whole request — the same way MindsHub bills it.
Failover order
Gemini 3.7 Flash Tried in this order if MindsHub Air is unavailable, so your work keeps moving.
Anthropic $2.00 $10.00 $0.20 $2.50
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$10.00 per 1,000 searches Anthropic's own search, built into the model.
Reasoning effort
lowmediumhighmax Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send high unless you ask for another.
Failover order
GPT 5.6 Sol Tried in this order if Claude Sonnet 5 is unavailable, so your work keeps moving.
Anthropic $5.00 $25.00 $0.50 $6.25
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$10.00 per 1,000 searches Anthropic's own search, built into the model.
Reasoning effort
lowmediumhighxhighmax Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send high unless you ask for another.
Failover order
GPT 5.6 Sol Tried in this order if Claude Opus 5 is unavailable, so your work keeps moving.
Anthropic $10.00 $50.00 $1.00 $12.50
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$10.00 per 1,000 searches Anthropic's own search, built into the model.
Reasoning effort
lowmediumhighxhighmax Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send high unless you ask for another.
Failover order
GPT 5.6 Sol Tried in this order if Claude Fable 5 is unavailable, so your work keeps moving.
Anthropic $1.00 $5.00 $0.10 $1.25
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$10.00 per 1,000 searches Anthropic's own search, built into the model.
Failover order
Gemini 3.7 Flash DeepSeek V4-Pro-0813 Tried in this order if Claude Haiku 4.5 is unavailable, so your work keeps moving.
OpenAI $5.00 $30.00 $0.50 $6.25
Long prompts — 272K+ tokens
$10.00 input $45.00 output $1.00 cached input $12.50 cache write Per 1M tokens. Once a prompt crosses 272K tokens, these rates replace the standard ones for the whole request — the same way OpenAI bills it.
Web search
$10.00 per 1,000 searches OpenAI's own search, built into the model.
Reasoning effort
nonelowmediumhighxhighmax Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send medium unless you ask for another.
Failover order
Claude Opus 5 Gemini 3.1 Pro Preview Tried in this order if GPT 5.6 Sol is unavailable, so your work keeps moving.
OpenAI $2.00 $12.00 $0.20 $2.50
Long prompts — 272K+ tokens
$4.00 input $18.00 output $0.40 cached input $5.00 cache write Per 1M tokens. Once a prompt crosses 272K tokens, these rates replace the standard ones for the whole request — the same way OpenAI bills it.
Web search
$10.00 per 1,000 searches OpenAI's own search, built into the model.
Reasoning effort
nonelowmediumhighxhighmax Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send medium unless you ask for another.
Failover order
Claude Sonnet 5 Gemini 3.1 Pro Preview Tried in this order if GPT 5.6 Terra is unavailable, so your work keeps moving.
OpenAI $0.20 $1.20 $0.02 $0.25
Long prompts — 272K+ tokens
$0.40 input $1.80 output $0.04 cached input $0.50 cache write Per 1M tokens. Once a prompt crosses 272K tokens, these rates replace the standard ones for the whole request — the same way OpenAI bills it.
Web search
$10.00 per 1,000 searches OpenAI's own search, built into the model.
Reasoning effort
nonelowmediumhighxhighmax Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send medium unless you ask for another.
Failover order
Gemini 3.7 Flash Claude Haiku 4.5 Tried in this order if GPT 5.6 Luna is unavailable, so your work keeps moving.
OpenAI $1.75 $14.00 $0.18 $1.75
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$10.00 per 1,000 searches OpenAI's own search, built into the model.
Reasoning effort
lowmediumhighxhigh Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send medium unless you ask for another.
Failover order
Claude Sonnet 5 Tried in this order if GPT 5.3 Codex is unavailable, so your work keeps moving.
OpenAI $0.75 $4.50 $0.08 $0.75
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$10.00 per 1,000 searches OpenAI's own search, built into the model.
Reasoning effort
nonelowmediumhighxhigh Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send none unless you ask for another.
Failover order
Gemini 3.7 Flash Claude Haiku 4.5 Tried in this order if GPT 5.4 Mini is unavailable, so your work keeps moving.
OpenAI $0.20 $1.25 $0.02 $0.20
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$10.00 per 1,000 searches OpenAI's own search, built into the model.
Reasoning effort
nonelowmediumhighxhigh Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send none unless you ask for another.
Failover order
Gemini 3.7 Flash Tried in this order if GPT 5.4 Nano is unavailable, so your work keeps moving.
Google $2.00 $12.00 $0.20 Free
Long prompts — 200K+ tokens
$4.00 input $18.00 output $0.40 cached input Per 1M tokens. Once a prompt crosses 200K tokens, these rates replace the standard ones for the whole request — the same way Google bills it.
Web search
$14.00 per 1,000 searches Google's own search, built into the model.
Cache writes
Free on this model — you're billed only when cached content is read back, at the cached-input rate above.
Reasoning effort
lowmediumhigh Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send high unless you ask for another.
Failover order
Claude Sonnet 5 GPT 5.6 Sol Tried in this order if Gemini 3.1 Pro Preview is unavailable, so your work keeps moving.
Google $0.75 $3.75 $0.08 Free
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$14.00 per 1,000 searches Google's own search, built into the model.
Cache writes
Free on this model — you're billed only when cached content is read back, at the cached-input rate above.
Reasoning effort
lowmediumhigh Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send medium unless you ask for another.
Failover order
GPT 5.4 Mini DeepSeek V4-Pro-0813 Tried in this order if Gemini 3.7 Flash is unavailable, so your work keeps moving.
Google $0.75 $3.75 $0.08 Free
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$14.00 per 1,000 searches Google's own search, built into the model.
Cache writes
Free on this model — you're billed only when cached content is read back, at the cached-input rate above.
Reasoning effort
minimallowmediumhigh Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send medium unless you ask for another.
Failover order
GPT 5.4 Mini DeepSeek V4-Pro-0813 Tried in this order if Gemini 3.6 Flash is unavailable, so your work keeps moving.
Google $1.50 $9.00 $0.15 Free
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$14.00 per 1,000 searches Google's own search, built into the model.
Cache writes
Free on this model — you're billed only when cached content is read back, at the cached-input rate above.
Reasoning effort
minimallowmediumhigh Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send medium unless you ask for another.
Failover order
GPT 5.4 Mini DeepSeek V4-Pro-0813 Tried in this order if Gemini 3.5 Flash is unavailable, so your work keeps moving.
Google $0.50 $3.00 $0.05 Free
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$14.00 per 1,000 searches Google's own search, built into the model.
Cache writes
Free on this model — you're billed only when cached content is read back, at the cached-input rate above.
Reasoning effort
minimallowmediumhigh Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send high unless you ask for another.
Failover order
GPT 5.4 Mini DeepSeek V4-Pro-0813 Tried in this order if Gemini 3 Flash Preview is unavailable, so your work keeps moving.
Google $0.25 $1.50 $0.03 Free
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$14.00 per 1,000 searches Google's own search, built into the model.
Cache writes
Free on this model — you're billed only when cached content is read back, at the cached-input rate above.
Reasoning effort
minimallowmediumhigh Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send minimal unless you ask for another.
Failover order
GPT 5.4 Nano DeepSeek V4-Pro-0813 Tried in this order if Gemini 3.1 Flash-Lite is unavailable, so your work keeps moving.
Fireworks AI $3.00 $15.00 $0.30 $3.00
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$7.00 per 1,000 searches $1.00 per 1,000 page fetches Served by Exa, which bills page fetches separately.
Failover order
DeepSeek V4-Pro-0813 Tried in this order if Kimi K3 is unavailable, so your work keeps moving.
Fireworks AI $1.32 $3.96 $0.05 $1.32
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$7.00 per 1,000 searches $1.00 per 1,000 page fetches Served by Exa, which bills page fetches separately.
Reasoning effort
lowhighmax Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send high unless you ask for another.
Failover order
Kimi K3 Tried in this order if DeepSeek V4-Pro-0813 is unavailable, so your work keeps moving.
Fireworks AI $1.74 $3.48 $0.15 $1.74
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$7.00 per 1,000 searches $1.00 per 1,000 page fetches Served by Exa, which bills page fetches separately.
Reasoning effort
lowhighmax Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send high unless you ask for another.
Failover order
Kimi K3 Tried in this order if DeepSeek V4-Pro-0813 is unavailable, so your work keeps moving.
Fireworks AI $2.00 $6.00 $0.25 $2.00
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$7.00 per 1,000 searches $1.00 per 1,000 page fetches Served by Exa, which bills page fetches separately.
Reasoning effort
lowmediumxhigh Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send xhigh unless you ask for another.
Failover order
DeepSeek V4-Pro-0813 Tried in this order if Qwen3.8-2.4T-A95B is unavailable, so your work keeps moving.
Fireworks AI $0.40 $1.60 $0.08 $0.40
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$7.00 per 1,000 searches $1.00 per 1,000 page fetches Served by Exa, which bills page fetches separately.
Failover order
DeepSeek V4-Pro-0813 Tried in this order if Qwen3.7 Plus is unavailable, so your work keeps moving.
Fireworks AI $1.40 $4.40 $0.14 $1.40
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$7.00 per 1,000 searches $1.00 per 1,000 page fetches Served by Exa, which bills page fetches separately.
Reasoning effort
highmax Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send max unless you ask for another.
Failover order
DeepSeek V4-Pro-0813 Tried in this order if GLM 5.2 is unavailable, so your work keeps moving.
Meta $1.25 $4.25 $0.15 $1.25
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$7.00 per 1,000 searches $1.00 per 1,000 page fetches Served by Exa, which bills page fetches separately.
Failover order
Claude Sonnet 5 Kimi K3 Tried in this order if Muse Spark 1.2 is unavailable, so your work keeps moving.
Meta $1.25 $4.25 $0.15 $1.25
Long prompts
No surcharge — the full context window bills at the standard rates above.
Web search
$7.00 per 1,000 searches $1.00 per 1,000 page fetches Served by Exa, which bills page fetches separately.
Failover order
Claude Sonnet 5 Kimi K3 Tried in this order if Muse Spark 1.1 is unavailable, so your work keeps moving.
SpaceXAI $2.00 $6.00 $0.50 $2.00
Long prompts — 200K+ tokens
$4.00 input $12.00 output $1.00 cached input $4.00 cache write Per 1M tokens. Once a prompt crosses 200K tokens, these rates replace the standard ones for the whole request — the same way SpaceXAI bills it.
Web search
$5.00 per 1,000 searches SpaceXAI's own search, built into the model.
Reasoning effort
lowmediumhighxhigh Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send high unless you ask for another.
Failover order
Claude Sonnet 5 Kimi K3 Tried in this order if Grok 4.6 is unavailable, so your work keeps moving.
SpaceXAI $2.00 $6.00 $0.30 $2.00
Long prompts — 200K+ tokens
$4.00 input $12.00 output $0.60 cached input $4.00 cache write Per 1M tokens. Once a prompt crosses 200K tokens, these rates replace the standard ones for the whole request — the same way SpaceXAI bills it.
Web search
$5.00 per 1,000 searches SpaceXAI's own search, built into the model.
Reasoning effort
lowmediumhigh Dial thinking up or down per task. Higher effort spends more output tokens, billed at the same rate. We send high unless you ask for another.
Failover order
Claude Sonnet 5 Kimi K3 Tried in this order if Grok 4.5 is unavailable, so your work keeps moving.
Units
Token prices are USD per 1M tokens. Web search and page fetches are USD per 1,000 calls.
Cached input
The discounted rate for prompt content the provider already has cached.
Cache write
What it costs to put content into that cache in the first place.
More per model
Expand any row for its long-prompt rates, search pricing and failover order.
MindsHub Foundry

Double your first 3 months of credit.

Selected teams test new features with our product team and receive a 100% match on inference-credit purchases for three months, up to $5,000.

About the Foundry

Cowork runs in your browser and in the desktop app — the same workspace either way. Get the app for macOS or Windows.

Ask questions and share ideas in the community Discord, or send a private support request about a specific issue.

FAQ

Common questions.

What does it cost to start?
Nothing. Accounts start with included monthly usage on MindsHub Air without a card. Input, output, cached reads, and cache writes all count toward it.
How do I use a model other than MindsHub Air?
Add credits and pay the published per-model rate. You can top up manually or use auto-recharge with a monthly cap.
Can I use my own provider API keys instead?
Yes. Add your provider keys and pay the provider directly.
Is there a subscription or a minimum?
No. There is no subscription, monthly fee, or minimum spend. You pay only for hosted usage beyond the included monthly tokens, or use your own keys.
What is MindsHub Air?
MindsHub Air is an alias that can move to a different model as the catalog changes. It carries the included monthly usage available to your organization. Air does not support web search or reasoning-effort controls.
What happens after I use the included monthly tokens?
It refreshes at the start of the next month. Until then, add credits or use your own provider API key.
Do Cowork and Unified Inference share this pricing?
Yes. Both use the same rate card and prepaid balance.
Which models are available?
The catalog includes MindsHub Air, Claude, GPT, Gemini, Grok, Muse Spark, Kimi, DeepSeek, Qwen, and GLM families. Check the live rate card for current versions, prices, and failover order.
Can I see what a task cost?
Yes. Your account shows usage, and the rate card lists each model price.
How do I get support?
Ask the team and community in Discord, or submit a private support request at mindshub.ai/support for a specific issue.
Can I pay with cryptocurrency?
Standard top-ups use a card. Volume customers can ask the team about settling invoices in cryptocurrency.
How do models stay current?
Family aliases can move to newer model versions. The rate card shows the fallback order for each model that has one.
What is MindsHub?
MindsHub includes Unified Inference, the API, and Cowork, the workspace. They share one model catalog and balance.
Can I run MindsHub myself?
Yes. You can run the open-source project on your computer or in your VPC with your own endpoints and keys. This page prices the hosted service.
Which open-source agents power MindsHub?
Cowork supports Anton and Hermes. Both are open-source agent harnesses, and switching does not remove your workspace data.