Models for octomind.
Zero keys.
Skip collecting API keys from five providers. One login — every model works, in the CLI and in the cloud. Free models forever; the good ones in the plan. Open to everyone today — no invite, no waitlist.
$ curl -fsSL https://octomind.run/install.sh | bash
$ octomind login # approve in the browser — no keys to paste
$ octomind # models just work — pick any with -m octohub:…
Transparent prices
Published per-token prices, metering that matches the table exactly. Caps you can see live — never a surprise bill.
Free tier, forever
A capable free model with a daily allowance — even when caps or credits run out, your agent degrades gracefully instead of dying.
One key for models everywhere
The same hub key serves model calls in the CLI, inside cloud machines, and from any OpenAI-compatible tool pointed at the endpoint. (Scripting machines and sessions is the separate Developer API, documented in the panel.)
The table
11 open coding models included in the plan · 12 premium models via prepaid credits · 10 embedding models. Same table the panel and the meter use.
| Model | Input /1M | Output /1M | Context |
|---|---|---|---|
| Included in every paid plan | |||
| DeepSeek V4 Flash free tier | $0.14 | $0.28 | 1M |
| DeepSeek V4 Pro | $0.43 | $0.87 | 1M |
| Qwen3.7 Plus | $0.32 | $1.28 | 256k |
| Qwen3.7 Max | $1.25 | $3.75 | 256k |
| Gemma 4 31B | $0.20 | $0.20 | 128k |
| MiniMax M3 | $0.60 | $2.40 | 1M |
| Kimi 2.6 | $0.60 | $2.50 | 256k |
| Kimi 2.7 Code | $0.95 | $4.00 | 256k |
| GLM-5.2 | $1.40 | $4.40 | 200k |
| Inkling | $1.87 | $4.68 | 1M |
| Kimi K3 | $3.00 | $15.00 | 1M |
| Claude Sonnet 5 | $2.00 | $10.00 | 200k |
| Claude Opus 4.8 | $5.00 | $25.00 | 1M |
| Claude Fable 5 | $10.00 | $50.00 | 1M |
| GPT 5.6 Luna | $1.00 | $6.00 | 1M |
| GPT 5.6 Terra | $2.50 | $15.00 | 1M |
| GPT 5.6 Sol | $5.00 | $30.00 | 1M |
| GPT 5.5 | $5.00 | $30.00 | 1M |
| GPT 5.4 Mini | $0.75 | $4.50 | 400k |
| GPT 5.4 Nano | $0.20 | $1.25 | 400k |
| GPT 5.5 Pro | $30.00 | $180.00 | 1M |
| Gemini 3.5 Flash | $1.50 | $9.00 | 1M |
| Gemini 3.1 Pro | $2.00 | $12.00 | 1M |
| Voyage 4 Large | $0.12 | — | 32k |
| Voyage 4 | $0.06 | — | 32k |
| Voyage 4 Lite | $0.02 | — | 32k |
| Voyage Code 3 | $0.18 | — | 32k |
| Jina v5 Omni Small | $0.02 | — | 32k |
| Jina v5 Omni Nano | $0.02 | — | 8k |
| Jina v5 Text Small | $0.02 | — | 32k |
| Jina v5 Text Nano | $0.02 | — | 8k |
| OpenAI Embedding 3 Small | $0.02 | — | 8k |
| OpenAI Embedding 3 Large | $0.13 | — | 8k |
Pro caps: $20 per 4-hour window · $60/week · $120/month of usage for $20/mo. Prices track the octomind version we run.
Embeddings route through the hub too — Voyage, Jina and OpenAI models from $0.02/1M tokens, billed from prepaid credits on any plan. Local CPU embeddings stay free on every plan, hub or not.
Fair use is per account, not per key: 1 model call at a time on Free, 2 in parallel on Pro, 5 on Max, 15 pooled on Team — past that, requests queue briefly and then 429.
Prefer to self-host?
The gateway is octohub — open source, Rust, one binary. Run your own with your own provider keys; the hosted hub is the same code with the keys problem solved.
One subscription. Models and machines.
Hub isn't a separate bill — the same plan powers hosted models here and full cloud machines. Use it straight from the panel, or drop the key into your own CLI as a provider. $10 your first month, then $20/mo. Free tier stays free.