Skip to content

Models & pricing

15 models, one balance

Prices are each lab's list price in USD per 1M tokens, verified on 2026-10-04. One credit buys $1 of compute at these prices; your tier's gateway fee is added on top.

Catalog

ModelContextMax outputInputCached inputOutputCapabilities

Claude Fable 5.1

anthropic/claude-fable-5.1

Anthropic · via Anthropic → Vercel AI Gateway

Not offered with zero data retention through the AI Gateway: served only via a direct Anthropic key.

1M128K$10.00$0.25$50.00tools, vision, reasoning, json, messages-api

Claude Opus 5.5

anthropic/claude-opus-5.5

Anthropic · via Anthropic → Vercel AI Gateway

1M128K$4.00$0.20$20.00tools, vision, reasoning, json, messages-api

Claude Sonnet 5.5

anthropic/claude-sonnet-5.5

Anthropic · via Anthropic → Vercel AI Gateway

1M128K$2.00$0.20$10.00tools, vision, reasoning, json, messages-api

Claude Haiku 4.5

anthropic/claude-haiku-4.5

Anthropic · via Anthropic → Vercel AI Gateway

200K64K$1.00$0.10$5.00tools, vision, reasoning, json, messages-api

GPT-6 Astra

openai/gpt-6-astra

OpenAI · via OpenAI → Vercel AI Gateway

1.05M128K$10.00$1.00$50.00tools, vision, reasoning, json

GPT-6.1 Sol

openai/gpt-6.1-sol

OpenAI · via OpenAI → Vercel AI Gateway

1.05M128K$2.00$0.10$10.00tools, vision, reasoning, json

GPT-6 Luna

openai/gpt-6-luna

OpenAI · via OpenAI → Vercel AI Gateway

1.05M128K$0.10$0.01$0.50tools, vision, json

Gemini 3.1 Pro (Preview) Preview

google/gemini-3.1-pro-preview

Google · via Google Gemini → Vercel AI Gateway

1M64K$2.00$0.20$12.00tools, vision, reasoning, json

Gemini 3.8 Flash

google/gemini-3.8-flash

Google · via Google Gemini → Vercel AI Gateway

1M66K$0.75$0.075$3.75tools, vision, reasoning, json

Gemini 3.5 Flash-Lite

google/gemini-3.5-flash-lite

Google · via Google Gemini → Vercel AI Gateway

1M65K$0.30$0.03$2.50tools, vision, json

DeepSeek V4 Pro Open

deepseek/deepseek-v4-pro

DeepSeek · via Vercel AI Gateway

1M384K$0.66$0.022$1.98tools, reasoning, json

DeepSeek V4 Flash Open

deepseek/deepseek-v4-flash

DeepSeek · via Vercel AI Gateway

1M384K$0.13$0.028$0.26tools, json

Llama 4 Maverick Open

meta/llama-4-maverick

Meta · via Vercel AI Gateway

128K8K$0.24—$0.97tools, vision, json

gpt-oss-120b Open

openai/gpt-oss-120b

OpenAI · via Vercel AI Gateway

131K131K$0.10$0.10$0.50tools, reasoning, json

Kimi K2.6 Open

moonshotai/kimi-k2.6

Moonshot AI · via Vercel AI Gateway

262K262K$0.95$0.16$4.00tools, json

Routes are tried in order; if a provider errors or its circuit breaker is open, the request fails over to the next route automatically. Aliases (each lab's native model id) resolve to the same model and price.

Long-context pricing

ModelApplies aboveInputOutput
Gemini 3.1 Pro (Preview)200K prompt tokens$4.00$18.00

Announced price changes

ModelFrom (UTC)InputOutput
Gemini 3.8 Flash2027-01-01$1.50$7.50

Gateway fee by tier

The gateway fee is a percentage of metered compute. It falls as your matured stake grows.

TierMatured stakeGateway fee
Standard≥ 0 tokens5.0%
Satellite≥ 10,000 tokens3.5%
Planet≥ 100,000 tokens2.0%
Star≥ 1,000,000 tokens1.0%

How a request is billed

  1. Before forwarding, the gateway reserves the most the request could cost: the prompt plus max_tokens (or the model's default output allowance), at list price plus your fee.
  2. If your balance can't cover that but can cover the prompt, the request still runs with a cap and stops cleanly with insufficient_credits if it reaches it.
  3. When the response completes, you are charged for the exact tokens used — uncached input, cached input, cache writes and output each at their own rate — rounded up to the micro-credit. The unused reservation is released immediately.
  4. Some open-weight models are priced per serving provider, and some providers charge regional or peak-hour rates. When the upstream reports that a request cost more than list price, compute is charged at that reported cost instead, so every credit stays backed by real compute.
  5. Spend is drawn from developer and referral credits first, then staker credits, then purchased credits.

Example

1,000 input + 500 output tokens on a model priced $3 / $15 per 1M costs 0.003 + 0.0075 = 0.0105 credits of compute. With the 5.0% Standard fee the total is 0.011025 credits.