BlockRun
Get started

Intelligence Pricing

Pay only for what you use. Per-token chat is billed at provider cost — no platform margin.

Pricing Formula

Your cost = Provider cost          (per-token chat — no margin, matches OpenRouter)
              + $0.001 / request     (flat x402 transaction fee)

Media generation and Live Search still carry a 5% platform margin, which covers:

  • x402 settlement infrastructure
  • Smart routing and reliability
  • No API key management
  • Instant on-chain payments

What $1 Gets You

ModelApproximate Usage
GPT-5.5~200K input tokens
DeepSeek V4 Flash Chat~5M input tokens
Gemini 3.5 Flash~635K input tokens
Image generation~10–65 images
Free tier (8 models — reasoning, coding, and vision)Unlimited (FREE)
Start with the free tier

The free tier costs $0 — 10 reasoning, coding, and vision models with no per-token charge. You still need a funded wallet for the x402 handshake, but these calls don't draw it down.

Full Price List

OpenAI

ModelInput (per 1M)Output (per 1M)
GPT-5.6 Sol (flagship)$5.00$30.00
GPT-5.6 Sol Pro$5.00$30.00
GPT-5.6 Terra$2.00$12.00
GPT-5.6 Terra Pro$1.00$6.00
GPT-5.6 Luna$0.20$1.20
GPT-5.6 Luna Pro$0.10$0.60
GPT-5.5$5.00$30.00
GPT-5.5 Pro$30.00$180.00
GPT-5.4$2.50$15.00
GPT-5.4 Pro$30.00$180.00
GPT-5.2$1.75$14.00

Anthropic

ModelInput (per 1M)Output (per 1M)
Claude Fable 5 (most capable)$10.00$50.00
Claude Opus 5 (flagship)$5.00$25.00
Claude Opus 4.8 (previous flagship)$5.00$25.00
Claude Opus 4.7$5.00$25.00
Claude Opus 4.5$5.00$25.00
Claude Sonnet 5$3.00$15.00
Claude Sonnet 4.6$3.00$15.00
Claude Haiku 4.5$1.00$5.00

Google

ModelInput (per 1M)Output (per 1M)
Gemini 3.1 Pro$2.00$12.00
Gemini 3.6 Flash$1.50$7.50
Gemini 3.5 Flash$1.50$9.00
Gemini 3.5 Flash Lite$0.30$2.50

Gemini Pro models double the input rate and add 50% to the output rate above 200K prompt tokens (the whole request reprices), mirroring Google's official long-context pricing — e.g. Gemini 2.5 Pro is $2.63 in · $15.75 out above the threshold. Flash tiers are flat.

xAI Grok

ModelInput (per 1M)Output (per 1M)Context
Grok 4.5 (flagship)$2.50$9.00500K
Grok 4.3$1.50$4.001M
Grok Build 0.1$1.50$3.00256K

Grok doubles the per-token rates above 200K prompt tokens (the whole request reprices — e.g. Grok 4.5 is $5.25 in · $18.90 out above the threshold), mirroring xAI's official long-context tier. Live Search adds $0.025 per source used.

Z.AI

ModelInput (per 1M)Output (per 1M)Context
GLM-5.2 (flagship)$1.40$4.401M
GLM-5.1$1.40$4.40200K
GLM-5$1.00$3.20200K
GLM-5 Turbo$1.20$4.00200K

Moonshot

ModelInput (per 1M)Output (per 1M)Context
Kimi K3$3.00$15.001M
Kimi K2.7$0.95$4.00256K

MiniMax

ModelInput (per 1M)Output (per 1M)
MiniMax M3$0.30$1.20

Qwen

ModelInput (per 1M)Output (per 1M)
Qwen3.7 Max$1.48$4.42
Qwen3.7 Plus$0.32$1.28
Qwen3.7 Flash$0.03$0.13

DeepSeek

ModelInput (per 1M)Output (per 1M)
DeepSeek V4 Flash Chat$0.14$0.28
DeepSeek V4 Pro$0.43$0.87

Free Tier (8 models)

The free tier is 10 reasoning, coding, and vision models with no per-token charge. The lineup is kept current by a self-healing health gate that routes around any model whose upstream is temporarily unavailable and auto-recovers it, so existing calls keep working. All free-tier models are FREE for both input and output — call GET /api/v1/models for the current live list.

Image Generation

ModelPrice per Image
GPT Image 1 (1024x1024)$0.02
GPT Image 1 (wide/tall)$0.04
ChatGPT Images 2.0 (1024x1024)$0.06
ChatGPT Images 2.0 (wide/tall)$0.12
Nano Banana$0.05
Nano Banana Pro (up to 2048²)$0.10
Nano Banana Pro (4K)$0.15
CogView-4$0.015
Grok Imagine$0.02
Grok Imagine Pro$0.07

Other media: video from $0.05/sec, music $0.15/track, text-to-speech $0.05–$0.10 per 1k characters, sound effects $0.0535/generation.

Cost Comparison: BlockRun vs Direct

ProviderDirect PricingBlockRunDifference
OpenAI GPT-5.4$2.50/$15.00$2.63/$15.75+5%
Anthropic Claude Sonnet 5$3.00/$15.00$3.15/$15.75+5%
Anthropic Claude Sonnet 4.6$3.00/$15.00$3.15/$15.75+5%
DeepSeek V4 Flash Chat$0.14/$0.28$0.14/$0.28+0%

You pay 5% more, but you get:

  • No API key management
  • No monthly invoices
  • No prepaid credits
  • One wallet for all providers
  • Instant per-request settlement

Budget Management

Session Budgets

from blockrun_llm import LLMClient

# Limit spending per session
client = LLMClient(session_budget=5.00)

Check Balance

balance = client.get_balance()
print(f"${balance} USDC remaining")

Track Spending

# Get usage stats
usage = client.get_usage()
print(f"Spent: ${usage['total_spent']}")
print(f"Requests: {usage['request_count']}")

Cost Optimization Tips

0. Use ClawRouter for Automatic Savings

Save 88% on average with ClawRouter — it automatically routes each request to the cheapest model that can handle it.

/model blockrun/auto

ClawRouter does all the optimization below automatically.

1. Use Cheaper Models for Routine Tasks

# Expensive
response = client.chat("openai/gpt-5.5", "Summarize this text")

# 50x cheaper, similar quality
response = client.chat("deepseek/deepseek-chat", "Summarize this text")

2. Use Flash Models for Speed

# For quick, simple tasks
response = client.chat("google/gemini-3.5-flash", prompt)

3. Match Model to Task

TaskRecommended ModelWhy
Bulk processingDeepSeekCheapest
Quick responsesGemini 3.5 FlashFast + cheap
Complex reasoningDeepSeek Reasoner, Claude Opus 5Best quality
Code generationGPT-5.4, Claude Sonnet 4.6Good balance
Real-time dataGrokWeb & news access

4. Optimize Prompts

Shorter prompts = fewer input tokens = lower cost.

What you don't pay

  • No subscriptions
  • No prepaid credits
  • No minimum top-up
  • No overage charges
  • No rate limit fees

The rates on this page are the providers' own published prices; our margin is added on top of them at settlement.

Payment Details

  • Currency: USDC on Base
  • Settlement: Instant, on-chain
  • Verification: Basescan

What's next?