BlockRun
Get started

Mainnet Models

Base

Production models on Base mainnet. Pay with real USDC.

69 models

GPT-5.6 Sol

openai

OpenAI flagship tier — deepest reasoning for complex coding, agentic workflows, and long-horizon problems. 1M context

Input: $5.00/MOutput: $30.00/M

GPT-5.6 Terra

openai

Balanced GPT-5.6 tier — everyday coding, reasoning, and agentic tasks at half the flagship price. 1M context

Input: $2.00/MOutput: $12.00/M

GPT-5.6 Luna

openai

Cost-efficient GPT-5.6 tier for high-volume, latency-sensitive chat and lightweight agentic workflows. 1M context

Input: $0.20/MOutput: $1.20/M

GPT-5.6 Sol Pro

openai

Highest-capability GPT-5.6 — Sol with pro reasoning mode for the hardest problems and long-running agentic work. 1M context

Input: $5.00/MOutput: $30.00/M

GPT-5.6 Terra Pro

openai

GPT-5.6 Terra with pro reasoning mode — deeper responses on complex tasks at half the standard Terra rate. 1M context

Input: $1.00/MOutput: $6.00/M

GPT-5.6 Luna Pro

openai

GPT-5.6 Luna with pro reasoning mode — budget tier with deeper reasoning for high-volume workloads. 1M context

Input: $0.10/MOutput: $0.60/M

GPT-5.5

openai

First fully retrained base since GPT-4.5. 1M context, 128K output, native agent + computer use

Input: $5.00/MOutput: $30.00/M

GPT-5.5 Pro

openai

Premium GPT-5.5 with maximum compute for the hardest problems

Input: $30.00/MOutput: $180.00/M

ChatGPT Instant (GPT-5.5)

openai

ChatGPT's default model — the rolling `chat-latest` alias, currently GPT-5.5 Instant. Tuned for speed and concision, same price as GPT-5.5

Input: $5.00/MOutput: $30.00/M

GPT-5.4

openai

Most capable and efficient frontier model with 1M context, native computer use, and thinking mode

Input: $2.50/MOutput: $15.00/M

GPT-5.4 Pro

openai

Premium GPT-5.4 with maximum compute for the hardest problems

Input: $30.00/MOutput: $180.00/M

GPT-5.3

openai

High intelligence with medium speed. Multimodal with vision, function calling, and structured outputs

Input: $1.75/MOutput: $14.00/M

GPT-5.2

openai

Frontier model with 400K context and adaptive reasoning

Input: $1.75/MOutput: $14.00/M

GPT-5.4 Mini

openai

Strongest mini model for coding, computer use, and subagents with GPT-5.4 capabilities

Input: $0.75/MOutput: $4.50/M

GPT-5 Mini

openai

Cost-optimized reasoning and chat

Input: $0.25/MOutput: $2.00/M

GPT-5.4 Nano

openai

Fastest and most affordable GPT-5.4 model for high-throughput tasks

Input: $0.20/MOutput: $1.25/M

GPT-5.2 Pro

openai

Uses more compute for consistently better answers

Input: $21.00/MOutput: $168.00/M

GPT-5.3 Codex

openai

Industry-leading agentic coding model. 400K context, reasoning, tool use, and complex execution

Input: $1.75/MOutput: $14.00/M

GPT-4.1

openai

Latest GPT-4 generation model

Input: $2.00/MOutput: $8.00/M

GPT-4.1 Mini

openai

Fast and affordable GPT-4.1 model

Input: $0.40/MOutput: $1.60/M

GPT-4.1 Nano

openai

Ultra-fast and cost-effective GPT-4.1

Input: $0.10/MOutput: $0.40/M

GPT-4o

openai

Multimodal model with vision and audio

Input: $2.50/MOutput: $10.00/M

GPT-4o Mini

openai

Fast and affordable GPT-4o model

Input: $0.15/MOutput: $0.60/M

o1

openai

Advanced reasoning model for complex tasks

Input: $15.00/MOutput: $60.00/M

o3

openai

Latest reasoning model with improved performance

Input: $2.00/MOutput: $8.00/M

o3-mini

openai

Efficient reasoning model for STEM tasks

Input: $1.10/MOutput: $4.40/M

o4-mini

openai

Latest generation efficient reasoning model

Input: $1.10/MOutput: $4.40/M

Claude Haiku 4.5

anthropic

Fastest and most efficient Claude, near-frontier intelligence

Input: $1.00/MOutput: $5.00/M

Claude Sonnet 5

anthropic

Newest Sonnet — near-Opus coding/agentic quality at Sonnet cost. 1M context, 128k output, adaptive thinking, vision

Input: $3.00/MOutput: $15.00/M

Claude Sonnet 4.6

anthropic

Best balance of intelligence, speed, and cost

Input: $3.00/MOutput: $15.00/M

Claude Sonnet 4.5

anthropic

Sonnet 4.5 — strong coding and agentic performance, vision

Input: $3.00/MOutput: $15.00/M

Claude Opus 4.5

anthropic

Latest Anthropic flagship with enhanced reasoning and creativity

Input: $5.00/MOutput: $25.00/M

Claude Opus 4.7

anthropic

Powerful Claude Opus for complex reasoning and agentic coding. 1M context, 128k output, adaptive thinking

Input: $5.00/MOutput: $25.00/M

Claude Fable 5

anthropic

Anthropic's most capable model — Mythos-class tier above Opus, for the most demanding reasoning and long-horizon agentic work. 1M context, 128K output, always-on thinking

Input: $10.00/MOutput: $50.00/M

Claude Opus 4.8

anthropic

Most capable Claude 4-series Opus for complex reasoning and agentic coding. 1M context, 128k output, adaptive thinking

Input: $5.00/MOutput: $25.00/M

Claude Opus 5

anthropic

Newest Opus — step-change over Opus 4.8 for deep reasoning and agentic coding at the same price. 1M context, 128k output, adaptive thinking

Input: $5.00/MOutput: $25.00/M

Gemini 3.1 Pro

google

Latest Gemini with improved thinking, token efficiency, and agentic capabilities. Optimized for software engineering (requires new SDK)

Input: $2.00/MOutput: $12.00/M

Gemini 3 Flash Preview

google

Frontier-class performance with Pro-level intelligence at Flash speed and pricing. Includes thinking mode (requires new SDK)

Input: $0.50/MOutput: $3.00/M

Gemini 3.6 Flash

google

Newest-generation Flash with built-in thinking mode — frontier-class quality at Flash speed

Input: $1.50/MOutput: $7.50/M

Gemini 3.5 Flash

google

Latest-generation Flash with built-in thinking mode — frontier-class quality at Flash speed

Input: $1.50/MOutput: $9.00/M

Gemini 2.5 Pro

google

State-of-the-art for reasoning, coding, and mathematics

Input: $1.25/MOutput: $10.00/M

Gemini 2.5 Flash

google

Fast and efficient Gemini model with vision support

Input: $0.30/MOutput: $2.50/M

Gemini 3.5 Flash Lite

google

Latest Flash Lite — ultra-fast, lightweight Gemini with thinking mode for high-throughput tasks

Input: $0.30/MOutput: $2.50/M

Gemini 3.1 Flash Lite

google

Ultra-fast and lightweight Gemini 3.1 model with thinking mode for high-throughput tasks

Input: $0.25/MOutput: $1.50/M

Gemini 2.5 Flash Lite

google

Most economical Gemini model - ultra-fast and lightweight (requires new SDK)

Input: $0.10/MOutput: $0.40/M

DeepSeek V4 Pro

deepseek

DeepSeek V4 flagship — 1.6T MoE / 49B active, 1M context. Strongest open-weight reasoner. Thinking mode default.

Input: $0.43/MOutput: $0.87/M

DeepSeek V4 Flash Chat

deepseek

Paid V4 Flash in non-thinking mode (1.6T-class quality at $0.14 in / $0.28 out). Production-grade reliability and 5MB request bodies.

Input: $0.14/MOutput: $0.28/M

DeepSeek V4 Flash Reasoner

deepseek

Paid V4 Flash in thinking mode for reasoning tasks. Same upstream as deepseek/deepseek-chat but with thinking enabled by default.

Input: $0.14/MOutput: $0.28/M

Kimi K3

moonshot

Moonshot's flagship — a 2.8-trillion-parameter open MoE with 1M context, image + text input, returning reasoning_content. Live-verified 2026-07-17 (chat, tools, vision).

Input: $3.00/MOutput: $15.00/M

GLM-5.3

zai

Z.AI's flagship — 1M-token context with always-on reasoning, strong at long-horizon coding. Verified live on Z.AI.

Input: $1.40/MOutput: $4.40/M

GLM-5.2

zai

Z.AI GLM-5.2 — 1M-token context, strong open-source long-horizon coding. Verified live on Z.AI.

Input: $1.40/MOutput: $4.40/M

GLM-5.1

zai

Z.AI flagship — #1 open source on SWE-Bench Pro, 8-hour autonomous execution. 200K context

Input: $1.40/MOutput: $4.40/M

GLM-5

zai

Z.AI's foundation model with 200K context. Strong reasoning and agentic capabilities

Input: $1.00/MOutput: $3.20/M

GLM-5 Turbo

zai

Optimized GLM-5 variant with faster inference

Input: $1.20/MOutput: $4.00/M

Grok 4.3

xai

xAI's Grok 4.3 reasoning model. 1M context, vision-capable, tuned for agentic workflows and instruction-following.

Input: $1.50/MOutput: $4.00/M

Grok Build 0.1

xai

xAI's fast agentic coding model, trained for interactive software-engineering workflows. 256K context, text + image input.

Input: $1.50/MOutput: $3.00/M

Grok 4.5

xai

xAI's flagship Grok 4.5 — their most intelligent and fastest model. 500K context, vision-capable, chain-of-thought reasoning. Supports Live Search (+$0.025/source)

Input: $2.50/MOutput: $9.00/M

MiniMax M2.7

minimax

MiniMax's flagship reasoning model with recursive self-improvement. Great value for complex tasks (~60 tps)

Input: $0.30/MOutput: $1.20/M

MiniMax M3

minimax

MiniMax's M3 flagship — 1M context, strong reasoning + coding.

Input: $0.30/MOutput: $1.20/M

Qwen3.7 Max

qwen

Alibaba's Qwen flagship — the Max tier. 1M context, strong reasoning, coding, and agentic tool use. Live-verified 2026-07-20.

Input: $1.48/MOutput: $4.42/M

Qwen3.7 Plus

qwen

Alibaba's balanced Qwen tier — 1M context with reasoning, coding, and agentic tool use at a fraction of the Max price

Input: $0.32/MOutput: $1.28/M

Qwen3.7 Flash

qwen

Alibaba's fastest Qwen tier — 1M context reasoning for high-volume, latency-sensitive workloads

Input: $0.03/MOutput: $0.13/M

Tencent Hy3

tencent

Tencent's Hy3 — fast, inexpensive reasoning at 262K context. One of the most-used open models of 2026.

Input: $0.13/MOutput: $0.53/M

Xiaomi MiMo-V2.5 Pro

xiaomi

Xiaomi's MiMo-V2.5 Pro — 1M context reasoning model, priced well below the frontier tier.

Input: $0.43/MOutput: $0.87/M

Nemotron 3 Nano Omni (Free)

nvidiaFree

NVIDIA's multimodal reasoning Nemotron Nano Omni hosted free by NVIDIA. 31B / 3.2B active MoE. Accepts text, images, video, audio. ChartQA 90.3, DocVQA 95.6, MMMU 70.8 — the only vision-capable free model in our catalog

Input: Free/MOutput: Free/M

Mistral Nemotron (Free)

nvidiaFree

Mistral × NVIDIA Nemotron instruction model hosted free by NVIDIA. Fast (~0.2s), strong instruction following.

Input: Free/MOutput: Free/M

StepFun Step 3.7 Flash (Free)

nvidiaFree

StepFun Step 3.7 Flash hosted free by NVIDIA. Fast lightweight reasoning, 131K context.

Input: Free/MOutput: Free/M

Nemotron Nano 9B v2 (Free)

nvidiaFree

NVIDIA Nemotron Nano 9B v2 hosted free by NVIDIA. Compact + fast (~0.7s), good for high-volume light tasks.

Input: Free/MOutput: Free/M

Nemotron Nano 12B v2 VL (Free)

nvidiaFree

NVIDIA Nemotron Nano 12B v2 Vision-Language hosted free by NVIDIA. Accepts images; compact + fast.

Input: Free/MOutput: Free/M