Available Models
Pay with USDC on Base. No account needed.
GPT-5.6 is live — three tiers: Sol, Terra, and Luna.
Chat: OpenAI openai/gpt-5.6-sol — flagship tier, deepest reasoning for complex coding and long-horizon agentic work, at $5 in / $30 out per 1M. openai/gpt-5.6-terra covers everyday coding and agentic tasks at $2 in / $12 out, and openai/gpt-5.6-luna handles high-volume, latency-sensitive work at $0.20 in / $1.20 out — all with 1M context, 128K output, and vision. Anthropic anthropic/claude-sonnet-5 — near-Opus coding and agentic quality at Sonnet pricing ($3 in / $15 out) — and the new flagship anthropic/claude-opus-5, a step-change over Opus 4.8 at the same price ($5 in / $25 out), round out the lineup. Video: OpenAI sora-2 (via Azure) at $0.10/sec — 720p with synced audio, 4/8/12s clips — alongside ByteDance Seedance (seedance-1.5-pro ≈ $0.493/5s · seedance-2.0-fast ≈ $1.276/5s · seedance-2.0 ≈ $1.595/5s) and xAI Grok Imagine at $0.05/sec. For consistent characters across clips, enroll a Seedance Virtual Portrait ($0.011, AI character) or Seedance RealFace ($0.011, real person + 1-min on-phone liveness check, no KYC).
Generated assets are mirrored to BlockRun's storage so URLs don't expire. Responses return a permanent url plus the upstream source_url. See the Image Generation and Video Generation docs.
92 models
GPT-5.6 Sol
openaiOpenAI flagship tier — deepest reasoning for complex coding, agentic workflows, and long-horizon problems. 1M context
Price: $5.00/M in · $30.00/M outGPT-5.6 Terra
openaiBalanced GPT-5.6 tier — everyday coding, reasoning, and agentic tasks at half the flagship price. 1M context
Price: $2.00/M in · $12.00/M outGPT-5.6 Luna
openaiCost-efficient GPT-5.6 tier for high-volume, latency-sensitive chat and lightweight agentic workflows. 1M context
Price: $0.20/M in · $1.20/M outGPT-5.6 Sol Pro
openaiHighest-capability GPT-5.6 — Sol with pro reasoning mode for the hardest problems and long-running agentic work. 1M context
Price: $5.00/M in · $30.00/M outGPT-5.6 Terra Pro
openaiGPT-5.6 Terra with pro reasoning mode — deeper responses on complex tasks at half the standard Terra rate. 1M context
Price: $1.00/M in · $6.00/M outGPT-5.6 Luna Pro
openaiGPT-5.6 Luna with pro reasoning mode — budget tier with deeper reasoning for high-volume workloads. 1M context
Price: $0.10/M in · $0.60/M outGPT-5.5
openaiFirst fully retrained base since GPT-4.5. 1M context, 128K output, native agent + computer use
Price: $5.00/M in · $30.00/M outGPT-5.5 Pro
openaiPremium GPT-5.5 with maximum compute for the hardest problems
Price: $30.00/M in · $180.00/M outChatGPT Instant (GPT-5.5)
openaiChatGPT's default model — the rolling `chat-latest` alias, currently GPT-5.5 Instant. Tuned for speed and concision, same price as GPT-5.5
Price: $5.00/M in · $30.00/M outGPT-5.4
openaiMost capable and efficient frontier model with 1M context, native computer use, and thinking mode
Price: $2.50/M in · $15.00/M outGPT-5.4 Pro
openaiPremium GPT-5.4 with maximum compute for the hardest problems
Price: $30.00/M in · $180.00/M outGPT-5.3
openaiHigh intelligence with medium speed. Multimodal with vision, function calling, and structured outputs
Price: $1.75/M in · $14.00/M outGPT-5.2
openaiFrontier model with 400K context and adaptive reasoning
Price: $1.75/M in · $14.00/M outGPT-5.4 Mini
openaiStrongest mini model for coding, computer use, and subagents with GPT-5.4 capabilities
Price: $0.75/M in · $4.50/M outGPT-5 Mini
openaiCost-optimized reasoning and chat
Price: $0.25/M in · $2.00/M outGPT-5.4 Nano
openaiFastest and most affordable GPT-5.4 model for high-throughput tasks
Price: $0.20/M in · $1.25/M outGPT-5.2 Pro
openaiUses more compute for consistently better answers
Price: $21.00/M in · $168.00/M outGPT-5.3 Codex
openaiIndustry-leading agentic coding model. 400K context, reasoning, tool use, and complex execution
Price: $1.75/M in · $14.00/M outGPT-4.1
openaiLatest GPT-4 generation model
Price: $2.00/M in · $8.00/M outGPT-4.1 Mini
openaiFast and affordable GPT-4.1 model
Price: $0.40/M in · $1.60/M outGPT-4.1 Nano
openaiUltra-fast and cost-effective GPT-4.1
Price: $0.10/M in · $0.40/M outGPT-4o
openaiMultimodal model with vision and audio
Price: $2.50/M in · $10.00/M outGPT-4o Mini
openaiFast and affordable GPT-4o model
Price: $0.15/M in · $0.60/M outo1
openaiAdvanced reasoning model for complex tasks
Price: $15.00/M in · $60.00/M outo3
openaiLatest reasoning model with improved performance
Price: $2.00/M in · $8.00/M outo3-mini
openaiEfficient reasoning model for STEM tasks
Price: $1.10/M in · $4.40/M outo4-mini
openaiLatest generation efficient reasoning model
Price: $1.10/M in · $4.40/M outGPT-OSS 20B
openaiTestnetOpen-weight 20B model (Apache 2.0), similar performance to o3-mini. Available on testnet for developer testing.
Price: $0.0020/requestGPT-OSS 120B
openaiTestnetOpen-weight 120B model (Apache 2.0), flagship open model. Available on testnet for developer testing.
Price: $0.0030/requestClaude Haiku 4.5
anthropicFastest and most efficient Claude, near-frontier intelligence
Price: $1.00/M in · $5.00/M outClaude Sonnet 5
anthropicNewest Sonnet — near-Opus coding/agentic quality at Sonnet cost. 1M context, 128k output, adaptive thinking, vision
Price: $3.00/M in · $15.00/M outClaude Sonnet 4.6
anthropicBest balance of intelligence, speed, and cost
Price: $3.00/M in · $15.00/M outClaude Sonnet 4.5
anthropicSonnet 4.5 — strong coding and agentic performance, vision
Price: $3.00/M in · $15.00/M outClaude Opus 4.5
anthropicLatest Anthropic flagship with enhanced reasoning and creativity
Price: $5.00/M in · $25.00/M outClaude Opus 4.7
anthropicPowerful Claude Opus for complex reasoning and agentic coding. 1M context, 128k output, adaptive thinking
Price: $5.00/M in · $25.00/M outClaude Fable 5
anthropicAnthropic's most capable model — Mythos-class tier above Opus, for the most demanding reasoning and long-horizon agentic work. 1M context, 128K output, always-on thinking
Price: $10.00/M in · $50.00/M outClaude Opus 4.8
anthropicMost capable Claude 4-series Opus for complex reasoning and agentic coding. 1M context, 128k output, adaptive thinking
Price: $5.00/M in · $25.00/M outClaude Opus 5
anthropicNewest Opus — step-change over Opus 4.8 for deep reasoning and agentic coding at the same price. 1M context, 128k output, adaptive thinking
Price: $5.00/M in · $25.00/M outGemini 3.1 Pro
googleLatest Gemini with improved thinking, token efficiency, and agentic capabilities. Optimized for software engineering (requires new SDK)
Price: $2.00/M in · $12.00/M outGemini 3 Flash Preview
googleFrontier-class performance with Pro-level intelligence at Flash speed and pricing. Includes thinking mode (requires new SDK)
Price: $0.50/M in · $3.00/M outGemini 3.6 Flash
googleNewest-generation Flash with built-in thinking mode — frontier-class quality at Flash speed
Price: $1.50/M in · $7.50/M outGemini 3.5 Flash
googleLatest-generation Flash with built-in thinking mode — frontier-class quality at Flash speed
Price: $1.50/M in · $9.00/M outGemini 2.5 Pro
googleState-of-the-art for reasoning, coding, and mathematics
Price: $1.25/M in · $10.00/M outGemini 2.5 Flash
googleFast and efficient Gemini model with vision support
Price: $0.30/M in · $2.50/M outGemini 3.5 Flash Lite
googleLatest Flash Lite — ultra-fast, lightweight Gemini with thinking mode for high-throughput tasks
Price: $0.30/M in · $2.50/M outGemini 3.1 Flash Lite
googleUltra-fast and lightweight Gemini 3.1 model with thinking mode for high-throughput tasks
Price: $0.25/M in · $1.50/M outGemini 2.5 Flash Lite
googleMost economical Gemini model - ultra-fast and lightweight (requires new SDK)
Price: $0.10/M in · $0.40/M outDeepSeek V4 Pro
deepseekDeepSeek V4 flagship — 1.6T MoE / 49B active, 1M context. Strongest open-weight reasoner. Thinking mode default.
Price: $0.43/M in · $0.87/M outDeepSeek V4 Flash Chat
deepseekPaid V4 Flash in non-thinking mode (1.6T-class quality at $0.20 in / $0.40 out). Same model as the free nvidia/deepseek-v4-flash but on a paid endpoint with higher reliability and 5MB request bodies.
Price: $0.20/M in · $0.40/M outDeepSeek V4 Flash Reasoner
deepseekPaid V4 Flash in thinking mode for reasoning tasks. Same upstream as deepseek/deepseek-chat but with thinking enabled by default.
Price: $0.20/M in · $0.40/M outKimi K3
moonshotMoonshot's flagship — a 2.8-trillion-parameter open MoE with 1M context, image + text input, returning reasoning_content. Live-verified 2026-07-17 (chat, tools, vision).
Price: $3.00/M in · $15.00/M outGLM-5.2
zaiZ.AI's newest flagship — 1M-token context, top open-source on long-horizon coding. Verified live on Z.AI.
Price: $1.40/M in · $4.40/M outGLM-5.1
zaiZ.AI flagship — #1 open source on SWE-Bench Pro, 8-hour autonomous execution. 200K context
Price: $1.40/M in · $4.40/M outGLM-5
zaiZ.AI's foundation model with 200K context. Strong reasoning and agentic capabilities
Price: $0.60/M in · $1.92/M outGLM-5 Turbo
zaiOptimized GLM-5 variant with faster inference
Price: $1.20/M in · $4.00/M outGrok 4.3
xaixAI's Grok 4.3 reasoning model. 1M context, vision-capable, tuned for agentic workflows and instruction-following.
Price: $1.50/M in · $4.00/M outGrok Build 0.1
xaixAI's fast agentic coding model, trained for interactive software-engineering workflows. 256K context, text + image input.
Price: $1.50/M in · $3.00/M outGrok 4.5
xaixAI's flagship Grok 4.5 — their most intelligent and fastest model. 500K context, vision-capable, chain-of-thought reasoning. Supports Live Search (+$0.025/source)
Price: $2.50/M in · $9.00/M outMiniMax M2.7
minimaxMiniMax's flagship reasoning model with recursive self-improvement. Great value for complex tasks (~60 tps)
Price: $0.30/M in · $1.20/M outMiniMax M3
minimaxMiniMax's M3 flagship — 1M context, strong reasoning + coding.
Price: $0.30/M in · $1.20/M outQwen3.7 Max
qwenAlibaba's Qwen flagship — the Max tier. 1M context, strong reasoning, coding, and agentic tool use. Live-verified 2026-07-20.
Price: $1.48/M in · $4.42/M outQwen3.7 Plus
qwenAlibaba's balanced Qwen tier — 1M context with reasoning, coding, and agentic tool use at a fraction of the Max price
Price: $0.32/M in · $1.28/M outQwen3.7 Flash
qwenAlibaba's fastest Qwen tier — 1M context reasoning for high-volume, latency-sensitive workloads
Price: $0.03/M in · $0.13/M outTencent Hy3
tencentTencent's Hy3 — fast, inexpensive reasoning at 262K context. One of the most-used open models of 2026.
Price: $0.13/M in · $0.53/M outXiaomi MiMo-V2.5 Pro
xiaomiXiaomi's MiMo-V2.5 Pro — 1M context reasoning model, priced well below the frontier tier.
Price: $0.43/M in · $0.87/M outDeepSeek V4 Flash (Free)
nvidiaFreeDeepSeek V4 Flash hosted free by NVIDIA. 1M context. Currently capacity-constrained — requests may be answered by an equivalent free model.
Price: Free/M in · Free/M outNemotron 3 Nano Omni (Free)
nvidiaFreeNVIDIA's multimodal reasoning Nemotron Nano Omni hosted free by NVIDIA. 31B / 3.2B active MoE. Accepts text, images, video, audio. ChartQA 90.3, DocVQA 95.6, MMMU 70.8 — the only vision-capable free model in our catalog
Price: Free/M in · Free/M outMistral Nemotron (Free)
nvidiaFreeMistral × NVIDIA Nemotron instruction model hosted free by NVIDIA. Fast (~0.2s), strong instruction following.
Price: Free/M in · Free/M outStepFun Step 3.7 Flash (Free)
nvidiaFreeStepFun Step 3.7 Flash hosted free by NVIDIA. Fast lightweight reasoning, 131K context.
Price: Free/M in · Free/M outNemotron Nano 9B v2 (Free)
nvidiaFreeNVIDIA Nemotron Nano 9B v2 hosted free by NVIDIA. Compact + fast (~0.7s), good for high-volume light tasks.
Price: Free/M in · Free/M outNemotron Nano 12B v2 VL (Free)
nvidiaFreeNVIDIA Nemotron Nano 12B v2 Vision-Language hosted free by NVIDIA. Accepts images; compact + fast.
Price: Free/M in · Free/M outGPT Image 1
openaiImageNative image generation in GPT-4o
Price: $0.020/imageChatGPT Images 2.0
openaiImageNewOpenAI's GPT Image 2 — reasoning-driven image generation with multilingual text rendering, character consistency, and high-fidelity edits
Price: $0.060/imageNano Banana
googleImageGoogle's Gemini 2.5 Flash image generation - fast and efficient
Price: $0.050/imageNano Banana 2
googleImageGoogle's Gemini 3.1 Flash image generation - pro-level quality at Flash speed
Price: $0.090/imageNano Banana Pro
googleImageGoogle's Gemini 3 Pro image generation - highest quality up to 4K
Price: $0.100/imageGrok Imagine
xaiImageNewxAI's Grok Imagine image generation. Fast, 300 RPM.
Price: $0.020/imageGrok Imagine Pro
xaiImageNewxAI's premium Grok Imagine image generation (quality tier). Higher quality, 30 RPM.
Price: $0.070/imageSeedream 5.0 Pro
bytedanceImageByteDance's Seedream 5.0 Pro — flagship image generation and editing, up to 4K-class resolution with reference-image support
Price: $0.045/imageCogView-4
zaiImageZhipu AI's CogView-4 image generation model — high quality, supports up to 1440x1440
Price: $0.015/imageMiniMax Music 2.5+
minimaxMusicMiniMax's flagship music generation model. Supports lyrics, instrumental, and style prompts. ~3 min output.
Price: $0.150/trackElevenLabs Flash v2.5
elevenlabsSpeechUltra-low-latency (~75ms) speech synthesis for real-time voice agents. 32 languages.
Price: $0.05/1k charsElevenLabs Turbo v2.5
elevenlabsSpeechBalanced quality and latency (~250ms) for interactive use cases. 32 languages.
Price: $0.05/1k charsElevenLabs Multilingual v2
elevenlabsSpeechHighest-consistency voice for long-form narration, audiobooks, and voiceover. 29 languages.
Price: $0.10/1k charsElevenLabs v3
elevenlabsSpeechMaximum expressiveness and emotional range for creative applications. 70+ languages.
Price: $0.10/1k charsSeed Audio 1.0
bytedanceSpeechByteDance's Seed Audio 1.0 — prompt-directed audio creation: describe the voice, emotion, and sound staging in natural language. Up to 120s output, mp3/wav. Billed by audio duration ($0.003/second, estimated from input length).
Price: $0.30/1k charsElevenLabs Sound Effects
elevenlabsSpeechGenerate cinematic sound effects and audio textures from a text prompt (up to 22s).
Price: $0.050/clipGrok Imagine Video
xaiVideoNewxAI's Grok Imagine video generation. Text or image to video, configurable 1–15s clips at $0.05/sec.
Price: $0.050/secSeedance 1.5 Pro
bytedanceVideoNewByteDance Seedance 1.5 Pro — budget text/image-to-video at 720p with synced audio (t2v), 5s default. Does NOT support RealFace assets.
Price: $0.098/secSeedance 2.0 Fast
bytedanceVideoNewByteDance Seedance 2.0 Fast — fast video at 720p with synced audio (t2v), 5s default. ~60-80s to generate. Supports BytePlus RealFace assets — see /docs/video/real-person-ip for enrollment.
Price: $0.255/secSeedance 2.0 Pro
bytedanceVideoNewByteDance Seedance 2.0 Pro — premium quality text/image-to-video at 720p with synced audio (t2v), 5s default. Supports BytePlus RealFace assets — see /docs/video/real-person-ip for enrollment.
Price: $0.319/secSora 2
azureVideoOpenAI Sora 2 via Azure AI Foundry — text-to-video AND image-to-video at 720p with synchronized audio. 4s default; 4, 8, or 12s. Portrait or landscape. Image-to-video takes a non-human reference image (human faces are rejected upstream — use Seedance + RealFace for real people). $0.10/sec.
Price: $0.100/sec
How pricing works on BlockRun
Every model above is billed per request, in USDC, settled on-chain through the x402 protocol. There is no subscription, no monthly minimum, and no API key to provision — a request arrives unpaid, the gateway answers with a signed price quote, your client signs it, and the retry returns the completion. Chat models are priced per million input and output tokens; image, video, music and speech models are priced per image, per second, per track, and per thousand characters respectively.
Token prices track each upstream lab's public list price, plus a flat platform margin and a small per-transaction fee that covers on-chain settlement. Because you pay per call, an agent that makes ten requests a day costs cents a month, and one that makes ten thousand pays exactly ten thousand times the per-call price — the unit economics do not change with volume, and nothing expires unused.
Choosing a model
The catalog spans 71 chat and reasoning models alongside image, video, music and speech generation, for 92 in total. Use the Reasoning filter for long-horizon planning and hard analytical work, Coding for repository-scale edits and agentic tool use, and Vision when requests carry images. The Free filter lists models that need no payment header at all — useful for prototyping an integration before you fund a wallet.
Every model is reachable through the same OpenAI-compatible /v1/chat/completions endpoint, so switching between them is a one-line change to the model field. See the documentation for request shapes and the pricing page for a full per-model rate card.