Available Models
Pay with USDC on Base. No account needed.
GLM-5.3 Flash — a 1M-context model that reads images, at value-tier prices.
Chat: Z.AI zai/glm-5.3-flash is the first natively multimodal GLM-5 — 320B/18B MoE, image input, 1M context, always-on reasoning — at $0.15 in / $0.50 out per 1M. Its flagship sibling zai/glm-5.3 (text-only, 1M context) is $1.40 / $4.40. Both are quoted in dollars before the call runs, like everything here. Video: ByteDance seedance-2.5 — the longest Seedance clips (30s) with synced audio — and xAI grok-imagine-video-1.5at xAI's official resolution tiers, including a real 1080p. For consistent characters across clips, enroll a Seedance Virtual Portrait ($0.011, AI character) or Seedance RealFace ($0.011, real person + 1-min on-phone liveness check, no KYC).
Still recent: claude-opus-5, gpt-5.6-sol · terra · luna, claude-sonnet-5 and sora-2 — see the changelogfor dates and prices. Generated assets are mirrored to BlockRun's storage so URLs don't expire.
100 models
GPT-5.6 Sol
openaiOpenAI flagship tier — deepest reasoning for complex coding, agentic workflows, and long-horizon problems. 1M context
Price: $5.00/M in · $30.00/M outGPT-5.6 Terra
openaiBalanced GPT-5.6 tier — everyday coding, reasoning, and agentic tasks at half the flagship price. 1M context
Price: $2.00/M in · $12.00/M outGPT-5.6 Luna
openaiCost-efficient GPT-5.6 tier for high-volume, latency-sensitive chat and lightweight agentic workflows. 1M context
Price: $0.20/M in · $1.20/M outGPT-5.6 Sol Pro
openaiHighest-capability GPT-5.6 — Sol with pro reasoning mode for the hardest problems and long-running agentic work. 1M context
Price: $5.00/M in · $30.00/M outGPT-5.6 Terra Pro
openaiGPT-5.6 Terra with pro reasoning mode — deeper responses on complex tasks at half the standard Terra rate. 1M context
Price: $2.00/M in · $12.00/M outGPT-5.6 Luna Pro
openaiGPT-5.6 Luna with pro reasoning mode — budget tier with deeper reasoning for high-volume workloads. 1M context
Price: $0.20/M in · $1.20/M outGPT-5.5
openaiFirst fully retrained base since GPT-4.5. 1M context, 128K output, native agent + computer use
Price: $5.00/M in · $30.00/M outGPT-5.5 Pro
openaiPremium GPT-5.5 with maximum compute for the hardest problems
Price: $30.00/M in · $180.00/M outChatGPT Instant (GPT-5.5)
openaiChatGPT's default model — the rolling `chat-latest` alias, currently GPT-5.5 Instant. Tuned for speed and concision, same price as GPT-5.5
Price: $5.00/M in · $30.00/M outGPT-5.4
openaiMost capable and efficient frontier model with 1M context, native computer use, and thinking mode
Price: $2.50/M in · $15.00/M outGPT-5.4 Pro
openaiPremium GPT-5.4 with maximum compute for the hardest problems
Price: $30.00/M in · $180.00/M outGPT-5.2
openaiFrontier model with 400K context and adaptive reasoning
Price: $1.75/M in · $14.00/M outGPT-5.4 Mini
openaiStrongest mini model for coding, computer use, and subagents with GPT-5.4 capabilities
Price: $0.75/M in · $4.50/M outGPT-5 Mini
openaiCost-optimized reasoning and chat
Price: $0.25/M in · $2.00/M outGPT-5.4 Nano
openaiFastest and most affordable GPT-5.4 model for high-throughput tasks
Price: $0.20/M in · $1.25/M outGPT-5.2 Pro
openaiUses more compute for consistently better answers
Price: $21.00/M in · $168.00/M outGPT-5.3 Codex
openaiIndustry-leading agentic coding model. 400K context, reasoning, tool use, and complex execution
Price: $1.75/M in · $14.00/M outGPT-4.1
openaiLatest GPT-4 generation model
Price: $2.00/M in · $8.00/M outGPT-4.1 Mini
openaiFast and affordable GPT-4.1 model
Price: $0.40/M in · $1.60/M outGPT-4.1 Nano
openaiUltra-fast and cost-effective GPT-4.1
Price: $0.10/M in · $0.40/M outGPT-4o
openaiMultimodal model with vision and audio
Price: $2.50/M in · $10.00/M outGPT-4o Mini
openaiFast and affordable GPT-4o model
Price: $0.15/M in · $0.60/M outo1
openaiAdvanced reasoning model for complex tasks
Price: $15.00/M in · $60.00/M outo3
openaiLatest reasoning model with improved performance
Price: $2.00/M in · $8.00/M outo3-mini
openaiEfficient reasoning model for STEM tasks
Price: $1.10/M in · $4.40/M outo4-mini
openaiLatest generation efficient reasoning model
Price: $1.10/M in · $4.40/M outGPT-OSS 20B
openaiTestnetOpen-weight 20B model (Apache 2.0), similar performance to o3-mini. Available on testnet for developer testing.
Price: $0.0020/requestGPT-OSS 120B
openaiTestnetOpen-weight 120B model (Apache 2.0), flagship open model. Available on testnet for developer testing.
Price: $0.0030/requestClaude Haiku 4.5
anthropicFastest and most efficient Claude, near-frontier intelligence
Price: $1.00/M in · $5.00/M outClaude Sonnet 5
anthropicNewest Sonnet — near-Opus coding/agentic quality at Sonnet cost. 1M context, 128k output, adaptive thinking, vision
Price: $3.00/M in · $15.00/M outClaude Sonnet 4.6
anthropicBest balance of intelligence, speed, and cost
Price: $3.00/M in · $15.00/M outClaude Sonnet 4.5
anthropicSonnet 4.5 — strong coding and agentic performance, vision
Price: $3.00/M in · $15.00/M outClaude Opus 4.5
anthropicLatest Anthropic flagship with enhanced reasoning and creativity
Price: $5.00/M in · $25.00/M outClaude Opus 4.7
anthropicPowerful Claude Opus for complex reasoning and agentic coding. 1M context, 128k output, adaptive thinking
Price: $5.00/M in · $25.00/M outClaude Fable 5
anthropicAnthropic's most capable model — Mythos-class tier above Opus, for the most demanding reasoning and long-horizon agentic work. 1M context, 128K output, always-on thinking
Price: $10.00/M in · $50.00/M outClaude Opus 4.8
anthropicMost capable Claude 4-series Opus for complex reasoning and agentic coding. 1M context, 128k output, adaptive thinking
Price: $5.00/M in · $25.00/M outClaude Opus 5
anthropicNewest Opus — step-change over Opus 4.8 for deep reasoning and agentic coding at the same price. 1M context, 128k output, adaptive thinking
Price: $5.00/M in · $25.00/M outGemini 3.1 Pro
googleLatest Gemini with improved thinking, token efficiency, and agentic capabilities. Optimized for software engineering (requires new SDK)
Price: $2.00/M in · $12.00/M outGemini 3 Flash Preview
googleFrontier-class performance with Pro-level intelligence at Flash speed and pricing. Includes thinking mode (requires new SDK)
Price: $0.50/M in · $3.00/M outGemini 3.6 Flash
googleNewest-generation Flash with built-in thinking mode — frontier-class quality at Flash speed
Price: $1.50/M in · $7.50/M outGemini 3.5 Flash
googleLatest-generation Flash with built-in thinking mode — frontier-class quality at Flash speed
Price: $1.50/M in · $9.00/M outGemini 2.5 Pro
googleState-of-the-art for reasoning, coding, and mathematics
Price: $1.25/M in · $10.00/M outGemini 2.5 Flash
googleFast and efficient Gemini model with vision support
Price: $0.30/M in · $2.50/M outGemini 3.5 Flash Lite
googleLatest Flash Lite — ultra-fast, lightweight Gemini with thinking mode for high-throughput tasks
Price: $0.30/M in · $2.50/M outGemini 3.1 Flash Lite
googleUltra-fast and lightweight Gemini 3.1 model with thinking mode for high-throughput tasks
Price: $0.25/M in · $1.50/M outGemini 2.5 Flash Lite
googleMost economical Gemini model - ultra-fast and lightweight (requires new SDK)
Price: $0.10/M in · $0.40/M outDeepSeek V4 Flash Vision
deepseekDeepSeek V4 Flash Vision — 1M context with text and image input. Experimental multimodal tier of the V4 Flash family.
Price: $0.44/M in · $1.32/M outDeepSeek V4 Pro
deepseekDeepSeek V4 flagship — 1.6T MoE / 49B active, 1M context. Strongest open-weight reasoner. Thinking mode default.
Price: $1.32/M in · $3.96/M outDeepSeek V4 Flash Chat
deepseekPaid V4 Flash in non-thinking mode (1.6T-class quality at $0.14 in / $0.28 out). Production-grade reliability and 5MB request bodies.
Price: $0.14/M in · $0.28/M outDeepSeek V4 Flash Reasoner
deepseekPaid V4 Flash in thinking mode for reasoning tasks. Same upstream as deepseek/deepseek-chat but with thinking enabled by default.
Price: $0.14/M in · $0.28/M outKimi K3
moonshotMoonshot's flagship — a 2.8-trillion-parameter open MoE with 1M context, image + text input, returning reasoning_content. Live-verified 2026-07-17 (chat, tools, vision).
Price: $3.00/M in · $15.00/M outGLM-5.3
zaiZ.AI's flagship — 1M-token context with always-on reasoning, strong at long-horizon coding. Verified live on Z.AI.
Price: $1.40/M in · $4.40/M outGLM-5.3 Flash
zaiZ.AI's first natively multimodal GLM-5 — 320B/18B MoE with image input, 1M-token context, and always-on reasoning at value-tier pricing. Verified live on Z.AI.
Price: $0.15/M in · $0.50/M outGLM-5.2
zaiZ.AI GLM-5.2 — 1M-token context, strong open-source long-horizon coding. Verified live on Z.AI.
Price: $1.40/M in · $4.40/M outGLM-5.1
zaiZ.AI flagship — #1 open source on SWE-Bench Pro, 8-hour autonomous execution. 200K context
Price: $1.40/M in · $4.40/M outGLM-5
zaiZ.AI's foundation model with 200K context. Strong reasoning and agentic capabilities
Price: $1.00/M in · $3.20/M outGLM-5 Turbo
zaiOptimized GLM-5 variant with faster inference
Price: $1.20/M in · $4.00/M outGrok 4.3
xaixAI's Grok 4.3 reasoning model. 1M context, vision-capable, tuned for agentic workflows and instruction-following.
Price: $1.50/M in · $4.00/M outGrok Build 0.1
xaixAI's fast agentic coding model, trained for interactive software-engineering workflows. 256K context, text + image input.
Price: $1.50/M in · $3.00/M outGrok 4.5
xaixAI's flagship Grok 4.5 — their most intelligent and fastest model. 500K context, vision-capable, chain-of-thought reasoning. Supports Live Search (+$0.025/source)
Price: $2.50/M in · $9.00/M outMiniMax M2.7
minimaxMiniMax's flagship reasoning model with recursive self-improvement. Great value for complex tasks (~60 tps)
Price: $0.30/M in · $1.20/M outMiniMax M3
minimaxMiniMax's M3 flagship — 1M context, strong reasoning + coding.
Price: $0.30/M in · $1.20/M outQwen3.7 Max
qwenAlibaba's Qwen flagship — the Max tier. 1M context, strong reasoning, coding, and agentic tool use. Live-verified 2026-07-20.
Price: $1.48/M in · $4.42/M outQwen3.7 Plus
qwenAlibaba's balanced Qwen tier — 1M context with reasoning, coding, and agentic tool use at a fraction of the Max price
Price: $0.32/M in · $1.28/M outQwen3.7 Flash
qwenAlibaba's fastest Qwen tier — 1M context reasoning for high-volume, latency-sensitive workloads
Price: $0.03/M in · $0.13/M outQwen3.8 Flash
qwenAlibaba's Qwen3.8 Flash — 125B MoE with hybrid attention, 1M context, image input. Outperforms the Qwen3.7 Plus tier at a lower price.
Price: $0.15/M in · $0.47/M outTencent Hy3
tencentTencent's Hy3 — fast, inexpensive reasoning at 262K context. One of the most-used open models of 2026.
Price: $0.13/M in · $0.53/M outMiMo V2.5
xiaomiXiaomi's MiMo V2.5 — 310B sparse MoE, natively multimodal, 1M context. Accepts text and images.
Price: $0.14/M in · $0.28/M outXiaomi MiMo-V2.5 Pro
xiaomiXiaomi's MiMo-V2.5 Pro — 1M context reasoning model, priced well below the frontier tier.
Price: $0.43/M in · $0.87/M outNemotron 3 Nano Omni (Free)
nvidiaFreeNVIDIA's multimodal reasoning Nemotron Nano Omni, free. 31B / 3.2B active MoE. Accepts text, images, video, audio. ChartQA 90.3, DocVQA 95.6, MMMU 70.8 — the only vision-capable free model in our catalog
Price: Free/M in · Free/M outNemotron 3.5 Lightning (Free)
nvidiaFreeNVIDIA Nemotron 3.5 Lightning 30B-A3B, free. Thinking-mode reasoning with a 1M-token context.
Price: Free/M in · Free/M outNemotron 3 Nano 30B (Free)
nvidiaFreeNVIDIA Nemotron 3 Nano 30B-A3B hosted free by NVIDIA. Compact MoE, ~121 tok/s — the fastest free model in the catalog.
Price: Free/M in · Free/M outLlama 3.2 11B Vision (Free)
nvidiaFreeMeta's Llama 3.2 11B Vision hosted free by NVIDIA. Accepts images; 128K context.
Price: Free/M in · Free/M outNemotron 3 Ultra 550B (Free)
nvidiaFreeNVIDIA Nemotron 3 Ultra 550B-A55B, free. 550B total / 55B active MoE with a 1M-token context — the largest free model in the catalog.
Price: Free/M in · Free/M outCohere North Mini Code (Free)
cohereFreeCohere's North Mini Code, free. Compact coding model, 256K context, sub-second responses.
Price: Free/M in · Free/M outPoolside Laguna XS 2.1 (Free)
nvidiaFreePoolside Laguna XS 2.1 hosted free by NVIDIA. Fast compact coding model (~161 tok/s).
Price: Free/M in · Free/M outGPT Image 1
openaiImageNative image generation in GPT-4o
Price: $0.020/imageChatGPT Images 2.0
openaiImageNewOpenAI's GPT Image 2 — reasoning-driven image generation with multilingual text rendering, character consistency, and high-fidelity edits
Price: $0.060/imageNano Banana
googleImageGoogle's Gemini 2.5 Flash image generation - fast and efficient
Price: $0.050/imageNano Banana 2
googleImageGoogle's Gemini 3.1 Flash image generation - pro-level quality at Flash speed
Price: $0.090/imageNano Banana Pro
googleImageGoogle's Gemini 3 Pro image generation - highest quality up to 4K
Price: $0.100/imageGrok Imagine
xaiImageNewxAI's Grok Imagine image generation. Fast, 300 RPM.
Price: $0.020/imageGrok Imagine Pro
xaiImageNewxAI's premium Grok Imagine image generation (quality tier). Higher quality, 30 RPM.
Price: $0.070/imageSeedream 5.0 Pro
bytedanceImageByteDance's Seedream 5.0 Pro — flagship image generation and editing, up to 4K-class resolution with reference-image support
Price: $0.045/imageCogView-4
zaiImageZhipu AI's CogView-4 image generation model — high quality, supports up to 1440x1440
Price: $0.015/imageMiniMax Music 2.5+
minimaxMusicMiniMax's flagship music generation model. Supports lyrics, instrumental, and style prompts. ~3 min output.
Price: $0.150/trackElevenLabs Flash v2.5
elevenlabsSpeechUltra-low-latency (~75ms) speech synthesis for real-time voice agents. 32 languages.
Price: $0.05/1k charsElevenLabs Turbo v2.5
elevenlabsSpeechBalanced quality and latency (~250ms) for interactive use cases. 32 languages.
Price: $0.05/1k charsElevenLabs Multilingual v2
elevenlabsSpeechHighest-consistency voice for long-form narration, audiobooks, and voiceover. 29 languages.
Price: $0.10/1k charsElevenLabs v3
elevenlabsSpeechMaximum expressiveness and emotional range for creative applications. 70+ languages.
Price: $0.10/1k charsSeed Audio 1.0
bytedanceSpeechByteDance's Seed Audio 1.0 — prompt-directed audio creation: describe the voice, emotion, and sound staging in natural language. Up to 120s output, mp3/wav. Billed by audio duration ($0.003/second, estimated from input length).
Price: $0.30/1k charsElevenLabs Sound Effects
elevenlabsSpeechGenerate cinematic sound effects and audio textures from a text prompt (up to 22s).
Price: $0.050/clipGrok Imagine Video
xaiVideoNewxAI's Grok Imagine video generation. Text or image to video, configurable 1–15s clips. $0.05/sec at 480p (default), $0.07/sec at 720p — official per-second rates, plus a flat $0.001 per generation.
Price: $0.050/secGrok Imagine Video 1.5
xaiVideoNewxAI's flagship Grok Imagine Video 1.5 — text or image to video with native synced audio, 1–15s clips. $0.08/sec at 480p (default), $0.14/sec at 720p, $0.25/sec at 1080p — official per-second rates, plus a flat $0.001 per generation.
Price: $0.080/secSeedance 1.5 Pro
bytedanceVideoNewByteDance Seedance 1.5 Pro — budget text/image-to-video at 720p with synced audio (t2v), 5s default. Does NOT support RealFace assets.
Price: $0.070/secSeedance 2.0 Fast
bytedanceVideoNewByteDance Seedance 2.0 Fast — fast video at 720p with synced audio (t2v), 5s default. ~60-80s to generate. Supports BytePlus RealFace assets — see /docs/video/real-person-ip for enrollment.
Price: $0.165/secSeedance 2.0 Mini
bytedanceVideoNewByteDance Seedance 2.0 Mini — 720p video with synced audio at half the flagship rate, 5s default. 480p and 720p (no 1080p/4K). Supports BytePlus RealFace assets — see /docs/video/real-person-ip for enrollment.
Price: $0.080/secSeedance 2.0 Pro
bytedanceVideoNewByteDance Seedance 2.0 Pro — premium quality text/image-to-video at 720p with synced audio (t2v), 5s default. Supports BytePlus RealFace assets — see /docs/video/real-person-ip for enrollment.
Price: $0.227/secSeedance 2.5
bytedanceVideoByteDance Seedance 2.5 — long-form text/image-to-video at 720p with synced audio, up to 30s. Multilingual, multi-asset. For 1080p or 4K use Seedance 2.0 Pro.
Price: $0.315/secSora 2
azureVideoOpenAI Sora 2 via Azure AI Foundry — text-to-video AND image-to-video at 720p with synchronized audio. 4s default; 4, 8, or 12s. Portrait or landscape. Image-to-video takes a non-human reference image (human faces are rejected upstream — use Seedance + RealFace for real people). $0.10/sec.
Price: $0.100/sec
How pricing works on BlockRun
Every model above is billed per request, in USDC, settled on-chain through the x402 protocol. There is no subscription, no monthly minimum, and no API key to provision — a request arrives unpaid, the gateway answers with a signed price quote, your client signs it, and the retry returns the completion. Chat models are priced per million input and output tokens; image, video, music and speech models are priced per image, per second, per track, and per thousand characters respectively.
Token prices track each upstream lab's public list price with no markup on chat tokens — only a $0.001 per-request fee that covers on-chain settlement (media models carry a 5% margin). Because you pay per call, an agent that makes ten requests a day costs cents a month, and one that makes ten thousand pays exactly ten thousand times the per-call price — the unit economics do not change with volume, and nothing expires unused.
Choosing a model
The catalog spans 76 chat and reasoning models alongside image, video, music and speech generation, for 100 in total. Use the Reasoning filter for long-horizon planning and hard analytical work, Coding for repository-scale edits and agentic tool use, and Vision when requests carry images. The Free filter lists models that need no payment header at all — useful for prototyping an integration before you fund a wallet.
Every model is reachable through the same OpenAI-compatible /v1/chat/completions endpoint, so switching between them is a one-line change to the model field. See the documentation for request shapes and the pricing page for a full per-model rate card.