AI Model Pricing
Every model available on OpenClaw Launch. Switch models anytime, even after deployment.
GPT-5.6 Sol
OpenAIOpenAI's maximum-capability GPT-5.6 model for demanding reasoning, coding, and multimodal agent workflows.
GPT-5.6 Terra
OpenAIThe balanced GPT-5.6 tier for everyday agents, with frontier capability at half the Sol token price.
GPT-5.6 Luna
OpenAIThe efficient GPT-5.6 tier for high-volume workflows, classification, drafting, and tool routing.
GPT-5.5 (Legacy)
OpenAIOpenAI's previous-generation flagship, retained for compatibility with existing agents. GPT-5.6 is recommended for new setups.
GPT-5.4 Mini
OpenAIFast and efficient version of GPT-5.4 with vision support. Great for high-throughput workloads and computer use tasks.
Claude Sonnet 5
AnthropicAnthropic's current Sonnet. The balanced default for most work, and priced below the Sonnet 4.6 it succeeds.
Claude Opus 5
AnthropicAnthropic's most capable model. Best for nuanced analysis, agentic coding, and complex multi-step tasks.
Claude Haiku 4.5
AnthropicThe smallest, fastest Claude. Well suited to high-volume work and quick turns where cost matters most.
Claude Opus 4.8
AnthropicAnthropic's newest and most capable model. Best for nuanced analysis, agentic coding, and complex multi-step tasks.
Claude Opus 4.7
AnthropicFrontier-tier reasoning and coding with deep agentic reliability. A step below Opus 4.8 at the same price point.
Claude Opus 4.6
AnthropicPrevious-generation Opus flagship. Still excellent for nuanced analysis, creative writing, and complex multi-step tasks.
Claude Sonnet 4.6
AnthropicThe best balance of intelligence, speed, and cost. Ideal default for most use cases.
Gemini 3.8 Flash
GoogleGoogle's current frontier Flash model, built for long-horizon agent work. Vision and video input, 1M context.
Gemini 3.6 Flash
GoogleThe previous Flash generation, at the same price as 3.8. Vision and video input, 1M context.
Gemini 3.1 Pro
GoogleGoogle's frontier reasoning model with enhanced software engineering performance and improved agentic reliability.
Gemini 3.5 Flash
GoogleGoogle's newest mid-tier multimodal model. Bigger reasoning step up from Gemini 3 Flash, with vision and video input.
Gemini 3 Flash
GoogleGoogle's fast and affordable multimodal model. Great for vision tasks, quick reasoning, and high-throughput workloads.
Gemini 3.5 Flash Lite
GoogleLow-latency, cost-effective subagent option with vision and video input.
Gemini 3.1 Flash Lite
GoogleUltra-efficient and fast. Great for high-volume tasks, quick responses, and the most cost-efficient workflows.
Gemini 2.5 Flash
GoogleThe previous-generation GA Flash. Kept selectable for cost-conscious setups already running it.
Gemini 2.5 Flash Lite
GoogleThe cheapest Gemini on this list. Best for very high-volume, simple turns.
Grok 4.6
xAIxAI's flagship since August 2026. 500K context, built for long-running agents, with X search available as a tool.
Grok 4.3
xAIThe cheapest capable Grok on this list, with a 1M-token context. A good everyday default.
Grok 4.20
xAIA previous xAI flagship, and cheaper than 4.6 at base rates. Kept for agents already pinned to it.
DeepSeek V4 Flash
DeepSeekUltra-low-cost flash model with 1M context. A strong choice for budget-conscious chat and agent workflows.
DeepSeek V4 Pro
DeepSeekDeepSeek's frontier model. 1M context, strong coding and reasoning at a fraction of the cost of other premium models.
DeepSeek V4.1 Flash
Hunyuan 3
TencentTencent's Hunyuan 3 preview — a remarkably cheap large model with a 256K context. One of the lowest-cost capable models on OpenRouter.
Kimi K2.6
Moonshot AINext-gen multimodal model for long-horizon coding, UI/UX generation, and multi-agent orchestration.
GLM 5.3 Flash
GLM 5.2
Z.aiZ.ai's latest coding and agentic model with a 1M context window. Cheaper than GLM 5.1 with a much longer context.
GLM 5.1
Z.aiCoding-focused model with 202K context. Strong at long-horizon agentic tasks and tool use.
MiniMax M3
MiniMax M2.7
MiniMaxSelf-evolving model with strong coding and reasoning. Affordable pricing with 200K context window.
Step 3.7 Flash
StepFunStepFun's newest flash model with vision and video input. Stronger reasoning than Step 3.5 Flash while staying low-cost.
Step 3.5 Flash
StepFunSparse MoE with 196B total params, 11B active per token. StepFun's most capable open-source model at an ultra-low price.
MiMo V2.5 Pro
XiaomiXiaomi's newest flagship, tuned for agentic workflows. Strong coding and orchestration with a 1M context window.
MiMo V2.5
XiaomiMultimodal MiMo with vision and video input at an ultra-low price. Great for high-volume agent and chat workloads.
MiMo V2 Pro
XiaomiXiaomi's flagship 1T-parameter model optimized for agentic workflows. Strong at coding, reasoning, and complex orchestration tasks.
Qwen 3.6 Plus
Alibaba CloudHybrid linear-attention + sparse MoE flagship. Strong at agentic coding and front-end tasks (78.8 on SWE-bench Verified) with 1M context and native reasoning mode.
Gemma 4 26B A4B
GoogleOpen-weight multimodal Gemma model with vision and video input. MoE architecture for efficient inference.
Gemma 4 31B
GoogleLarger Gemma variant with vision and video support. Great balance of quality and cost for multimodal tasks.
Free Models Router
OpenRouterAuto-routes each request to a working free model on OpenRouter. When one free model goes dark, the router picks another — the most reliable way to stay on free tier.
Nemotron 3 Super
NVIDIAHybrid Mamba-Transformer MoE model — 120B params, 12B active. 1M context window with zero cost. Great for complex agent tasks.
custom/Qwen3.6 Plus
custom/Qwen3.7 Plus
custom/MiniMax M2.7
custom/MiniMax M3
custom/Kimi K2.6
custom/DeepSeek V4 Pro
custom/DeepSeek V4 Flash
custom/GLM 5.2
custom/GLM 5.3 Flash
custom/DeepSeek V4.1 Flash
custom/Qwen3.8 Flash
custom/MiMo V2.6 Flash
custom/GPT 5.6 Luna
custom/MiMo V2.6 Pro
custom/MiMo V2.5
custom/MiMo V2.5 Pro
custom/Kimi K2.7 Code
custom/LongCat 2.0
custom/MiniMax M2.5
custom/DeepSeek V4 Flash Vision Exp
custom/Hy4 Preview
custom/Hy3
custom/Space Bunny Free
custom/Qwen3.8 Max
GPT-5.6 on Your ChatGPT Subscription
Connect a ChatGPT Plus or Pro login to your bot and run OpenAI's newest GPT-5.6 family at no per-token cost — usage counts against your ChatGPT plan's weekly allowance instead. Available on OpenClaw v2026.7.1+ instances. Read the GPT-5.6 guide.
GPT-5.6 Sol
OpenAIOpenAI's flagship frontier model — their most capable for agentic work, coding, and hard reasoning. 272k context on the subscription route.
GPT-5.6 Terra
OpenAIBalanced GPT-5.6 tier for everyday work — strong reasoning at lower latency than Sol.
GPT-5.6 Luna
OpenAIFast, lighter GPT-5.6 tier for quick replies and high-volume chats.
Search Models
Perplexity-powered web search with real-time citations. Used automatically when your bot needs current information.
Sonar
PerplexityFast web search with citations. Good for quick factual lookups.
Image Generation Models
Generate and edit images directly from chat. Bring your own provider key (OpenAI, Google, or ByteDance) to unlock.
GPT-5.4 Image 2
OpenAIGPT-5.4 reasoning paired with GPT Image 2 generation. Best-in-class prompt adherence and text rendering.
Gemini 3.1 Flash Image
GoogleFast Nano-Banana-class image generation and editing. Strong at photoreal output and low-latency workflows.
Seedream 4.5
ByteDance SeedByteDance Seed's latest text-to-image model. Sharp detail, native Chinese prompt support, and competitive per-image pricing.
Intelligence Ranking
Ranked by the Artificial Analysis Intelligence Index, with detailed scores from major independent benchmarks
| # | Model | AA Index | GDPval-AA | Arena Elo | GPQA | SWE-bench | AIME | BrowseComp | tau2-Bench |
|---|---|---|---|---|---|---|---|---|---|
| Intelligence (0–100) | Real-world tasks | Human preference | Science reasoning | Coding & bugs | Math | Web research | Agent tool use | ||
| 1 | Claude Opus 4.8 | 61 | — | — | — | — | — | — | — |
| 2 | GPT-5.5 | 60 | 1782 | 1485 | 92.8 | — | — | 82.7 | — |
| 3 | Claude Opus 4.7 | 57 | — | — | — | — | — | — | — |
| 4 | Gemini 3.1 Pro | 57 | 1314 | 1500 | 94.3 | 80.6% | — | — | — |
| 5 | Claude Opus 4.6 | — | 1619 | 1504 | 91.3 | 80.8% | — | 86.8 | — |
| 6 | Gemini 3.5 Flash | 55 | — | — | — | — | — | — | — |
| 7 | Kimi K2.6 | 54 | 1486 | — | — | — | — | — | — |
| 8 | MiMo V2.5 Pro | 54 | — | — | — | — | — | — | — |
| 9 | Claude Sonnet 4.6 | 52 | 1676 | 1446 | 89.9 | 79.6% | — | — | — |
| 10 | DeepSeek V4 Pro | 52 | 1558 | — | — | — | — | — | — |
| 11 | GLM 5.1 | 51 | 1535 | — | — | — | — | — | — |
| 12 | MiniMax M2.7 | 50 | 1514 | — | — | 78.0% | — | — | — |
| 13 | Qwen 3.6 Plus | 50 | 1298 | — | — | 78.8% | — | — | — |
| 14 | MiMo V2.5 | 49 | — | — | — | — | — | — | — |
| 15 | GPT-5.4 Mini | 49 | 1435 | — | — | 54.4% | — | — | — |
| 16 | MiMo V2 Pro | — | 1418 | — | — | — | — | — | — |
| 17 | DeepSeek V4 Flash | 47 | 1414 | — | — | — | — | — | — |
| 18 | Step 3.7 Flash | 44 | — | — | — | — | — | — | — |
| 19 | Hunyuan 3 | 42 | — | — | — | — | — | — | — |
| 20 | Gemini 3 Flash | — | 1119 | 1473 | 90.4 | — | — | — | — |
| 21 | Gemma 4 31B | 39 | 1117 | 1452 | 84.3 | — | 89.2 | — | — |
| 22 | Step 3.5 Flash | 38 | 1073 | — | — | 74.4% | 97.3 | — | 88.2% |
| 23 | Gemma 4 26B A4B | — | 1012 | — | — | — | — | — | — |
| 24 | Nemotron 3 Super | — | 1004 | — | 79.2 | 60.5% | 90.2 | — | — |
| 25 | Gemini 3.1 Flash Lite | — | 927 | 1432 | 86.9 | — | — | — | — |
| 26 | Free Models Router | — | — | — | — | — | — | — | — |
Ready to deploy?
Pick any model and launch your AI chatbot in under 30 seconds.
Deploy Now