Skip to content

Models & Pricing

Prices are configured in src/config/site.ts and are shown in USD per million tokens.

Zhipu

Long-context coding

GLM-5.2

glm-5.2

A long-context reasoning model for coding agents, tool use, and large document workflows that need stronger state retention.

1M contextCodingTool use
Input
$0.90
Output
$3.10

Moonshot

Agentic coding & reasoning

Kimi K3

kimi-k3

Moonshot's newest-generation Kimi model with an internal reasoning budget for agentic coding, tool use, and long-context problem solving.

Reasoning1M contextAgents
Input
$2.30
Output
$12.3

Moonshot

Agentic coding

Kimi K2.7 Code

kimi-k2.7-code

A coding-focused model for agentic programming, repository-level edits, and multi-step developer workflows.

CodeAgentsRepositories
Input
$0.72
Output
$3

Alibaba

Complex automation

Qwen3.7-Max

qwen3.7-max

The premium Qwen route for complex planning, multimodal automation, and higher-stakes product workflows.

FlagshipVisionPlanning
Input
$1.24
Output
$3.72

DeepSeek

Reasoning and debugging

DeepSeek V4 Pro

deepseek-v4-pro

A stronger DeepSeek tier for heavier reasoning, debugging, and multi-step product work with 1M context.

Reasoning1M contextDebugging
Input
$0.31
Output
$0.62

Alibaba

Multimodal agents

Qwen3.7-Plus

qwen3.7-plus-2026-05-26

A balanced multimodal model for agents that read screens, reason through product tasks, and write code without premium frontier pricing.

VisionCodingAgents
Input
$0.25
Output
$1.20

xAI

Reasoning & coding

Grok 4.5

grok-4.5

xAI's advanced reasoning and coding model for agents, analysis, and tool-assisted developer workflows.

ReasoningCodingAgents
Input
$1.40
Output
$4.20

Anthropic

Frontier coding & agents

Claude Opus 5

claude-opus-5

Anthropic's newest frontier flagship for the hardest coding, long-horizon agent runs, and deep reasoning, with up to 1M tokens of context.

FlagshipReasoningAgents
Input
$0.85
Output
$4.25

Anthropic

Hard coding & long-horizon agents

Claude Opus 4.8

claude-opus-4-8

Anthropic's frontier flagship for the hardest coding, agents, and long-horizon work through one OpenAI-compatible gateway.

FlagshipCodingAgents
Input
$3
Output
$15

Anthropic

Real-time agents & tool use

Claude Sonnet 5

claude-sonnet-5

Fast frontier intelligence for real-time agents, tool use, and computer use — the everyday Claude workhorse for low-latency workflows.

AgentsTool use1M context
Input
$1.20
Output
$6

Anthropic

Frontier reasoning

Claude Fable 5

claude-fable-5

Anthropic's most capable flagship for frontier reasoning, complex tool use, and long-horizon agents.

FlagshipFrontierAgents
Input
$7
Output
$35

Google

Multimodal reasoning & coding

Gemini 3.1 Pro

gemini-3.1-pro-preview

Google's high-capability multimodal model for complex reasoning, coding, and long-context work.

MultimodalReasoningLong context
Input
$1.20
Output
$7.20

OpenAI

Frontier agentic coding

GPT-5.5

gpt-5.5

OpenAI's frontier model for agentic coding, computer use, and knowledge work through one OpenAI-compatible gateway.

FlagshipCodingAgents
Input
$3
Output
$18

OpenAI

Balanced frontier + long context

GPT-5.6 Terra

gpt-5.6-terra

A balanced GPT-5.6 tier for frontier agents, tool use, and long-context coding.

FrontierAgentsLong context
Input
$1.75
Output
$10.5
ModelProviderUse caseInputOutput
GLM-5.2glm-5.2ZhipuLong-context coding
Kimi K3kimi-k3MoonshotAgentic coding & reasoning
Kimi K2.7 Codekimi-k2.7-codeMoonshotAgentic coding
Kimi K2.6kimi-k2.6MoonshotLong-form reasoning & coding
Qwen3.7-Maxqwen3.7-maxAlibabaComplex automation
DeepSeek V4 Prodeepseek-v4-proDeepSeekReasoning and debugging
Qwen3.7-Plusqwen3.7-plus-2026-05-26AlibabaMultimodal agents
LongCat-2.0meituan-longcat/LongCat-2.0Meituan LongCatAgentic coding
$1.35/ 1M
$5.40/ 1M
MiniMax M3minimax-m3MiniMaxLong-context product work
$0.55/ 1M
$2.20/ 1M
Qwen3.5-Plusqwen3.5-plus-2026-04-20AlibabaProduction value
Grok 4.5grok-4.5xAIReasoning & coding
Grok 4.3grok-4.3xAIFast reasoning
Claude Opus 5claude-opus-5AnthropicFrontier coding & agents
Claude Opus 4.8claude-opus-4-8AnthropicHard coding & long-horizon agents
Claude Sonnet 5claude-sonnet-5AnthropicReal-time agents & tool use
Claude Fable 5claude-fable-5AnthropicFrontier reasoning
Claude Opus 4.7claude-opus-4-7AnthropicHard coding & agents
Claude Sonnet 4.6claude-sonnet-4-6AnthropicBalanced coding & agents
Claude Haiku 4.5claude-haiku-4-5-20251001AnthropicFast, high-volume Claude
Gemini 3.1 Progemini-3.1-pro-previewGoogleMultimodal reasoning & coding
Gemini 2.5 Progemini-2.5-proGoogleMultimodal reasoning
Gemini 3.5 Flashgemini-3.5-flashGoogleFast multimodal reasoning
Gemini 2.5 Flashgemini-2.5-flashGoogleHigh-volume multimodal
GPT-5.5gpt-5.5OpenAIFrontier agentic coding
GPT-5.6 Terragpt-5.6-terraOpenAIBalanced frontier + long context
GPT-5.4gpt-5.4OpenAIGeneral coding & agents
GPT-5gpt-5OpenAIGeneral assistants & coding
GPT-5 Minigpt-5-miniOpenAIFast, high-volume GPT
GPT Image 2gpt-image-2OpenAIPosters, packaging & illustration
$2.25/ 1M
$13.5/ 1M
Gemini 2.5 Flash Imagegemini-2.5-flash-imageGoogleEveryday image generation & editing
$0.09/ 1M
$9/ 1M
Gemini 3.1 Flash Imagegemini-3.1-flash-image-previewGoogleFast, higher-fidelity images
$0.15/ 1M
$18/ 1M
Seedance 2.0doubao-seedance-2-0-260128ByteDanceAd, film & social video
$2.32/ 1M
$2.32/ 1M
Seedance 2.0 Fastdoubao-seedance-2-0-fast-260128ByteDanceHigh-volume video generation
$3.20/ 1M
$3.20/ 1M
Seedance 2.0 Minidoubao-seedance-2-0-mini-260615ByteDanceDay-to-day video runs
$2.04/ 1M
$2.04/ 1M
DeepSeek V4 Flashdeepseek-v4-flash-202605DeepSeekBudget · General · Long context
$0.20/ 1M
$0.40/ 1M
GLM-5-Turboglm-5-turboZhipuAgent · Tool use
$1.50/ 1M
$5.20/ 1M
GLM-5V-Turboglm-5v-turboZhipuVision · Coding
$1.60/ 1M
$5.40/ 1M
Kimi K2.7 Code HighSpeedkimi-k2.7-code-highspeedMoonshotCoding · Fast · Agentic
MiniMax M2.5minimax-m2.5MiniMaxValue · Agents
$0.50/ 1M
$2/ 1M
MiniMax M2.7minimax-m2.7MiniMaxCoding · Productivity
$0.52/ 1M
$2.10/ 1M
Qwen3.5-Flashqwen3.5-flashAlibabaUltra-budget · Fast
$0.12/ 1M
$0.95/ 1M
Qwen3.6 Open 27Bqwen3.6-27bAlibabaOpen · Advanced
$0.95/ 1M
$4.95/ 1M
Qwen3.6-Flashqwen3.6-flashAlibabaFast · Budget · Long context
$0.17/ 1M
$1/ 1M
Qwen3.6-Plusqwen3.6-plus-2026-04-02AlibabaCompatibility · Multimodal
Qwen3.5-35B-A3Bqwen3.5-35b-a3bAlibabaFast · Budget · MoE
Claude Opus 4.5claude-opus-4-5-20251101AnthropicFlagship · Coding · Agents
Claude Opus 4.5 Thinkingclaude-opus-4-5-20251101-thinkingAnthropicFlagship · Reasoning · Thinking
Claude Opus 4.6claude-opus-4-6AnthropicFlagship · Coding · Agents
Claude Opus 4.6 Thinkingclaude-opus-4-6-thinkingAnthropicFlagship · Reasoning · Thinking
Claude Sonnet 4.5 (0929)claude-sonnet-4-5-20250929AnthropicBalanced · Coding · Agents
Claude Sonnet 4.5 Thinkingclaude-sonnet-4-5-20250929-thinkingAnthropicBalanced · Reasoning · Thinking
Claude Sonnet 4.6 Thinkingclaude-sonnet-4-6-thinkingAnthropicBalanced · Reasoning · Thinking
GPT-4o minigpt-4o-miniOpenAIReasoning · Math
Kimi K2.5kimi-k2.5MoonshotGeneral · Multimodal
GLM-5glm-5ZhipuAgents · Value
GLM-5.1glm-5.1ZhipuReasoning · Agents
Gemini 2.5 Flash Litegemini-2.5-flash-liteGoogleFast · Low cost · Multimodal
Gemini 2.5 Flash Thinkinggemini-2.5-flash-thinkingGoogleReasoning · Multimodal · Thinking
Gemini 3 Flash (preview)gemini-3-flash-previewGoogleFast · Multimodal · Preview
Gemini 3.1 Flash Litegemini-3.1-flash-liteGoogleFast · Low cost · Multimodal
Gemini 3.1 Flash Lite (preview)gemini-3.1-flash-lite-previewGoogleFast · Low cost · Preview
$0.19/ 1M
$1.16/ 1M
Gemini 3.1 Pro Thinkinggemini-3.1-pro-preview-thinkingGoogleReasoning · Multimodal · Thinking
GPT-5.4 Minigpt-5.4-miniOpenAIFast · Low cost
GPT-5.6 Lunagpt-5.6-lunaOpenAIFrontier · Efficient · Long context
GPT-5.6 Solgpt-5.6-solOpenAIFrontier · Agents · Long context

Sell prices may change as upstream prices move. Linked public references open their source in a new tab.