AI Model Pricing

Published rates are in USD per 1M tokens. Availability, pricing and limits may vary by provider, region and account.

DeepSeek-V4-Flash

DeepSeek V4 Flash

Fast, economical general-purpose model for high-volume chat and code tasks.

Input: $0.18 / 1M tokens · Output: $0.36 / 1M tokens

Capabilities: text, code, fast

DeepSeek-V4-Pro

DeepSeek V4 Pro

Frontier reasoning and coding model for complex analysis and production code.

Input: $0.67 / 1M tokens · Output: $1.34 / 1M tokens

Capabilities: text, code, reasoning, structured-output

DeepSeek-V4.1-Flash

DeepSeek V4.1 Flash

Fast next-generation DeepSeek model for high-volume everyday tasks.

Input: $0.18 / 1M tokens · Output: $0.54 / 1M tokens

Capabilities: text, code

GLM-5.2

GLM 5.2

Previous-generation Zhipu general model. Reliable Chinese dialogue and coding tasks at a lower cost.

Input: $1.16 / 1M tokens · Output: $4.05 / 1M tokens

Capabilities: text, code, reasoning, structured-output

GLM-5.3

GLM 5.3

Zhipu flagship general model. Balanced Chinese dialogue, analysis, and structured output for production workloads.

Input: $1.61 / 1M tokens · Output: $5.64 / 1M tokens

Capabilities: text, code, reasoning, structured-output

GLM-5.3-flash

GLM 5.3 Flash

Ultra-low-cost GLM model for high-volume lightweight tasks.

Input: $0.01 / 1M tokens · Output: $0.035 / 1M tokens

Capabilities: text, code

Kimi-K2.5

Kimi K2.5

Previous-generation general model for cost-sensitive long-text tasks.

Input: $0.54 / 1M tokens · Output: $2.84 / 1M tokens

Capabilities: text, code

Kimi-K2.6

Kimi K2.6

Moonshot general model balancing dialogue and long-text processing.

Input: $0.87 / 1M tokens · Output: $3.66 / 1M tokens

Capabilities: text, code

Kimi-K2.7-Code

Kimi K2.7 Code

Optimized for programming: code generation, completion, and refactoring.

Input: $1.16 / 1M tokens · Output: $4.81 / 1M tokens

Capabilities: text, code, tool-use, reasoning

Kimi-K3

Kimi K3

Moonshot's latest flagship with strong long-context understanding and complex reasoning.

Input: $3.34 / 1M tokens · Output: $16.7 / 1M tokens

Capabilities: text, code, long-context, reasoning

MiniMax-M2.5

MiniMax M2.5

Previous-generation general model for routine Chinese tasks.

Input: $0.32 / 1M tokens · Output: $1.28 / 1M tokens

Capabilities: text, code

MiniMax-M2.7

MiniMax M2.7

Latest-generation general model with excellent Chinese writing and analysis.

Input: $0.38 / 1M tokens · Output: $1.5 / 1M tokens

Capabilities: text, code, reasoning

Qwen3.5-397B-A17B

Qwen3.5 397B A17B

Large MoE model for high-throughput general reasoning.

Input: $0.16 / 1M tokens · Output: $0.96 / 1M tokens

Capabilities: text, code, reasoning

gpt-5.6-sol

GPT-5.6 Sol

OpenAI-family model tuned for long-form reasoning and coding.

Input: $4.53 / 1M tokens · Output: $27.18 / 1M tokens

Capabilities: text, code, reasoning

gpt-5.6-terra

GPT-5.6 Terra

Cost-efficient OpenAI-family model for everyday coding and chat.

Input: $1.81 / 1M tokens · Output: $10.86 / 1M tokens

Capabilities: text, code

gpt-6-astra

GPT-6 Astra

Frontier OpenAI-family flagship for complex reasoning and agentic workloads.

Input: $9.07 / 1M tokens · Output: $45.35 / 1M tokens

Capabilities: text, code, reasoning

gpt-6-sol

GPT-6 Sol

Next-generation OpenAI-family model for general reasoning and coding.

Input: $4.53 / 1M tokens · Output: $27.18 / 1M tokens

Capabilities: text, code, reasoning

gpt-6.1-sol

GPT-6.1 Sol

Latest OpenAI-family model for general reasoning and coding.

Input: $4.53 / 1M tokens · Output: $27.18 / 1M tokens

Capabilities: text, code, reasoning

grok-4.6

Grok 4.6

xAI Grok model for reasoning, coding, and real-time knowledge tasks.

Input: $4.89 / 1M tokens · Output: $24.45 / 1M tokens

Capabilities: text, code, reasoning

hy3

Hy3

Hunyuan third-generation model with stable Chinese dialogue and content generation.

Input: $0.23 / 1M tokens · Output: $0.89 / 1M tokens

Capabilities: text, code, fast

qwen3-vl-plus

Qwen3 VL Plus

Qwen vision-language model for image understanding and multimodal tasks.

Input: $0.11 / 1M tokens · Output: $1.1 / 1M tokens

Capabilities: text, code, vision

qwen3.6-plus

Qwen3.6 Plus

Enhanced Qwen model balancing capability and cost for general workloads.

Input: $0.36 / 1M tokens · Output: $2.14 / 1M tokens

Capabilities: text, code, fast

qwen3.7-max

Qwen3.7 Max

Largest Qwen model with top-tier Chinese understanding and generation.

Input: $1.34 / 1M tokens · Output: $4.01 / 1M tokens

Capabilities: text, code, reasoning, tool-use

Read the docs · Compare plans · Start building free

Only models with a published price are purchasable. Models without a published price are not offered for sale.