DeepSeek-V4-Flash
DeepSeek V4 Flash
Fast, economical general-purpose model for high-volume chat and code tasks.
Input: $0.18 / 1M tokens · Output: $0.36 / 1M tokens
Capabilities: text, code, fast
Published rates are in USD per 1M tokens. Availability, pricing and limits may vary by provider, region and account.
DeepSeek V4 Flash
Fast, economical general-purpose model for high-volume chat and code tasks.
Input: $0.18 / 1M tokens · Output: $0.36 / 1M tokens
Capabilities: text, code, fast
DeepSeek V4 Pro
Frontier reasoning and coding model for complex analysis and production code.
Input: $0.67 / 1M tokens · Output: $1.34 / 1M tokens
Capabilities: text, code, reasoning, structured-output
DeepSeek V4.1 Flash
Fast next-generation DeepSeek model for high-volume everyday tasks.
Input: $0.18 / 1M tokens · Output: $0.54 / 1M tokens
Capabilities: text, code
GLM 5.2
Previous-generation Zhipu general model. Reliable Chinese dialogue and coding tasks at a lower cost.
Input: $1.16 / 1M tokens · Output: $4.05 / 1M tokens
Capabilities: text, code, reasoning, structured-output
GLM 5.3
Zhipu flagship general model. Balanced Chinese dialogue, analysis, and structured output for production workloads.
Input: $1.61 / 1M tokens · Output: $5.64 / 1M tokens
Capabilities: text, code, reasoning, structured-output
GLM 5.3 Flash
Ultra-low-cost GLM model for high-volume lightweight tasks.
Input: $0.01 / 1M tokens · Output: $0.035 / 1M tokens
Capabilities: text, code
Kimi K2.5
Previous-generation general model for cost-sensitive long-text tasks.
Input: $0.54 / 1M tokens · Output: $2.84 / 1M tokens
Capabilities: text, code
Kimi K2.6
Moonshot general model balancing dialogue and long-text processing.
Input: $0.87 / 1M tokens · Output: $3.66 / 1M tokens
Capabilities: text, code
Kimi K2.7 Code
Optimized for programming: code generation, completion, and refactoring.
Input: $1.16 / 1M tokens · Output: $4.81 / 1M tokens
Capabilities: text, code, tool-use, reasoning
Kimi K3
Moonshot's latest flagship with strong long-context understanding and complex reasoning.
Input: $3.34 / 1M tokens · Output: $16.7 / 1M tokens
Capabilities: text, code, long-context, reasoning
MiniMax M2.5
Previous-generation general model for routine Chinese tasks.
Input: $0.32 / 1M tokens · Output: $1.28 / 1M tokens
Capabilities: text, code
MiniMax M2.7
Latest-generation general model with excellent Chinese writing and analysis.
Input: $0.38 / 1M tokens · Output: $1.5 / 1M tokens
Capabilities: text, code, reasoning
Qwen3.5 397B A17B
Large MoE model for high-throughput general reasoning.
Input: $0.16 / 1M tokens · Output: $0.96 / 1M tokens
Capabilities: text, code, reasoning
GPT-5.6 Sol
OpenAI-family model tuned for long-form reasoning and coding.
Input: $4.53 / 1M tokens · Output: $27.18 / 1M tokens
Capabilities: text, code, reasoning
GPT-5.6 Terra
Cost-efficient OpenAI-family model for everyday coding and chat.
Input: $1.81 / 1M tokens · Output: $10.86 / 1M tokens
Capabilities: text, code
GPT-6 Astra
Frontier OpenAI-family flagship for complex reasoning and agentic workloads.
Input: $9.07 / 1M tokens · Output: $45.35 / 1M tokens
Capabilities: text, code, reasoning
GPT-6 Sol
Next-generation OpenAI-family model for general reasoning and coding.
Input: $4.53 / 1M tokens · Output: $27.18 / 1M tokens
Capabilities: text, code, reasoning
GPT-6.1 Sol
Latest OpenAI-family model for general reasoning and coding.
Input: $4.53 / 1M tokens · Output: $27.18 / 1M tokens
Capabilities: text, code, reasoning
Grok 4.6
xAI Grok model for reasoning, coding, and real-time knowledge tasks.
Input: $4.89 / 1M tokens · Output: $24.45 / 1M tokens
Capabilities: text, code, reasoning
Hy3
Hunyuan third-generation model with stable Chinese dialogue and content generation.
Input: $0.23 / 1M tokens · Output: $0.89 / 1M tokens
Capabilities: text, code, fast
Qwen3 VL Plus
Qwen vision-language model for image understanding and multimodal tasks.
Input: $0.11 / 1M tokens · Output: $1.1 / 1M tokens
Capabilities: text, code, vision
Qwen3.6 Plus
Enhanced Qwen model balancing capability and cost for general workloads.
Input: $0.36 / 1M tokens · Output: $2.14 / 1M tokens
Capabilities: text, code, fast
Qwen3.7 Max
Largest Qwen model with top-tier Chinese understanding and generation.
Input: $1.34 / 1M tokens · Output: $4.01 / 1M tokens
Capabilities: text, code, reasoning, tool-use
Read the docs · Compare plans · Start building free
Only models with a published price are purchasable. Models without a published price are not offered for sale.