Models
Every model, one key.
39 models from 8 makers, priced per token on the default plan and callable from the SDK you already use.
- $10 / $50
Anthropic: Claude Fable 5
A Mythos-class model for demanding reasoning and long-running autonomous work, with always-on adaptive thinking.
- $2 / $10
Anthropic: Claude Haiku 4.5
The fastest and cheapest model in the family, for high-volume work where latency matters more than depth.
- $10 / $50
Anthropic: Claude Opus 4.8
The previous Opus generation for complex reasoning and agentic work, superseded by Opus 5.
- $5 / $25
Anthropic: Claude Opus 5
Anthropic's flagship model for complex agentic coding and long-horizon professional work, with adaptive reasoning across a million-token window.
- $6 / $30
Anthropic: Claude Sonnet 4.6
The previous Sonnet generation for general coding and assistant workloads, superseded by Sonnet 5.
- $4 / $20
Anthropic: Claude Sonnet 5
A mid-tier model balancing latency against reasoning quality for everyday application and assistant work.
- $0.88 / $2.64
DeepSeek: DeepSeek V4 Flash
The faster, cheaper V4 tier with the same window and output ceiling as V4 Pro.
- $2.64 / $7.92
DeepSeek: DeepSeek V4 Pro
The higher-capability V4 text model, with thinking mode on by default and a million-token window.
- $1 / $6
Google: Gemini 3 Flash Preview
The preview-channel Flash model, with the same multimodal input envelope as 3.5 Flash.
- $1 / $6
Google: Gemini 3.5 Flash
A multimodal Flash model for agentic and coding work, taking text, images, audio, video, and documents.
- $0.4 / $2.20
Z.ai: GLM-4.5 Air
The lightweight, lower-priced member of the GLM-4.5 family.
- $1.20 / $3.60
Z.ai: GLM-4.5V
A vision-language model taking images, video, and documents alongside text.
- $1.20 / $4.40
Z.ai: GLM-4.7
A prior-generation text model for code generation and long-context understanding.
- $2 / $6.40
Z.ai: GLM-5
The first GLM-5 generation for agentic engineering, with thinking mode.
- $2.40 / $8
Z.ai: GLM-5 Turbo
A GLM-5 variant tuned for agent scenarios, emphasising tool calls over long execution chains.
- $2.80 / $8.80
Z.ai: GLM-5.1
A text model for autonomous multi-step execution, with planning and iterative refinement.
- $2.80 / $8.80
Z.ai: GLM-5.2
A long-context text model aimed at long-horizon coding and engineering tasks.
- $1.40 / $4.40
Z.ai: GLM-5.3
The current GLM flagship for software engineering and agent workloads, with configurable reasoning effort.
- $4 / $16
OpenAI: GPT-4.1
A non-reasoning model oriented toward instruction following and tool calling at low latency.
- $5 / $30
OpenAI: GPT-5.4
An earlier flagship reasoning model, priced below GPT-5.5.
- $5 / $30
OpenAI: GPT-5.5
A prior flagship reasoning model for coding and professional work.
- $0.2 / $1.20
OpenAI: GPT-5.6 Luna
The cost-optimised GPT-5.6 tier, built for high-volume workloads.
- $5 / $30
OpenAI: GPT-5.6 Sol
The flagship GPT-5.6 tier for complex professional work, with reasoning effort configurable from none to maximum.
- $2 / $12
OpenAI: GPT-5.6 Terra
The middle GPT-5.6 tier, trading some capability for a lower price.
- $2 / $6
xAI: Grok 4.6
xAI's frontier model for coding, agentic tasks, and knowledge work, with four reasoning-effort levels.
- $1.90 / $8
Moonshot AI: Kimi K2.6
A general-purpose multimodal model offering both thinking and non-thinking modes.
- $1.90 / $8
Moonshot AI: Kimi K2.7 Code
A coding-oriented multimodal model supporting extended reasoning.
- $3.80 / $16
Moonshot AI: Kimi K2.7 Code Highspeed
The higher-throughput serving variant of K2.7 Code, at twice the price.
- $6 / $30
Moonshot AI: Kimi K3
Moonshot's flagship multimodal model, with always-on reasoning and a selectable effort level.
- $1 / $2.01
Qwen: Qwen Coder Plus
An earlier-generation programming and code-generation model.
- $0.264 / $2.12
Qwen: Qwen Flash
A speed- and cost-oriented model with switchable thinking mode.
- $6.34 / $31.60
Qwen: Qwen Max
An earlier-generation flagship text model with a short window.
- $2.12 / $6.34
Qwen: Qwen Plus
The balanced tier, supporting function calling, structured output, and batch inference.
- $1.60 / $6.40
Qwen: Qwen VL Max
The higher-capability vision-language tier, taking text, images, and video.
- $0.42 / $1.26
Qwen: Qwen VL Plus
The balanced vision-language tier, taking text, images, and video.
- $0.6 / $3
Qwen: Qwen3 Coder Flash
The smaller Coder tier, with the same window and multi-turn tool use.
- $2 / $10
Qwen: Qwen3 Coder Plus
A code model with agentic tool calling and repository-level understanding.
- $0.2 / $0.8
Qwen: Qwen3.5 Flash
A lower-cost Qwen 3.5 tier taking text, images, and video.
- $5 / $15
Qwen: Qwen3.7 Max
The Qwen flagship, with hybrid thinking and non-thinking modes.
