Usage-based pricing · start free
Pricing
- Per-token
- usage-based billing
- No limits
- practically no rate limits
- 200 req
- free every month
Open Source Models
Frontier open weights, served on custom kernels for codegen.
Chat
GLM-5.2 744BFeatured
morph-glm52-744bSpeed~80 tok/sec
Input$1.10/1M
Cache read$0.22/1M
Output$4.10/1M
Context1M
Qwen 3.5 397B
morph-qwen35-397bSpeed~180 tok/sec
Input$0.50/1M
Cache read$0.30/1M
Output$3.50/1M
Context262k
Qwen 3.6 27B
morph-qwen36-27bSpeed~100 tok/sec
Input$0.29/1M
Output$2.40/1M
Context131k
MiniMax M2.7
morph-minimax27-230bSpeed~90 tok/sec
Input$0.28/1M
Output$1.20/1M
Context196k
MiniMax M3
morph-minimax3-428bSpeed~90 tok/sec
Input$0.60/1M
Output$2.40/1M
Context256k
DeepSeek V4 Flash
morph-dsv4flashSpeed~150 tok/sec
Input$0.14/1M
Output$0.28/1M
Context1M
Gemma 4 31B
morph-gemma4-31bSpeed~120 tok/sec
Input$0.20/1M
Cache read$0.12/1M
Output$0.55/1M
Context175k
| Model | Speed | Input | Cache Read | Output | Context | |
|---|---|---|---|---|---|---|
GLM-5.2 744BFeatured morph-glm52-744b | ~80 tok/sec | $1.10/1M | $0.22/1M | $4.10/1M | 1M | |
Qwen 3.5 397B morph-qwen35-397b | ~180 tok/sec | $0.50/1M | $0.30/1M | $3.50/1M | 262k | |
Qwen 3.6 27B morph-qwen36-27b | ~100 tok/sec | $0.29/1M | — | $2.40/1M | 131k | |
MiniMax M2.7 morph-minimax27-230b | ~90 tok/sec | $0.28/1M | — | $1.20/1M | 196k | |
MiniMax M3 morph-minimax3-428b | ~90 tok/sec | $0.60/1M | — | $2.40/1M | 256k | |
DeepSeek V4 Flash morph-dsv4flash | ~150 tok/sec | $0.14/1M | — | $0.28/1M | 1M | |
Gemma 4 31B morph-gemma4-31b | ~120 tok/sec | $0.20/1M | $0.12/1M | $0.55/1M | 175k |
Specialized Models
Purpose-built APIs that offload the bottlenecks: code editing, search, context management, and per-turn classification.
Fast Apply
Fastest model
morph-v3-fastSpeed10,500+ tok/sec
Price
$0.80/1M in$1.20/1M out
Context262k
Most diversePopular
morph-v3-largeSpeed5,000+ tok/sec
Price
$0.90/1M in$1.90/1M out
Context262k
Code Search
Fast context for agentsNew
morph-warp-grep-v2Price
$0.80/100K
Context100K (1M for Pro)
Compaction
Verbatim context compaction
morph-compactSpeed< 2s P99
Price
$0.20/1M in$0.50/1M out
Context1M
| Model | Speed | Price | Context | |
|---|---|---|---|---|
Fast Apply | ||||
Fastest model morph-v3-fast | 10,500+ tok/sec | $0.80/1M in$1.20/1M out | 262k | |
Most diversePopular morph-v3-large | 5,000+ tok/sec | $0.90/1M in$1.90/1M out | 262k | |
Code Search | ||||
Fast context for agentsNew morph-warp-grep-v2 | — | $0.80/100K | 100K (1M for Pro) | |
Compaction | ||||
Verbatim context compaction morph-compact | < 2s P99 | $0.20/1M in$0.50/1M out | 1M | |
Reflex | ||||
Realtime per-turn classifiers morph-reflex-v1 | < 90ms | $0.00/event$0.00 over 1M/mo | 64K | |
Batch (offline) morph-reflex-v1 | — | $0.00/event$0.00 over 1M/mo | 64K | |
Router
Difficulty-based model routing
morph-routerPrice$0.01/request
Context—
| Model | Price | Context | |
|---|---|---|---|
Difficulty-based model routing morph-router | $0.01/request | — |
Subscriptions
Prepaid credits with volume discounts. Credits apply to all models above.
Free
For testing and personal projects
$0/month
Credits250K($2.50)
LimitsLow rate limits
Starter
For individuals with moderate usage
$20/month
Credits3M($30)
LimitsGenerous rate limits
Pro
For individuals who code all day
$60/month
Credits10M($100)
LimitsGenerous rate limits
Scale
For individuals who can't stop coding
$400/month
Credits80M($800)
LimitsPractically no rate limits
Our fastest endpoints are private deployments
Over 100 billion tokens per day served on dedicated capacity with custom speculators and caching, at large discounts over the public pricing above. Includes SSO, custom rate limits, and priority support.
Get in touch