Phala
US · serverless
OpenAI-compatible
- Category
- serverless
- Free tier
- No
- OpenAI API
- Compatible
- Adds on top
- -
Margin over the underlying model price
What Phala charges
Live prices for widely-hosted open-weight models, sampled from the shared catalogue. Compare the same row against other hosts on each model page.
| Model | Input / 1M | Output / 1M | Context | Throughput |
|---|---|---|---|---|
| Nemotron 3.5 Lightning | $0.080 | $0.200 | 262K | - |
| GLM 5.3 Flash | $0.150 | $0.500 | 1.0M | - |
| Hy3 | $0.150 | $0.640 | 262K | - |
| Qwen3.8 27B | $0.240 | $2.20 | 262K | - |
| Muse Glimmer 30B | $0.300 | $1.10 | 131K | - |
| DeepSeek V4 Flash 0731 | $0.308 | $0.924 | 1.0M | - |
| DeepSeek V4.1 Flash | $0.345 | $1.38 | 1.0M | - |
| GLM 5.3 | $0.910 | $2.86 | 1.0M | - |
| DeepSeek V4 Pro 0813 | $1.02 | $3.05 | 1.0M | - |
| Kimi K3 | $2.85 | $14.25 | 1.0M | - |
Phala alternatives
Other services in the same category. These are genuine substitutes - services in a different category solve a different problem.
Frequently asked
Is Phala OpenAI-compatible?
Yes. Phala accepts the OpenAI chat-completions request shape, so most SDKs work by changing the base URL and API key.
Does Phala have a free tier?
Not currently. Phala bills from the first request.
How does Phala charge?
Per token consumed, priced per model.