Skip to content
LLMs
Phala logo

Phala

US · serverless

OpenAI-compatible
Category
serverless
Free tier
No
OpenAI API
Compatible
Adds on top
-

Margin over the underlying model price

What Phala charges

Live prices for widely-hosted open-weight models, sampled from the shared catalogue. Compare the same row against other hosts on each model page.

Models served by Phala
ModelInput / 1MOutput / 1MContextThroughput
Nemotron 3.5 Lightning$0.080$0.200262K-
GLM 5.3 Flash$0.150$0.5001.0M-
Hy3$0.150$0.640262K-
Qwen3.8 27B$0.240$2.20262K-
Muse Glimmer 30B$0.300$1.10131K-
DeepSeek V4 Flash 0731$0.308$0.9241.0M-
DeepSeek V4.1 Flash$0.345$1.381.0M-
GLM 5.3$0.910$2.861.0M-
DeepSeek V4 Pro 0813$1.02$3.051.0M-
Kimi K3$2.85$14.251.0M-

Phala alternatives

Other services in the same category. These are genuine substitutes - services in a different category solve a different problem.

Frequently asked

Is Phala OpenAI-compatible?

Yes. Phala accepts the OpenAI chat-completions request shape, so most SDKs work by changing the base URL and API key.

Does Phala have a free tier?

Not currently. Phala bills from the first request.

How does Phala charge?

Per token consumed, priced per model.