Skip to content
LLMs
Wafer logo

Wafer

US · serverless

OpenAI-compatible
Category
serverless
Free tier
No
OpenAI API
Compatible
Adds on top
-

Margin over the underlying model price

What Wafer charges

Live prices for widely-hosted open-weight models, sampled from the shared catalogue. Compare the same row against other hosts on each model page.

Models served by Wafer
ModelInput / 1MOutput / 1MContextThroughput
GLM 5.3 Flash$0.100$0.3501.0M-
DeepSeek V4 Flash 0731$0.100$0.2501.0M-
DeepSeek V4.1 Flash$0.300$1.201.0M-
GLM 5.3$1.19$4.401.0M-
Kimi K3$3.00$12.751.0M-

Wafer alternatives

Other services in the same category. These are genuine substitutes - services in a different category solve a different problem.

Frequently asked

Is Wafer OpenAI-compatible?

Yes. Wafer accepts the OpenAI chat-completions request shape, so most SDKs work by changing the base URL and API key.

Does Wafer have a free tier?

Not currently. Wafer bills from the first request.

How does Wafer charge?

Per token consumed, priced per model.