BaseTen
US · serverless
Model hosting with dedicated deployments and autoscaling for production traffic.
OpenAI-compatible
- Category
- serverless
- Free tier
- No
- OpenAI API
- Compatible
- Adds on top
- None - sets its own prices
Margin over the underlying model price
What BaseTen charges
Live prices for widely-hosted open-weight models, sampled from the shared catalogue. Compare the same row against other hosts on each model page.
| Model | Input / 1M | Output / 1M | Context | Throughput |
|---|---|---|---|---|
| DeepSeek V4 Flash 0731 | $0.130 | $0.260 | 1.0M | - |
| GLM 5.3 Flash | $0.150 | $0.500 | 1.0M | - |
| DeepSeek V4.1 Flash | $0.300 | $1.20 | 1.0M | - |
| Inkling Small | $0.500 | $1.20 | 1.0M | - |
| Inkling | $1.00 | $4.05 | 1.0M | - |
| DeepSeek V4 Pro 0813 | $1.32 | $3.96 | 1.0M | - |
| GLM 5.3 | $1.40 | $4.40 | 1.0M | - |
| Kimi K3 | $3.00 | $15.00 | 1.0M | - |
BaseTen alternatives
Other services in the same category. These are genuine substitutes - services in a different category solve a different problem.
Frequently asked
Is BaseTen OpenAI-compatible?
Yes. BaseTen accepts the OpenAI chat-completions request shape, so most SDKs work by changing the base URL and API key.
Does BaseTen have a free tier?
Not currently. BaseTen bills from the first request.
How does BaseTen charge?
Per token or per GPU-minute Margin over the underlying model price: none - sets its own prices.