Venice
US · serverless
Privacy-first inference that does not retain prompts or completions.
Free tierOpenAI-compatibleUS
- Category
- serverless
- Free tier
- Yes
- OpenAI API
- Compatible
- Adds on top
- None - subscription
Margin over the underlying model price
What Venice charges
Live prices for widely-hosted open-weight models, sampled from the shared catalogue. Compare the same row against other hosts on each model page.
| Model | Input / 1M | Output / 1M | Context | Throughput |
|---|---|---|---|---|
| GLM 5.3 Flash | $0.150 | $0.500 | 1.0M | - |
| DeepSeek V4 Flash 0731 | $0.175 | $0.350 | 1M | - |
| DeepSeek V4.1 Flash | $0.375 | $1.50 | 1M | - |
| Qwen3.8 27B | $0.450 | $3.20 | 262K | - |
| GLM 5.3 | $1.40 | $4.40 | 1M | - |
| DeepSeek V4 Pro 0813 | $1.65 | $4.95 | 1M | - |
| Qwen3.8 2.4T A95B | $2.00 | $6.00 | 262K | - |
Venice alternatives
Other services in the same category. These are genuine substitutes - services in a different category solve a different problem.
Frequently asked
Is Venice OpenAI-compatible?
Yes. Venice accepts the OpenAI chat-completions request shape, so most SDKs work by changing the base URL and API key.
Does Venice have a free tier?
Yes - Venice offers free usage, though limits and eligible models change frequently. Check the pricing page before relying on it.
How does Venice charge?
Subscription or per token Margin over the underlying model price: none - subscription.