Skip to content
LLMs
Magnum v4 72B logo

Magnum v4 72B

by Magnum v4 72B

Magnum v4 72B is a large language model from Magnum v4 72B. It costs $2.50 per million input tokens and $5.00 per million output tokens. Its context window is 33K tokens.

Structured outputOpen weights
Input / 1M tokens
$2.50
Output / 1M tokens
$5.00
Cached input / 1M
-

Not supported

Context window
33K

tokens

Who serves it cheapest

1 hosts serve Magnum v4 72B. Same weights, same API - the price difference is pure margin and routing.

Providers serving Magnum v4 72B, cheapest first
ProviderInput / 1MOutput / 1MContextThroughputUptime 24h
Mancer 2Cheapestfp8$2.50$5.0033K-100.0%

About Magnum v4 72B

This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).

Specifications

Magnum v4 72B specifications
Model IDanthracite-org/magnum-v4-72b
ProviderMagnum v4 72B
Context window33K tokens
Max output4K tokens
Input modalitiestext
Output modalitiestext
Knowledge cutoff2024-06-30
Open weightsYes - anthracite-org/magnum-v4-72b
ReleasedOctober 22, 2024

Cheaper alternatives

Models that cost less than Magnum v4 72B while keeping at least half its context window and every input modality it supports.

Frequently asked

How much does Magnum v4 72B cost?

$2.50 per million input tokens and $5.00 per million output tokens.

What is the context window of Magnum v4 72B?

33K tokens, with up to 4K tokens of output per request.