Magnum v4 72B
Decision snapshot from the global rankings; composite scoring matches the on-site model board.
Data updated:
Model
1 verified sources
Aggregator quote only — no official sync yet
Next steps on AI Hippo
About this model
Magnum v4 72B ranks #60 on our global model leaderboard. It is listed as a LLM with 33k ctx and a typical snapshot price around $3.75 per 1M tokens. Use Compare for side-by-side picks or the Token hub for verified multi-provider pricing from the same snapshot.
You can also explore more models from Anthracite Org , and browse more options from 🇺🇸 United States .
Capabilities & specs
- Modality
- text->text
- Input modalities
- text
- Knowledge cutoff
- 2024-06-30
- Supported API parameters
- frequency_penalty, logit_bias, logprobs, max_tokens, min_p, presence_penalty, repetition_penalty, response_format, seed, stop, structured_outputs, temperature, top_a, top_k, top_logprobs, top_p
Catalog description: This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).
Token pricing by provider
Compare per-provider token prices for this model across available platforms.
| Provider | Input / 1M tokens | Output / 1M tokens | Latency | Status |
|---|---|---|---|---|
| OpenRouter | $2.50 | $5.00 | — | Verified · 2026-09-30 |
Provider prices are sourced from the token comparison dataset and may change between snapshots.
Alternative picks
-
Meituan: LongCat 2.0
Compare now -
DeepSeek: DeepSeek V4 Flash Latest
Compare now -
Poolside: Laguna S 2.1
Compare now
Pick one or two more models on global rankings and use Compare to view them side by side.
FAQ
Why is the official price missing?
Official rows require a verified direct vendor or cloud API price. When only aggregators list a model, the Official column shows — and verified aggregator quotes appear in the token table.
How is context different from max output tokens?
Context window is how much input the model can accept in a single request. Max output tokens is the per-response generation cap from the primary listing—often much smaller than the context window.
Where do multi-provider prices come from?
Prices are crawled or synced from token providers (OpenRouter, Groq, Together, and others) and merged into the infrastructure comparison dataset. Verified rows show a fetch date in the status column.
Data source & methodology
- Source
- OpenRouter catalog, verified vendor crawls, and Hugging Face hub snapshots.
- Metrics
- Context window, blended 1M-token price, max output tokens, and capability fields from the OpenRouter models API.
- Update cadence
- Daily pipeline refresh; provider prices may change between snapshots.
- Fetched at
- 2026-09-30
- Method
- Composite rank uses on-site context × price weighting (not LMSYS Arena ELO).