Magnum V4 72B is Anthracite's language model with a 33K context window and up to 4K output tokens, starting at $3.00 / 1M input and $5.00 / 1M output. A 72B fine-tuned LLM designed to replicate the high-quality prose and creative writing style of Claude Sonnet and Opus.
Specifications
Canonical IDanthracite-org-magnum-4-72b
TypeLanguage
StatusActive
CreatorAnthracite
Providers
Context Window33K tokens
Max Output4K tokens
Input ModalitiesText
Output ModalitiesText
Parameters72B
HuggingFace Likes51
HuggingFace Downloads (30d)26,027
HuggingFace Downloads (all-time)42,131
Release Date · 2 years ago
Knowledge Cutoff · 2 years ago

Capabilities

Input1/5
Text
Image·
Audio·
Video·
PDF·
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities2/13
Reasoning·
Adaptive Reasoning·
Function Calling·
Parallel Function Calling·
Structured Outputs
Native JSON Schema
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching·
Assistant Prefill·

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandard
Input
$ / 1M
Output
$ / 1M
OpenRouter logo
OpenRouter
anthracite-org/magnum-v4-72b
$3.00$5.00

Cost Calculator

US Dollar ($)
Preset:

Cheapest Instances to Run It

Cloud GPU instances that can host Magnum V4 72B, ranked by cheapest on-demand price. The model needs about 173 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
Standard_NP40sAzure4× AMD Alveo U250 FPGA (64GB)256 GB$6.60/hr
g2-standard-96GCP8× nvidia-l4192 GB$7.98/hr
g7e.12xlargeAWS2× RTX PRO Server 6000192 GB$8.29/hr
7 more instances can run Magnum V4 72B
Unlock the full ranked list and FP8 / INT4 quantization with a CloudPrice subscription.

Model IDs

anthracite-org-magnum-4-72b
anthracite-org/magnum-v4-72b