Currently listed through these providers:
Model details
MS3.2 24B Magnum Diamond
MS3.2 24B Magnum Diamond is a fine-tuned variant built on the Mistral Small 3.2 24B family, repositioning a larger assistant-tuned base into a more compact creative-writing package. The approach uses rsLoRA fine-tuning on a text-only conversion of Mistral Small 3.2 24B Instruct 2506, keeping the workflow lean while retaining the prose quality associated with the broader Magnum mix. A distinctive step in the pipeline involves pre-tokenization combined with custom loss masking, which shapes how the model attends to narrative cues, character names, and prefill segments. The result is a model tuned to behave more like a literary collaborator than a general-purpose assistant, with attention paid to stylistic consistency across longer passages.
On the deployment side, the model supports a generous context window suitable for multi-chapter drafts, dialogue-heavy scenes, or extended role-play sequences, and it is offered as an open-weight release that developers can self-host or route through inference providers. Independent merged derivatives, such as the arcee-fusion LazyMergekit combination of this model with ReadyArt's Omega Directive variant, suggest that the fine-tune serves as a useful substrate for further experimentation and model merging. For writers and prompt engineers seeking a smaller-footprint alternative to larger Magnum-family models, Magnum Diamond fits naturally into workflows that value character-driven prose, controlled prefill behavior, and flexible integration with standard transformer pipelines running in bfloat16 or float16 precision.
Quick Info
Powered by- Provider
- NanoGPT
- Model key
- Doctor-Shotgun/MS3.2-24B-Magnum-Diamond
- Release date
- Nov 24, 2025
- Last updated
- Nov 24, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.493
- Output token cost
- $0.493
Limits
- Input tokens
- 16,384 tokens
- Output tokens
- 32,768 tokens
- Context window
- 16,384 tokens
Transparent token rates
Compare mistral pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about MS3.2 24B Magnum Diamond
No articles yet. Fetch the latest news to show it here.