Sulat.com
AI models
NanoGPT logo

Model details

MS3.2 24B Magnum Diamond

MS3.2 24B Magnum Diamond is a fine-tuned variant built on the Mistral Small 3.2 24B family, repositioning a larger assistant-tuned base into a more compact creative-writing package. The approach uses rsLoRA fine-tuning on a text-only conversion of Mistral Small 3.2 24B Instruct 2506, keeping the workflow lean while retaining the prose quality associated with the broader Magnum mix. A distinctive step in the pipeline involves pre-tokenization combined with custom loss masking, which shapes how the model attends to narrative cues, character names, and prefill segments. The result is a model tuned to behave more like a literary collaborator than a general-purpose assistant, with attention paid to stylistic consistency across longer passages.

On the deployment side, the model supports a generous context window suitable for multi-chapter drafts, dialogue-heavy scenes, or extended role-play sequences, and it is offered as an open-weight release that developers can self-host or route through inference providers. Independent merged derivatives, such as the arcee-fusion LazyMergekit combination of this model with ReadyArt's Omega Directive variant, suggest that the fine-tune serves as a useful substrate for further experimentation and model merging. For writers and prompt engineers seeking a smaller-footprint alternative to larger Magnum-family models, Magnum Diamond fits naturally into workflows that value character-driven prose, controlled prefill behavior, and flexible integration with standard transformer pipelines running in bfloat16 or float16 precision.

NanoGPTDoctor-Shotgun/MS3.2-24B-Magnum-Diamondmistral

Quick Info

Powered by
Provider
NanoGPT
Model key
Doctor-Shotgun/MS3.2-24B-Magnum-Diamond
Release date
Nov 24, 2025
Last updated
Nov 24, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.493
Output token cost
$0.493

Limits

Input tokens
16,384 tokens
Output tokens
32,768 tokens
Context window
16,384 tokens

Transparent token rates

Compare mistral pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about MS3.2 24B Magnum Diamond

No articles yet. Fetch the latest news to show it here.

Videos about MS3.2 24B Magnum Diamond

More models around MS3.2 24B Magnum Diamond