Currently listed through these providers:
Model details
Llama 3.1 Euryale 70B v2.2
Llama 3.1 Euryale 70B v2.2 is a Sao10K fine-tune in the Llama 3.1 70B lineage, distributed as an open-weight text model and routed through the OpenRouter unified API alongside related Sao10K variants such as Llama 3.3 Euryale 70B and Llama 3.1 70B Hanami x1. Independent listings characterize it as a text-only conversational model aimed at general chat, analysis, and production workloads, leaning on the Llama 3 tokenizer family. It sits in the broader Euryale series of community-tuned 70B releases, where each iteration aims to refine instruction-following and dialogue quality on top of the base architecture rather than introduce new pretraining scale.
In practice the model is positioned for medium-to-long context tasks, with a roughly 131K token memory window that comfortably fits long documents, extended transcripts, or multi-file code reviews without aggressive truncation, while replies are capped at a 16K token ceiling suitable for detailed multi-section answers. Its pricing is symmetric for input and output at $the listed price per million tokens, with discounted batch and cached tiers, making it attractive for steady, high-volume production pipelines where predictable per-call cost matters more than peak reasoning depth. Compared with sibling Sao10K models, usage telemetry shows it holding a consistent middle-tier share of tokens served, suggesting it is a dependable generalist rather than a flagship reasoning release.
Quick Info
Powered by- Provider
- OpenRouter
- Model key
- sao10k/l3.1-euryale-70b
- Release date
- Aug 28, 2024
- Last updated
- Aug 28, 2024
- Knowledge cutoff
- 2023-12-31
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.85
- Output token cost
- $0.85
Limits
- Output tokens
- 16,384 tokens
- Context window
- 131,072 tokens