Currently listed through these providers:
Model details
Shisa V2 Llama 3.3 70B
Shisa V2 Llama 3.3 70B is a bilingual chat model built on Meta's Llama-3.3-70B-Instruct foundation, designed to prioritize Japanese language performance while retaining strong English capabilities. Unlike earlier Shisa releases that relied on tokenizer modifications or extended pretraining, this version takes a leaner approach, focusing its optimization entirely on post-training. The model is intended for high-performance bilingual chat, instruction following, and translation tasks across Japanese and English, with native-level fluency as the north star for its Japanese output.
The training recipe centers on a refined mix of supervised fine-tuning and Direct Preference Optimization datasets, encompassing regenerated ShareGPT-style data, translation tasks, roleplaying conversations, and instruction-following prompts. This combination enables the model to handle diverse conversational styles while excelling at structured tasks. On Japanese benchmarks including JA MT Bench, ELYZA 100, and Rakuda, the model achieves leading task performance, demonstrating its strength in nuanced Japanese understanding. It inherits safety characteristics from its base model without additional alignment, and integrates smoothly with inference frameworks like vLLM and SGLang, making it practical for production bilingual applications that demand high-context understanding.
Quick Info
Powered by- Provider
- NanoGPT
- Model key
- shisa-ai/shisa-v2-llama3.3-70b
- Release date
- Jul 26, 2025
- Last updated
- Jul 26, 2025
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.50
- Output token cost
- $0.50
Limits
- Input tokens
- 128,000 tokens
- Output tokens
- 16,384 tokens
- Context window
- 128,000 tokens
Transparent token rates
Compare llama pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Shisa V2 Llama 3.3 70B
No articles yet. Fetch the latest news to show it here.