Currently listed through these providers:
Model details
Qwen2.5 72B Instruct
Qwen2.5 72B Instruct is an instruction-tuned large language model in the Qwen2.5 family, positioned as a flagship model for general-purpose text tasks, with particular strength in multilingual reasoning, coding, and mathematical problem solving. Independent descriptive sources frame it as a flexible conversational and analytical assistant aimed at language understanding and practical interactive tasks, where it can serve as a single workhorse for chat, drafting, code generation, and structured problem solving across many languages.
At the architectural level, the model is built on a dense Transformer decoder that combines Grouped Query Attention, SwiGLU activation, and modified Rotary Positional Embeddings to improve efficiency and scalability. Post-training follows a two-stage alignment pipeline that pairs supervised fine-tuning with reinforcement learning, a recipe intended to produce robust instruction following and stable behavior on complex prompts. This blend of attention and normalization refinements with explicit preference-based tuning gives the model a balance of broad knowledge and controlled, task-aware responses that suit enterprise assistants, developer tooling, and research workflows.
Quick Info
Powered by- Provider
- Kilo Gateway
- Model key
- qwen/qwen-2.5-72b-instruct
- Release date
- Sep 19, 2024
- Last updated
- Sep 19, 2024
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.36
- Output token cost
- $0.40
Limits
- Output tokens
- 16,384 tokens
- Context window
- 32,768 tokens
Transparent token rates
Compare Qwen2.5 72B Instruct pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Qwen2.5 72B Instruct
No articles yet. Fetch the latest news to show it here.