Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Kilo Gateway logo

Model details

Qwen: Qwen3 8B (retires Oct 9)

Qwen3 8B is a dense causal language model built with 8.2 billion parameters, designed to serve as a flexible solution for both complex analytical tasks and everyday conversation. Its architecture features 36 layers and utilizes Grouped Query Attention to balance performance and efficiency. A defining characteristic of this model is its ability to toggle between a specialized thinking mode—optimized for mathematics, coding, and logical inference—and a non-thinking mode for standard dialogue. This dual-mode design allows users to adapt the model's behavior to the specific demands of a task, ensuring high-quality outputs whether the goal is creative writing, role-playing, or solving intricate agent-based problems.

Developed through extensive pre-training and post-training, the model demonstrates significant advancements in instruction following and human preference alignment. It is engineered to excel in multilingual environments, providing robust support for over 100 languages and dialects. Beyond its core conversational strengths, the model is built for practical agent integration, allowing it to interact effectively with external tools. With a native context window of 32,768 tokens that can be extended to 131,072 tokens using YaRN scaling, the model is well-positioned for long-context applications, making it a powerful tool for developers and users seeking a balance of reasoning depth and operational versatility.

Kilo Gatewayqwen/qwen3-8bqwen

Quick Info

Powered by
Provider
Kilo Gateway
Model key
qwen/qwen3-8b
Release date
Apr 28, 2025
Last updated
Apr 28, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.117
Output token cost
$0.455

Limits

Output tokens
8,192 tokens
Context window
131,072 tokens

Transparent token rates

Compare Qwen: Qwen3 8B (retires Oct 9) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Qwen: Qwen3 8B (retires Oct 9)

No articles yet. Fetch the latest news to show it here.

Videos about Qwen: Qwen3 8B (retires Oct 9)

More models around Qwen: Qwen3 8B (retires Oct 9)