Currently listed through these providers:
Model details
Meta: Llama 3.2 1B Instruct
The Llama 3.2 1B Instruct is a compact, instruction-tuned language model from Meta's 3.2 generation, built with multilingual dialogue as a core design priority. Rather than pursuing raw scale, this model targets efficient instruction-following and text generation that works well across languages, with particular strength in agentic retrieval and summarization tasks. The focus on lightweight deployment makes it suitable for scenarios where a smaller footprint matters without sacrificing meaningful capability.
As part of the broader Llama 3.2 collection, this model has undergone instruction-tuning to sharpen its responses for chat and task-oriented use. The architecture supports fine-tuning through methods like LoRA, allowing developers to adapt the base model with their own data while preserving its efficiency. Sources indicate it outperforms many competing open-source and closed chat models on standard benchmarks, positioning it as a strong choice for developers who need reliable, commercially-ready performance in a compact form factor that can be accelerated on modern GPU infrastructure.
Quick Info
Powered by- Provider
- Kilo Gateway
- Model key
- meta-llama/llama-3.2-1b-instruct
- Release date
- Sep 25, 2024
- Last updated
- Sep 25, 2024
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.027
- Output token cost
- $0.201
Limits
- Output tokens
- 54,000 tokens
- Context window
- 60,000 tokens
Transparent token rates
Compare Meta: Llama 3.2 1B Instruct pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Meta: Llama 3.2 1B Instruct
No articles yet. Fetch the latest news to show it here.