Currently listed through these providers:
Model details
Gemma 4 E4B Instruct (MLX 4-bit)
This deployment belongs to the Gemma 4 family and is distributed as an MLX 4-bit quantized build that pairs the Instruct-tuned variant with an MLX runtime, an Apple Silicon-focused inference path known for running compact, open-weights language models efficiently on local hardware. The model is purpose-built for general natural language generation and conversational text tasks where keeping tooling dependencies minimal matters more than long-form reasoning chains. Because the MLX 4-bit quant preserves the Instruct-tuned behavior of the underlying Gemma 4 E4B checkpoint, it is well suited to instruction following, drafting, summarization, and other text-in, text-out workflows that benefit from a lightweight open-weights footprint without the overhead of larger dense runs.
Practical use cases align with short to mid-length text generation: the model accepts roughly 32.8K tokens of context and returns up to 8K tokens per response, making it a reasonable choice for document-aware drafting, Q&A, and chat sessions that stay within that envelope. Tool use, structured output, reasoning mode, and attachment handling are not exposed in this deployment, so the model fits best as a plain text generation engine rather than an agent orchestrator. Developers can integrate it through an OpenAI-compatible endpoint with adjustable temperature, and the open-weights packaging lets teams self-host or experiment further, while zero input and output pricing through Atomic Chat removes cost as a constraint for high-volume prototyping and evaluation.
Quick Info
Powered by- Provider
- Atomic Chat
- Model key
- gemma-4-E4B-it-MLX-4bit
- Release date
- Apr 2, 2026
- Last updated
- Apr 2, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
A provider subscription or plan supersedes token-based pricing for this model.
Limits
- Output tokens
- 8,192 tokens
- Context window
- 32,768 tokens
Latest news about Gemma 4 E4B Instruct (MLX 4-bit)
No articles yet. Fetch the latest news to show it here.
