Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
Atomic Chat logo

Model details

Gemma 4 E4B Instruct (MLX 4-bit)

This deployment belongs to the Gemma 4 family and is distributed as an MLX 4-bit quantized build that pairs the Instruct-tuned variant with an MLX runtime, an Apple Silicon-focused inference path known for running compact, open-weights language models efficiently on local hardware. The model is purpose-built for general natural language generation and conversational text tasks where keeping tooling dependencies minimal matters more than long-form reasoning chains. Because the MLX 4-bit quant preserves the Instruct-tuned behavior of the underlying Gemma 4 E4B checkpoint, it is well suited to instruction following, drafting, summarization, and other text-in, text-out workflows that benefit from a lightweight open-weights footprint without the overhead of larger dense runs.

Practical use cases align with short to mid-length text generation: the model accepts roughly 32.8K tokens of context and returns up to 8K tokens per response, making it a reasonable choice for document-aware drafting, Q&A, and chat sessions that stay within that envelope. Tool use, structured output, reasoning mode, and attachment handling are not exposed in this deployment, so the model fits best as a plain text generation engine rather than an agent orchestrator. Developers can integrate it through an OpenAI-compatible endpoint with adjustable temperature, and the open-weights packaging lets teams self-host or experiment further, while zero input and output pricing through Atomic Chat removes cost as a constraint for high-volume prototyping and evaluation.

Atomic Chatgemma-4-E4B-it-MLX-4bitgemma

Quick Info

Powered by
Provider
Atomic Chat
Model key
gemma-4-E4B-it-MLX-4bit
Release date
Apr 2, 2026
Last updated
Apr 2, 2026
Input modalities
Output modalities
Capabilities

Cost

A provider subscription or plan supersedes token-based pricing for this model.

Limits

Output tokens
8,192 tokens
Context window
32,768 tokens

Latest news about Gemma 4 E4B Instruct (MLX 4-bit)

No articles yet. Fetch the latest news to show it here.

Videos about Gemma 4 E4B Instruct (MLX 4-bit)

More models around Gemma 4 E4B Instruct (MLX 4-bit)