Currently listed through these providers:
Model details
Gemma 4 E4B Instruct (IQ4_XS)
This listing delivers the Gemma 4 E4B Instruct model in an IQ4_XS quantization, a heavily compressed build designed to fit comfortably on consumer hardware while preserving instruct-tuned behavior. The variant is part of the broader Gemma family and is positioned as a compact instruction-following model aimed at developers who want to wire a lightweight text model into local coding agents, IDE plugins, or personal automation workflows. Because it runs through an OpenAI-style local endpoint, it can be plugged into existing developer tools that already speak the OpenAI API without routing traffic through a remote cloud provider.
The model is served through Atomic Chat, an open-source local AI chat application that runs on a user's own computer and exposes an OpenAI-compatible API for agents and developer tools. That local-runtime posture makes the IQ4_XS build well suited to on-device workloads where low memory footprint and easy integration matter more than raw scale, such as code assistance, rewriting, summarization, and conversational helpers inside personal tools. The listing pairs the model with an OpenCode integration and an OpenAI-compatible SDK path, reinforcing its practical fit for local-first developer environments that prioritize compact, quantized open-weight variants over large remote models.
Quick Info
Powered by- Provider
- Atomic Chat
- Model key
- gemma-4-E4B-it-IQ4_XS
- Release date
- Apr 2, 2026
- Last updated
- Apr 2, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
A provider subscription or plan supersedes token-based pricing for this model.
Limits
- Output tokens
- 8,192 tokens
- Context window
- 32,768 tokens
Latest news about Gemma 4 E4B Instruct (IQ4_XS)
No articles yet. Fetch the latest news to show it here.
