Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
LLM Gateway logo

Model details

Gemma 4 26B A4B IT (NovitaAI)

Gemma 4 26B A4B IT sits within Google's Gemma family of open-weights instruction-tuned models, and through NovitaAI it is positioned as a mid-tier option aimed at developers who want a permissive license without paying flagship prices. Compared with peers on the same NovitaAI price table, its input cost is roughly 1.86 times that of Qwen3 Coder 30B A3B Instruct, while its output rate matches the larger Llama 3.3 70B Instruct, signalling a balanced trade-off between spend and capability for routine assistant workloads.

Practically, the NovitaAI route advertises a 262K-token context window alongside tools and open-weights tags, making it well suited to long-document reasoning, multi-turn agentic flows, and code or retrieval pipelines that benefit from function calling at extended context. The same model identifier is also exposed through Google Vertex on Zenmux, which gives teams flexibility to compare latency, throughput, or cost across providers without changing application code.

LLM Gatewaynovita/gemma-4-26b-a4b-itgemma

Quick Info

Powered by
Provider
LLM Gateway
Model key
novita/gemma-4-26b-a4b-it
Release date
Apr 2, 2026
Last updated
Apr 2, 2026
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.13
Output token cost
$0.40

Limits

Output tokens
32,768 tokens
Context window
262,144 tokens

Transparent token rates

Compare Gemma 4 26B A4B IT (NovitaAI) pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemma 4 26B A4B IT (NovitaAI)

No articles yet. Fetch the latest news to show it here.

Videos about Gemma 4 26B A4B IT (NovitaAI)

More models around Gemma 4 26B A4B IT (NovitaAI)