Currently listed through these providers:
Model details
Gemma 4 26B A4B IT (NovitaAI)
Gemma 4 26B A4B IT sits within Google's Gemma family of open-weights instruction-tuned models, and through NovitaAI it is positioned as a mid-tier option aimed at developers who want a permissive license without paying flagship prices. Compared with peers on the same NovitaAI price table, its input cost is roughly 1.86 times that of Qwen3 Coder 30B A3B Instruct, while its output rate matches the larger Llama 3.3 70B Instruct, signalling a balanced trade-off between spend and capability for routine assistant workloads.
Practically, the NovitaAI route advertises a 262K-token context window alongside tools and open-weights tags, making it well suited to long-document reasoning, multi-turn agentic flows, and code or retrieval pipelines that benefit from function calling at extended context. The same model identifier is also exposed through Google Vertex on Zenmux, which gives teams flexibility to compare latency, throughput, or cost across providers without changing application code.
Quick Info
Powered by- Provider
- LLM Gateway
- Model key
- novita/gemma-4-26b-a4b-it
- Release date
- Apr 2, 2026
- Last updated
- Apr 2, 2026
- Input modalities
- Output modalities
- Capabilities
Cost
- Input token cost
- $0.13
- Output token cost
- $0.40
Limits
- Output tokens
- 32,768 tokens
- Context window
- 262,144 tokens
Transparent token rates
Compare Gemma 4 26B A4B IT (NovitaAI) pricing
Rates are shown per one million tokens. Combined means one million input plus one million output tokens.
Latest news about Gemma 4 26B A4B IT (NovitaAI)
No articles yet. Fetch the latest news to show it here.
Videos about Gemma 4 26B A4B IT (NovitaAI)
More models around Gemma 4 26B A4B IT (NovitaAI)
This exact model name is also listed by 19 other providers.
