Sulat.com
AI models
$10 off the fastest DeepSeek V4.1 Flash, Kimi K3 and GLM 5.3 from Synthetic
NovitaAI logo

Model details

Gemma 3 27B

Gemma 3 27B is Google's flagship open-weight entry in the Gemma 3 family, a multimodal vision-language model that accepts both text and image input while producing text output. It belongs to a lineup of five instruction-tuned variants ranging from 270M to 27B parameters, with the 4B, 12B, and 27B versions sharing multimodal capability and a 128K-token context window. The 27B variant in BF16 occupies roughly 54 GB of VRAM including its SigLIP vision encoder, making it feasible to run on a single high-end GPU such as an H100 without quantization, though production workloads benefit from around 80 GB of headroom to accommodate KV cache and longer contexts.

At launch, Gemma 3 27B ranked in the global Chatbot Arena top ten and was reported to outperform considerably larger open models, positioning it as a strong choice for high-quality production deployments that demand both language understanding and visual reasoning. Beyond raw capability, the model supports over 140 languages and offers structured outputs and function calling, broadening its applicability across diverse enterprise use cases. Its combination of open weights, multimodal input, a long context window, and single-GPU footprint makes it well suited for teams seeking a capable general-purpose model without the infrastructure overhead of much larger frontier systems.

NovitaAIgoogle/gemma-3-27b-itgemma

Quick Info

Powered by
Provider
NovitaAI
Model key
google/gemma-3-27b-it
Release date
Mar 25, 2025
Last updated
Mar 25, 2025
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.119
Output token cost
$0.20

Limits

Output tokens
16,384 tokens
Context window
98,304 tokens

Transparent token rates

Compare Gemma 3 27B pricing

Rates are shown per one million tokens. Combined means one million input plus one million output tokens.

Browse this family

Latest news about Gemma 3 27B

NovitaAI

CoverageBenchmark

Roboflow's Playground page provides independent specification corroboration for Gemma 3 27B: roughly 27 billion parameters, multimodal (text and image input, text output), a 128,000-token context window, 8,192-token max output, instruction-tuned for chat/reasoning/summarization, structured outputs and function calling Roboflow's vision-focused benchmark gives a 58.21% pass rate across 67 visual understanding tasks (ranked 51 of 77), an average response time of 33.60 seconds per task, and roughly $0.080 in inference cost across 65 inferences in the past 30 days. The page explicitly notes the Vision Evals section is legacy and the cur

NovitaAI

CoverageBenchmark

The OpenRouter marketplace listing for Google's Gemma 3 27B IT shows NovitaAI as one of five hosting providers, with provider-specific data: $0.119 per million input tokens, $0.20 per million output tokens, 0.55s latency, 32 tokens/sec throughput, 94.35% uptime, and roughly 9.9% token share. Effective 30-day weighted-a The listing confirms Gemma 3 27B's core capabilities: multimodal vision-language input with text output, up to a 262K context window as reported by OpenRouter (versus Google's documented 128K native, a noted discrepancy), 140+ language support, structured outputs, and function calling, with a March 12, 2025 release and

Videos about Gemma 3 27B

More models around Gemma 3 27B