Sulat.com
AI models
Inference logo

Model details

Google Gemma 3

Gemma 3 is a family of lightweight, instruction-tuned multimodal models released by Google DeepMind, drawing on the same research lineage as the Gemini models. The smaller variants in the family, such as the 1B parameter instruction-tuned build, accept both text and image inputs and produce text outputs, with images normalized to an 896 by 896 resolution and encoded as 256 tokens each. This compact multimodal design lets a single model handle document-style text alongside visual content, which is useful for tasks like image data extraction, visual question answering, and content summarization where a developer wants one API surface instead of separate vision and language pipelines.

Beyond raw inputs, Gemma 3 is positioned as a long-context, multilingual model. The 1B-it variant operates within a 32K token context window and is trained on web documents spanning more than 140 languages, making it relevant for cross-lingual assistants, translation-adjacent workflows, and global content generation. Reported evaluations cover HellaSwag, BoolQ, MMLU, and HumanEval, indicating competence in commonsense reasoning, factuality, STEM problem solving, and code generation. Practically, the combination of a small footprint, open distribution, and multimodal input support makes Gemma 3 a good fit for resource-constrained deployments such as laptops, desktops, or private cloud environments, as well as for educational tools, research baselines, and lightweight conversational applications where latency and cost matter more than top-tier frontier accuracy.

Inferencegoogle/gemma-3gemma

Quick Info

Powered by
Provider
Inference
Model key
google/gemma-3
Release date
Jan 1, 2025
Last updated
Jan 1, 2025
Knowledge cutoff
2024-12
Input modalities
Output modalities
Capabilities

Cost

Input token cost
$0.15
Output token cost
$0.30

Limits

Output tokens
4,096 tokens
Context window
125,000 tokens

Latest news about Google Gemma 3

Inference

CoverageBenchmark

This IoT Digital Twin PLM explainer dated July 24, 2026 consolidates Gemma 3's lineage and deployment profile for developer audiences: it frames Gemma 3 (released March 2025) as a 27-billion-parameter open-weights model with native vision, a 128K-token context window, and quantization-aware-trained checkpoints that the The article is a third-party editorial retrospective, not an official Google or Inference.net source, and several of its claims (including a reported LMArena Elo around 1338 for the instruction-tuned 27B variant and the assertion that "Gemma 4 has shipped") are not corroborated by the supplied evidence set, so readers

Videos about Google Gemma 3

More models around Google Gemma 3