Gemini 3.8 Flash is the newest entry in Google's Flash-tier lineup, described in the official model card as the next iteration after Gemini 3.7 Flash with focused gains in software engineering and agentic knowledge work. It continues the Flash design philosophy of a versatile workhorse model, pairing configurable reasoning effort levels with broad multimodal input support so the same deployment can handle text, image, video, audio, and PDF inputs while producing text output. Alongside the standard model, Google introduced a Cyber variant tuned for security workloads, sharing the same base intelligence but optimized for vulnerability discovery and patching, with access kept restricted. The release landed alongside updates to tooling such as the Antigravity coding environment, positioning the model as a drop-in engine for production agents and developer assistants.
Independent reviewers note that Gemini 3.8 Flash is reported as Google's most intelligent Flash model to date, with a particularly large jump on coding benchmarks where it scores 90.8% on Terminal-Bench 2.1 (up from 81.6% on the prior Flash version), 54.9% on HLE-Verified, and a top placement on the DeepSWE v1.1 leaderboard for long-horizon software engineering tasks. Beyond raw scores, the model pairs this coding strength with strong general reasoning and tool-calling, making it well-suited for autonomous agents that must read large documents, call external systems, and produce structured outputs. The combination of a one-million-token context window, native multimodal ingestion, and tunable reasoning effort gives teams a flexible default model for mixed workflows ranging from inline code assistance to multi-step research and planning tasks.