Skip to main content

Gemma 4

image-text-to-text

View on GitHub Official site
Llm Models gemma Easy setup 950,000 stars

Overview

Description

Gemma 4 is Google DeepMind's open-weights multimodal model family, engineered with an encoder-free architecture for fast text and image understanding on consumer hardware. Runs efficiently on laptops, edge devices, and local workstations without relying on cloud APIs.

Technical scorecard

License

gemma

Commercial use

No

OpenAI-compatible API

Yes

REST API

Yes

Fine-tuning support

Yes

Quantization support

safetensors, GGUF, ONNX

Docker available

No

GUI / no-code available

No

Telemetry

None

Offline after setup

Yes

Data & Privacy

Does it send data online?

After setup, this listing is marked as usable offline. Confirm network behavior against the upstream project before regulated deployment.

Does it store history?

Not verified in this directory yet. Review the upstream docs for persistence, logs, and workspace storage.

License checks?

The listed license needs review before commercial use.

Telemetry?

None

Last verified: Aug 24, 2026. Maintainer verification should be treated as directory guidance, not legal advice.

🧮

Planning to deploy Gemma 4?

Estimate exact GPU VRAM requirements, context limits, fine-tuning memory, and infrastructure costs for Gemma 4 before committing to hardware.

Calculate Hardware & GPU Sizing →

Setup & Installation

Easy

A developer can usually get this running with standard docs.

# Start with the official project repository
# https://github.com/google-deepmind/gemma

Hardware Requirements

RAM 24.8 GB minimum / 37.2 GB recommended
Hardware tagsGPU (NVIDIA), GPU (Apple Silicon), CPU, TPU
Model formats safetensors, GGUF, ONNX
Languagesen

Works Well With