← All modelsEVIDENCE-LABELLED PROFILE · REVIEWED 2026-08-24

Gemma 4 E2B QAT

Gemma 4 E2B QAT is a curated dense profile for general, writing, reasoning workloads.

Officially supported35/100 · Experimental coverage

The Ollama artifact and published model specifications were checked. Memory figures are conservative MamiLens planning estimates, not a speed result from your device.

  • Official model source
  • Official engine artifact
MAMILENS RAM FLOOR11 GB
ARTIFACT SIZE~4.3 GB
INPUTText + Image + Audio
LICENCEGemma Terms

Will it run?

MamiLens uses 11 GB as a conservative planning floor for this profile and its listed quantization. This is not measured speed. Processor, backend support, context length, memory bandwidth and other applications still affect the result.

Check your complete system →

What is it useful for?

GeneralWritingReasoningCodingVisionAudioTools

Suggested Ollama artifact

Confirm the tag on the official page before downloading, then run:

ollama run gemma4:e2b-it-qat

Start with a representative non-sensitive task. Record the exact runtime, quantization, context, load time, generation speed, memory use and answer quality.

Evidence and limits

Published context: 256K published max. The source establishes specifications; it does not guarantee that maximum context fits at the MamiLens minimum RAM floor.