← All modelsEVIDENCE-LABELLED PROFILE · REVIEWED 20 AUGUST 2026

EmbeddingGemma 300M

Small embedding model for private semantic search and RAG; not a chat model.

Official specifications + independent configuration linked

DeviceTerra has not lab-tested this exact artifact on every hardware combination. RAM and VRAM floors are LocalLens planning estimates, not publisher guarantees.

LOCALLENS RAM FLOOR3 GB
ARTIFACT SIZE~0.3 GB
INPUTText
LICENCEGemma Terms

Will it run?

LocalLens uses 3 GB as a conservative planning floor for this profile and its listed quantization. This is not measured speed. Processor, backend support, context length, memory bandwidth and other applications still affect the result.

Check your complete system →

What is it useful for?

EmbeddingsRAGSearch

Suggested Ollama artifact

Confirm the tag on the official page before downloading, then run:

ollama pull embeddinggemma

Start with a representative non-sensitive task. Record the exact runtime, quantization, context, load time, generation speed, memory use and answer quality.

Evidence and limits

Published context: Embedding context varies by runtime. The source establishes specifications; it does not guarantee that maximum context fits at the LocalLens minimum RAM floor.