← All modelsEVIDENCE-LABELLED PROFILE · REVIEWED 2026-08-24

mxbai Embed Large 335M

An embedding model for semantic search and RAG. It does not generate chat answers by itself.

Officially supported35/100 · Experimental confidence

The official artifact path was checked. Memory is a conservative planning estimate; speed and output quality require a test on the target device.

  • Official model source
  • Official engine artifact
LOCALLENS RAM FLOOR7 GB
ARTIFACT SIZE~0.67 GB
INPUTText
LICENCEApache 2.0

Will it run?

LocalLens uses 7 GB as a conservative planning floor for this profile and its listed quantization. This is not measured speed. Processor, backend support, context length, memory bandwidth and other applications still affect the result.

Check your complete system →

What is it useful for?

EmbeddingsRAGSearch

Suggested Ollama artifact

Confirm the tag on the official page before downloading, then run:

ollama pull mxbai-embed-large

Start with a representative non-sensitive task. Record the exact runtime, quantization, context, load time, generation speed, memory use and answer quality.

Evidence and limits

Published context: 512 token model input. The source establishes specifications; it does not guarantee that maximum context fits at the LocalLens minimum RAM floor.