← All modelsEVIDENCE-LABELLED PROFILE · REVIEWED 2026-08-24

Llama 4 Maverick 400B

Large multimodal MoE model intended for multi-GPU servers and very high-memory systems.

Officially supported35/100 · Experimental coverage

The official artifact path was checked. Memory is a conservative planning estimate; speed and output quality require a test on the target device.

  • Official model source
  • Official engine artifact
MAMILENS RAM FLOOR320 GB
ARTIFACT SIZE~245 GB
INPUTText + Image
LICENCELlama 4 Community

Will it run?

MamiLens uses 320 GB as a conservative planning floor for this profile and its listed quantization. This is not measured speed. Processor, backend support, context length, memory bandwidth and other applications still affect the result.

Check your complete system →

What is it useful for?

GeneralVisionCodingMultilingualResearch

Suggested Ollama artifact

Confirm the tag on the official page before downloading, then run:

ollama run llama4:maverick

Start with a representative non-sensitive task. Record the exact runtime, quantization, context, load time, generation speed, memory use and answer quality.

Evidence and limits

Published context: Very long context; practical limit is hardware-dependent. The source establishes specifications; it does not guarantee that maximum context fits at the MamiLens minimum RAM floor.