Llama 3.2 1B
Llama 3.2 1B is a curated dense profile for general, writing, multilingual workloads.
The Ollama artifact and published model specifications were checked. Memory figures are conservative LocalLens planning estimates, not a speed result from your device.
- Official model source
- Official engine artifact
Will it run?
LocalLens uses 8 GB as a conservative planning floor for this profile and its listed quantization. This is not measured speed. Processor, backend support, context length, memory bandwidth and other applications still affect the result.
Check your complete system →What is it useful for?
Suggested Ollama artifact
Confirm the tag on the official page before downloading, then run:
ollama run llama3.2:1bStart with a representative non-sensitive task. Record the exact runtime, quantization, context, load time, generation speed, memory use and answer quality.
Evidence and limits
Published context: 128K published max. The source establishes specifications; it does not guarantee that maximum context fits at the LocalLens minimum RAM floor.