Compared llama3.2:1b vs llama3.2:3b Memory Footprint Ollama's llama3.2:3b model consumes 2.5 GB of memory (2.47 GB RSS) versus 1.5 GB (1.24 GB RSS) for the 1b model, showing that tripling parameters roughly doubles memory footprint. The test also revealed that the larger model without a custom system prompt reverts to default kubectl-based answers, confirming that domain-specific behavior requires prompt engineering regardless of model size. Context: The number in a model name like 1b or 3b refers to parameters — roughly, the tunable values inside the model that encode what it’s learned. More parameters generally means better reasoning and more nuanced answers, at the cost of more memory and slower responses. One thing that trips people up early with Ollama: typing ollama run