cd /news/ai-tools/ollama-confusion-about-model-size · home topics ai-tools article
[ARTICLE · art-114411] src=forum.level1techs.com ↗ pub= topic=ai-tools verified=true sentiment=↓ negative

Ollama: Confusion about model size

A developer criticizes Ollama, a tool for running large language models locally, stating that its model size display ignores context and only accounts for weights, and recommends using llama.cpp directly instead due to Ollama's non-standard model hosting and patched version of llama.cpp with suboptimal flags.

read1 min views1 publishedAug 28, 2026

Yes, it’s not taking the context into account, just the weights alone. You can check it out at huggingface:

Other than that, I’d recommend against using ollama. It does not use huggingface directly and has its own weird model hosting, and it uses a version of llama.cpp with some weird patches of theirs and not optimal flags. If possible, I’d suggest you just to use llama.cpp directly.

── more in #ai-tools 4 stories · sorted by recency
── more on @ollama 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/ollama-confusion-abo…] indexed:0 read:1min 2026-08-28 ·