cd/entity/Ollama· home entities Ollama
grep -l @ollama /news/*.json | wc -l → 1016

Ollama

mentions 1016 type Organization page 51/51 feed RSS

// recent coverage 1016 mentions

04:51
2026-05-20
dev.to
artificial-intelligence

Your AI Memory Workspace

Memonia, a new AI memory workspace designed to solve the problem of lost project context between AI tool sessions. It features persistent project memory, task tracking, session reports, and automatic …

01:14
2026-05-20
dev.to
large-language-models

Ollama vs llama.cpp vs vLLM: Which Should You Use in 2026?

This article compares three dominant tools for local LLM inference in 2026: Ollama, llama.cpp, and vLLM. Ollama is recommended for personal, non-technical use due to its ease of setup, while llama.cpp…

01:00
2026-05-20
dev.to
large-language-models

Unload All llama.cpp Router Models Without Restarting

While llama.cpp's router mode allows loading and unloading individual models via HTTP API calls to `/models/unload`, there is no built-in "unload all" endpoint. The recommended approach for unloading …

06:50
2026-05-19
dev.to
artificial-intelligence

gemma4-safe-agent: a tool-using research agent on Gemma 4 e2b

Tool-using research agent built for the Gemma 4 DEV Challenge, which runs locally on the Gemma 4 e2b model via Ollama using roughly 200 lines of Node.js code. The agent accepts a question, selects bet…

00:00
2026-05-14
nobodywho.ooo
large-language-models

What's in a GGUF, besides the weights - and what's still missing?

The GGUF file format consolidates all necessary model components—including weights, chat templates, and special tokens—into a single file, offering a more ergonomic alternative to the scattered JSON f…

00:00
2026-05-11
loopholelabs.io
ai-infrastructure

Ollama Doesn't Know Its GPU Is on Another Machine

Ollama, an AI model server, ran on a MacBook with no NVIDIA GPU by using GTAP software to intercept CUDA calls and forward them to a remote DGX Spark workstation with a 128 GB Blackwell GPU. The setup…

00:00
2026-05-10
jola.dev
large-language-models

Running local models on an M4 with 24GB memory

The article describes the author's successful setup for running local AI models on an M4 Mac with 24GB of memory, specifically highlighting Qwen 3.5-9B (Q4 quantized) as the best performing model at ~…

22:00
2026-03-01
jeffgeerling.com
large-language-models

Expert Beginners and Lone Wolves will dominate this early LLM era

Based on the article, the author, a self-described "AI skeptic," successfully used local LLMs (GPT-OSS 20B and Qwen3 Coder 30B) to migrate blog comments, completing the task in a few evenings. The aut…

11:45
2025-08-01
alexselimov.com
artificial-intelligence

Misc thoughts on AI

The author, a software developer, finds generative AI tools like Claude highly useful for certain tasks, such as Java Spring Boot development, but notes that local models are only adequate and that th…

← prev page 51 / 51
// co-occurs with top 8 entities
// topics top 6 topics