cd/entity/Ollama· home› entities› Ollama
grep -l @ollama /news/*.json | wc -l → 1526

Ollama

mentions 1526 type Organization page 16/77 feed RSS

// recent coverage 1526 mentions

15:53
2026-09-11
blog.kilo.ai
large-language-models

How to Choose a Local LLM: Models, Hardware, and Quantization

Atomic Chat published a guest blog guide on selecting local large language models, recommending that users start with a GGUF Q4_K_M quantization if it fits their hardware. The guide provides memory-es…

12:20
2026-09-11
github.com
ai-agents

Pizza Bot: a local-first inbox for long-running AI agents

Amazon released Pizza Bot, an open-source, local-first inbox for long-running AI agents, under the Apache 2.0 license. The tool runs on a stateful DeepAgents/LangGraph runtime with an Electron desktop…

11:49
2026-09-11
ory.com
ai-tools

Make Claude Code Faster and Cheaper with Ory Lumen

Ory released Ory Lumen, an open-source local semantic code search engine that runs as an MCP server alongside Claude Code and cut Claude Code runtime by up to 53% and API costs by up to 39% in SWE-ben…

10:17
2026-09-11
dev.to
ai-tools

KPIAssembler: stop hand-picking KPIs, let AI propose them

Developer Akshat Srivastava released KPIAssembler, an open-source Ruby gem that uses an LLM to propose KPI metric SQL recipes while deterministic Ruby code certifies which ones are safe to publish. Th…

09:09
2026-09-11
dev.to
large-language-models

KV Cache on 16 GB GPUs: Making Long Context Actually Fit

A developer published a practical guide to fitting long-context LLM inference on 16 GB GPUs by budgeting VRAM for the KV cache, which grows with every active token and sequence. The guide provides a K…

01:57
2026-09-11
github.com
ai-agents

Show HN: Kern Agent – See inside your agent's brain

Developer Oguz Bilgic released Kern Agent, an open-source autonomous agent framework that runs on a user's own machine and pairs with the agent-kernel project to manage memory, tools, and multi-channe…

00:45
2026-09-11
tensorsandtokens.com
large-language-models

Setting up OpenCode with Ollama and sbx on Mac

A developer guide details running local LLMs with OpenCode, Ollama, and Docker Sandboxes (sbx) on an Apple MacBook Pro M5 with 48GB of memory, pulling Qwen 3.8 27B mxfp8 (32GB) and Gemma 4 31B mxfp8 (…

23:31
2026-09-10
blog.lewman.com
ai-infrastructure

A $537 Local LLM Machine (2025)

A refurbished Miniforum UM790 Pro with 64GB RAM and 1TB NVMe, powered by an AMD Ryzen 9 7940HS CPU with Radeon 780M GPU, was purchased for $537 and used to run local LLMs via Ollama and Open Web UI on…

23:18
2026-09-10
promptcube3.com
ai-tools

Pick the right AI coding tool for your stack

A head-to-head test of Cursor, Windsurf, and Claude Code on a 40k-line Next.js project found Cursor's Composer needed 4 tries to fix a React useEffect race condition, Windsurf's Flow took 2 attempts, …

14:56
2026-09-10
firethering.com
ai-products

LLM Wiki: An AI Knowledge Base That Builds Itself

Developer nash_su released LLM Wiki v0.6.11, a GPLv3 open-source desktop application for Windows, macOS and Linux that uses an LLM to build a persistent, structured wiki from ingested documents. The 4…

05:40
2026-09-10
dev.to
ai-tools

Open WebUI: Chat Interface & Context Options

A developer detailed the core features of Open WebUI v0.9.6, an open-source chat interface that began as a front end for local Ollama instances and has expanded to support any OpenAI API provider plus…

21:58
2026-09-09
promptcube3.com
ai-tools

Dify tutorial, best AI community to join

A developer's tutorial argues that Dify, an open-source AI workflow tool, is more efficient than building complex LangChain wrappers, claiming a RAG pipeline that took three days to code manually can …

20:29
2026-09-09
promptcube3.com
ai-tools

Build a local AI coding setup with Llama and Cursor

A developer reports setting up a local AI coding environment using Llama 3.1 8B via Ollama and Cursor in about 15 minutes, avoiding the $20 monthly subscription and preventing cloud data leaks. The se…

18:48
2026-09-09
promptcube3.com
developer-tools

Solving the Qwen Coder local setup lag in VS Code

A developer reports that running Qwen2.5-Coder-32B via Ollama in VS Code caused 2-3 second latency, fixed by switching to the q4_K_M quantized version and setting num_ctx to 16384, improving response …

← prev page 16 / 77 next →
// co-occurs with top 8 entities
// topics top 6 topics