cd/entity/ollama· home› entities› ollama
grep -l @ollama /news/*.json | wc -l → 15

ollama

mentions 15 type Organization feed RSS

// recent coverage 15 mentions

23:21
2026-09-21
forum.level1techs.com
ai-tools

I haven't been able to get ollama to recognize B390 on the 358H

A user running CachyOS reports that ollama will not recognize the integrated Intel B390 GPU on the 358H, despite the GPU being detected by the system, and cites the archived IPEX-LLM GitHub repository…

15:08
2026-09-15
thill.me
artificial-intelligence

What I Did at RC

A participant in the Recurse Center summer programming retreat in Brooklyn documented work across four study groups, including Agentic Adventures, where members built a remote sandbox for a "vibecodin…

21:34
2026-09-03
news.ycombinator.com
ai-agents

Show HN: Building AI agents client-side JavaScript

A developer has launched Buttercup, an open-source project demonstrating AI agents built entirely in client-side JavaScript that run in the browser, aiming to reduce infrastructure costs by eliminatin…

15:04
2026-09-03
dev.to
large-language-models

Your monitoring says healthy. Your agents are not.

A developer reported that their local LLM agent fleet experienced silent failures during a 58-day unattended production run, including process errors and timeouts that went unnoticed for days. The roo…

16:46
2026-09-01
promptcube3.com
artificial-intelligence

Stop chasing prompts and start building deterministic systems

AI engineering is shifting from prompt optimization to building deterministic systems with observability and local-first architectures, according to a technical article. The piece advocates for treati…

02:54
2026-08-31
forum.level1techs.com
artificial-intelligence

Ain’t the fastest or prettiest, but it’s still been fun

A hobbyist detailed a three-node home AI lab built around Ryzen CPUs, an RTX 5060 Ti 16GB, and a Radeon AI PRO R9700 32GB, running ComfyUI, Open WebUI, ollama, and Home Assistant, with plans to expand…

14:12
2026-08-26
forum.level1techs.com
large-language-models

Running qwen 3.6 / 3.8 on 3090+3080 over RPC?

A user running Qwen 3.8 27B on a 3090 reports that switching to ninfer, a Qwen-focused inference engine, doubled their token generation speed from 35-40 to 55-60 tokens per second. The user is conside…

22:23
2026-08-11
github.com
ai-tools

Claudish to English

A working prototype plugin for Claude Code, named Claudish to English, rewrites assistant messages into plain English using a local LLM via ollama, and is display-only, preserving the original text in…

07:42
2026-07-30
github.com
developer-tools

Graft

Graft, a tool from Nanonets, reduces coding agent tool calls by 46%, tokens by 42%, and time by 60% in a 162-run controlled benchmark by building a persistent graph of codebase context as linked markd…

19:17
2026-07-11
github.com
developer-tools

Show HN: Catch your local LLM falling back to CPU

Picchio, a new open-source Python tool, diagnoses whether local large language models (LLMs) are actually using the GPU or falling back to CPU by running three inference passes and reporting token spe…

01:13
2026-06-05
umrashrf.github.io
large-language-models

LLM AI Chatbots are letting me down every single day

A daily user of large language model (LLM) chatbots reports that the AI consistently fails to complete complex tasks, delivering only "half-baked" answers that require significant human effort to fini…

03:52
2026-05-21
dev.to
artificial-intelligence

Day 7 - Dense Embedding - RAG

Dense embeddings represent text as continuous numeric vectors (e.g., [0.3455566, 0.6777779]) plotted in a latent space, while sparse embeddings mostly contain zeros and focus on word frequency rather …

// co-occurs with top 8 entities
// topics top 6 topics