cd/entity/Ollama· home entities Ollama
grep -l @ollama /news/*.json | wc -l → 996

Ollama

mentions 996 type Organization page 14/50 feed RSS

// recent coverage 996 mentions

19:46
2026-07-25
promptcube3.com
large-language-models

DeepSeek-R1 Local Deployment: My Hardware Struggles

A user reports that deploying the full 671B parameter DeepSeek-R1 model locally requires over 100GB of VRAM and is impractical on consumer hardware, with CUDA out-of-memory errors occurring even at sm…

17:03
2026-07-25
promptcube3.com
artificial-intelligence

Open-Weight AI: Model Wars vs Ecosystem Wars

Open-weight AI models offer freedom but require significant effort to deploy, according to a technical guide that argues the real value lies in deployment pipelines and developer ecosystems rather tha…

15:45
2026-07-25
promptcube3.com
artificial-intelligence

AI Cyberdeck: Building a Local LLM Workstation

A local LLM workstation build using a Raspberry Pi 5, NVMe SSD, and Hailo-8 M.2 module achieves 26 TOPS of inference for real-time transcription and local RAG, according to the author. The setup, cont…

15:14
2026-07-25
dev.to
artificial-intelligence

I Used the OpenAI SDK—and Claude Answered. Here’s Why.

An engineer demonstrated that the OpenAI Python SDK can be used to call Anthropic's Claude model by pointing the base_url to Anthropic's compatibility endpoint. The post explains the conceptual separa…

12:16
2026-07-25
promptcube3.com
large-language-models

LLM Judges: Accuracy vs Model Size

A test of LLM judges across 20 scenarios found that models under 4B parameters are unreliable for grading other LLMs, with qwen3:0.5b achieving only 61.5% global accuracy versus 92.0% for gemma3 and d…

05:47
2026-07-25
promptcube3.com
artificial-intelligence

Local AI Setup: Coding, RAG, and Voice in 38 Minutes

A developer built a fully local AI stack combining coding assistance, a RAG system, and voice interface in 38 minutes using Ollama, AnythingLLM, and Open WebUI, with no cloud API reliance. The setup p…

00:53
2026-07-25
dev.to
artificial-intelligence

Weekend #2: Scafolding the 3-Way LLM Orchestration

A developer built a 3-way orchestration system between ChatGPT, Claude, and a local LLM to maintain coding momentum despite token limits. The system, called agent-orchestra, uses Git worktrees to run …

19:49
2026-07-24
promptcube3.com
large-language-models

Critical Thinking with LLMs: A Practical Workflow

A practical workflow for using large language models (LLMs) to improve critical thinking shifts from summary prompts to a 'Socratic AI' approach that forces critique, according to a guide on promptcub…

18:47
2026-07-24
promptcube3.com
large-language-models

Local LLM Selection for Mac Mini M4 Pro 24GB

A developer reports that Mistral NeMo 12B is the best local LLM for a Mac Mini M4 Pro with 24GB of unified memory, balancing reasoning and tool calling, but warns that adding a reranker like qllama/bg…

15:35
2026-07-24
dev.to
artificial-intelligence

JSON Schema Doesn't Prevent AI Hallucinations (And That's Okay)

A developer building ShapeCraft, an open-source structured output library, argues that JSON Schema validation prevents malformed responses but does not guarantee factual accuracy, as demonstrated by a…

15:05
2026-07-24
promptcube3.com
artificial-intelligence

Open-Weight Models: Why Big Tech is Fighting for Them

Open-weight models, which release trained parameters for public use, prevent a monopoly on AI intelligence by lowering barriers to entry for developers, according to a joint letter from major tech com…

14:50
2026-07-24
promptcube3.com
artificial-intelligence

Claude Code Workflow: Leveraging Open Weights for Local Dev

A developer reports that shifting to a hybrid AI workflow using Anthropic's Claude 3.5 Sonnet for architectural planning and a local Llama 3.1 8B model for unit test generation reduced token spend by …

← prev page 14 / 50 next →
// co-occurs with top 8 entities
// topics top 6 topics