cd/entity/Ollama· home entities Ollama
grep -l @ollama /news/*.json | wc -l → 996

Ollama

mentions 996 type Organization page 13/50 feed RSS

// recent coverage 996 mentions

08:52
2026-07-27
jcs.org
ai-tools

On AI

Developer jcs recounts his evolving relationship with AI-assisted coding, from initial skepticism to adopting Claude Code and local models like Ollama, while maintaining a vintage Macintosh programmin…

00:03
2026-07-27
promptcube3.com
artificial-intelligence

Open Source AI: Why the Hype is Real

Open source AI models like Llama 3 and Mistral offer users ownership and control over their AI workflow, eliminating the black box problem of proprietary APIs. Running models locally via tools like Ol…

22:46
2026-07-26
promptcube3.com
artificial-intelligence

KubeAura: A Complete Guide to AI-Powered K8s Management

KubeAura, an open-source tool from developer Ganesh Dev, enables AI-powered Kubernetes cluster management with a zero-deployment philosophy, reading existing kubeconfig to launch a dashboard at http:/…

21:00
2026-07-26
promptcube3.com
artificial-intelligence

Ollama vs LM Studio: Which is Better for Running Local LLMs?

Ollama and LM Studio both enable running local large language models, but Ollama offers better performance and developer integration as a headless background service with a robust API, while LM Studio…

14:46
2026-07-26
promptcube3.com
developer-tools

ThoughtDAG: Editing LLM Context via Graphs

A developer built ThoughtDAG, a prototype that replaces the standard scrollable LLM chat history with a directed acyclic graph (DAG) on an infinite canvas, making context fully editable. Each Q&A exch…

14:20
2026-07-26
promptcube3.com
artificial-intelligence

Llama local deployment for secure code reviews

A developer reports that running Llama 3.1 8B locally on an RTX 3090 enables secure code reviews with full control over context and system prompts, achieving 1.2-second response times on 50-line snipp…

13:46
2026-07-26
promptcube3.com
large-language-models

Local LLM Deployment: A Practical Guide

Deploying large language models locally requires matching hardware to model size, with quantization enabling massive models to run on consumer hardware. Ollama, LM Studio, and vLLM are recommended too…

05:18
2026-07-26
kraghavan.ca
large-language-models

Introduction to LLM Inference

A senior engineer with 11 years of distributed systems experience explains the full LLM inference pipeline, from request arrival to text output, detailing the GGUF file structure and the distinction b…

05:02
2026-07-26
promptcube3.com
ai-infrastructure

Ollama Scout: Testing for Exposed Endpoints

Ollama Scout, a tool from developer Ember2819, exposes a security risk where users bind Ollama to 0.0.0.0 without a reverse proxy or VPN, turning GPUs into free public APIs. The tool probes port 11434…

02:22
2026-07-26
trekhleb.dev
artificial-intelligence

Multiple LLMs trying to reach a consensus

Oleksii Trekhleb built Yes-Brainer, a web app that sends one question to multiple large language models simultaneously and offers three deliberation modes — Parallel, Trial, and Consensus — to surface…

02:11
2026-07-26
byteiota.com
ai-tools

GitHub Models Shuts Down July 30: Migration Guide

GitHub Models, the free AI inference playground launched in September 2024, will shut down permanently on July 30, with no grace period for users relying on its playground, model catalog, inference AP…

20:01
2026-07-25
promptcube3.com
large-language-models

Qwen 27B with 128k Context on 24GB VRAM?

Users report running Qwen 27B with a 128k context window on a single 24GB VRAM card by setting Ollama environment variables OLLAMA_FLASH_ATTENTION=1 and OLLAMA_KV_CACHE_TYPE=q4_0, which quantizes the …

← prev page 13 / 50 next →
// co-occurs with top 8 entities
// topics top 6 topics