cd/entity/Ollama· home entities Ollama
grep -l @ollama /news/*.json | wc -l → 1004

Ollama

mentions 1004 type Organization page 28/51 feed RSS

// recent coverage 1004 mentions

00:00
2026-07-07
huggingbay.xyz
large-language-models

Qwen/Qwen2.5-1.5B-Instruct

Qwen released Qwen2.5-1.5B-Instruct, an Apache-2.0 licensed 1.5B parameter instruct model for local chat and agent tasks, requiring 8-16 GB RAM/VRAM. The model is available on Hugging Bay with externa…

19:47
2026-07-06
huggingbay.xyz
large-language-models

mradermacher/sarashina2-70b-GGUF

Hugging Bay has hosted the mradermacher/sarashina2-70b-GGUF model, a 445.1 GB quantized version of the sbintuitions/sarashina2-70b base model under the MIT license, with 2 of 15 files verified and sca…

13:36
2026-07-06
vettedconsumer.com
large-language-models

Which Edge Chips Can Run an LLM (and Which Can't)?

A new analysis of edge AI hardware reveals that TOPS (trillions of operations per second) is a poor predictor of large language model (LLM) performance, with memory capacity and bandwidth being the cr…

12:00
2026-07-06
kdnuggets.com
large-language-models

5 Ways Small Language Models Are Powering Next-Gen Agents

Small language models (SLMs) are increasingly powering next-generation AI agents, challenging the assumption that larger models are always better. NVIDIA research shows SLMs excel at repetitive, speci…

00:00
2026-07-06
mbgsec.com
ai-safety

Attackers Don’t Buy Tokens. They Steal Yours.

Security researchers at MBG Security built a global network of honeypots with exposed AI inference and agent endpoints, observing attackers actively scanning for and exploiting vulnerabilities like CV…

00:00
2026-07-05
runagentrun.co.uk
large-language-models

A new workbench for running local AI models

Solo developer released Kivarro, an open-source local inference workbench for running AI models on personal hardware, built on Rust and Tauri and targeting GGUF models. The creator posted it to r/Loca…

12:00
2026-07-04
magnus919.com
artificial-intelligence

Vibe Coding Open Core Out of its Lockbox I: Use the Source

A developer is using AI coding agents to fork the open-core meeting assistant Meetily, replacing its paywalled Pro features with open-source alternatives. The fork, tentatively named LibreMeet, aims t…

10:32
2026-07-04
dev.to
large-language-models

Solving the GPU Pinning Saga and Gemma's Meta-Commentary

Glad Labs fixed a GPU pinning issue where LiteLLM 1.89.2's global api_base override prevented per-model routing, causing vision tasks to cold-load onto the wrong GPU. The team also hardened content gu…

09:55
2026-07-04
dev.to
large-language-models

Scaling LLMs: Why Deterministic Hashing Isn't Enough

A developer built a Go library for semantic LLM caching that combines deterministic hashing with vector similarity search to reduce costs from repeated but differently worded queries. The library supp…

09:00
2026-07-04
letsdatascience.com
ai-tools

PewDiePie Releases Open-Source Odysseus AI Workspace

Felix Kjellberg (PewDiePie) released Odysseus, an open-source, self-hosted AI workspace bundling chat, agents, research, and local model workflows under an AGPL-3.0 license. The project, launched in M…

06:59
2026-07-04
maloyan.xyz
large-language-models

Running Qwen 3.6 Locally on a Mac Mini M4 with 16GB RAM

Qwen open-sourced the 35-billion parameter Mixture of Experts model Qwen 3.6-35B-A3B, which activates only 3 billion parameters per token and runs on a $599 Mac Mini M4 with 16GB RAM at 17 tok/s with …

06:59
2026-07-04
loomcycle.dev
ai-infrastructure

Budgets, costs, and encrypted credentials (v1.9.0 to v1.11.1)

Loomcycle released versions 1.9.0 to 1.11.1 with four major arcs: a security hardening pass closing 17 findings, a new CredentialDef system for encrypted per-tenant secrets using AES-256-GCM, cost att…

06:59
2026-07-04
dev.to
developer-tools

OrinIDE v1.0.8 is here and it's a whole vibe upgrade 🚀

OrinIDE v1.0.8, an AI-powered code editor that runs entirely in the browser without cloud accounts or subscriptions, now supports offline AI models via Ollama and introduces a 4-agent workflow for pla…

← prev page 28 / 51 next →
// co-occurs with top 8 entities
// topics top 6 topics