cd/entity/Metal· home entities Metal
grep -l @metal /news/*.json | wc -l → 36

Metal

mentions 36 type Organization page 2/2 feed RSS

// recent coverage 36 mentions

05:32
2026-06-29
squish.run
large-language-models

Squish – The fastest way to run local LLMs on Apple Silicon

Squish, a new local AI agent runtime for Apple Silicon, claims to load models 54× faster than standard paths and serve them faster than Ollama, with full offline capability and no cloud dependencies. …

04:04
2026-06-26
devclubhouse.com
artificial-intelligence

Apple Fast-Tracks M7 Silicon to Rewrite On-Device AI Limits

Apple is skipping the high-end M6 Pro and Max chips to fast-track its M7 family, aiming to deliver a massive leap in memory bandwidth and Neural Engine capacity for on-device AI. The base M7 will offe…

17:36
2026-06-25
devclubhouse.com
large-language-models

Quantize and Run Llama 3.2 on Apple Silicon with llama.cpp

Mariana Souza published a tutorial on quantizing and running Meta's Llama 3.2 3B model on Apple Silicon using llama.cpp with Metal GPU acceleration, achieving local inference with Q4_K_M quantization.…

05:22
2026-06-25
dev.to
artificial-intelligence

How to Transcribe Meetings Locally in 2026 (Whisper, On-Device)

Off Grid AI Desktop is a free, open-source app that records and transcribes meetings locally on Mac or PC using OpenAI's Whisper model, ensuring audio never leaves the machine and eliminating per-minu…

05:10
2026-06-25
dev.to
artificial-intelligence

How to Run Local AI on Your Mac in 2026 (No Cloud, No Account)

Off Grid AI Desktop is a free, open-source app that runs chat, image generation, and voice directly on Apple Silicon Macs, using unified memory and Metal GPU acceleration. It supports models like llam…

14:52
2026-06-24
sipp.sh
large-language-models

Show HN: Sipp – Run small local LLMs in browser 3x faster

Sipp, an open-source AI inference library, enables running small local LLMs in browsers with up to 3x faster decode speeds than alternatives. Built by HCI and graphics programmers, it uses a unified A…

09:38
2026-06-23
en.andros.dev
developer-tools

I built a GPU back end for Emacs

A developer built a GPU display backend for Emacs, creating Metal and OpenGL drivers that offload text rendering from the CPU to the GPU, enabling video playback and animated cursor effects without mo…

09:36
2026-06-18
dev.to
developer-tools

llama-bench skipped FA on capable GPUs — b9437 corrects it

Build b9437 of llama.cpp fixes two default-value bugs in llama-bench that caused flash attention to be skipped on capable GPUs and GPU-layer count to use a legacy sentinel. The flash attention flag no…

06:55
2026-06-17
github.com
large-language-models

Native Inference Engine for macOS 14 or newer

Embershard, a macOS chat app with its own LLM inference engine, has been released in beta v0.1.1 for Apple Silicon devices running macOS 14 or newer. The app bypasses llama.cpp for inference, instead …

17:12
2026-06-03
github.com
ai-tools

Gooey: A GPU-accelerated UI framework for Zig

Gooey, a GPU-accelerated UI framework for the Zig programming language, has been released as an open-source project targeting macOS, Linux, and browser platforms. The framework provides declarative UI…

02:24
2026-05-28
dev.to
large-language-models

Quantizing Gemma 4 on Mac with llama.cpp

A developer successfully quantized Google's Gemma 4 model to 4-bit precision using llama.cpp on a Mac with Metal acceleration. The process involved converting the model to GGUF format and applying the…

17:15
2025-12-31
gist.github.com
developer-tools

Running Minecraft on macOS with Zink + KosmicKrisp

This article describes a method to run OpenGL 4.6 applications, specifically Minecraft, on macOS by translating OpenGL calls through a chain of graphics APIs: OpenGL → Vulkan → Metal. The process uses…

← prev page 2 / 2
// co-occurs with top 8 entities
// topics top 6 topics