cd/entity/Apple Metal· home entities Apple Metal
grep -l @apple metal /news/*.json | wc -l → 11

Apple Metal

mentions 11 type Organization feed RSS

// recent coverage 11 mentions

10:53
2026-09-11
dealign.ai
ai-safety

Empirical Research on Moe Safety Mechanisms

Independent researcher Jinho Jang published a cross-scale mechanistic study of safety training in Mixture of Experts reasoning models, reporting 40+ novel findings from 200+ controlled experiments acr…

22:00
2026-08-23
fratepietro.com
artificial-intelligence

Ferrox on Metal: at parity with llama.cpp, and past it

Ferrox, a pure-Rust GGUF inference engine, now runs mixture-of-experts (MoE) prefill on Apple Metal 2.4x faster, reaching 1402 tok/s on OLMoE-1B-7B and closing the gap to llama.cpp from 2.62x behind t…

22:00
2026-08-23
fratepietro.com
artificial-intelligence

Ferrox v0.9.1: a Rust GGUF engine, measured against llama.cpp

Ferrox v0.9.1, a pure-Rust GGUF inference engine, shipped today with MoE prefill on Apple Metal 2.4x faster, closing the gap to llama.cpp from 2.62x behind to 1.11x on OLMoE-1B-7B. The update also imp…

05:08
2026-08-11
sourcefeed.dev
artificial-intelligence

Local Video Generation Gets Its llama.cpp Moment

Salvatore Sanfilippo, creator of Redis, released h3.c, a native C inference engine for MiniMax's open-source H3 video model, rendering video with synchronized stereo audio on Apple Silicon in about 75…

14:09
2026-07-09
byteiota.com
artificial-intelligence

ZML LLMD: Run LLMs on Any Chip — No NVIDIA Required

ZML launched LLMD, a free inference server that runs LLaMA, Gemma, Qwen, and Mistral models on NVIDIA, AMD, Google TPU, Intel, and Apple hardware from a single Docker image. Built in Zig and compiled …

15:02
2026-06-30
theoric.com
machine-learning

We rewrote an ML Framework* in Lean, (and yes it is faster*)

A team rewrote a subset of the TinyGrad deep learning framework in Lean 4, creating TGrad, which outperforms the original in 4 of 5 benchmarks by up to 3x. The project demonstrates Lean 4's viability …

16:42
2026-06-16
flox.dev
machine-learning

Training NanoGPT on Slurm with a Nix-Pinned Environment

A researcher using nanoGPT on a MacBook Pro faces dependency failures when moving to a GPU cluster, prompting a solution using Nix and Flox to create reproducible, cross-platform runtime environments …

05:25
2026-06-16
anil.recoil.org
large-language-models

Language integrated LLMs as an OCaml function

Developer Anil Madhavapeddy released ocaml-deepseek, an OCaml library that integrates DeepSeek's open-weight LLM directly into applications via a native inference engine, enabling local, dependency-fr…

// co-occurs with top 8 entities
// topics top 6 topics