cd/entity/MLX· home entities MLX
grep -l @mlx /news/*.json | wc -l → 114

MLX

mentions 114 type Organization page 4/6 feed RSS

// recent coverage 114 mentions

08:16
2026-07-28
byteiota.com
artificial-intelligence

Neutrino-1 8B: 763 tok/s Without Standard Quantization

Fermion Research released Neutrino-1 8B, a 3.88 GB language model trained natively with ternary weights ({-1, 0, +1}) that achieves 763 tokens per second on an H100 with speculative decoding and 24.9 …

23:48
2026-07-23
promptcube3.com
artificial-intelligence

MoE Model Layout: A Deep Dive into Weight Reordering

A new weight reordering technique for Mixture-of-Experts (MoE) models optimizes physical file layout to minimize NVMe seek overhead, achieving massive gains in explicit-read inference engines like MLX…

18:26
2026-07-23
developer.apple.com
machine-learning

Explore distributed inference and training with MLX [video]

Apple's MLX team introduced distributed inference and training capabilities for machine learning workloads across multiple Macs at WWDC26, demonstrating how to scale large language models using RDMA o…

13:49
2026-07-23
dev.to
machine-learning

Take your benchmark to the people who can kill it

A developer's MoE model file reordering technique, mbolt, achieved +32.3% decode throughput and −26.3% time-to-first-token on a 235B-parameter model running on a 48 GB MacBook, as independently verifi…

21:10
2026-07-22
byteiota.com
artificial-intelligence

BaseRT: Run Local LLMs on Apple Silicon 6x Faster

A new LLM inference runtime called BaseRT achieves up to 6.4x faster local inference on Apple Silicon than llama.cpp by writing directly to Apple's Metal GPU API, skipping intermediate frameworks like…

14:22
2026-07-21
simonwillison.net
artificial-intelligence

Nativ: Run AI models locally on your Mac

Prince Canuma's Nativ is a new macOS desktop application that wraps Apple's MLX framework to run AI models locally, offering both a chat interface and a localhost API server. The app automatically det…

22:59
2026-07-15
eaon.dev
ai-tools

Eaon (Preview)- Private all in one AI super app

Eaon, a free and open-source native macOS app, lets users switch between 49 AI models from a single interface, supporting both cloud keys and fully offline local models via Ollama, llama.cpp, or MLX. …

19:04
2026-07-14
sourcefeed.dev
artificial-intelligence

Bonsai 27B Puts Real Agents on Phones

PrismML shipped Bonsai 27B, a 27B-parameter model based on Qwen3.6 27B, with a 1-bit variant packing to 3.9 GB that fits on an iPhone 17 Pro, clearing the memory gate that blocked prior builds of this…

← prev page 4 / 6 next →
// co-occurs with top 8 entities
// topics top 6 topics