ZML/LLMD 20260929.0
ZML released zml/llmd 20260929.0, adding support for deepseek-ai/DeepSeek-V4.1-Flash on NVIDIA and AMD GPUs, the project's first release supporting a frontier model. The release adds 2 model families,β¦
ZML released zml/llmd 20260929.0, adding support for deepseek-ai/DeepSeek-V4.1-Flash on NVIDIA and AMD GPUs, the project's first release supporting a frontier model. The release adds 2 model families,β¦
ZML, a Paris startup, released LLMD on July 8, a free, open-source inference server that runs LLaMA, Gemma, Qwen, and Mistral on NVIDIA CUDA, AMD ROCm, Google TPU, Intel oneAPI, and Apple Metal from aβ¦
ZML released ZML/LLMD, an inference server written in Zig that runs LLaMa, Gemma, Qwen, and Mistral LLMs on five architectures including NVIDIA CUDA, AMD ROCm, Google TPU, Intel oneAPI, and Apple Metaβ¦
ZML launched LLMD, a free inference server that runs LLaMA, Gemma, Qwen, and Mistral models on NVIDIA, AMD, Google TPU, Intel, and Apple hardware from a single Docker image. Built in Zig and compiled β¦
ZML released ZML/LLMD on July 8, 2026 as an alpha inference server for LLaMa, Gemma, Qwen and Mistral models across NVIDIA CUDA, AMD ROCm, Google TPU, Intel oneAPI and Apple Metal targets, aiming to rβ¦