ZML/LLMD 20260929.0 ZML released zml/llmd 20260929.0, adding support for deepseek-ai/DeepSeek-V4.1-Flash on NVIDIA and AMD GPUs, the project's first release supporting a frontier model. The release adds 2 model families, and the 763B-parameter DeepSeek model weighs 510GB, so ZML said it is only available on platforms with enough VRAM to host it natively and that custom quants are not planned in the near future. We’re happy to announce the release of zml/llmd 20260929.0 , which notably brings deepseek-ai/DeepSeek-V4.1-Flash https://hf.co/deepseek-ai/DeepSeek-V4.1-Flash support on NVIDIA and AMD GPUs. This release adds 2 model families: As stated, DeepSeek is only available on platforms with enough VRAM to host it natively. We do not plan to support custom quants in the near future. Obviously the highlight, this is our first release supporting a frontier model. ZML is now mature enough to support complex frontier models, and we wanted to bring it to you. Bear in mind this is a 763B model, weighing 510GB, so for now we only support it on platforms LLMD supports with enough VRAM, namely NVIDIA and AMD GPUs.