{"slug": "wait-are-open-weights-effectively-open-source", "title": "Wait, are open weights effectively open source?", "summary": "Open-weight AI models are not open source under the G7's four-tier openness spectrum, with Stephen O'Grady of RedMonk finding that zero of 40 open models qualify as open source AI. However, the ability to create derivative works through distillation means open weights may be more practically useful than full source code, as models like DeepSeek-V3, which required 2.7 million GPU hours to train, can be fine-tuned or distilled by others.", "body_md": "# Wait, are open weights effectively open source?\n\n*\n*\n\nIt annoys me when people call open weight models open source. You get a directory of safetensors files and a license, with no training data, no data pipeline, no training code, and no hyperparameters, so you can’t rebuild the model from what was published. The G7 recently proposed a four-tier openness spectrum, and Stephen O’Grady [checked 40 open models against it](https://redmonk.com/sogrady/2026/06/02/g7-open-weights/) and found that zero of them qualify as open source AI.\n\nThen I thought about distillation and mostly talked myself out of the complaint.\n\n[You can make new models out of them](#you-can-make-new-models-out-of-them)\n\nMaking derivative works is the freedom that actually matters, and open weights give you that completely. You can fine-tune, merge, prune, or distill, running the teacher on your own GPUs to generate training data for a new model that belongs to you. Qwen’s Apache 2.0 and DeepSeek’s MIT licenses explicitly allow it.\n\nThis is [how frontier models get built now](https://huggingface.co/blog/sergiopaniego/distillation-2026). Gemma, Qwen3, DeepSeek-V4, and Nvidia’s Nemotron 3 Ultra all distill from teachers, and Qwen3’s report puts it at roughly 1/10 the GPU hours of RL with better results. DeepSeek shipped the R1-Distill family by pouring R1’s reasoning traces into Qwen and Llama students, and a [recent paper on on-policy distillation](https://arxiv.org/abs/2604.13016) finds the student can pass the teacher when the teacher has genuinely new capabilities to offer.\n\n[The weights might be better than the source](#the-weights-might-be-better-than-the-source)\n\nDeepSeek-V3’s pretraining run took [2.7 million GPU hours](https://arxiv.org/abs/2412.19437). If they’d published the data and the training code instead of the weights, almost nobody could have done anything with it. What they did publish is the compressed result of all that compute, small enough to fit on a USB drive and cheap enough to modify on rented GPUs, so you get to skip the expensive part.\n\nProvenance is the real loss. You can’t see what went into the model, which matters for security and trust, and licenses vary more than people assume since derivatives inherit them (a distill of Llama is still stuck with the Llama community license). The [best objection](https://www.lesswrong.com/posts/c48aC7X2CZgMc9Ytq/open-weights-open-source) to my reasoning here is that me being unable to use the training data doesn’t mean nobody can, and a research group with real compute would get a lot out of having it.\n\nBut if the test is whether a stranger can take your artifact and build a better one without asking you, open weight models pass. The label still bugs me and I have to admit it’s mostly a semantic complaint now.", "url": "https://wpnews.pro/news/wait-are-open-weights-effectively-open-source", "canonical_source": "https://jacob.gold/posts/are-open-weights-effectively-open-source/", "published_at": "2026-07-25 03:52:00+00:00", "updated_at": "2026-08-19 17:55:16.246424+00:00", "lang": "en", "topics": ["artificial-intelligence", "ai-policy", "ai-research"], "entities": ["G7", "Stephen O'Grady", "RedMonk", "Qwen", "DeepSeek", "Nvidia", "Llama"], "alternates": {"html": "https://wpnews.pro/news/wait-are-open-weights-effectively-open-source", "markdown": "https://wpnews.pro/news/wait-are-open-weights-effectively-open-source.md", "text": "https://wpnews.pro/news/wait-are-open-weights-effectively-open-source.txt", "jsonld": "https://wpnews.pro/news/wait-are-open-weights-effectively-open-source.jsonld"}}