Which Open AI Models People Download: September 2026 Data Qwen repositories account for 27 of the top 60 most-downloaded open AI models on the Hugging Face Hub, according to a September 18, 2026 query of the Hub API sorted by rolling 30-day downloads, with Qwen/Qwen3-0.6B leading at 22,498,727 downloads. Counting community repacks, 36 of the 60 rows are Qwen-based and total 240 million downloads, roughly seven times the next family, Gemma. The dataset also flags three test fixtures in the top 15 — openai-community/gpt2 at 15,439,333 downloads, a tiny Qwen2 test model and OPT-125m — whose downloads come from CI pipelines rather than users, and notes 17 of the 60 rows are quantised repacks and 44 carry an Apache 2.0 licence. Ask which open AI models people actually use and most answers are opinions. Download counts are not a perfect answer, but they are a measured one. On September 18, 2026 we queried the Hugging Face Hub, the main public host for open model weights, for the most-downloaded models in its three language and vision-language categories and kept the top 60 by downloads over the previous 30 days. Qwen repositories hold 27 of the 60 rows. Qwen3-0.6B, a model small enough to run on a phone, leads at 22.5 million. This page is for anyone choosing an open model who wants to know what the rest of the ecosystem is pulling, and for writers who need a dated, reproducible adoption table rather than a vibe. It explains what a Hugging Face download does and does not count, prints all 60 rows, sums them by model family, and says plainly where downloads stop being evidence. The query is stated in the methodology so anyone can repeat it. 1. 01Qwen dominates the top 60.27 rows are Qwen's own repositories and 36 are Qwen-based once community repacks are counted. Those 36 rows total 240 million downloads, about seven times the next family, Gemma. 2. 02Small models win on raw downloads.Nine of the 17 non-fixture models in the top 20 are 9B parameters or fewer, and four more are mixture-of-experts models with 3 to 4B active. Downloads scale with how many machines a model fits on. 3. 03Test fixtures are in the top 15, and they are flagged.GPT-2, a tiny Qwen2 test model and OPT-125m are downloaded by CI pipelines, not users. We show them because the raw ranking includes them, and mark them so nobody quotes them as adoption. 4. 04Seventeen rows are repacks; 44 are Apache 2.0.GGUF, FP8, NVFP4, AWQ and MLX conversions of other models fill 17 of 60 rows. The licence mix is 44 Apache 2.0, 6 MIT, 4 custom, 3 Meta or Google model licences, 1 OpenRAIL and 2 unstated. 01 — DefinitionsWhat a download counts The Hugging Face Hub reports a "downloads" number per model repository. In the Hub API this figure is a rolling count over the last 30 days, which is why it moves every day and why the date of collection matters. It counts file fetches from the repository, which means a person downloading a model once, a server pulling weights on every fresh deployment, and a continuous-integration job fetching a tiny model to run a test all add to the same number. It does not count a model that was downloaded once and then copied around a company, and it does not count use through an API. The Hub's API documentation https://huggingface.co/docs/hub/api describes the endpoint and its sort options. Three consequences follow, and they shape how to read the table. Small models rank high because they fit in more places and are cheap to re-download. Test fixtures rank high because CI runs many times a day. And a quantised repack of a popular model counts separately from the original, so a model family's real reach is the sum of its rows, not its best one. We flag the fixtures with a dagger, keep every row otherwise, and sum by family in section 03. 02 — DatasetThe top 60, September 2026 Rank is by rolling 30-day downloads at collection time. Model is the repository name, organisation included, exactly as the Hub returns it. Licence is the tag the repository declares; "Other custom " is the Hub's own category for a licence file that is not a standard one. A dagger marks a repository we consider a test fixture: a model whose downloads come mainly from automated test suites. | Source: Hugging Face Hub API, /api/models sorted by downloads, pipeline tags text-generation, image-text-to-text and any-to-any, queried September 18, 2026 09:28 UTC. † = test fixture, shown but not adoption. | | | | |---|---|---|---| | | Model repository | 30-day downloads | Licence | |---|---|---|---| | 1 | Qwen/Qwen3-0.6B | 22,498,727 | Apache 2.0 | | 2 | Qwen/Qwen3-VL-8B-Instruct | 19,098,599 | Apache 2.0 | | 3 | openai-community/gpt2 † | 15,439,333 | MIT | | 4 | trl-internal-testing/tiny-Qwen2ForCausalLM-2.5 † | 14,140,672 | Not stated | | 5 | Qwen/Qwen3-8B | 12,988,756 | Apache 2.0 | | 6 | unsloth/Qwen3-Coder-30B-A3B-Instruct-GGUF | 12,752,716 | Apache 2.0 | | 7 | Qwen/Qwen3.6-35B-A3B-FP8 | 10,210,602 | Apache 2.0 | | 8 | google/gemma-4-26B-A4B-it | 9,794,984 | Apache 2.0 | | 9 | Qwen/Qwen2.5-7B-Instruct | 9,714,062 | Apache 2.0 | | 10 | Qwen/Qwen3.5-9B | 9,285,215 | Apache 2.0 | | 11 | google/gemma-4-31B-it | 9,045,393 | Apache 2.0 | | 12 | nvidia/Qwen3.6-35B-A3B-NVFP4 | 8,594,317 | Apache 2.0 | | 13 | Qwen/Qwen2.5-0.5B-Instruct | 8,490,193 | Apache 2.0 | | 14 | Qwen/Qwen3.8-27B-FP8 | 7,603,458 | Apache 2.0 | | 15 | facebook/opt-125m † | 7,486,472 | Other custom | | 16 | Qwen/Qwen3.8-27B | 7,456,257 | Apache 2.0 | | 17 | farbodtavakkoli/OTel-2.0-LLM-31B-IT | 7,407,454 | Apache 2.0 | | 18 | Qwen/Qwen3-4B | 7,242,404 | Apache 2.0 | | 19 | Qwen/Qwen2.5-1.5B-Instruct | 7,184,297 | Apache 2.0 | | 20 | Qwen/Qwen3.5-4B | 6,958,404 | Apache 2.0 | | 21 | meta-llama/Llama-3.2-1B-Instruct | 6,921,311 | Llama 3.2 | | 22 | Qwen/Qwen2.5-VL-7B-Instruct | 6,851,431 | Apache 2.0 | | 23 | openai/gpt-oss-20b | 6,675,065 | Apache 2.0 | | 24 | Qwen/Qwen3.6-27B-FP8 | 6,150,001 | Apache 2.0 | | 25 | meta-llama/Llama-3.1-8B-Instruct | 5,934,139 | Llama 3.1 | | 26 | ornith-ai/Ornith-1.5-9B-GGUF | 5,532,094 | MIT | | 27 | openai/gpt-oss-120b | 5,185,534 | Apache 2.0 | | 28 | Qwen/Qwen2.5-3B-Instruct | 5,114,042 | Other custom | | 29 | Qwen/Qwen3-32B | 4,958,092 | Apache 2.0 | | 30 | lmstudio-community/Qwen3.8-27B-MLX-4bit | 4,873,208 | Apache 2.0 | | 31 | dphn/dolphin-2.9.1-yi-1.5-34b | 4,820,618 | Apache 2.0 | | 32 | Qwen/Qwen3.5-2B | 4,811,106 | Apache 2.0 | | 33 | lmstudio-community/Qwen3.8-27B-MLX-8bit | 4,665,399 | Apache 2.0 | | 34 | lmstudio-community/Qwen3.8-27B-MLX-6bit | 4,620,470 | Apache 2.0 | | 35 | lmstudio-community/Qwen3.8-27B-MLX-5bit | 4,591,269 | Apache 2.0 | | 36 | ornith-ai/Ornith-1.5-35B-A3B-GGUF | 4,492,820 | MIT | | 37 | google/gemma-4-E4B-it | 4,473,231 | Apache 2.0 | | 38 | deepseek-ai/DeepSeek-V4-Flash-0731 | 4,394,898 | MIT | | 39 | Qwen/Qwen-72B | 4,014,339 | Other custom | | 40 | Qwen/Qwen3-4B-Instruct-2507 | 3,992,032 | Apache 2.0 | | 41 | Qwen/Qwen3-VL-4B-Instruct | 3,798,548 | Apache 2.0 | | 42 | Qwen/Qwen3.6-27B | 3,777,499 | Apache 2.0 | | 43 | Qwen/Qwen3-1.7B | 3,756,550 | Apache 2.0 | | 44 | Qwen/Qwen2.5-7B-Instruct-AWQ | 3,507,773 | Apache 2.0 | | 45 | google/gemma-4-E2B-it | 3,500,771 | Apache 2.0 | | 46 | nvidia/NVIDIA-Nemotron-3-Nano-4B-BF16 | 3,484,079 | Other custom | | 47 | EleutherAI/pythia-160m | 3,469,135 | Apache 2.0 | | 48 | Qwen/Qwen3.6-35B-A3B | 3,452,957 | Apache 2.0 | | 49 | ornith-ai/Ornith-1.0-9B-GGUF | 3,443,588 | MIT | | 50 | cdiamond/Qwen3.8-27B-iMatrix-NVFP4-MTP-GGUF | 3,399,484 | Apache 2.0 | | 51 | RadixArk/Kimi-K3-DSpark | 3,338,184 | Not stated | | 52 | Qwen/Qwen3-VL-2B-Instruct | 3,078,197 | Apache 2.0 | | 53 | google/gemma-3-1b-it | 3,018,334 | Gemma | | 54 | microsoft/Florence-2-base | 3,010,170 | MIT | | 55 | huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF | 2,859,860 | Apache 2.0 | | 56 | datalab-to/chandra-ocr-2 | 2,836,048 | OpenRAIL | | 57 | google/gemma-4-12B-it | 2,802,172 | Apache 2.0 | | 58 | Qwen/Qwen2.5-Coder-7B-Instruct | 2,672,084 | Apache 2.0 | | 59 | Qwen/Qwen3-14B-AWQ | 2,555,771 | Apache 2.0 | | 60 | JonathanColetti/Qwen3.8-27B-Uncensored-GGUF | 2,545,055 | Apache 2.0 | 03 — FamiliesDownloads by model family Summing rows by the model they derive from, and leaving out the three fixtures, gives the chart below. We assigned a family by the repository name and its declared base model, so a community GGUF of Qwen3.8-27B counts as Qwen and Nvidia's NVFP4 build of Qwen3.6-35B-A3B counts as Qwen. The 10 largest families are shown; the remaining three rows, Kimi K3, Florence-2 and Chandra OCR, are each a single repository under 3.4 million. 30-day downloads by model family, top 60 excluding fixtures Our sum of the table above, Hugging Face Hub API, September 18, 2026. Family assigned by repository name and declared base model. Two families deserve a note. Ornith's three GGUF repositories sum to 13.5 million, which puts a name most readers will not know above Llama and gpt-oss; we have not investigated the source of those downloads and record the number as the Hub reports it. And Llama has two rows in the top 60, both from the 3.x generation released in 2024, which fits the shift toward Qwen and DeepSeek described in our first-half 2026 open-weight retrospective https://www.digitalapplied.com/blog/open-weight-models-h1-2026-retrospective-deepseek-qwen-llama . 04 — FindingsFour things the table shows Rows that are Qwen's own repositories Google has 6, lmstudio-community 4, Ornith 3, and Nvidia, Meta and OpenAI 2 each. No other organisation has more than one row. Non-fixture top-20 models at 9B parameters or fewer Qwen3-0.6B, Qwen3-VL-8B, Qwen3-8B, Qwen2.5-7B, Qwen3.5-9B, Qwen2.5-0.5B, Qwen3-4B, Qwen2.5-1.5B and Qwen3.5-4B. Four more are MoE models with 3 to 4B active. Quantised or converted copies of another model Four are lmstudio-community MLX builds of Qwen3.8-27B alone, at 4, 5, 6 and 8 bits, totalling 18.8 million. The FP8 build of Qwen3.8-27B out-downloads the original, 7.6 to 7.5 million. Repositories created in 2026 Half the table is under nine months old. The oldest non-fixture rows are Pythia-160m and Qwen-72B from 2023; Qwen2.5 models from late 2024 and January 2025 still hold seven rows. The repack finding is the practical one. If you want to know which size and format of a model people actually run, the quantised rows answer it more directly than the originals: an FP8 or 4-bit file is what gets loaded onto a real machine. Our self-hosting guide for open coding models https://www.digitalapplied.com/blog/best-open-weight-coding-models-self-host-hardware-match-2026 pairs those formats with hardware, and this week's comparison of SSD streaming and ternary weights https://www.digitalapplied.com/blog/run-35b-model-24gb-mac-ssd-streaming-vs-ternary covers two newer ways to make a large model fit; the Edge0 model from that post was the Hub's top trending repository on the day we collected, with 37,131 downloads and 3,349 likes. 05 — LimitsWhat it cannot tell you Downloads are not users, and users are not production. A download count says how many times files were fetched from one host in 30 days. It cannot distinguish a person from a script, a trial from a deployment, or ten thousand developers from one company's build farm. It excludes every model served through an API, every mirror such as ModelScope, and every weight file that was fetched once and then cached inside an organisation. A model that is downloaded by a few hundred companies and run at scale can rank below a model that is fetched by a hundred thousand hobbyists once each. It also cannot tell you where use is. We have not computed a country share and do not print one: a repository's organisation says where a model was made, not where it is run. Treat the table as a measure of distribution, read it alongside benchmark and price data, and check the licence column before assuming any row is usable in a product. If you want help choosing an open model for a specific workload, our AI transformation service https://www.digitalapplied.com/services/ai-transformation includes that evaluation. openai-community/gpt2 15.4 million , trl-internal-testing/ tiny-Qwen2ForCausalLM-2.5 14.1 million and facebook/opt-125m 7.5 million are downloaded by test suites for training and inference libraries. Together they are 37 million of the table's 381 million downloads. EleutherAI/pythia-160m at row 47 is a research model that is also common in tests; we left it unmarked because we cannot separate its uses. 06 — How we collected itMethodology This is our own measurement from a public API. Nothing here comes from a vendor's report or a third-party ranking. - Query - Hugging Face Hub API, GET /api/models, with pipeline tag set to each of text-generation, image-text-to-text and any-to-any in turn, sort set to downloads, descending, top 100 per tag. The three lists were merged and the top 60 by downloads kept. The trending list was collected separately with the same endpoint's trending sort. - What the count is - The Hub's downloads field, a rolling count of the previous 30 days at query time. Likes, creation date, licence tag and declared base model were recorded from the same response. - Scope - Language and vision-language models only. Image, audio, video and embedding models are excluded by the tag filter. Models without one of the three tags, however popular, do not appear. - Fixture rule - A row is marked as a fixture when the repository is a known test model for a library, its organisation is an internal-testing account, or it is a 2022-era baseline whose current downloads are implausible as human use. Three rows meet the rule. Fixtures are shown in the table and excluded from the family chart. - As-of date - Collected September 18, 2026 at 09:28 UTC. This page is dated to the editorial day before; the collection date is stated here and in the dataset card. The next pull will be a new dated snapshot, not a revision of this one. - Known limitations - One host, one 30-day window, no de-duplication of repeated fetches, no API usage, no mirrors. Family assignment by name and base-model tag can misclassify a fine-tune that does not declare its base. Parameter counts in the findings are read from repository names. 07 — Next stepOne family, many small files Check the repack rows before you pick a size and format If you are choosing an open model this month, the table's most useful signal is not the winner but the shape: small models and quantised copies of mid-sized ones are what the ecosystem loads. Find the family you are considering, look at which of its formats people fetch, and test that file on your hardware. Come back next month; the next pull will be a new dated row in the same series.