cd /news/large-language-models/qwen-3-8-vs-ornith-vs-nemotron-vs-mu… · home topics large-language-models article
[ARTICLE · art-110046] src=dev.to ↗ pub= topic=large-language-models verified=true sentiment=· neutral

Qwen 3.8 vs Ornith vs Nemotron vs Muse-Glimmer

A developer benchmarked four open-source models—Qwen 3.8 27B, Ornith 1.5 35B-A3B, Nemotron 3.5 Lightning 30B-A3B, and Muse-Glimmer 30B—on coding and computer-use tasks. Qwen 3.8 proved the most accurate, matching GPT-5.6 Luna in computer use, while Ornith impressed with speed and efficiency, achieving 50 TPS on Strix Halo. Nemotron and Muse-Glimmer lagged in coding and reliability, respectively.

read2 min views5 publishedAug 25, 2026

Srovnání Qwen 3.8 27B, Nemotron Lightning 30B, Ornith 35B a Muse-Glimmer 30B v reálných testech. Qwen 3.8 ovládá coding, Ornith překvapuje rychlostí, Nemotron a Muse-Glimmer ztrácejí.

Uživatel provedl detailní srovnání čtyř populárních open-source modelů na coding úloze. Test běžel 25 minut a na Qwen 3.8 vyprodukoval 50 000 tokenů — jde o těžký thinking model, ale výsledek je fenomenální. Ornith se ukázal jako velmi schopný model, zejména vzhledem k dané rychlosti.

Výsledky detailně: llm-bench.io porovnání

Qwen 3.8 27B — jasný vítěz v codingu a architektuře. V testech computer use MCP byl bezchybný, na úrovni GPT-5.6 Luna. Jak uživatel popsal: Qwen3.8 was flawless, on par with gpt 5.6 luna. Ale je pomalý — na některých strojích sotva 15 TPS. Pro náročné úlohy, kde záleží na přesnosti, je to jasná volba. Podle dalšího uživatele: Glad to see i was justified in focusing mainly on 3.8 since its release.

Ornith 1.5 35B-A3B — překvapení soutěže. Ornith is seriously amazing especially given the fact it only has 3b active. Na Strix Halo dosahuje 50 TPS při 50k+ tokenech a zvládne plný kontext. Podle jednoho uživatele: After more testing Ornith is absolute sorcery, this is GPT oss20b levels of optimization. 50TPS on strix halo at 50k+ tokens and can load full context, compared to Qwen 3.8 that struggles to serve 15TPs, and I honestly can't notice the difference in front end work maybe 2% at best.

Nemotron 3.5 Lightning 30B-A3B — určen pro ne-technické agentic úlohy. Nemotron is for non technical agentic tasks. Solidní rychlost, ale v codingu zaostává.

Muse-Glimmer 30B — zaměřen na technické psaní. Muse Glimmer take care of 3 stages of technical writing. S dflash běží rychleji (150-200 TPS na 3090), ale v testech computer use dělal chyby: Muse glimmer was almost good, but would miss some step, misclick some button or forget to activate a window and that would doom it.

Zdroj: Reddit

── more in #large-language-models 4 stories · sorted by recency
── more on @qwen 3.8 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/qwen-3-8-vs-ornith-v…] indexed:0 read:2min 2026-08-25 ·