Srovnání Qwen 3.8 27B, Nemotron Lightning 30B, Ornith 35B a Muse-Glimmer 30B v reálných testech. Qwen 3.8 ovládá coding, Ornith překvapuje rychlostí, Nemotron a Muse-Glimmer ztrácejí. Uživatel provedl detailní srovnání čtyř populárních open...
Srovnání Qwen 3.8 27B, Nemotron Lightning 30B, Ornith 35B a Muse-Glimmer 30B v reálných testech. Qwen 3.8 ovládá coding, Ornith překvapuje rychlostí, Nemotron a Muse-Glimmer ztrácejí.
Co se stalo
Uživatel provedl detailní srovnání čtyř populárních open-source modelů na coding úloze. Test běžel 25 minut a na Qwen 3.8 vyprodukoval 50 000 tokenů — jde o těžký thinking model, ale výsledek je fenomenální. Ornith se ukázal jako velmi schopný model, zejména vzhledem k dané rychlosti.
Výsledky detailně: llm-bench.io porovnání
Jak modely obstály
Qwen 3.8 27B — jasný vítěz v codingu a architektuře. V testech computer use MCP byl bezchybný, na úrovni GPT-5.6 Luna. Jak uživatel popsal: Qwen3.8 was flawless, on par with gpt 5.6 luna. Ale je pomalý — na některých strojích sotva 15 TPS. Pro náročné úlohy, kde záleží na přesnosti, je to jasná volba.
Ornith 1.5 35B-A3B — překvapení soutěže. Ornith is seriously amazing especially given the fact it only has 3b active. Na Strix Halo dosahuje 50 TPS při 50k+ tokenech a zvládne plný kontext. Podle jednoho uživatele: After more testing Ornith is absolute sorcery, this is GPT oss20b levels of optimization. 50TPS on strix halo at 50k+ tokens and can load full context, compared to Qwen 3.8 that struggles to serve 15TPs.
Nemotron 3.5 Lightning 30B-A3B — určen pro ne-technické agentic úlohy. Nemotron is for non technical agentic tasks. Solidní rychlost, ale v codingu zaostává.
Muse-Glimmer 30B — zaměřen na technické psaní. Muse Glimmer take care of 3 stages of technical writing. S dflash běží rychleji (150-200 TPS na 3090), ale v testech computer use dělal chyby: Muse glimmer was almost good, but would miss some step, misclick some button or forget to activate a window.
Co říkají komentáře
Pozitivní
Ornith is seriously amazing especially given the fact it only has 3b active (DerDave, 29 bodů)
After more testing Ornith is absolute sorcery, this is GPT oss20b levels of optimization (vinis_artstreaks)
Glad to see i was justified in focusing mainly on 3.8 since its release. Qwen4 scheduled for Sept. Feels like a storm. (Ell2509)
Looks like Orinth is the winner here overall for speed to quality generation (Zennytooskin123)
Negativní / Kritici
Ornith's nowhere near qwen3.8 in my testing... Sometimes making glaring mistakes like gemma 4 12b (SeriousPanic34)
Is anybody still using Nemotron or Muse-glimmer? Those were DOA (Fun_Jaguar8231)
Tipy a rady
Pro coding: Qwen 3.8, ale počítejte s nižší rychlostí
Pro rychlost + solidní kvalitu: Ornith 35B
Pro agentic úlohy: Nemotron Lightning
Pro psaní: Muse-Glimmer
Qwen4 přijde v září — sledujte
Zdroj: Reddit