Note · 1 min read
Qwen3.6-35B base outperforms most fine‑tunes in Reddit benchmark
A Reddit user compared the Qwen3.6-35B model and five fine‑tuned variants on a coding benchmark, finding the base model generally superior and only Occamy‑1.0 close.
Oossa · About Oossa
On Sep 28 2026 a Reddit user posted results comparing the Qwen3.6-35B-A3B base model with five fine‑tuned variants using the Aider Polyglot coding benchmark. The base model achieved a 37.4 % first‑try pass rate and 71.0 % retry pass rate, beating most fine‑tunes.
Only the Occamy‑1.0 fine‑tune came close, with 29.0 % first‑try and 70.1 % retry, while Ornith‑1.5, KAT‑Coder, Tiel‑Coder and Nex‑N2.5 fell well behind. Token usage and solve time were also lower for the base model, suggesting it remains the best choice for most users.
Why it matters
It shows that, for now, using the unmodified Qwen3.6-35B gives better coding performance than most community fine‑tunes, saving users from unnecessary model switching.
Sources & references
| # | Source | Outlet | Date | Key takeaway |
|---|---|---|---|---|
| 1 | Searching for 3.8 35B: Qwen3.6-35B-A3B (Testing 5 Finetunes vs. Base) ↗ | Reddit r/LocalLLaMA | Sep 28, 2026 | TL;DR -- You should probably just use base Qwen3.6-35B, as only Occamy-1.0 is competitive with it. |
1 sources
Last updated:
Oossa · Newsletter
The week in AI, explained
Every Monday: the stories worth knowing, in plain language. Free, no spam.