I tested every uncensored Ternary Bonsai 2 27B on the Hub, all 11 builds from 6 uploaders, including my own. Same prompts, same judge, same GPUs, every file pinned by sha256.
• Best answers: @Hikari07jp and @dealignai (~0.94 answer quality with thinking off) • Only edit with no measurable MMLU cost: mine (±0.15 pp). Every other edit loses 0.64–2.68 pp, all p < 0.001 • Heretic and Blackfrost still refuse 11–12% of harmful prompts • Thinking mode at 4,096 tokens: 5–31% of harmful prompts get no answer. PrismML recommends 16,384+, and I'm rerunning at that budget
This is the first fine tune to exceed 730 "arc-c" ("735": 144 pts higher than Qwen 3.8 27B) AND 880 ARC-E (The OpenAI, Claude and Gemini "zone of intelligence") in 8 bit and over 718 arc-c in 4 bit.
This version is called TURBO because it drastically reduces thinking tokens (by 1/2 to as high as 1/10), yet maintains output detail and quality.
In other words while "reg" Qwen3.8 27B is thinking about "formatting" for a few 1000 tokens, this model is already done and waiting for more.
This repo contains both "regular" and "MTP" Neo-CODER MAX DI-MATRIX (duel imatrix) GGUF quants.
PS: There are 29 additional quant repos as of this writing too, as well NVFP4 and many more as well.
This is one of 10+ Qwen 3.8 27B at or above ARC-C of 717 (all 10 exceed all core benchmarks of Qwen 3.8, 3.6 and 3.5 27B and 35B-A3B versions) - you can see the complete project and some of the training here :