Join the conversation

Join the community of Machine Learners and AI enthusiasts.

Sign Up
DavidAUΒ 
posted an update 6 days ago
Post
9114
Qwen 3.5 9B - The Defiant, 27B power ; now with Qwen 3.8 Reasoning modes.

640 ARC-C for both 8bit and 4bit. Model exceeds 7 of 7 benchmarks for Qwen 3.5 9B, Qwen3.5 27B, Qwen3.6 35B-A3B, and meets Qwen 3.6 27B in some cases... and it does so in 4bit and 8bit. Regular and MTP (fast) NEO IMATRIX GGUFs provided. (this model is part of the Qwen 3.6 27B Fable Fusion 711 pipelines: 2200+ likes, 3 million + downloads)

NEW - Qwen 3.8 Reasoning Modes: 2 MTP quants (Q6/Q8) Now with 5 reasoning modes (2 new - Spoon / Einstein), and 5 instruct modes (2 new - Spoon / Einstein, all use ZERO REASONING TOKENS) all switchable on the fly via API, direct and "in chat" (yes - model control at the chat/message level). Model name has "plusIQ" in the name.

(there is also a extra robust "tools" version too.)

DavidAU/Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF

PS: I have posted a 22k output example using "spoon" mode too.

well done

best 9b on the planet imho

Β·

At least two of the models in the merge have space station coding curriculum with characters on DS9, so technically it is also the best 9B in Space :)

Do you plan to update the base model files? Also, what changes did you make weight-wise? If it's just chat template(s), that'd be good to know so I can update my quant by extracting from a GGUF.

Β·

At the moment, only the chat template -> modified specifically for 3.5.
Future version[s] will have both tuning and chat template alignments.