David, do you accept contributions (buy me a beer, etc.) and does huggingface allow that?
Shortly going to add this, in part for upgrades as well as some training ($$) I can't do on local hardware.
David, do you accept contributions (buy me a beer, etc.) and does huggingface allow that?
Shortly going to add this, in part for upgrades as well as some training ($$) I can't do on local hardware.
This one will be slightly smarter, whereas Twin is slightly faster [fewer tokens] , and not as creative.
At the moment, only the chat template -> modified specifically for 3.5.
Future version[s] will have both tuning and chat template alignments.
Make sure you get the updated ggufs, with tool fixes. ;
You can find additional quants here:
https://huggingface.co/models?other=base_model:quantized:DavidAU/Qwen3.8-27B-Cold-Fusion-GAIN-V1.1
Some are a lot smaller / will work better on 16GB cards.
Look up LLFAN or other top Heretic'ers ; search "heretic" in models tab at hugging face.
They can help with this ; I don't have the compute avail ATM. ;
First ; thank you!
Based on early checking/Qwen's own statements (uses same Qwen 3.5 arch as 3.5,.6) -> it is doable.
But need to verify everything.
RE: deepseek V4 ; sorry don't have the VRAM to do it.
Yes. It is on the list.
We are still revising the 9-14B pipeline.
We do have an interm Qwen 3.5 9B from the experimental pipeline here:
It matches/exceeds 27B Qwen 3.5 ; and meets in some cases 27B Qwen 3.6 performance.
There is still a lot of optimizations to do at this time.
Try the Q6, and/or IQ4_XS.
Even IQ3_M will be very strong.
Try this one:
https://huggingface.co/DavidAU/Qwen3.5-9B-The-Defiant-Fable-Uncensored-Heretic-NEO-IMATRIX-MAX-MTP-GGUF
This is the "smaller version" of the Fable Fusion 711.
It operates at almost 27B power at only 9B parameters.
This will far exceed Qwen 2.5 and Qwen 3 performance.
Excellent. The MOE variants require a lot more VRAM/time for training.
Please take a moment to visit the repo where some of your concerns are addressed on the repo card itself.
2nd; publishing all the metrics at each step would be both exhausting and worse confusing.
I don't follow what you mean by "cheap" ; as a heretic [step] you usually lose 2-4 points on some metrics.
So the comparison of "heretic" vs "non-heretic" is even STRONGER ; the fairer one would be "heretic base" to "heretic tuned" which would likely show even greater change / improvement.
Source/MLX here:
https://huggingface.co/nightmedia/Qwen3.5-9B-DS9-USS-Defiant
This is on my partner's repo.