Inside Qwen3.6 Genesis Hermes V9

https://hackernoon.imgix.net/images/1787150534896_33itj8vg.png

Overview

Qwen3.6-35B-A3B-Uncensored-Genesis-Hermes-V9-GGUF is a 35-billion parameter mixture-of-experts language model created by LuffyTheFox that applies a proprietary post-training tensor repair technique called Genesis to reduce accumulated training noise and improve output consistency. The model runs 3 billion parameters active per forward pass through 256 experts (8 routed + 1 shared per token) and uses a hybrid architecture combining Gated DeltaNet linear attention with full softmax attention in a 3:1 ratio across 40 layers. It supports 262,000 native context tokens (extendable to 1 million with YaRN), handles 248,000 vocabulary tokens across 201 languages, and includes native multimodal capabilities for text, image, and video. The critical distinction is that Genesis does not retrain or finetune—instead, it performs "numerical surgery" on GGUF-format model weights to eliminate saturated weights, scale mismatches, mean drift, zero blocks, and training noise without removing learned knowledge or gradients. Built on the HauhauCS uncensored base (0 refusals on 465 test...

Copyright of this story solely belongs to hackernoon.com. To see the full text click HERE

Read more