SAVRN
Search Contact SAVRN

Independent publisher

CS

Bahushruth

Models in Library1
Datasets in Library0
Models on Hugging Face8
Followers22

Models

Model · Text generation

Qwen3.6-35B-A3B-abliterated-v4

CS

Uncensored version of Qwen/Qwen3.6-35B-A3B with refusal behavior removed via abliteration (norm-preserving orthogonalization). Zero refusals on harmful prompts. No false refusals on harmless prompts. Abliteration identifies the "refusal direction" in the model's residual stream — the linear direction that activates when the model decides to refuse — and surgically removes it from all output projection weights using norm-preserving orthogonalization. 1. Collect residual stream activations (last token position) for 512 harmful + 512 harmless prompts across all 40 layers 2. Compute mean difference vector per layer → this is the "refusal direction" candidate 3. Score layers by…

Open weights apache-2.0 34.7B parameters 262,144 tokens transformers