A headless four-layer text backbone distilled from Qwen3.5-0.8B, with a vocabulary cut to English and Korean. It has no language-model head, no classification head, no labels and no thresholds. You attach a head and train it. This card is the record of how it was made. It walks the whole path — prompting the original model, turning it into a classifier, removing layers, changing precision, cutting the vocabulary, and cutting it again by rule — and shows what each step measured. We are not arguing that this model is better than anything. We built it, we measured it beside other models, and where we have a guess about why a number moved we say it is a guess. Every arm below was trained…
Independent publisher
Junsung Kim
mp-juuuns
Efficient and deployable NLP: compact transformers, layer reduction, quantization, multilingual classification, propaganda technique detection, and reproducible edge-device evaluation.
Models in Library1
Datasets in Library0
Models on Hugging Face4
Followers—