This is an unfinished training run, published as it goes. Checkpoints appear here every half epoch of a planned six. Nothing here is a final result, and the numbers below will move. A DFlash draft head for Qwen/Qwen3.6-27B, trained as a reproduction of DFlash (arXiv:2602.06036) on a model the paper does not cover. Full method, scripts and measurement records: It is not a standalone model. It cannot generate text by itself. It writes a 16-token block in one forward pass, conditioned on the target's hidden states at layers 1/16/31/46/61, and the target then verifies that block in a single pass and commits the leading run that matches. Greedy verification means the committed tokens are exactly…
Independent publisher
MINSIK CHOI
CHOI0126
Models in Library1
Datasets in Library0
Models on Hugging Face8
Followers—