Local hardware: Mechrevo Kuangshi GM7AG0M — RTX 3060 Laptop 6GB GDDR6, 64GB DDR5, i7-12700H (14C/20T, 4.7GHz), Windows 11, Samsung 990 Pro. Good for imatrix and 0.6–35B-class work in RAM. 9B+ and searches need rented H200/Blackwell, typically $100 per quant. Released as part of the NOESIS Professional Multilingual Dubbing Automation Platform (framework: Deterministic Hybrid Control Framework for Frozen Neural Operators — DHCF-FNO). Repack (BF16 safetensors, shards ≤50 GB) and community GGUF quantizations of (modeltype: aliceai, 80B total / 3B active MoE with KDA layers). Исходные 49 шардов Yandex слиты в один стриминговый файл и заново нарезаны SHA256 каждого тензора, 0 расхождений (см.…
Open weights
apache-2.0
81.3B parameters
262,144 tokens
transformers
Local hardware: Mechrevo Kuangshi GM7AG0M — RTX 3060 Laptop 6GB GDDR6, 64GB DDR5, i7-12700H (14C/20T, 4.7GHz), Windows 11, Samsung 990 Pro. Good for imatrix and 0.6–35B-class work in RAM. 9B+ and searches need rented H200/Blackwell, typically $100 per quant. Released as part of the NOESIS Professional Multilingual Dubbing Automation Platform (framework: Deterministic Hybrid Control Framework for Frozen Neural Operators — DHCF-FNO). GGUF-конвертация и кванты текстового бэкбона Moondream 3.1 — vision-language модели с MoE-архитектурой (9B всего / 2B активных параметров). Конвертация и квантование выполнены AMAImedia; vision-часть (mmproj) в репозиторий не входит. Все файлы…
Open weights
other
llama.cpp