SAVRN
Search Contact SAVRN

Independent publisher

IsValorum

IsValorum

I'm tired of generic quantizations that consume too many resources when they could be even higher quality and consume much less, but nobody was doing it then I thought Why don't I do it myself? And here I am, the first person to dedicate everything to quality over quantity.

Models in Library3
Datasets in Library0
Models on Hugging Face15
Followers7

Models

Model · Image and text to text

Nex-N2.5-mini-APEX-I-MiniPlus-GGUF

IsValorum

Also, don't confuse APEX-I-MiniPlus (Standard) with a generic baseline APEX-I-Mini. Traditional APEX-I-Mini drops core experts aggressively to 2-bit IQ2S and leaves output.weight at 3-bit Q3KM, which creates a noticeable perplexity hit on complex reasoning tasks. Standard MiniPlus avoids that degradation floor while keeping boundary layers in linear Q3K for single-cycle vectorized AVX2 CPU dequantization (hitting 23 to 26+ tok/s on DDR4 laptops), while protecting output in Q6K and routers in F32. To put the numbers in perspective: this cuts nearly 2 GB off a flat 3-bit quant (approx. 15.6 GB), and weighs only about approx. 1 GB more than a generic APEX-I-Mini (approx. 12.5 GB). For that…

Open weights apache-2.0 gguf