ONNX conversion of gliner-community/glinersmall-v2.5, dynamically quantised to INT8, packaged as a self-contained bundle for offline NER. This is a re-serialisation, not a fine-tune: the weights are the upstream ones. Only the format (PyTorch → ONNX) and the precision (fp32 → INT8) are ours. The fp32 reference graph (model.onnx, sha256 5245733ccb2b75072cce0b4bbb14424988f92f9daf775d97bdf0de74be28df63) is not shipped — it is only needed to reproduce the INT8 graph. Its hash is recorded in NOTICE. The graph has six inputs, fed per span-encoded prompt: spanmask is bool (not int64) — the graph declares tensor(bool). The prompt follows the GLiNER label format, using the model's own special tokens…
Independent publisher
Yevhenii
GG-QandV
Models in Library1
Datasets in Library0
Models on Hugging Face1
Followers—