This model is a fine-tuned version of the DistilBERT model to classify toxic comments. You can use the model with the following code. This model is intended to use for classify toxic online classifications. However, one limitation of the model is that it performs poorly for some comments that mention a specific identity subgroup, like Muslim. The following table shows a evaluation score for different identity group. You can learn the specific meaning of this metrics here. But basically, those metrics shows how well a model performs for a specific group. The larger the number, the better. The table above shows that the model performs poorly for the muslim and jewish group. In fact, you pass…
Independent publisher
Martin Pan
martin-ha
Models in Library1
Datasets in Library0
Models on Hugging Face4
Followers2