Hate Speech and Toxic Comment Detector
MSc Data Science | CMP-L016 Deep Learning Applications | University of Roehampton
This tool uses HateBERT fine-tuned on the Jigsaw Toxic Comment dataset to classify text as toxic or non-toxic. Token-level SHAP explanations show which words most influenced the prediction.
Example inputs - click to load
Research Disclaimer: This tool is for academic research purposes only. It does not constitute a content moderation decision. False positives and false negatives are possible.