Team Ai
Datasetpublic

SEACrowd/tgl_profanity

This dataset contains 13.8k Tagalog sentences containing profane words, together with binary labels denoting whether or not the sentence conveys profanity / abuse / hate speech. The data was scraped from Twitter using a Python library called SNScrape and annotated manually by a panel of native Filipino speakers.

sourceHugging Faceunknownupdated 2y agoView on Hugging Face
0likes35downloads

SEACrowd/tgl_profanity · main · files are served by the source, never re-hosted here