Team Ai
Datasetpublic

SEACrowd/tgl_profanity

This dataset contains 13.8k Tagalog sentences containing profane words, together with binary labels denoting whether or not the sentence conveys profanity / abuse / hate speech. The data was scraped from Twitter using a Python library called SNScrape and annotated manually by a panel of native Filipino speakers.

sourceHugging Faceunknownupdated 2y agoView on Hugging Face
0likes35downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

Team Ai shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
SEACrowd/tgl_profanity · Team Ai