Team Ai
Datasetpublic

NLPC-UOM/Sinhala-English-Code-Mixed-Code-Switched-Dataset

Sinhala-English-Code-Mixed-Code-Switched-Dataset This dataset contains 10,000 comments that have been annotated at the sentence level for sentiment analysis, humor detection, hate speech detection, aspect identification, and language identification. The following is the tag scheme. Sentiment - Positive, Negative, Neutral, Conflict Humor - Humorous, Non humorous Hate Speech - Hate-Inducing, Abusive, Not offensive Aspect - Network, Billing or Price, Package, Customer Service… See the full description on the dataset page: https://huggingface.co/datasets/NLPC-UOM/Sinhala-English-Code-Mixed-Code-Switched-Dataset.

sourceHugging Facemitupdated 2y agoView on Hugging Face
6likes187downloads

NLPC-UOM/Sinhala-English-Code-Mixed-Code-Switched-Dataset · main · files are served by the source, never re-hosted here