Team Ai
Datasetpublic

NLPC-UOM/Sinhala-English-Code-Mixed-Code-Switched-Dataset

Sinhala-English-Code-Mixed-Code-Switched-Dataset This dataset contains 10,000 comments that have been annotated at the sentence level for sentiment analysis, humor detection, hate speech detection, aspect identification, and language identification. The following is the tag scheme. Sentiment - Positive, Negative, Neutral, Conflict Humor - Humorous, Non humorous Hate Speech - Hate-Inducing, Abusive, Not offensive Aspect - Network, Billing or Price, Package, Customer Service… See the full description on the dataset page: https://huggingface.co/datasets/NLPC-UOM/Sinhala-English-Code-Mixed-Code-Switched-Dataset.

sourceHugging Facemitupdated 2y agoView on Hugging Face
6likes187downloads
settings

This repository belongs to NLPC-UOM on Hugging Face.

Team Ai never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameSinhala-English-Code-Mixed-Code-Switched-Dataset
visibilitypublic
licencemit
gatedno
ownerNLPC-UOM
Account settings
NLPC-UOM/Sinhala-English-Code-Mixed-Code-Switched-Dataset · Team Ai