fathan/autotrain-data-code-mixed-language-identification
AutoTrain Dataset for project: code-mixed-language-identification Dataset Description This dataset has been automatically processed by AutoTrain for project code-mixed-language-identification. Languages The BCP-47 code for the dataset's language is unk. Dataset Structure Data Instances A sample from this dataset looks as follows: [ { "feat_Unnamed: 0": 1104, "tokens": [ "@user", "salah"… See the full description on the dataset page: https://huggingface.co/datasets/fathan/autotrain-data-code-mixed-language-identification.
Conversations for this repository live on Hugging Face.
Team Ai shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face