multi-label-classification
multi-label-class-classification-on-github-issuesxlm-roberta-finance-multi-label-classificationmulti-label-emotion-classification-reddit-comments-robertaMulti-Label-Classification-of-PubMed-Articlesdistilbert-base-pwc-task-multi-label-classificationmultilabel_classificationaift-model-review-multiple-label-classificationmultilabel-classification-bert-ontonotes5
multi-label-class-github-issues-text-classification
Dataset Card for "multi-label-class-github-issues-text-classification"
More Information needed
wos_hierarchical_multi_label_text_classificationIntroduced by du Toit and Dunaiski (2024) Introducing Three New Benchmark Datasets for Hierarchical Text Classification.
The WOS Hierarchical Text Classification are three dataset variants created from Web of Science (WOS) title and abstract data categorised into a hierarchical, multi-label class structure. The aim of the sampling and filtering methodology used was to create well-balanced class distributions (at chosen hierarchical levels). Furthermore, the WOS_JTF variant was also created… See the full description on the dataset page: https://huggingface.co/datasets/marcelsun/wos_hierarchical_multi_label_text_classification.synthetic-text-classification-news-multi-label
Dataset Card for synthetic-text-classification-news-multi-label
This dataset has been created with distilabel.
Dataset Summary
This dataset contains a pipeline.yaml which can be used to reproduce the pipeline that generated it in distilabel using the distilabel CLI:
distilabel pipeline run --config "https://huggingface.co/datasets/davidberenstein1957/synthetic-text-classification-news-multi-label/raw/main/pipeline.yaml"
or explore the configuration:… See the full description on the dataset page: https://huggingface.co/datasets/argilla/synthetic-text-classification-news-multi-label.prachathai67k-tha-multilabelclassificationref: https://github.com/PyThaiNLP/prachathai-67k
prachathai67k-tha-multilabelclassification
Prachathai67k_tha_MultiLabelClassification
Deduplicated copy of kornwtp/prachathai67k-tha-multilabelclassification.
Splits
split
rows
train
67,488
AURA-Multi_Label_Classification
AURA-Classification (Multi-Label Version)
Dataset Description
The AURA (App User Review in Arabic) Classification dataset is a collection of 2,900 Arabic-language app reviews collected from various mobile applications. This dataset is designed for multi-label text classification, where each review can belong to multiple classes simultaneously.
Each review in the dataset was independently annotated by five different annotators. To construct the multi-label version of the… See the full description on the dataset page: https://huggingface.co/datasets/irfan-ahmad/AURA-Multi_Label_Classification.
