datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
enrun-emails-token-classificationtask388_torque_token_classification
Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task388_torque_token_classification
Additional Information
Citation Information
The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it:
@misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions,
title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP Tasks}… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task388_torque_token_classification.token_classification_19thJulyclaims_token_classificationtoken_classification_datasetclassification_token_propagandatoken_classification_datasetstoken-classification-brand
Dataset Card for "token-classification-brand"
More Information needed
sangapac-token-classificationinappropriateness-token-classification-multi-reftoken-classification-checkpoint-downloadstoken-classification-japanese-search-local-cuisine料理を検索するための質問文と、質問文に含まれる検索検索用キーワードの情報を持ったデータセットです
固有表現の種類は以下の4つです。
AREA: 都道府県/地方
TYPE: 種類
SZN: 季節
INGR: 食材
GitHub
untokenized_dataset_list.ipynb(データセットの作成に使ったノートブック)
このデータセットを使った言語モデルのファインチューニングと、ファインチューニングした言語モデルを使ったアプリのコードもこのリポジトリにあります
詳細情報
Qiita
mock_token_classification_datasetflan_combined_task388_torque_token_classificationbert_token_classificationinappropriateness-token-classification-binarized-multi-refinappropriateness-token-classification-binarizedglobalise_NER_token_classification_dataset
Dataset Card for Dataset Name
The globalise_NER_token_classification dataset is a fine-grained dataset for the training of token-classification NER models on Dutch East-India Company texts (17th to 18th century).
Dataset Details
Dataset Description
The dataset provides 15 fine-grained labels detailing activities and people of the Dutch East-India Company (VOC), and can be used to train NER token-classification models for the
period 17th-18th century and the… See the full description on the dataset page: https://huggingface.co/datasets/globalise/globalise_NER_token_classification_dataset.cord-v2-token-classificationinappropriateness-token-classificationbert_token_classification_augmentedFAKEINVEST_TOKENCLASSIFICATION
