Team Ai
28 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Lots-of-LoRAs /task388_torque_token_classification Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task388_torque_token_classification Additional Information Citation Information The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it: @misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions, title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP Tasks}… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task388_torque_token_classification.texttext-generation1K<n<10K0 likes86 downloads2y agoHugging Face02hossein20s /enrun-emails-token-classificationtext10K<n<100K2 likes84 downloads4y agoHugging Face03bikashpatra /token_classification_19thJulytextn<1K1 likes62 downloads2y agoHugging Face04anismahmahi /classification_token_propagandatexttoken-classificationn<1K0 likes49 downloads2y agoHugging Face05Ailaysa-MTPE /token_classification_datasettext100K<n<1M1 likes48 downloads3y agoHugging Face06bikashpatra /claims_token_classificationtextn<1K1 likes48 downloads2y agoHugging Face07open-source-metrics /token-classification-checkpoint-downloadstabular1K<n<10K1 likes34 downloads4y agoHugging Face08KIT-RoboInfo /token_classification_datasetstext1K<n<10K1 likes31 downloads2y agoHugging Face09dumyy /token-classification-brand Dataset Card for "token-classification-brand" More Information needed textn<1K0 likes26 downloads3y agoHugging Face10hemangjoshi37a /token_classification_ratnakar_13000 likes24 downloads4y agoHugging Face11wolf4032 /token-classification-japanese-search-local-cuisine料理を検索するための質問文と、質問文に含まれる検索検索用キーワードの情報を持ったデータセットです 固有表現の種類は以下の4つです。 AREA: 都道府県/地方 TYPE: 種類 SZN: 季節 INGR: 食材 GitHub untokenized_dataset_list.ipynb(データセットの作成に使ったノートブック) このデータセットを使った言語モデルのファインチューニングと、ファインチューニングした言語モデルを使ったアプリのコードもこのリポジトリにあります 詳細情報 Qiita texttoken-classification1K<n<10K0 likes22 downloads2y agoHugging Face12timonziegenbein /inappropriateness-token-classification-binarized-multi-reftabular1K<n<10K0 likes20 downloads8mo agoHugging Face13PhaniManda /autotrain-data-demo-on-token-classification AutoTrain Dataset for project: demo-on-token-classification Dataset Description This dataset has been automatically processed by AutoTrain for project demo-on-token-classification. Languages The BCP-47 code for the dataset's language is unk. Dataset Structure Data Instances A sample from this dataset looks as follows: [ { "tokens": [ "I", "will", "be", "traveling", "to", "Tokyo", "next"… See the full description on the dataset page: https://huggingface.co/datasets/PhaniManda/autotrain-data-demo-on-token-classification.token-classification0 likes19 downloads3y agoHugging Face14timonziegenbein /inappropriateness-token-classification-multi-reftabular1K<n<10K0 likes19 downloads8mo agoHugging Face15timonziegenbein /inappropriateness-token-classification-binarizedtabularn<1K0 likes18 downloads8mo agoHugging Face16supergoose /flan_combined_task388_torque_token_classificationtext1K<n<10K0 likes16 downloads2y agoHugging Face17globalise /globalise_NER_token_classification_dataset Dataset Card for Dataset Name The globalise_NER_token_classification dataset is a fine-grained dataset for the training of token-classification NER models on Dutch East-India Company texts (17th to 18th century). Dataset Details Dataset Description The dataset provides 15 fine-grained labels detailing activities and people of the Dutch East-India Company (VOC), and can be used to train NER token-classification models for the period 17th-18th century and the… See the full description on the dataset page: https://huggingface.co/datasets/globalise/globalise_NER_token_classification_dataset.texttoken-classificationn<1K1 likes16 downloads1y agoHugging Face18Cleanlab /token-classification-tutorial Token Classification Tutorial Dataset Dataset Description This dataset contains predicted probabilities for token classification used in the cleanlab tutorial: Token Classification. The dataset demonstrates how to use cleanlab to identify and correct label issues in token classification datasets, such as Named Entity Recognition (NER) tasks where each token in a sequence is assigned a class label. Dataset Summary Task: Token classification / Named Entity… See the full description on the dataset page: https://huggingface.co/datasets/Cleanlab/token-classification-tutorial.token-classificationn<1K0 likes16 downloads10mo agoHugging Face19PhaniManda /autotrain-data-test-token-classification AutoTrain Dataset for project: test-token-classification Dataset Description This dataset has been automatically processed by AutoTrain for project test-token-classification. Languages The BCP-47 code for the dataset's language is en. Dataset Structure Data Instances A sample from this dataset looks as follows: [ { "tokens": [ "I", "will", "be", "traveling", "to", "Tokyo", "next"… See the full description on the dataset page: https://huggingface.co/datasets/PhaniManda/autotrain-data-test-token-classification.token-classification0 likes15 downloads3y agoHugging Face20exonics /bert_token_classificationtext1K<n<10K1 likes15 downloads10mo agoHugging Face21s3h /gec-token-classification0 likes13 downloads5y agoHugging Face22gonzalo-santamaria-iic /mock_token_classification_datasettextn<1K0 likes10 downloads11mo agoHugging Face23QNN /autotrain-data-token-classificationgated AutoTrain Dataset for project: token-classification Dataset Description This dataset has been automatically processed by AutoTrain for project token-classification. Languages The BCP-47 code for the dataset's language is unk. Dataset Structure Data Instances A sample from this dataset looks as follows: [ { "tokens": [ "Pd", "has", "been", "regarded", "as", "one", "of", "the"… See the full description on the dataset page: https://huggingface.co/datasets/QNN/autotrain-data-token-classification.token-classification2 likes9 downloads3y agoHugging Face24Pisethan /sangapac-token-classificationtextn<1K0 likes8 downloads2y agoHugging Face25timonziegenbein /inappropriateness-token-classificationtabularn<1K0 likes8 downloads8mo agoHugging Face26exonics /bert_token_classification_augmentedtext1K<n<10K1 likes7 downloads8mo agoHugging Face27seandi /cord-v2-token-classificationimage1K<n<10K0 likes6 downloads2y agoHugging Face28WaritMahitti /FAKEINVEST_TOKENCLASSIFICATIONtextn<1K0 likes4 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.