tasksource/tasksource-instruct
tasksource-instruct Instruction-tuning data recast from the ~480 English classification, multiple-choice and token-classification tasks of tasksource. Every example comes from a human-built dataset (NLI, logical reasoning, sentiment, hate speech, discourse, argumentation, ...), not from a teacher model. Each task is capped at 30k training examples, so no task dominates. Many tasks aren't in FLAN v2, for example DynaSent, DynaHate, discriminative bAbI, epistemic logic, RuleTaker… See the full description on the dataset page: https://huggingface.co/datasets/tasksource/tasksource-instruct.
242.6k
../
test-00000-of-00001.parquetdownload
train-00000-of-00013.parquetdownload
train-00001-of-00013.parquetdownload
train-00002-of-00013.parquetdownload
train-00003-of-00013.parquetdownload
train-00004-of-00013.parquetdownload
train-00005-of-00013.parquetdownload
train-00006-of-00013.parquetdownload
train-00007-of-00013.parquetdownload
train-00008-of-00013.parquetdownload
train-00009-of-00013.parquetdownload
train-00010-of-00013.parquetdownload
train-00011-of-00013.parquetdownload
train-00012-of-00013.parquetdownload
validation-00000-of-00001.parquetdownload
