datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
annotation-pack-a
Annotation pack A — does a response genuinely follow an instruction?
50 rows. Each row: a user prompt, one instruction from it, and a model response that an automatic checker
marks as satisfying that instruction. Annotators judge whether it is satisfied genuinely.
For annotators / 标注人:
Read RUBRIC_HUMAN.md (English + 中文).
Annotator A downloads annotator_A.csv; annotator B downloads annotator_B.csv (same items).
Fill label (GENUINE / LOOPHOLE / GARBLED), helpfulness (1–5)… See the full description on the dataset page: https://huggingface.co/datasets/LawrenceYin/annotation-pack-a.annotation-pack-b
Annotation pack B — are any words forced into the text?
100 short texts (60 web-style excerpts, 40 one-sentence news summaries) written by small language models.
Annotators judge whether any word looks forced in.
For annotators / 标注人:
Read RUBRIC_HUMAN.md.
Annotator A downloads annotator_A.csv; annotator B downloads annotator_B.csv (same items).
Fill label (GENUINE / LOOPHOLE / GARBLED), wrong_sense (Y/N), fluency (1–5), flagged_words, optional notes. Work alone.
Send the… See the full description on the dataset page: https://huggingface.co/datasets/LawrenceYin/annotation-pack-b.
