Team Ai
14 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01RicardoRei /wmt-mqm-human-evaluation Dataset Summary This dataset contains all MQM human annotations from previous WMT Metrics shared tasks and the MQM annotations from Experts, Errors, and Context. The data is organised into 8 columns: lp: language pair src: input text mt: translation ref: reference translation score: MQM score system: MT Engine that produced the translation annotators: number of annotators domain: domain of the input text (e.g. news) year: collection year You can also find the original data here.… See the full description on the dataset page: https://huggingface.co/datasets/RicardoRei/wmt-mqm-human-evaluation.tabular100K<n<1M1 likes420 downloads4y agoHugging Face02ymoslem /wmt-da-human-evaluation-long-context Dataset Summary Long-context / document-level dataset for Quality Estimation of Machine Translation. It is an augmented variant of the sentence-level WMT DA Human Evaluation dataset. In addition to individual sentences, it contains augmentations of 2, 4, 8, 16, and 32 sentences, among each language pair lp and domain. The raw column represents a weighted average of scores of augmented sentences using character lengths of src and mt as weights. The code used to apply the augmentation… See the full description on the dataset page: https://huggingface.co/datasets/ymoslem/wmt-da-human-evaluation-long-context.tabular1M<n<10M9 likes352 downloads2y agoHugging Face03RicardoRei /wmt-da-human-evaluation Dataset Summary This dataset contains all DA human annotations from previous WMT News Translation shared tasks. The data is organised into 8 columns: lp: language pair src: input text mt: translation ref: reference translation score: z score raw: direct assessment annotators: number of annotators domain: domain of the input text (e.g. news) year: collection year You can also find the original data for each year in the results section https://www.statmt.org/wmt{YEAR}/results.html… See the full description on the dataset page: https://huggingface.co/datasets/RicardoRei/wmt-da-human-evaluation.tabular1M<n<10M10 likes217 downloads4y agoHugging Face04RicardoRei /wmt-sqm-human-evaluation Dataset Summary In 2022, several changes were made to the annotation procedure used in the WMT Translation task. In contrast to the standard DA (sliding scale from 0-100) used in previous years, in 2022 annotators performed DA+SQM (Direct Assessment + Scalar Quality Metric). In DA+SQM, the annotators still provide a raw score between 0 and 100, but also are presented with seven labeled tick marks. DA+SQM helps to stabilize scores across annotators (as compared to DA). The data is… See the full description on the dataset page: https://huggingface.co/datasets/RicardoRei/wmt-sqm-human-evaluation.tabular100K<n<1M1 likes59 downloads4y agoHugging Face05ymoslem /Human-Evaluation Human Evaluation Dataset The dataset includes human evaluation for General and Health domains. It was created as part of my two papers: “Domain-Specific Text Generation for Machine Translation” (Moslem et al., 2022) "Adaptive Machine Translation with Large Language Models" (Moslem et al., 2023) The evaluators were asked to assess the acceptability of each translation using a scale ranging from 1 to 4, where 4 is ideal and 1 is unacceptable translation. For the paper Moslem et al.… See the full description on the dataset page: https://huggingface.co/datasets/ymoslem/Human-Evaluation.tabulartranslation1K<n<10K1 likes15 downloads2y agoHugging Face06rsepulvedat /human_evaluation_agreement_subset_eutextn<1K0 likes12 downloads7mo agoHugging Face07rsepulvedat /human_evaluation_agreement_subset_catextn<1K0 likes10 downloads7mo agoHugging Face08Lakshan2003 /customerservice-Human-evaluation-results-evaluator_2tabularn<1K0 likes6 downloads9mo agoHugging Face09rsepulvedat /human_evaluation_agreement_subset_estextn<1K0 likes6 downloads7mo agoHugging Face10Lakshan2003 /customerservice-Human-evaluation-results-evaluator_1tabularn<1K0 likes4 downloads9mo agoHugging Face11Lakshan2003 /customerservice-Human-evaluation-results-evaluator_3tabularn<1K0 likes4 downloads9mo agoHugging Face12Lakshan2003 /customerservice-Human-evaluation-results-overalltabularn<1K0 likes4 downloads9mo agoHugging Face13zwhe99 /mt-human-evaluation-dagatedtabular1M<n<10M0 likes3 downloads3y agoHugging Face14rsepulvedat /human_evaluation_agreement_subset_vatextn<1K0 likes3 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. Team Ai does not host these files.