datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
text-2-video-Rich-Human-Feedback
Rapidata Video Generation Rich Human Feedback Dataset
If you get value from this dataset and would like to see more in the future, please consider liking it.
This dataset was collected in ~4 hours total using the Rapidata Python API, accessible to anyone and ideal for large scale data annotation.
Overview
In this dataset, ~22'000 human annotations were collected to evaluate AI-generated videos (using Sora) in 5 different categories.
Prompt - Video… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/text-2-video-Rich-Human-Feedback.lave-human-feedback
LAVE human judgments
This repository contains the human judgment data for Improving Automatic VQA Evaluation Using Large Language Models. Details about the data collection process and crowdworker population can be found in our paper, specifically in section 5.2 and appendix A.1.
Fields:
dataset: VQA dataset of origin for this example (vqav2, vgqa, okvqa).
model: VQA model that generated the predicted answer (blip2, promptcap, blip_vqa, blip_vg).
qid: question ID coming from the… See the full description on the dataset page: https://huggingface.co/datasets/mair-lab/lave-human-feedback.agile-cymru-synthetic-dataset-with-human-feedbackhuman_feedback
