bigscience/P3
Dataset Card for P3 Dataset Summary P3 (Public Pool of Prompts) is a collection of prompted English datasets covering a diverse set of NLP tasks. A prompt is the combination of an input template and a target template. The templates are functions mapping a data example into natural language for the input and target sequences. For example, in the case of an NLI dataset, the data example would include fields for Premise, Hypothesis, Label. An input template would be… See the full description on the dataset page: https://huggingface.co/datasets/bigscience/P3.
Convert to no-code dataset (#18)
Fix TypeError when loading some subsets (#12)
Streaming support (#11)
Fix tensorflow UnimplementedError (#10)
Optimize downloading (#9)
Fix TimeoutError (#8)
Fix `license` metadata (#1)
[README.md] Fix yaml metadata
hardcoded data_split_sizes and split_infos to avoid downloading
update citation bibtex
Remove TODO tag
Update readme with citation bibtex
add read_from_url logging
fix data path
final touches
converging
data_split_sizes
Update README.md
cleaning
cleaning
oooops
fix load_cahced_task
smaller thing to debug
remove dataset_info for now
fix split_name in data_dir url directory
fix data_dir dictionnary task_name
breaking down download of files
t
test
test
test
test
update link to csv
update readme for data_splits
update point of contact + data splits sizes
update model card
add dataset infos
update script and data card
upload data
update README + gitignore
rename data loading script
add .gitattributes and .gitignore
first attempt
First draft - data card
initial commit
